Compensation
This employer didn't list pay. Model a likely range with the calculator.
Model this offer in the calculatorJob description
Worth AI is hiring a Senior Agentic AI Engineer to design and ship production agent systems that automate KYB underwriting and risk decisions on regulated financial data. You’ll own agents end-to-end architecture retrieval tools evals and production deployment and partner closely with our Chief AI Officer applied scientists and platform teams.
Responsibilities- Design and ship multi-step agentic systems (planner/executor tool-using multi-agent human-in-the-loop) for onboarding underwriting case review and continuous monitoring.
- Architect agent graphs in LangGraph (or comparable — CrewAI AutoGen Claude Agent SDK) with explicit state durable execution retries and safe fallbacks.
- Build the retrieval layer powering our agents — chunking hybrid search reranking and grounded citation.
- Own the eval stack: golden sets offline regression suites LLM-as-judge online A/B and shadow evals and red-teaming for jailbreaks prompt injection and PII leakage.
- Expose agents to production systems via well-typed tools and MCP servers. Treat tool surface area as a product.
- Drive production MLOps: deployment versioning traffic shaping cost/latency budgets tracing and on-call playbooks for agent incidents.
- Partner with security and compliance to keep agents inside SOC 2 GDPR CCPA and fair-lending posture — auditability and explainability built in not bolted on.
- Mentor engineers on agent patterns prompt hygiene eval discipline and LLM failure modes.
- Technology Stack
- Languages: Python Node.js TypeScript
- Agent / LLM frameworks: LangGraph LangChain Claude Agent SDK MCP OpenAI SDK
- Models: Anthropic Claude OpenAI open-weight where appropriate
- Retrieval & Data: PostgreSQL pgvector OpenSearch Kafka Redshift Redis
- Infra: AWS Kubernetes (EKS) ArgoCD Terraform
- Evals & Observability: LangSmith / Langfuse / Braintrust-style tooling DataDog
Requirements
- 5+ years of software engineering experience with 2+ years building production LLM or agentic systems (not just notebooks or demos).
- Hands-on experience with a modern agent framework (LangGraph strongly preferred) and a track record of shipping agents that run fail gracefully and recover.
- Strong RAG fundamentals chunking embeddings hybrid retrieval reranking grounding — and judgment about when RAG isn’t the right answer.
- Real eval experience golden sets offline and online evaluations used to make ship/no-ship calls.
- Production MLOps fluency: deployed LLM workloads under real latency cost and reliability constraints.
- Strong Python; comfortable in TypeScript / Node.js.
- Solid systems engineering instincts APIs async patterns queues databases distributed system failure modes.
- Calibrated communicator; thrives in ambiguous fast-moving environments.
- Prior experience in fintech lending payments KYB/KYC fraud or AML.
- Experience building MCP servers or other structured tool interfaces for LLMs.
- Background in classical ML (ranking scoring calibration).
- Experience designing explainable / auditable AI workflows for regulated environments.
- Open-source contributions to agent frameworks eval tooling or retrieval libraries.
- AWS depth (EKS MSK RDS S3 Lambda) and IaC with Terraform.
- Agent Quality: Measurable improvements in task success rate grounding accuracy and hallucination rate on our eval suites.
- Production Reliability: Agents you own meet defined SLOs for latency (P90/P99) tool-call success and cost per task.
- Velocity: New agent capabilities go from prototype to production in weeks without skipping evals or guardrails.
- Risk Posture: Zero material incidents tied to prompt injection PII leakage or unsafe tool use on agents you own.
- Force Multiplier: Patterns tools and eval scaffolding you build get adopted across engineering.
All Remote Hires will be required to travel to Orlando Florida at least twice per year for Town Halls and team collaboration in addition to orientation in Orlando.
Benefits
- Health Care Plan (Medical Dental & Vision)
- Retirement Plan (401k IRA)
- Life Insurance
- Flexible Paid Time Off
- 9 paid Holidays
- Family Leave
- Remote
- Hybrid work (for Orlando Associates)
- Free Food & Snacks (Orlando)
- Wellness Resources
Skills Required
- 5+ years of software engineering experience with 2+ years building production LLM or agentic systems
- Hands-on experience with a modern agent framework (LangGraph strongly preferred) and track record of shipping resilient agents
- Strong RAG fundamentals: chunking embeddings hybrid retrieval reranking grounding
- Real evaluation experience: golden sets offline and online evaluations making ship/no-ship calls
- Production MLOps fluency: deployed LLM workloads under latency cost and reliability constraints
- Strong Python; comfortable in TypeScript / Node.js
- Solid systems engineering instincts: APIs async patterns queues databases distributed failure modes
- Calibrated communicator; thrives in ambiguous fast-moving environments
- Prior experience in fintech lending payments KYB/KYC fraud or AML
- Experience building MCP servers or other structured tool interfaces for LLMs
- Background in classical ML (ranking scoring calibration)
- Experience designing explainable / auditable AI workflows for regulated environments (SOC 2 GDPR CCPA fair-lending)
- Open-source contributions to agent frameworks eval tooling or retrieval libraries
- AWS depth (EKS MSK RDS S3 Lambda) and IaC with Terraform
Worth Compensation & Benefits Highlights
- Healthcare Strength—Medical dental and vision coverage appear consistently across company materials and third‑party profiles with HSA/FSA and life insurance also cited. Employer‑verified benefits listings indicate core health coverage is formally in place.
- Parental & Family Support—Parental leave is highlighted as generous on external profiles and shows up in employer‑verified benefits. Family‑oriented offerings complement the broader health and time‑off package.
- Leave & Time Off Breadth—Unlimited/flexible PTO and paid holidays are repeatedly listed across postings and profiles. Flexible vacation language and family leave references point to broad time‑off availability.
Worth Insights
What We Do
Worth is the AI-powered platform that consolidates onboarding underwriting and risk monitoring for fintechs lenders payment processors and financial institutions. Founded in 2023 we built Worth to replace slow manual underwriting with a single system that verifies scores and monitors small and medium-sized businesses (SMBs) in real time.At the center of the platform is Crosswalking Technology. Our proprietary AI/ML models intelligently match businesses across disparate data sources ensuring the highest level of accuracy and reliability in SMB entity resolution. By integrating multiple first- and third-party authoritative data sources into our crosswalk-matching logic Worth ensures that businesses are correctly identified even in cases of duplicate addresses name variations or incomplete records. This data moat spans 186 integrations and 25 global and local partners across 200+ countries and territories resolving fragmented SMB signals into a database of 350M+ SMBs with a 98% data match rate.Our product suite — Worth Pre-Fill Custom Onboarding Case Management Decisioning Engine Perpetual Risk Monitoring and Worth Wallet — is available via API SDK or fully white-labeled enabling financial institutions to consolidate their entire onboarding and underwriting stack into one platform.Customers using Worth have increased approval rates by 37%+ reduced application abandonment by 43%+ cut vendor costs by 25% and reduced time to revenue by 55%+. We're SOC 2 Type II certified and GDPR and CCPA compliant and have raised $55M in funding to date. Today 50+ customers rely on Worth to onboard and underwrite their SMB customers faster and more accurately.
Why Work With Us
We're solving a genuinely hard problem: turning fragmented SMB data into one durable explainable identity that banks and lenders can trust. Backed by $55M in funding and already live with 50+ customers we're a tight-knit team with real traction where your work would help shape the roadmap.
Worth Offices
Hybrid Workspace
Employees engage in a combination of remote and on-site work.
Related roles