Senior Agentic (AI) Engineer

Posted 23 Days Ago
4 Locations
In-Office or Remote
Senior level
Artificial Intelligence • Fintech • Software • Financial Services
Worth is the underwriting & onboarding platform that helps financial institutions say yes to small businesses, faster.
The Role
Design, build, and productionize multi-step agentic systems for onboarding, underwriting, and monitoring regulated financial workflows. Own agent architecture, retrieval, evals, tooling, MLOps, observability, and compliance; partner with AI, platform, security teams; mentor engineers and ensure explainability, auditability, and reliability in production.
Summary Generated by Built In

Worth AI is hiring a Senior Agentic AI Engineer to design and ship production agent systems that automate KYB, underwriting, and risk decisions on regulated financial data. You’ll own agents end-to-end architecture, retrieval, tools, evals, and production deployment and partner closely with our Chief AI Officer, applied scientists, and platform teams.

Responsibilities
  • Design and ship multi-step agentic systems (planner/executor, tool-using, multi-agent, human-in-the-loop) for onboarding, underwriting, case review, and continuous monitoring.
  • Architect agent graphs in LangGraph (or comparable — CrewAI, AutoGen, Claude Agent SDK) with explicit state, durable execution, retries, and safe fallbacks.
  • Build the retrieval layer powering our agents — chunking, hybrid search, reranking, and grounded citation.
  • Own the eval stack: golden sets, offline regression suites, LLM-as-judge, online A/B and shadow evals, and red-teaming for jailbreaks, prompt injection, and PII leakage.
  • Expose agents to production systems via well-typed tools and MCP servers. Treat tool surface area as a product.
  • Drive production MLOps: deployment, versioning, traffic shaping, cost/latency budgets, tracing, and on-call playbooks for agent incidents.
  • Partner with security and compliance to keep agents inside SOC 2, GDPR, CCPA, and fair-lending posture — auditability and explainability built in, not bolted on.
  • Mentor engineers on agent patterns, prompt hygiene, eval discipline, and LLM failure modes.
  • Technology Stack
    • Languages: Python, Node.js, TypeScript
    • Agent / LLM frameworks: LangGraph, LangChain, Claude Agent SDK, MCP, OpenAI SDK
    • Models: Anthropic Claude, OpenAI, open-weight where appropriate
    • Retrieval & Data: PostgreSQL, pgvector, OpenSearch, Kafka, Redshift, Redis
    • Infra: AWS, Kubernetes (EKS), ArgoCD, Terraform
    • Evals & Observability: LangSmith / Langfuse / Braintrust-style tooling, DataDog

Requirements
  • 5+ years of software engineering experience, with 2+ years building production LLM or agentic systems (not just notebooks or demos).
  • Hands-on experience with a modern agent framework (LangGraph strongly preferred) and a track record of shipping agents that run, fail gracefully, and recover.
  • Strong RAG fundamentals chunking, embeddings, hybrid retrieval, reranking, grounding — and judgment about when RAG isn’t the right answer.
  • Real eval experience golden sets, offline and online evaluations, used to make ship/no-ship calls.
  • Production MLOps fluency: deployed LLM workloads under real latency, cost, and reliability constraints.
  • Strong Python; comfortable in TypeScript / Node.js.
  • Solid systems engineering instincts APIs, async patterns, queues, databases, distributed system failure modes.
  • Calibrated communicator; thrives in ambiguous, fast-moving environments.
  • Prior experience in fintech, lending, payments, KYB/KYC, fraud, or AML.
  • Experience building MCP servers or other structured tool interfaces for LLMs.
  • Background in classical ML (ranking, scoring, calibration).
  • Experience designing explainable / auditable AI workflows for regulated environments.
  • Open-source contributions to agent frameworks, eval tooling, or retrieval libraries.
  • AWS depth (EKS, MSK, RDS, S3, Lambda) and IaC with Terraform.
Success Metrics
  • Agent Quality: Measurable improvements in task success rate, grounding accuracy, and hallucination rate on our eval suites.
  • Production Reliability: Agents you own meet defined SLOs for latency (P90/P99), tool-call success, and cost per task.
  • Velocity: New agent capabilities go from prototype to production in weeks, without skipping evals or guardrails.
  • Risk Posture: Zero material incidents tied to prompt injection, PII leakage, or unsafe tool use on agents you own.
  • Force Multiplier: Patterns, tools, and eval scaffolding you build get adopted across engineering.

All Remote Hires will be required to travel to Orlando, Florida at least twice per year for Town Halls and team collaboration, in addition to orientation in Orlando.


Benefits
  • Health Care Plan (Medical, Dental & Vision)
  • Retirement Plan (401k, IRA)
  • Life Insurance
  • Flexible Paid Time Off
  • 9 paid Holidays
  • Family Leave
  • Remote
  • Hybrid work (for Orlando Associates)
  • Free Food & Snacks (Orlando)
  • Wellness Resources

Skills Required

  • 5+ years of software engineering experience with 2+ years building production LLM or agentic systems
  • Hands-on experience with a modern agent framework (LangGraph strongly preferred) and track record of shipping resilient agents
  • Strong RAG fundamentals: chunking, embeddings, hybrid retrieval, reranking, grounding
  • Real evaluation experience: golden sets, offline and online evaluations, making ship/no-ship calls
  • Production MLOps fluency: deployed LLM workloads under latency, cost, and reliability constraints
  • Strong Python; comfortable in TypeScript / Node.js
  • Solid systems engineering instincts: APIs, async patterns, queues, databases, distributed failure modes
  • Calibrated communicator; thrives in ambiguous, fast-moving environments
  • Prior experience in fintech, lending, payments, KYB/KYC, fraud, or AML
  • Experience building MCP servers or other structured tool interfaces for LLMs
  • Background in classical ML (ranking, scoring, calibration)
  • Experience designing explainable / auditable AI workflows for regulated environments (SOC 2, GDPR, CCPA, fair-lending)
  • Open-source contributions to agent frameworks, eval tooling, or retrieval libraries
  • AWS depth (EKS, MSK, RDS, S3, Lambda) and IaC with Terraform

Worth Compensation & Benefits Highlights

  • Healthcare Strength Medical, dental, and vision coverage appear consistently across company materials and third‑party profiles, with HSA/FSA and life insurance also cited. Employer‑verified benefits listings indicate core health coverage is formally in place.
  • Parental & Family Support Parental leave is highlighted as generous on external profiles and shows up in employer‑verified benefits. Family‑oriented offerings complement the broader health and time‑off package.
  • Leave & Time Off Breadth Unlimited/flexible PTO and paid holidays are repeatedly listed across postings and profiles. Flexible vacation language and family leave references point to broad time‑off availability.

Worth Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Winter Park, Florida
70 Employees
Year Founded: 2023

What We Do

Worth is the AI-powered platform that consolidates onboarding, underwriting, and risk monitoring for fintechs, lenders, payment processors, and financial institutions. Founded in 2023, we built Worth to replace slow, manual underwriting with a single system that verifies, scores, and monitors small and medium-sized businesses (SMBs) in real time. At the center of the platform is Crosswalking Technology. Our proprietary AI/ML models intelligently match businesses across disparate data sources, ensuring the highest level of accuracy and reliability in SMB entity resolution. By integrating multiple first- and third-party authoritative data sources into our crosswalk-matching logic, Worth ensures that businesses are correctly identified, even in cases of duplicate addresses, name variations, or incomplete records. This data moat spans 186 integrations and 25 global and local partners across 200+ countries and territories, resolving fragmented SMB signals into a database of 350M+ SMBs with a 98% data match rate. Our product suite — Worth Pre-Fill, Custom Onboarding, Case Management, Decisioning Engine, Perpetual Risk Monitoring, and Worth Wallet — is available via API, SDK, or fully white-labeled, enabling financial institutions to consolidate their entire onboarding and underwriting stack into one platform. Customers using Worth have increased approval rates by 37%+, reduced application abandonment by 43%+, cut vendor costs by 25%, and reduced time to revenue by 55%+. We're SOC 2 Type II certified and GDPR and CCPA compliant, and have raised $55M in funding to date. Today, 50+ customers rely on Worth to onboard and underwrite their SMB customers faster and more accurately.

Why Work With Us

We're solving a genuinely hard problem: turning fragmented SMB data into one durable, explainable identity that banks and lenders can trust. Backed by $55M in funding and already live with 50+ customers, we're a tight-knit team with real traction, where your work would help shape the roadmap.

Worth Offices

Hybrid Workspace

Employees engage in a combination of remote and on-site work.

Typical time on-site: Not Specified
HQOrlando

Similar Jobs

Worth AI Logo Worth AI

Technical Product Manager

Artificial Intelligence • Fintech • Software • Financial Services
In-Office or Remote
4 Locations
32 Employees

Worth AI Logo Worth AI

Solutions Engineer

Artificial Intelligence • Fintech • Software • Financial Services
In-Office or Remote
4 Locations
32 Employees

Worth AI Logo Worth AI

Senior Software Engineer

Artificial Intelligence • Fintech • Software • Financial Services
In-Office or Remote
4 Locations
32 Employees

Worth AI Logo Worth AI

Development Engineer

Artificial Intelligence • Fintech • Software • Financial Services
In-Office or Remote
4 Locations
32 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account