Senior Agentic (AI) Engineer

Reposted 24 Days Ago
4 Locations
In-Office or Remote
Senior level
Artificial Intelligence • Fintech • Software • Financial Services
Worth is the underwriting & onboarding platform that helps financial institutions say yes to small businesses, faster.
The Role
Design, build, and productionize multi-step agentic systems for onboarding, underwriting, and monitoring regulated financial workflows. Own agent architecture, retrieval, evals, tooling, MLOps, observability, and compliance; partner with AI, platform, security teams; mentor engineers and ensure explainability, auditability, and reliability in production.
Summary Generated by Built In

Worth AI is hiring a Senior Agentic AI Engineer to design and ship production agent systems that automate KYB, underwriting, and risk decisions on regulated financial data. You’ll own agents end-to-end architecture, retrieval, tools, evals, and production deployment and partner closely with our Chief AI Officer, applied scientists, and platform teams.

Responsibilities
  • Design and ship multi-step agentic systems (planner/executor, tool-using, multi-agent, human-in-the-loop) for onboarding, underwriting, case review, and continuous monitoring.
  • Architect agent graphs in LangGraph (or comparable — CrewAI, AutoGen, Claude Agent SDK) with explicit state, durable execution, retries, and safe fallbacks.
  • Build the retrieval layer powering our agents — chunking, hybrid search, reranking, and grounded citation.
  • Own the eval stack: golden sets, offline regression suites, LLM-as-judge, online A/B and shadow evals, and red-teaming for jailbreaks, prompt injection, and PII leakage.
  • Expose agents to production systems via well-typed tools and MCP servers. Treat tool surface area as a product.
  • Drive production MLOps: deployment, versioning, traffic shaping, cost/latency budgets, tracing, and on-call playbooks for agent incidents.
  • Partner with security and compliance to keep agents inside SOC 2, GDPR, CCPA, and fair-lending posture — auditability and explainability built in, not bolted on.
  • Mentor engineers on agent patterns, prompt hygiene, eval discipline, and LLM failure modes.
  • Technology Stack
    • Languages: Python, Node.js, TypeScript
    • Agent / LLM frameworks: LangGraph, LangChain, Claude Agent SDK, MCP, OpenAI SDK
    • Models: Anthropic Claude, OpenAI, open-weight where appropriate
    • Retrieval & Data: PostgreSQL, pgvector, OpenSearch, Kafka, Redshift, Redis
    • Infra: AWS, Kubernetes (EKS), ArgoCD, Terraform
    • Evals & Observability: LangSmith / Langfuse / Braintrust-style tooling, DataDog

Requirements
  • 5+ years of software engineering experience, with 2+ years building production LLM or agentic systems (not just notebooks or demos).
  • Hands-on experience with a modern agent framework (LangGraph strongly preferred) and a track record of shipping agents that run, fail gracefully, and recover.
  • Strong RAG fundamentals chunking, embeddings, hybrid retrieval, reranking, grounding — and judgment about when RAG isn’t the right answer.
  • Real eval experience golden sets, offline and online evaluations, used to make ship/no-ship calls.
  • Production MLOps fluency: deployed LLM workloads under real latency, cost, and reliability constraints.
  • Strong Python; comfortable in TypeScript / Node.js.
  • Solid systems engineering instincts APIs, async patterns, queues, databases, distributed system failure modes.
  • Calibrated communicator; thrives in ambiguous, fast-moving environments.
  • Prior experience in fintech, lending, payments, KYB/KYC, fraud, or AML.
  • Experience building MCP servers or other structured tool interfaces for LLMs.
  • Background in classical ML (ranking, scoring, calibration).
  • Experience designing explainable / auditable AI workflows for regulated environments.
  • Open-source contributions to agent frameworks, eval tooling, or retrieval libraries.
  • AWS depth (EKS, MSK, RDS, S3, Lambda) and IaC with Terraform.
Success Metrics
  • Agent Quality: Measurable improvements in task success rate, grounding accuracy, and hallucination rate on our eval suites.
  • Production Reliability: Agents you own meet defined SLOs for latency (P90/P99), tool-call success, and cost per task.
  • Velocity: New agent capabilities go from prototype to production in weeks, without skipping evals or guardrails.
  • Risk Posture: Zero material incidents tied to prompt injection, PII leakage, or unsafe tool use on agents you own.
  • Force Multiplier: Patterns, tools, and eval scaffolding you build get adopted across engineering.

All Remote Hires will be required to travel to Orlando, Florida at least twice per year for Town Halls and team collaboration, in addition to orientation in Orlando.


Benefits
  • Health Care Plan (Medical, Dental & Vision)
  • Retirement Plan (401k, IRA)
  • Life Insurance
  • Flexible Paid Time Off
  • 9 paid Holidays
  • Family Leave
  • Remote
  • Hybrid work (for Orlando Associates)
  • Free Food & Snacks (Orlando)
  • Wellness Resources

Skills Required

  • 5+ years of software engineering experience with 2+ years building production LLM or agentic systems
  • Hands-on experience with a modern agent framework (LangGraph strongly preferred) and track record of shipping resilient agents
  • Strong RAG fundamentals: chunking, embeddings, hybrid retrieval, reranking, grounding
  • Real evaluation experience: golden sets, offline and online evaluations, making ship/no-ship calls
  • Production MLOps fluency: deployed LLM workloads under latency, cost, and reliability constraints
  • Strong Python; comfortable in TypeScript / Node.js
  • Solid systems engineering instincts: APIs, async patterns, queues, databases, distributed failure modes
  • Calibrated communicator; thrives in ambiguous, fast-moving environments
  • Prior experience in fintech, lending, payments, KYB/KYC, fraud, or AML
  • Experience building MCP servers or other structured tool interfaces for LLMs
  • Background in classical ML (ranking, scoring, calibration)
  • Experience designing explainable / auditable AI workflows for regulated environments (SOC 2, GDPR, CCPA, fair-lending)
  • Open-source contributions to agent frameworks, eval tooling, or retrieval libraries
  • AWS depth (EKS, MSK, RDS, S3, Lambda) and IaC with Terraform

Worth Compensation & Benefits Highlights

  • Healthcare Strength Core medical, dental, and vision coverage plus FSA/HSA and life insurance are explicitly listed and employer-verified, indicating solid baseline health support.
  • Leave & Time Off Breadth Unlimited PTO with paid holidays and parental leave are highlighted, with parental leave also shown as employer-verified.
  • Wellbeing & Lifestyle Benefits Hybrid/remote flexibility, work-from-home support, and onsite snacks/perks are consistently mentioned as part of the day-to-day experience.

Worth Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Winter Park, Florida
70 Employees
Year Founded: 2023

What We Do

Worth is the AI-powered platform that consolidates onboarding, underwriting, and risk monitoring for fintechs, lenders, payment processors, and financial institutions. Founded in 2023, we built Worth to replace slow, manual underwriting with a single system that verifies, scores, and monitors small and medium-sized businesses (SMBs) in real time. At the center of the platform is Crosswalking Technology. Our proprietary AI/ML models intelligently match businesses across disparate data sources, ensuring the highest level of accuracy and reliability in SMB entity resolution. By integrating multiple first- and third-party authoritative data sources into our crosswalk-matching logic, Worth ensures that businesses are correctly identified, even in cases of duplicate addresses, name variations, or incomplete records. This data moat spans 186 integrations and 25 global and local partners across 200+ countries and territories, resolving fragmented SMB signals into a database of 350M+ SMBs with a 98% data match rate. Our product suite — Worth Pre-Fill, Custom Onboarding, Case Management, Decisioning Engine, Perpetual Risk Monitoring, and Worth Wallet — is available via API, SDK, or fully white-labeled, enabling financial institutions to consolidate their entire onboarding and underwriting stack into one platform. Customers using Worth have increased approval rates by 37%+, reduced application abandonment by 43%+, cut vendor costs by 25%, and reduced time to revenue by 55%+. We're SOC 2 Type II certified and GDPR and CCPA compliant, and have raised $55M in funding to date. Today, 50+ customers rely on Worth to onboard and underwrite their SMB customers faster and more accurately.

Why Work With Us

We're solving a genuinely hard problem: turning fragmented SMB data into one durable, explainable identity that banks and lenders can trust. Backed by $55M in funding and already live with 50+ customers, we're a tight-knit team with real traction, where your work would help shape the roadmap.

Gallery

Gallery
Gallery
Gallery

Worth Offices

Hybrid Workspace

Employees engage in a combination of remote and on-site work.

Typical time on-site: Not Specified
HQOrlando

Similar Jobs

Worth Logo Worth

Security Engineer

Artificial Intelligence • Fintech • Software • Financial Services
In-Office or Remote
4 Locations
70 Employees

Worth Logo Worth

Customer Success Manager

Artificial Intelligence • Fintech • Software • Financial Services
In-Office or Remote
Orlando, FL, USA
70 Employees

Worth Logo Worth

Senior Software Engineer

Artificial Intelligence • Fintech • Software • Financial Services
In-Office or Remote
5 Locations
70 Employees

Worth Logo Worth

Solutions Engineer

Artificial Intelligence • Fintech • Software • Financial Services
In-Office or Remote
4 Locations
70 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account