Senior Agentic, AI Engineer
Posted 52mins ago
Employment Information
Report this job
Job expired or something wrong with this job?
Job Description
Senior Agentic AI Engineer building production agents for Worth AI’s regulated-finance automation. Owning retrieval, evaluations, tools, MLOps, and compliant deployment.
Responsibilities:
- Design and ship multi-step agentic systems for onboarding, underwriting, case review, and continuous monitoring
- Architect agent graphs in LangGraph or comparable frameworks with explicit state, durable execution, retries, and safe fallbacks
- Build the retrieval layer powering agents, including chunking, hybrid search, reranking, and grounded citation
- Own the evaluation stack, including golden sets, offline regression suites, LLM-as-judge, online A/B and shadow evaluations, and red-teaming
- Expose agents to production systems via well-typed tools and MCP servers
- Drive production MLOps, including deployment, versioning, traffic shaping, cost and latency budgets, tracing, and on-call playbooks
- Partner with security and compliance to maintain SOC 2, GDPR, CCPA, and fair-lending posture with auditability and explainability
- Mentor engineers on agent patterns, prompt hygiene, evaluation discipline, and LLM failure modes
- Improve agent task success, grounding accuracy, reliability, velocity, and risk posture
- Travel to Orlando, Florida at least twice per year for Town Halls and team collaboration, in addition to orientation in Orlando
Requirements:
- 5+ years of software engineering experience, with 2+ years building production LLM or agentic systems (not just notebooks or demos)
- Hands-on experience with a modern agent framework (LangGraph strongly preferred) and a track record of shipping agents that run, fail gracefully, and recover
- Strong RAG fundamentals: chunking, embeddings, hybrid retrieval, reranking, grounding — and judgment about when RAG isn’t the right answer
- Real eval experience: golden sets, offline and online evaluations, used to make ship/no-ship calls
- Production MLOps fluency: deployed LLM workloads under real latency, cost, and reliability constraints
- Strong Python; comfortable in TypeScript / Node.js
- Solid systems engineering instincts: APIs, async patterns, queues, databases, distributed system failure modes
- Calibrated communicator; thrives in ambiguous, fast-moving environments
- Prior experience in fintech, lending, payments, KYB/KYC, fraud, or AML
- Experience building MCP servers or other structured tool interfaces for LLMs
- Background in classical ML (ranking, scoring, calibration)
- Experience designing explainable / auditable AI workflows for regulated environments
- Open-source contributions to agent frameworks, eval tooling, or retrieval libraries
- AWS depth (EKS, MSK, RDS, S3, Lambda) and IaC with Terraform
Benefits:
- Health Care Plan (Medical, Dental & Vision)
- Retirement Plan (401k, IRA)
- Life Insurance
- Flexible Paid Time Off
- 9 paid Holidays
- Family Leave
- Remote
- Hybrid work (for Orlando Associates)
- Free Food & Snacks (Orlando)
- Wellness Resources



















