About the role
We’re seeking a mid-level AI Engineer to join a core product team building agentic systems that automate complex, multi-step workflows across regulated and enterprise domains. You’ll work across the stack to ship production LLM-based services, ensure reliability and safety, and collaborate with leadership, product, and design to deliver measurable user impact.
What you’ll do
- Design, build, and maintain agentic systems that automate multi-step workflows across industries such as healthcare, legal, fintech, logistics, and compliance.
- Own production retrieval-augmented generation (RAG) pipelines and retrieval infrastructure including vector databases, embeddings, and domain-specific search at scale.
- Implement multi-agent orchestration, tool-calling, memory, and reasoning components to deliver robust AI-driven user experiences.
- Develop evaluation and safety infrastructure to measure model performance, surface regressions, and ensure enterprise-level trust and reliability.
- Ship full-stack AI products from MVP to enterprise-grade by designing APIs and data models, implementing frontend and backend code, and operating production systems with CI/CD, monitoring, and testing.
- Collaborate with product and design to prioritize work, define success metrics, and iterate based on user feedback and telemetry.
What we’re looking for
- 2–8 years of software engineering experience delivering shipped user-facing or backend products.
- Practical experience deploying LLMs or LLM-based services in production, including prompt design, orchestration, and tool integration.
- Proficiency across the stack (Python and TypeScript/React or equivalent); experience with cloud platforms (AWS or GCP) and relational or NoSQL databases.
- Working knowledge of retrieval-augmented generation (RAG) patterns, vector databases, embeddings, and retrieval pipelines with sound judgment on approaches.
- Experience building automated tests, evaluations, and monitoring for AI systems to ensure reliability beyond demos.
- Experience with API-driven, high-throughput systems and real-time product features.
- Background building multi-tenant or enterprise-ready systems, or experience in regulated industries (healthcare, fintech, legal).
- Familiarity with fine-tuning, parameter-efficient tuning, or multi-modal model integration.
- Experience with agent or workflow frameworks (e.g., LangGraph, CrewAI) and orchestration tools (e.g., Temporal, Trigger).
Compensation & benefits
- Salary range: 180,000 – 400,000 USD per year.
- Full-time employment. Visa sponsorship not provided.
Location
Primary location: San Francisco, California, United States. This role supports a remote-friendly setup with the option to work from the office as needed.