AI Engineer – Production LLM Systems
Location: New York, NY — Hybrid, 2–3 days/week
Employment: Full-time
Base Salary: $200,000–$350,000 + competitive equity
Visa: Open to visa transfers, including H-1B
About the Role
We’re hiring an experienced AI Engineer to join a fast-growing, venture-backed AI startup building production-grade AI/LLM technology for large regulated enterprises.
This is a highly hands-on, high-ownership engineering role. You’ll take ownership of the technical build-out of a custom AI/LLM platform for a top-5 U.S. bank, working across architecture, backend development, AI/LLM systems, enterprise integrations, infrastructure, deployment, and production reliability.
This is not a research or prototype-focused position. We’re looking for an engineer who enjoys shipping real AI systems into production and solving complex technical problems end-to-end.
What You’ll Do
Own the end-to-end technical build and deployment of an enterprise AI/LLM platform.
Customize the core AI platform for a major financial institution’s environment.
Build and deploy production-grade LLM, GenAI and RAG systems.
Develop scalable backend services and APIs primarily using Python.
Own enterprise integrations, deployment, troubleshooting, and production reliability.
Work with Kubernetes, CI/CD and cloud infrastructure (AWS/GCP/Azure).
Collaborate closely with the core engineering team on architecture and product development.
Work directly with enterprise engineering stakeholders as a technical point of contact.
Diagnose complex production issues and independently drive them through resolution.
What We’re Looking For
6–12+ years of software, backend, ML, or AI engineering experience.
Strong hands-on Python/backend engineering experience.
Recent experience building and deploying LLM/Generative AI systems in production.
Experience with technologies such as RAG, LLM infrastructure, model serving, vector databases, or AI agents.
Strong experience with Kubernetes, Docker, cloud infrastructure, APIs and CI/CD.
Experience owning systems from architecture through production deployment.
Experience working directly with enterprise customers/stakeholders, ideally within fintech, banking, or another regulated industry.
Recent experience at a venture-backed AI startup (Series A–D) actively building innovative AI products.
Comfortable operating independently in a fast-moving startup environment.
Tech Stack
Python | LLMs | GenAI | RAG | Kubernetes | Docker | REST APIs | CI/CD | AWS/GCP/Azure
Why Join?
Own a high-profile AI deployment for a top-5 U.S. bank from day one.
Build real production AI rather than research projects or prototypes.
Join a small, deeply technical and execution-focused startup team.
Work directly with senior leadership with significant technical ownership.
Opportunity to grow into broader technical and team leadership as the company scales.
$200K–$350K base salary + competitive equity.
Hybrid working environment in New York City, 2–3 days per week.
Generative AI Engineer
BeaconFire Inc.
Release Train Engineer
Ford Motor Company
Applied AI Engineer - Onsite - SF / NYC
RS Global Services
Agentic AI Engineer (Python)
BeaconFire Inc.
Senior AI / ML Engineer (LLMs)
EvolutionIQ
Senior AI Engineer — Voice Systems
Jack