AI Field Engineer, AI Infrastructure (Remote - US)
TLDR
Own the technical win from discovery through production, shipping customer code and deploying open-model LLMs with vLLM, SGLang, and TensorRT-LLM.
Lavendo partners with startups and high‑growth companies to help them hire top‑tier sales, GTM, and technical talent. This role is with one of our clients; we’ll share full details about the company and interview process as we get to know you and confirm mutual fit.
About the Company
Our client is a fast-scaling AI infrastructure company that helps enterprises and AI-native startups build, fine-tune, and serve their own specialized AI models — no more relying on someone else's black-box API. Their platform runs in production for some of the most recognizable names in tech, spanning text, image, embedding, audio, and multimodal workloads. Freshly backed by a Series D at a $17.5B valuation, this is a company with real scale, real customers, and real momentum behind it.
The Mission
They believe the next generation of great products will be built by companies that own their AI stack, not just rent it. Their mission is making open-model AI infrastructure genuinely fast, flexible, and production-grade for the teams building at the frontier.
The Opportunity
This is a forward-deployed engineering role where you're never far from either a terminal or a whiteboard with a VP on it. You'll work directly with VP Engineering, Head of AI/ML, and CTO-level partners at Fortune 500 enterprises and cutting-edge AI-native startups, owning the technical win from the first discovery call all the way through production. Two tracks are open — Enterprise and AI-Native — same role, different customer base, and we'll talk through which fits your background and interests best.
What You'll Do
Run the full pre-sales field cycle yourself: discovery, POC scoping, model evals, and final model selection
Ship real production code inside customer environments — this is hands-on building, not advisory slides
Deploy and fine-tune open-model LLMs using frameworks like vLLM, SGLang, and TensorRT-LLM (SFT baseline, DPO/RFT a big plus)
Act as a true partner to your sales counterpart, shaping deal strategy rather than just supporting it
Build relationships across a customer's org, from engineers to executives, and know how to move a deal forward on both levels
What You Bring
3+ years in a client-facing AI/ML role — Solutions Architect, Forward-Deployed Engineer, Customer Success Engineer, Applied AI Engineer, Sales Engineer, or an AI/ML-focused SWE with genuine customer exposure
Real hands-on experience with open-model LLM inference and/or fine-tuning
Hyperscaler experience in an AI context (AWS, Azure, or GCP)
Strong Python skills, comfort with GPU infrastructure and Kubernetes
A background at an AI-native company or a SaaS company genuinely building AI features (not bolting them on)
Full-time work history, and openness to travel to enterprise customers as needed
Key Success Drivers
You're the rare person equally comfortable debugging inference trade-offs with an ML engineer in the morning and presenting the same trade-offs to a VP by afternoon. You bring genuine customer obsession alongside technical depth — this role rewards people who've shipped things themselves, not just advised on them. Strong candidates own outcomes end-to-end and know how to build trust across both engineering and executive stakeholders.
Why Join?
Compensation & Equity
Base salary $176K–$224K, OTE $220K–$280K
Performance-based variable compensation, paid quarterly
Meaningful equity in a fast-growing, well-funded startup — Series D at a $17.5B valuation
Location
Remote-friendly with hubs in New York and San Mateo — flexibility to work from anywhere in the US with regular travel to enterprise customers
Visa Sponsorship Details
Open to visa transfers (e.g. OPT, H1B transfers)
Open to visa sponsorships (e.g. new H1B, TN)
Benefits
Comprehensive benefits package
Interviewing Process
Take-home assignment (self-paced, 24-hour window once scheduled): build a working text-to-SQL system and show it works
Recruiter screen (30 min): logistics, motivation, fit
Culture + live coding (1 hr): product discussion, then extend your take-home live
Discovery + hiring manager (45 min): live discovery role-play
On-site final loop (~2 hrs): customer demo/presentation plus an executive values conversation
Executive interview (30 min)
Offer
We are proud to be an equal opportunity workplace and consider all qualified applicants without regard to race, color, religion, national origin, age, sex, marital status, ancestry, disability, genetic information, veteran or military status, gender identity or expression, sexual orientation, or any other characteristic protected by law.
Benefits
Equity Compensation
Meaningful equity in a fast-growing, well-funded startup — Series D at a $17.5B valuation
Health Insurance
Comprehensive benefits package
Remote-Friendly
Remote-friendly with hubs in New York and San Mateo — flexibility to work from anywhere in the US with regular travel to enterprise customers
Visa Sponsorship
Open to visa sponsorships (e.g. new H1B, TN)
Lavendo builds a robust Enterprise SaaS platform focused on data privacy governance and security, leveraging patented AI technology to help Fortune 1000 companies manage compliance and protect sensitive information. Our solutions stand out by seamlessly integrating data security protocols, ensuring that enterprises can confidently navigate the complexities of data management and regulatory requirements.