Senior ML Engineer (Systems)
About the role
๐ Senior ML Engineer (Systems) | AI Agents & Distributed Systems
Location: Sunnyvale, CA โ On-site
Employment Type: Full-time
Experience: 5+ years
Compensation: $150Kโ$230K + up to 1% equity
Visa: H-1B transfers, new H-1B applications & TN visas supported
Weโre hiring a Senior ML Engineer (Systems) to join an early-stage AI company building infrastructure for the next generation of multi-agent enterprise workflows.
This is a highly hands-on role for an engineer who can bridge ML systems research, agentic AI, distributed infrastructure, full-stack development, and product engineering.
Youโll work closely with the CEO and Chief Architect to turn research-grade ML systems ideas into products that developers and enterprise customers genuinely enjoy using.
๐ฅ What Youโll Do
Build and productionize multi-agentic AI systems
Design scalable agent orchestration and infrastructure
Develop full-stack applications primarily using Python
Work with agent frameworks such as LangGraph, LangChain, AutoGen, CrewAI, Semantic Kernel, Google ADK, or equivalent
Deploy and optimize model-serving infrastructure using technologies such as vLLM, SGLang, Ray, NVIDIA Triton, or NVIDIA Dynamo
Build systems involving APIs, distributed systems, asynchronous jobs, queues, containers, deployment platforms, and cloud infrastructure
Develop intuitive developer-facing products around complex ML infrastructure
Create visualizations and product experiences using tools such as Tableau and Grafana
Work extensively with open-source software, with opportunities to contribute upstream
Translate research-grade concepts into documentation, examples, onboarding experiences, and product language
Collaborate directly with technical leadership in an ambiguous, fast-moving startup environment
Help shape architecture, engineering practices, and the product itself
๐ง Ideal Candidate
Youโll be a strong fit if you have:
(academic years can substitute if PhD from top institution in relevant ML systems field)
5+ years of professional software engineering experience
5+ years of ML systems engineering experience in production
Strong experience building multi-agent systems
Experience deploying multi-agent systems into live production environments
Strong understanding of agent scaling, distributed systems, or AI infrastructure
Hands-on experience with one or more agent frameworks:
LangGraph, LangChain, AutoGen, CrewAI, Semantic Kernel, Google ADK, or custom agent frameworks
Experience with model-serving platforms such as vLLM, SGLang, Ray, NVIDIA Triton, or NVIDIA Dynamo
Strong full-stack engineering capabilities
Experience with containers, cloud infrastructure, deployment systems, APIs, async jobs, queues, and distributed systems
Experience developing with open-source software
Strong product instincts and the ability to make sophisticated backend capabilities understandable and useful to developers
Excellent written communication skills
Ability to operate independently in an early-stage, ambiguous, rapidly changing environment
โญ Strong Plus
Candidates with any of the following will stand out:
Experience as a Solutions Architect
Experience as a Forward Deployed Engineer
Contributions to open-source projects
Experience working at an AI agent development company or inference provider
Experience building products sold to enterprise CTO/CIO buyers
Experience with products combining an open-source core + managed cloud/service layer
Experience scaling AI compute or agent orchestration systems
Experience in startup or high-growth technical environments
PhD or MS in ML Systems / Computer Science / a closely related field from a strong program
๐ ๏ธ Technology Environment
AI / Agentic Systems:
LangGraph โข LangChain โข AutoGen โข CrewAI โข Semantic Kernel โข Google ADK โข MCP
ML Infrastructure:
vLLM โข SGLang โข Ray โข NVIDIA Triton โข NVIDIA Dynamo
Systems & Infrastructure:
Distributed Systems โข Software-Defined Networking โข Docker โข Kubernetes โข Cloud Infrastructure โข Deployment Systems โข APIs โข Async Jobs โข Queues
Engineering:
Python โข Full-Stack Development โข Open Source
Observability / Visualization:
Tableau โข Grafana
๐ซ This Role Is NOT a Good Fit If You Are:
Primarily from a traditional enterprise/non-technical background
Focused only on the application layer without systems, scaling, or infrastructure experience
Looking for a role where you are primarily managing rather than coding and building hands-on
Too far removed from day-to-day technical implementation
๐ Education / Experience Flexibility
Professional experience is highly valued. A PhD or MS from a top institution in a relevant ML systems field may substitute for some professional experience.
๐ Work Arrangement
On-site in Sunnyvale, California, with limited flexibility considered on a case-by-case basis.
๐ฐ Compensation & Benefits
Base Salary: $150,000โ$230,000
Equity: Up to 1%
Position: Full-time
Hiring: 1โ2 engineers
๐ Visa Support
The company is open to:
H-1B transfers
New H-1B applications
TN visas
OPT / eligible visa transfers
๐ Why This Opportunity?
This is an opportunity to work at the intersection of agentic AI, ML infrastructure, distributed systems, and enterprise software at an early-stage company.
You wonโt simply maintain an existing platformโyouโll help design, build, scale, and productize the systems that power the next generation of enterprise AI workflows.
If youโre an engineer who enjoys going deep technically, working directly with technical leadership, solving ambiguous systems problems, and turning cutting-edge AI research into production software, weโd love to hear from you.
๐ฉ Apply directly or message me with your resume/profile.
#Hiring #MachineLearning #MLEngineer #MLSystems #AI #AgenticAI #GenerativeAI #ArtificialIntelligence #DistributedSystems #AIInfrastructure #Python #LangGraph #LangChain #AutoGen #CrewAI #vLLM #SGLang #Ray #NVIDIA #Kubernetes #Docker #OpenSource #SoftwareEngineering #SiliconValley #Sunnyvale #SanFranciscoBayArea #TechJobs
Responsibilities
- Build and productionize multi-agentic AI systems
- Design scalable agent orchestration and infrastructure
- Develop full-stack applications primarily using Python
- Work with agent frameworks such as LangGraph, LangChain, AutoGen, CrewAI, Semantic Kernel, Google ADK, or equivalent
- Deploy and optimize model-serving infrastructure using technologies such as vLLM, SGLang, Ray, NVIDIA Triton, or NVIDIA Dynamo
- Build systems involving APIs, distributed systems, asynchronous jobs, queues, containers, deployment platforms, and cloud infrastructure
- Develop intuitive developer-facing products around complex ML infrastructure
- Create visualizations and product experiences using tools such as Tableau and Grafana
Qualifications
- 5+ years of professional software engineering experience
- 5+ years of ML systems engineering experience in production
- Strong experience building multi-agent systems
- Experience deploying multi-agent systems into live production environments
- Strong understanding of agent scaling, distributed systems, or AI infrastructure
- Hands-on experience with one or more agent frameworks
- Experience with model-serving platforms
- Strong full-stack engineering capabilities
Benefits
- Base Salary: $150,000โ$230,000
- Equity: Up to 1%
- Visa support for H-1B transfers and new applications
Skills mentioned
About Carnaby Fox
๐๐ฎ๐ฟ๐ป๐ฎ๐ฏ๐ ๐๐ผ๐ A premier executive search firm uniting visionary organisations with elite leadership and specialised talent. Spanning the US and Europe, we bridge the gap between leading businesses and the high-impact professionals who drive them forward. We are deep-domain experts operating across the most dynamic, fast-paced, and regulated sectors in the global economy. ๐ข๐๐ฟ ๐๐ผ๐ฟ๐ฒ ๐ ๐ฎ๐ฟ๐ธ๐ฒ๐๐ โข ๐๐ฒ๐ด๐ฎ๐น: We facilitate strategic lateral hires, from Partners and Senior Associates to General Counsel, for elite law firms and corporate legal departments. ย ย โข ๐ง๐ฒ๐ฐ๐ต & ๐๐ถ๐ป๐๐ฒ๐ฐ๐ต: We source visionary technologists, product leaders, and engineers who are building the future of digital finance and software. ย ย โข ๐๐ฒ๐ฎ๐น๐๐ต๐ง๐ฒ๐ฐ๐ต & ๐ ๐ฒ๐ฑ๐ง๐ฒ๐ฐ๐ต: We identify the rare talent capable of navigating the complex intersection of healthcare compliance, life sciences, and cutting-edge innovation. โข ๐-๐ฆ๐๐ถ๐๐ฒ ๐๐ฒ๐ฎ๐ฑ๐ฒ๐ฟ๐๐ต๐ถ๐ฝ: We place transformational CEOs, CTOs, Founders, and board-level executives equipped with the strategic foresight to scale businesses and disrupt markets. ๐ช๐ต๐ ๐ฃ๐ฎ๐ฟ๐๐ป๐ฒ๐ฟ ๐ช๐ถ๐๐ต ๐จ๐? Transatlantic Reach: With a robust network across the US and European markets, we fluidly source top-tier talent on a global scale to ensure the perfect technical and cultural fit. Industry Insiders: Our consultants are sector specialists. We understand the precise nuances of your business because we speak your language and intimately know your commercial ecosystem. Bespoke Executive Search: We look far beyond the CV. Through rigorous market mapping, targeted headhunting, and discreet engagement, we deliver exceptional candidates who elevate your entire organisation. Whether you are scaling a Fintech startup in London, expanding a MedTech enterprise in New York, or executing strategic legal hires across continents, Carnaby Fox is your trusted partner. Let's build the future, together.