C

Senior Machine Learning Systems Engineer (AI Agents & Distributed Infrastructure)

carnaby fox β€’ United State
Visa Sponsorship
Apply
AI Summary

Join an early-stage AI company to design, build, and scale multi-agent enterprise workflow systems. Bridge ML research, distributed infrastructure, and full-stack development to create production-grade AI agent platforms for developers and enterprise customers. Requires 5+ years of ML systems engineering experience with hands-on expertise in agent orchestration, model serving, and scalable infrastructure.

Key Highlights
Hands-on role bridging ML systems research, agentic AI, and distributed infrastructure
Direct collaboration with CEO and Chief Architect to productize cutting-edge AI systems
Full-stack development in Python with focus on agent frameworks, model serving, and cloud infrastructure
Key Responsibilities
Build and productionize multi-agentic AI systems for enterprise workflows
Design scalable agent orchestration and distributed infrastructure for AI platforms
Develop full-stack applications in Python, integrating agent frameworks and model-serving infrastructure
Deploy and optimize model-serving systems using vLLM, SGLang, Ray, or NVIDIA Triton
Create developer-friendly products and experiences around complex ML infrastructure
Collaborate with technical leadership to shape architecture, engineering practices, and product direction
Translate research-grade ML concepts into production-ready documentation and onboarding materials
Technical Skills Required
Python Distributed Systems Machine Learning Systems
Benefits & Perks
$150,000–$230,000 base salary
Up to 1% equity
Visa sponsorship (H-1B transfers, new H-1B applications, TN visas)
Nice to Have
Experience as a Solutions Architect or Forward Deployed Engineer
Contributions to open-source projects in ML/AI systems
PhD or MS in ML Systems/Computer Science from a top institution
Experience scaling AI compute or agent orchestration systems
Experience building enterprise products with open-source + managed cloud layers

Job Description


πŸš€ Senior ML Engineer (Systems) | AI Agents & Distributed Systems


Location: Sunnyvale, CA β€” On-site

Employment Type: Full-time

Experience: 5+ years

Compensation: $150K–$230K + up to 1% equity

Visa: H-1B transfers, new H-1B applications & TN visas supported


We’re hiring a Senior ML Engineer (Systems) to join an early-stage AI company building infrastructure for the next generation of multi-agent enterprise workflows.

This is a highly hands-on role for an engineer who can bridge ML systems research, agentic AI, distributed infrastructure, full-stack development, and product engineering.

You’ll work closely with the CEO and Chief Architect to turn research-grade ML systems ideas into products that developers and enterprise customers genuinely enjoy using.

πŸ”₯ What You’ll Do

  • Build and productionize multi-agentic AI systems
  • Design scalable agent orchestration and infrastructure
  • Develop full-stack applications primarily using Python
  • Work with agent frameworks such as LangGraph, LangChain, AutoGen, CrewAI, Semantic Kernel, Google ADK, or equivalent
  • Deploy and optimize model-serving infrastructure using technologies such as vLLM, SGLang, Ray, NVIDIA Triton, or NVIDIA Dynamo
  • Build systems involving APIs, distributed systems, asynchronous jobs, queues, containers, deployment platforms, and cloud infrastructure
  • Develop intuitive developer-facing products around complex ML infrastructure
  • Create visualizations and product experiences using tools such as Tableau and Grafana
  • Work extensively with open-source software, with opportunities to contribute upstream
  • Translate research-grade concepts into documentation, examples, onboarding experiences, and product language
  • Collaborate directly with technical leadership in an ambiguous, fast-moving startup environment
  • Help shape architecture, engineering practices, and the product itself


🧠 Ideal Candidate

You’ll be a strong fit if you have:

(academic years can substitute if PhD from top institution in relevant ML systems field)


  • 5+ years of professional software engineering experience
  • 5+ years of ML systems engineering experience in production
  • Strong experience building multi-agent systems
  • Experience deploying multi-agent systems into live production environments
  • Strong understanding of agent scaling, distributed systems, or AI infrastructure
  • Hands-on experience with one or more agent frameworks:
  • LangGraph, LangChain, AutoGen, CrewAI, Semantic Kernel, Google ADK, or custom agent frameworks
  • Experience with model-serving platforms such as vLLM, SGLang, Ray, NVIDIA Triton, or NVIDIA Dynamo
  • Strong full-stack engineering capabilities
  • Experience with containers, cloud infrastructure, deployment systems, APIs, async jobs, queues, and distributed systems
  • Experience developing with open-source software
  • Strong product instincts and the ability to make sophisticated backend capabilities understandable and useful to developers
  • Excellent written communication skills
  • Ability to operate independently in an early-stage, ambiguous, rapidly changing environment

⭐ Strong Plus

Candidates with any of the following will stand out:

  • Experience as a Solutions Architect
  • Experience as a Forward Deployed Engineer
  • Contributions to open-source projects
  • Experience working at an AI agent development company or inference provider
  • Experience building products sold to enterprise CTO/CIO buyers
  • Experience with products combining an open-source core + managed cloud/service layer
  • Experience scaling AI compute or agent orchestration systems
  • Experience in startup or high-growth technical environments
  • PhD or MS in ML Systems / Computer Science / a closely related field from a strong program

πŸ› οΈ Technology Environment

AI / Agentic Systems:

LangGraph β€’ LangChain β€’ AutoGen β€’ CrewAI β€’ Semantic Kernel β€’ Google ADK β€’ MCP

ML Infrastructure:

vLLM β€’ SGLang β€’ Ray β€’ NVIDIA Triton β€’ NVIDIA Dynamo

Systems & Infrastructure:

Distributed Systems β€’ Software-Defined Networking β€’ Docker β€’ Kubernetes β€’ Cloud Infrastructure β€’ Deployment Systems β€’ APIs β€’ Async Jobs β€’ Queues

Engineering:

Python β€’ Full-Stack Development β€’ Open Source

Observability / Visualization:

Tableau β€’ Grafana

🚫 This Role Is NOT a Good Fit If You Are:

  • Primarily from a traditional enterprise/non-technical background
  • Focused only on the application layer without systems, scaling, or infrastructure experience
  • Looking for a role where you are primarily managing rather than coding and building hands-on
  • Too far removed from day-to-day technical implementation

πŸŽ“ Education / Experience Flexibility

Professional experience is highly valued. A PhD or MS from a top institution in a relevant ML systems field may substitute for some professional experience.

πŸ“ Work Arrangement

On-site in Sunnyvale, California, with limited flexibility considered on a case-by-case basis.

πŸ’° Compensation & Benefits

Base Salary: $150,000–$230,000

Equity: Up to 1%

Position: Full-time

Hiring: 1–2 engineers

πŸ›‚ Visa Support

The company is open to:

  • H-1B transfers
  • New H-1B applications
  • TN visas
  • OPT / eligible visa transfers

🌟 Why This Opportunity?

This is an opportunity to work at the intersection of agentic AI, ML infrastructure, distributed systems, and enterprise software at an early-stage company.

You won’t simply maintain an existing platformβ€”you’ll help design, build, scale, and productize the systems that power the next generation of enterprise AI workflows.

If you’re an engineer who enjoys going deep technically, working directly with technical leadership, solving ambiguous systems problems, and turning cutting-edge AI research into production software, we’d love to hear from you.

πŸ“© Apply directly or message me with your resume/profile.

#Hiring #MachineLearning #MLEngineer #MLSystems #AI #AgenticAI #GenerativeAI #ArtificialIntelligence #DistributedSystems #AIInfrastructure #Python #LangGraph #LangChain #AutoGen #CrewAI #vLLM #SGLang #Ray #NVIDIA #Kubernetes #Docker #OpenSource #SoftwareEngineering #SiliconValley #Sunnyvale #SanFranciscoBayArea #TechJobs


Similar Jobs

Explore other opportunities that match your interests

Senior Cloud Security Software Engineer (Golang, AI-Augmented Development)

Devops
β€’
2h ago

Premium Job

Sign up is free! Login or Sign up to view full details.

β€’β€’β€’β€’β€’β€’ β€’β€’β€’β€’β€’β€’ β€’β€’β€’β€’β€’β€’
Job Type β€’β€’β€’β€’β€’β€’
Experience Level β€’β€’β€’β€’β€’β€’

Palo Alto Networks

United State
Visa Sponsorship Relocation Remote
Job Type Full-time
Experience Level Entry level

Veridian Tech Solutions, Inc.

United State
Visa Sponsorship Relocation Remote
Job Type Full-time
Experience Level Not Applicable

inferact

United State

Subscribe our newsletter

New Things Will Always Update Regularly