AI summary
Design and build AI infrastructure for large language models on Kubernetes, focusing on deployment, scaling, and model lifecycle management, while integrating with APIs and ensuring GPU inference observability.
Job Description
Mirantis is building a new enterprise AI infrastructure product that lets organizations run and govern large language models on their own Kubernetes clusters. You will join a small senior team early, with broad ownership of the model-serving layer and its path to production.
What you'll do
- Design and build LLM serving infrastructure on Kubernetes: deployment, GPU scheduling, scaling, and model lifecycle management.
- Package the platform for enterprise environments: Helm-based installs, upgrades, and restricted/offline networks.
- Integrate the serving layer with the platform's API gateway, identity, and metering services.
- Build the observability for operating GPU inference in production (serving metrics, GPU telemetry).
- Contribute across a multi-service codebase and help set engineering direction through design docs and reviews.
Requirements
What we're looking for
- 5+ years of software engineering experience in infrastructure, platform, or distributed systems.
- Deep hands-on Kubernetes experience: building and operating production workloads and Helm charts, not just consuming managed clusters.
- Experience with GPU workloads or LLM inference, or strong adjacent systems experience and a track record of learning fast.
- Strong Go programming skills; solid CI/CD and infrastructure-as-code skills.
- Fluency with AI-assisted development tools (Claude Code, OpenAI Codex) as part of your daily engineering workflow.
- Comfortable with high autonomy on a small, remote-first, written-culture team.
Nice to have
- Inference performance work (quantization, batching, caching) or distributed serving frameworks.
- Enterprise deployment experience: air-gapped installs, SSO/OIDC, supply-chain security.
- UI development experience (e.g. React/TypeScript), useful as the product's management surfaces grow.
- Open-source contributions in the Kubernetes or ML-infrastructure ecosystems.
About the job
- Posted on
- Sep 15, 2026
- Job type
- Full-time
- Location
- Remote, USA, usRemote
Keep looking
Related roles you might like
Technical Product Manager, Observability – remote in the US
Mirantis
AI Infrastructure & Platform Operations Engineer (remote in the EU)
Mirantis
Senior Storage Systems Engineer - remote in the US
Mirantis
AI Storage Infrastructure Engineer - remote in the US
Mirantis
Senior Technical Product Marketing Manager, k0rdent AI - Remote
Mirantis
Head of Solutions Architecture - AI Infrastructure
Mirantis
Head of Solutions Architecture - AI Infrastructure
Mirantis
Director, Presales Solution Architecture - NeoCloud
Mirantis
Business Analyst – Cloud Commerce & SaaS (Remote in Poland)
Mirantis
Disclaimer: Real Jobs From Anywhere is an independent platform dedicated to providing information about job openings. We are not affiliated with, nor do we represent, any company, agency, or agent mentioned in the job listings. Please refer to our Terms of Services for further details.
