Member of Technical Staff, Platform - AI Infrastructure
1736186
Posted: 17/09/2026
- $250,000 Base Salary
- United States
- Permanent
- 250000
- Artificial Intelligence
- AI Data Center
Keen to join a company that champions growth and development?
Join a high-growth infrastructure company that scales specialized hardware within data center environments for advanced artificial intelligence workloads. The organization builds robust cloud platforms that expose high-performance hardware as elastic compute via developer-friendly interfaces. Its technology enables software teams to develop agentic capabilities and deploy complex, long-running AI models.
If you would like to learn more about this opportunity, feel free to reach out and apply today!
Responsibilities:
- Design, build, and operate managed Kubernetes and other customer-facing infrastructure services.
- Build control-plane services, Kubernetes operators, and controllers.
- Automate managed-service lifecycle, including cluster creation, version upgrades, configuration changes, and safe deletion. Coordinate with Fleet for the underlying capacity.
- Build APIs and tools that customers use to deploy, manage, and monitor workloads.
- Integrate orchestration with compute, networking, storage, identity, and security systems.
- Improve service availability, isolation, performance, observability, and upgrade safety.
- Work directly with customers and Product to define and deliver new platform capabilities.
- Use coding agents throughout design, implementation, testing, debugging, and operations.
- Participate in on-call and own the systems you build in production.
Skills/Must Have:
- Built customer-facing infrastructure products or managed cloud services.
- Experience building distributed systems, control planes, operators, or controllers.
- Deep knowledge of Kubernetes internals and cluster lifecycle management.
- Strong programming skills in Go, Python, Rust, or a similar language.
- Strong knowledge of Linux, containers, networking, storage, and security.
- Experience building reliable multi-tenant services and public APIs.
- Led complex infrastructure work from design through production.
- Experience using coding agents to build production software.
Desirable Skills:
- Built a managed Kubernetes, Slurm, inference, database, or container service.
- Built Kubernetes schedulers, operators, admission controllers, CNI plugins, or CSI plugins.
- Worked on bare-metal, HPC, or AI compute infrastructure.
- Built private networking, workload isolation, or enterprise identity features.
- Operated macOS or Apple Silicon infrastructure at scale.
- Built infrastructure that software agents can operate safely.
Benefits:
- Equity, Health Care etc
Salary:
- $250,000 Base Salary
Sam Merchant
Principal AI Infrastructure Consultant