Featured Job

Staff Infrastructure Engineer

New York, NY Full-time On-site $175k — $240k per year 09/30/2026 Job ID: 000278
Apply Now
Kubernetes Cluster scaling Resource management Operational improvements Infrastructure tooling

Summary

What you’ll impact

The role is a Member of Technical Staff, Infrastructure Engineer at the organization, responsible for designing, scaling, and operating the Kubernetes-based platform that powers AI scientist agents. The engineer will work closely with backend, ML, and research teams to ensure reliable, observable, and secure infrastructure for autonomous scientific discovery.

Responsibilities

What you'll do

  • Build, maintain, and improve Kubernetes-based infrastructure that supports agent workloads, research services, and internal platforms.
  • Contribute to cluster scaling, resource management, and operational improvements as platform usage grows.
  • Develop and maintain infrastructure tooling, automation, and deployment workflows that improve reliability and developer productivity.
  • Help implement monitoring, observability, and alerting systems to ensure platform health and performance.
  • Support storage, networking, and security initiatives within our Kubernetes environments.
  • Troubleshoot infrastructure and production issues across distributed systems and participate in incident response efforts.
  • Collaborate with backend, ML, and research teams to understand workload requirements and implement reliable infrastructure solutions.
  • Contribute to infrastructure best practices, documentation, and operational processes.

Requirements

What you’ll bring

  • 4+ years of software, infrastructure, platform, or DevOps engineering experience.
  • Experience working with Kubernetes in development or production environments.
  • Familiarity with cloud platforms such as AWS, GCP, or Azure.
  • Proficiency in at least one programming language such as Python, Go, Java, or TypeScript.
  • Experience with infrastructure-as-code tools such as Terraform, Pulumi, or similar technologies.
  • Understanding of containerized applications, networking fundamentals, and distributed systems concepts.
  • Experience using CI/CD pipelines, version control systems, and automated testing practices.
  • Ability to work independently while collaborating effectively across teams.
  • Curiosity, strong problem-solving skills, and a desire to learn new technologies.

Ready to Move Forward?

Apply now and our recruiting team will reach out with next steps, interview guidance, and client insights tailored to this role.