Anthropic logo
Anthropic Posted 30+ days ago

Staff Infrastructure Engineer, Cluster Infrastructure

Location

London

Contract

Permanent

Full-time

Salary

Not disclosed

Experience

10+ years

Lead level

Job summary

Responsibilities

Anthropic's Infrastructure organization is foundational to our mission of developing AI systems that are reliable, interpretable, and steerable. The systems we build determine how quickly we can train new models, how reliably we can run safety experiments, and how effectively we can scale Claude…

  • Own the technical strategy and roadmap for agent-driven cluster lifecycle management - provisioning, updates and decommissioning
  • Partner across teams to ensure new compute capacity is ingested on time
  • Align with partner teams on physical build-out and leverage cloud solutions to deliver high-bandwidth inter-cluster connectivity
  • Collaborate with security owners to ensure clusters are provisioned secure-by-default
  • Define and drive strategy on cluster scalability, homogeneity and fault tolerance

What they're looking for

  • Deep expertise in distributed systems, reliability, and cloud platforms (e.g., Kubernetes, IaC, AWS/GCP/Azure)
  • Strong proficiency in at least one systems language (e.g., Rust, Go, or Python), IaC proficiency with Terraform.
  • Track record of leading complex, multi-quarter technical initiatives spanning multiple teams or systems
  • Ability to build alignment across senior stakeholders and communicate effectively at all levels
  • 10+ years of software engineering experience, including time as a technical lead setting direction for a team

Tech stack

Read the full job posting

This summary is written by Tokn from the original posting. Only the posting published by Anthropic is authoritative. Request a removal

job-boards.greenhouse.io ↗