Anthropic
AI Infrastructure Operations, Demand Planning
About this role
Lead infrastructure capacity planning and delivery operations at Anthropic, managing large-scale accelerator fleet deployments across multiple clouds and regions. You'll convert demand forecasts into concrete tranche requirements, own the delivery schedule from contract through production, and drive the systems and reporting that keep Anthropic's rapidly growing compute infrastructure on track.
What you'll do
- Convert demand forecasts into per-tranche accelerator and infrastructure requirements, representing them in sourcing and data center negotiations
- Qualify infrastructure tranches for deliverability before contract signature, validating shape, region, supporting resources, and timing
- Track and close forecast-versus-actual variance on shape, region, and timing; publish results and feed back into planning
- Define and maintain the canonical state machine and system of record for capacity bring-up from contract through production
- Orchestrate parallel bring-ups across cloud regions, on-prem sites, and neocloud blocks with integrated scheduling
- Build automation and reporting for capacity readiness, time-to-production, and idle-cost metrics with executive visibility
What they're looking for
- Large-scale infrastructure delivery (multi-region cloud, HPC, accelerator clusters, or bare-metal fleets at 10k+ scale)
- Cluster orchestration, node health, and fleet telemetry systems
- SQL and Python for data analysis and custom reporting
- Demand planning, capacity forecasting, and variance analysis
- Cloud and neocloud reserved capacity, private offers, and commitment negotiations
- Data center operations: power, space, networking, site acceptance, vendor management
- Systems of record and infrastructure lifecycle management
- Hardware onboarding and health burn-in processes
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
Anthropic
Anthropic builds Claude, an AI assistant, and is hiring for engineering roles across infrastructure, data systems, and security that support both AI research operations and the company's internal technology needs. The company seeks infrastructure engineers, systems integrators, data scientists, and security specialists to build production-scale systems for training data pipelines, financial operations, developer productivity measurement, research infrastructure, and server firmware security.
- Website
- anthropic.com
Likely interview questions
- Tell us about the largest infrastructure deployment you've managed—what was the scale, how many regions or sites, and what was your specific role in bringing it to production?
- Walk us through a time when actual infrastructure delivery diverged significantly from forecast. How did you identify it, and what did you feed back into planning to prevent recurrence?