A

Backend Engineer, Cloud Inference Scaling

anthropic • United State
Visa Sponsorship
Apply Now

Design and build backend services to scale Claude across multiple cloud providers, ensuring reliable and cost-effective inference at massive scale. Collaborate with internal teams and CSP partners to manage capacity, routing, and operational issues. Requires high-performance distributed systems experience and cross-functional collaboration skills.

Key Highlights
Scale Claude across AWS, GCP, Azure, and future CSPs
Design cross-cloud abstractions for cost-effective inference management
Build CI/CD pipelines for reliable model deployment
Analyze observability data to optimize performance and cost
Work 25% hybrid onsite in San Francisco office
Key Responsibilities
Design, build, and own backend services and infrastructure that serve Claude across multiple CSPs, accounting for differences in compute hardware, networking, APIs, and operational models
Work cross-functionally with internal inference, product API, systems, and security teams, and with CSP partners to stand up the full serving stack on new cloud platforms, resolve operational issues, and influence provider roadmaps
Build and evolve CI/CD automation systems, including validation and deployment pipelines, that reliably ship new model versions to millions of users across cloud platforms without regressions
Design interfaces and tooling abstractions across CSPs that enable cost-effective inference management, scale across providers, and reduce per-platform complexity
Contribute to capacity planning, autoscaling, and workload routing strategies that match supply with demand and direct requests to the most cost-effective accelerator and region
Analyze observability data across providers to identify performance bottlenecks, cost anomalies, and regressions, and drive remediation based on real-world production workloads
Technical Skills Required
Python Rust Kubernetes Infrastructure as Code
Benefits & Perks
Competitive compensation
Generous vacation and parental leave
Flexible working hours
Visa sponsorship available
Nice to Have
Hands-on experience with capacity management, cost optimization, or resource planning at scale across heterogeneous environments
Solid understanding of multi-region deployments, geographic routing, and global traffic management

Job Description

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.The Cloud Inference team scales and optimizes Claude to serve the massive audiences of developers and enterprise companies across AWS, GCP, Azure, and future cloud service providers (CSPs). We own the end-to-end product of Claude on each cloud platform, from API integration and intelligent request routing to inference execution, capacity management, and day-to-day operations.
Want the full job description? Read the complete details on LinkedIn, the original posting.
Continue on LinkedIn

This is a short excerpt. All rights to the full description belong to its original publisher.

See Jaabz jobs first on Google 1 tap · free · in AI Overviews Jaabz is on your Google Manage Preferred Sources

Similar Jobs

Explore other opportunities that match your interests

Visa Sponsorship Relocation Remote
Job Type Contract
Experience Level Mid-Senior level

Focuz Mindz Inc.

United State

Senior Cloud Inference Engineer (Model & Inference Launch) - AI Systems

Devops
•
57m ago

Premium Job

Sign up is free! Login or Sign up to view full details.

•••••• •••••• ••••••
Job Type ••••••
Experience Level ••••••

anthropic

United State
Visa Sponsorship Relocation Remote
Job Type Full-time
Experience Level Not Applicable

Addison Group

United State

Subscribe our newsletter

New Things Will Always Update Regularly