H

Senior Site Reliability Engineer (SRE) – Cloud Infrastructure & Kubernetes

happyglobal service • United State
Visa Sponsorship
Apply Now

Lead the design, operation, and optimization of Bambu Lab’s U.S.-based cloud infrastructure, ensuring reliability, security, and scalability. Responsible for Kubernetes management, incident response, observability, and automation while collaborating with engineering and security teams. Requires 3+ years of SRE/DevOps experience with hands-on cloud (AWS/GCP) and IaC expertise.

Key Highlights
Build and maintain cloud infrastructure across AWS and GCP with Kubernetes reliability
Lead incident response, root-cause analysis, and postmortem reviews with on-call rotation
Implement DevSecOps practices, observability systems, and security governance for production environments
Key Responsibilities
Design, build, and maintain cloud infrastructure across AWS and GCP environments, including compute, networking, and storage services
Operate and optimize Kubernetes clusters to ensure platform stability, availability, and scalability in production
Lead Tier 2 incident response, conduct root-cause analysis, and drive post-incident reviews and remediation initiatives
Develop and optimize observability systems, including monitoring, alerting, and operational metrics, while defining SLIs and SLOs
Implement DevSecOps practices and enforce U.S.-based infrastructure security policies, collaborating with security teams
Automate cloud infrastructure and operational workflows using Terraform and scripting languages (Python, Go, Shell)
Collaborate with engineering, security, and infrastructure teams to improve system reliability and troubleshoot production issues
Technical Skills Required
Kubernetes Amazon Web Services / Google Cloud Platform Terraform
Benefits & Perks
Medical, dental, and vision insurance
401(k) retirement plan
Annual performance bonus and incentive compensation
Paid time off and professional development opportunities
Nice to Have
Bachelor's degree in Computer Science or related field
Experience defining SLIs/SLOs and reliability metrics
Experience with large-scale, highly available cloud environments
Familiarity with incident management and postmortem processes

Job Description

SRE Engineer – Cloud Environment

About the Job

The SRE Engineer – Cloud Environment will be responsible for building, operating, and improving Bambu Lab’s U.S.-based cloud infrastructure and production reliability environment. This role will work closely with engineering, security, and infrastructure teams and will play a key role in maintaining cloud stability, Kubernetes reliability, observability, incident response, and secure infrastructure operations.

Location: Santa Clara, CA

Work Model: Onsite

Employment Type: Full-Time

Client Details

Want the full job description? Read the complete details on LinkedIn, the original posting.
Continue on LinkedIn

This is a short excerpt. All rights to the full description belong to its original publisher.

See Jaabz jobs first on Google 1 tap · free · in AI Overviews Jaabz is on your Google Manage Preferred Sources

Similar Jobs

Explore other opportunities that match your interests

Applied Researcher 5 (AI Foundations - LLM, Optimization and Finetuning)

Devops
•
1h ago

Premium Job

Sign up is free! Login or Sign up to view full details.

•••••• •••••• ••••••
Job Type ••••••
Experience Level ••••••

construction job board usa

United State

Senior Data Engineer

Devops
•
2h ago

Premium Job

Sign up is free! Login or Sign up to view full details.

•••••• •••••• ••••••
Job Type ••••••
Experience Level ••••••

GEICO

United State
Visa Sponsorship Relocation Remote
Job Type Full-time
Experience Level Not Applicable

intellibee inc

United State

Subscribe our newsletter

New Things Will Always Update Regularly