Staff DevOps Engineer
[Build the next generation of an AWS platform while modernizing a mission-critical estate]
Compensation: Up to $200K base + bonus in the US
Location: Remote — US or Canada
The Company
People in AI is partnering with an established global SaaS business hiring a Staff DevOps Engineer into its infrastructure organization.
The business operates a large-scale AWS platform used by customers around the world and is currently running two infrastructure worlds in parallel: a mature, business-critical AWS estate and a modern platform built around containers, EKS, GitOps and infrastructure-as-code.
This creates a genuinely interesting engineering problem: keep an established platform secure and reliable while helping design and build what replaces it.
The Opportunity
This is a senior individual-contributor role for an engineer who enjoys building, not simply operating systems designed by somebody else.
You’ll work across both mature infrastructure and modern cloud-native projects, with significant scope to influence architecture, automation, developer tooling and the ongoing modernization of the platform.
The team is specifically looking for someone who can take an ambiguous infrastructure problem, design the solution and then remain hands-on through implementation.
The Role
You’ll operate at Staff level across AWS infrastructure, reliability, platform engineering and DevOps architecture.
The environment combines long-running EC2/Linux infrastructure with newer services running on EKS, giving you the chance to work on both complex modernization projects and greenfield platform development.
What You’ll Do
Architect and build scalable, secure and highly available infrastructure on AWS.
Design and implement new platform capabilities for services running on EKS.
Modernize existing EC2/Linux infrastructure and help containerize legacy workloads over time.
Build and improve Terraform/Terragrunt modules, automation and internal infrastructure tooling.
Improve CI/CD and deployment workflows across Jenkins, GitHub Actions and GitOps-based environments.
Troubleshoot complex production issues, contribute to RCA and implement long-term preventative fixes.
Work closely with software engineers and other infrastructure specialists on architecture, security, reliability and scalability.
What You’ll Bring
7+ years of hands-on experience across DevOps, SRE, Cloud Infrastructure or Platform Engineering.
Deep and recent experience with AWS and Terraform in production environments.
Strong Linux, EC2, networking, security and troubleshooting fundamentals.
Evidence that you have personally designed and built infrastructure or platforms from scratch.
Experience with CI/CD tooling such as Jenkins and GitHub Actions.
Strong communication skills and the ability to explain technical decisions, trade-offs and failures in detail.
What This Role Requires
You should be comfortable working with both modern cloud-native infrastructure and older production systems.
You need to be someone who creates patterns and solutions rather than simply follows existing runbooks.
You should enjoy technical ownership and solving infrastructure problems without a predefined playbook.
You must be comfortable participating in an on-call environment and working through complex production incidents.
Tech Stack
AWS, EC2, EKS, Linux, Terraform, Terragrunt, Chef, Jenkins, GitHub Actions, GitOps, Bash, Python, MongoDB, Postgres, Datadog, Prometheus and Vercel.
Why Join
Work on genuine infrastructure modernization rather than a superficial “cloud transformation.”
Combine greenfield platform engineering with technically challenging production systems.
Take Staff-level ownership while remaining deeply hands-on.
Influence architecture and tooling used across the broader engineering organization.
About People in AI
People in AI is a specialist search firm focused on AI, software engineering, infrastructure and data. We partner with high-growth technology companies across the US and Canada to hire exceptional technical talent.