hiringremote.Get job alerts
Still open

Cloud / DevOps Engineer (Infra & IaC)

AI expert network · Remote · 40 hrs/week

Pay
$75–110/hr
Where
US
Degree
See who gets hired
Posted
43d ago
Slots left
10
Checked
today

An older role: first posted 43d ago, and AI expert network still listed it when we checked today. Newer roles tend to fill faster. See the jobs hiring now.

What the work is

Join a leading AI lab's cutting-edge GenAI team to be at the core of the AI revolution, where your expertise fuels the development of the most advanced Large Language Models.

Overview

Join a leading AI lab's cutting-edge GenAI team and help build foundational AI models from the ground up. We're seeking talented Cloud and DevOps Engineers with deep, hands-on expertise in Kubernetes operations, AWS service integration, and Infrastructure-as-Code to bring hands-on technical excellence and elevate the quality of our AI training and inference infrastructure data. This is a full-time commitment of 40 hours per week.

Key Responsibilities

  • Guide research and engineering teams to close knowledge gaps and improve AI model performance in cloud infrastructure, Kubernetes operations, and Infrastructure-as-Code topics.
  • Design challenging, domain-relevant tasks, and write accurate and well-structured solutions to cluster diagnosis, AWS service integration, and infrastructure automation problems.
  • Evaluate infrastructure engineering tasks and solutions and provide clear, written technical feedback.
  • Develop guidelines and detailed rubrics/evaluation frameworks to assess cluster failure diagnosis, IaC design quality, and CI/CD pipeline reasoning across tasks.
  • Collaborate with other subject matter experts to ensure consistency and accuracy in training data.

Who gets hired

Core Qualifications

  • 4+ years of dedicated professional experience in cloud infrastructure, DevOps, site reliability, or platform engineering at a recognized, top-tier organization.
  • Hands-on production experience operating Kubernetes, including diagnosing and repairing cluster failures — not solely authoring manifests or using a managed control plane.
  • Production experience with Infrastructure-as-Code, specifically Terraform and/or AWS CDK.
  • Direct experience integrating AWS services in production, including Lambda, API Gateway, and DynamoDB.
  • Experience building and owning CI/CD pipelines.
  • Demonstrable career progression.
  • Ability to engage reliably for at least 40 hours/week during weekdays.
  • Strong written communication skills and the ability to explain complex technical decisions clearly.

Pay

$75–110/hr, fully remote.

Check history

How we score →
today10 slotsOpen
Apply now$75–110/hr