
Job Overview
Location
San Francisco
Job Type
Full-time
Category
Engineering
Date Posted
August 18, 2026
Full Job Description
📋 Description
- • We're a fast-growing AI/ML platform startup building infrastructure for training, evaluating, and aligning AI models within reinforcement learning environments.
- • Our engineering team of ~15 includes competitive programming medalists, serial AI startup founders, and researchers published at top venues — and we're looking for a Platform Engineer to own the reliability, scale, performance, and developer experience of our core infrastructure.
- • This is a backend-architecture-heavy role with high ownership. Your work will directly determine how fast, reliable, and cost-effective our platform is to build on and run.
- • Own production uptime, latency, provisioning speed, infrastructure cost, and incident response for core platform services.
- • Build and maintain AWS infrastructure using Terraform, Kubernetes/EKS, Helm, Docker, EC2, CodeBuild, ECR, S3, IAM, networking, and secrets management.
- • Design and improve backend and platform systems for scale — capacity planning, autoscaling, queueing, backpressure, cleanup jobs, retries, and rollback paths.
- • Define and improve dashboards, alerts, logs, traces, SLOs, runbooks, and on-call workflows so failures are detected, debugged, and resolved quickly.
- • Build reliable CI/CD pipelines, release automation, environment management, and deployment workflows that improve developer productivity and reduce production risk.
- • Write clean, maintainable production code to automate systems, improve backend services, and create internal developer tooling.
🎯 Requirements
- • 2–4 years of experience owning production cloud infrastructure for a high-availability, user-facing platform, with accountability for uptime, performance, deployment safety, and cost.
- • Deep hands-on experience with AWS and containerized systems; strong familiarity with Terraform, Kubernetes/EKS, Docker, EC2, load balancers, networking, and secrets management.
- • Track record of building or operating CI/CD, release automation, observability, alerting, and incident response systems.
- • Strong backend engineering judgment — able to reason about service architecture, APIs, databases, async systems, queues, scaling limits, and production failure modes.
- • Ability to write production-quality code to automate infrastructure, improve backend services, and build internal tooling.
🏖️ Benefits
- • Salary: $150,000 – $250,000 USD annually (for full-time roles)
- • Opportunity to have significant ownership and direct impact at an early-stage, well-funded AI infrastructure company.
- • Work alongside a world-class technical team building foundational infrastructure for AI alignment and post-training data.
Skills & Technologies
See exactly how your profile matches this role — strengths, skill gaps, and what to do about them.
About Clera
Clera, Inc. operates as an AI Talent Agent, leveraging artificial intelligence to enhance and streamline various aspects of the talent acquisition and management process. The company's core offering appears to center on an AI-powered platform designed to connect skilled individuals with suitable opportunities or to assist organizations in efficiently sourcing and evaluating candidates. While comprehensive details regarding specific features, target industries, or the scale of its operations are not yet publicly available, the current online presence indicates that Clera's services are actively being prepared for launch. This suggests an upcoming introduction of innovative AI solutions aimed at optimizing the ecosystem where talent meets demand.
Subscribe to the weekly newsletter for similar remote roles and curated hiring updates.
Newsletter
Weekly remote jobs and featured talent.
No spam. Only curated remote roles and product updates. You can unsubscribe anytime.
Similar Opportunities

NETGEAR, Inc.
2 months ago

Unilever PLC
2 months ago

Silver.com LLC
2 months ago

Latamcent
3 months ago