
Job Overview
Location
Remote, USA
Job Type
Full-time
Category
Engineering
Date Posted
July 16, 2026
Full Job Description
đź“‹ Description
- • As a Staff Site Reliability Engineer at TENEX, you will be a key technical driver responsible for ensuring the scalability, reliability, and performance of our AI-driven cybersecurity platform.
- • You will play a crucial role in designing resilient infrastructure, automating operational workflows, and shaping the future of our production environments while collaborating across engineering teams to drive technical excellence.
- • System Resilience: Design, build, and maintain highly available, scalable, and secure infrastructure to support our AI-native cybersecurity platform.
- • Automation & Tooling: Develop internal tooling and automation to streamline deployment processes, incident response, and capacity planning.
- • Performance Engineering: Monitor system performance and proactively identify bottlenecks, optimizing infrastructure for low-latency, high-throughput AI workloads.
- • Incident Management: Lead incident response efforts, conduct post-mortems, and implement long-term solutions to prevent recurring reliability issues.
- • Infrastructure as Code (IaC): Manage infrastructure via code, driving consistency, auditability, and scalability across our cloud environments (e.g., AWS, GCP).
- • Cross-Functional Collaboration: Partner with sibling Engineering teams, Product, and Security teams to ensure reliability is baked into our development lifecycle from concept to production.
🎯 Requirements
- • SRE & INFRASTRUCTURE EXPERTISE
- • Core Engineering: 10+ years of experience in SRE, DevOps, or Software/Systems Engineering, particularly in managing production systems at scale.
- • Cloud Infrastructure: Deep expertise in public cloud environments (AWS, GCP, or Azure) and managing services such as Kubernetes (EKS/GKE), networking, and storage.
- • Infrastructure as Code: Extensive experience with tools like Terraform, Pulumi, or similar technologies to manage complex infrastructure deployments.
- • Observability: Hands-on experience with monitoring, logging, and tracing stacks (e.g., Prometheus, Grafana, ELK, Datadog) to drive data-informed reliability decisions.
- • Distributed Systems: Solid understanding of microservices architecture, distributed databases, and event-driven systems.
🏖️ Benefits
- • Opportunity to work with cutting-edge AI-driven cybersecurity technologies and Google SecOps solutions.
- • Collaborate with a talented and innovative team focused on continuously improving security operations and system reliability.
- • Competitive salary and benefits package.
- • A culture of growth and development, with opportunities to expand your knowledge in AI, cybersecurity, and emerging technologies.
Skills & Technologies
See exactly how your profile matches this role — strengths, skill gaps, and what to do about them.
About Tenex.AI, Inc
Tenex.AI is a cybersecurity company that offers an AI-native managed detection and response (MDR) platform. It combines automated threat detection, risk management, and incident response capabilities with human oversight to identify, contain, and remediate security incidents in real time. Tenex integrates with cloud and security stacks from providers like Google, Microsoft, and others to reduce response times and streamline operations.
Subscribe to the weekly newsletter for similar remote roles and curated hiring updates.
Newsletter
Weekly remote jobs and featured talent.
No spam. Only curated remote roles and product updates. You can unsubscribe anytime.
Similar Opportunities

NETGEAR, Inc.
8 days ago

Unilever PLC
19 days ago

Silver.com LLC
10 days ago

Latamcent
2 months ago