Staff Site Reliability Engineer
Design resilient infrastructure for AI-driven cybersecurity platform scalability
As a Staff Site Reliability Engineer at TENEX.AI, you will design, build, and maintain highly available, scalable, and secure infrastructure to support the AI-native cybersecurity platform. You will develop internal tooling and automation to streamline deployment processes, incident response, and capacity planning while monitoring system performance and optimizing infrastructure for low-latency, high-throughput AI workloads.
Why This Role?
Play a meaningful role in defining and building company culture as an early employee
Key Responsibilities
- Design, build, and maintain highly available, scalable, and secure infrastructure for the AI-native cybersecurity platform
- Develop internal tooling and automation to streamline deployment processes, incident response, and capacity planning
- Monitor system performance and proactively identify bottlenecks to optimize infrastructure for low-latency, high-throughput AI workloads
- Lead incident response efforts, conduct post-mortems, and implement long-term solutions to prevent recurring issues
Requirements
- Experience designing resilient infrastructure for scalable platforms
- Background in developing internal tooling and automation for deployment and incident response
- Expertise in performance engineering and monitoring for AI workloads
- Proven ability to lead incident response and conduct post-mortems
Required Skills
Indonesia Context
- Working Hours Overlap:
- Flexible — work your own hours
Keywords
View Original Description from Ashby Job Boards
Original description from Ashby Job Boards
Company Overview TENEX is an AI-native, automation-first, built-for-scale Managed Detection and Response (MDR) provider. We are a force multiplier for defenders, helping organizations enhance their cybersecurity posture through advanced threat detection, rapid response, and continuous protection. Our team is composed of industry experts with deep experience in cybersecurity, automation, and AI-driven solutions. Backed by leading investors, we are rapidly growing and seeking top talent to join our mission of revolutionizing the AI-Native MDR landscape. We’re a fast-growing startup backed by industry experts and top-tier investors led by Crosspoint Capital Partners and also backed by Shield Capital, DTCP (formerly Deutsche Telekom Capital Partners), Deepwork Capital, and the Florida Opportunity Fund. Seed round led by Andreessen Horowitz (a16z). As an early employee, you’ll play a meaningful role in defining and building our culture. Get in on the ground floor. We’re a small but well-funded team that just raised a substantial round – joining now comes with limited risk and unlimited upside. As a Staff Site Reliability Engineer at TENEX, you will be a key technical driver responsible for ensuring the scalability, reliability, and performance of our AI-driven cybersecurity platform. You will play a crucial role in designing resilient infrastructure, automating operational workflows, and shaping the future of our production environments while collaborating across engineering teams to drive technical excellence. Culture is one of the most important things at TENEX.AI http://TENEX.AI—explore our culture deck at culture.tenex.ai http://culture.tenex.ai to witness how we embody it, prioritizing the irreplaceable collaboration and community of in-person work. Location: This role will require Monday - Thursday onsite in any of our locations. WFH Friday. Job Responsibilities - System Resilience: Design, build, and maintain highly available, scalable, and secure infrastructure to support our AI-native cybersecurity platform. - Automation & Tooling: Develop internal tooling and automation to streamline deployment processes, incident response, and capacity planning. - Performance Engineering: Monitor system performance and proactively identify bottlenecks, optimizing infrastructure for low-latency, high-throughput AI workloads. - Incident Management: Lead incident response efforts, conduct post-mortems, and implement long-term solutions to prevent recurring reliability issues. - Infrastructure as Code (IaC): Manage infrastructure via code, driving consistency, auditability, and scalability across our cloud environments (e.g., AWS, GCP). - Cross-Functional Collaboration: Partner with sibling Engineering teams, Product, and Security teams to ensure reliability is baked into our development lifecycle from concept to production. REQUIRED SKILLS & QUALIFICATIONS SRE & INFRASTRUCTURE EXPERTISE - Core Engineering: 10+ years of experience in SRE, DevOps, or Software/Systems Engineering, particularly in managing production systems at scale. - Cloud Infrastructure: Deep expertise in public cloud environments (AWS, GCP, or Azure) and managing services such as Kubernetes (EKS/GKE), networking, and storage. - Infrastructure as Code: Extensive experience with tools like Terraform, Pulumi, or similar technologies to manage complex infrastructure deployments. - Observability: Hands-on experience with monitoring, logging, and tracing stacks (e.g., Prometheus, Grafana, ELK, Datadog) to drive data-informed reliability decisions. - Distributed Systems: Solid understanding of microservices architecture, distributed databases, and event-driven systems. - SOFT SKILLS - Communication: Clear, concise communication skills and a bias for collaborative problem-solving. - Leadership Alignment: Proven track record of guiding multi-stakeholder initiatives and influencing engineering practices across teams. - Analytical Rigor: Strong problem-solving, debugging, and analytical skills, especially in high-pressure environments. NICE-TO-HAVE - Domain Background: Prior work in cybersecurity, specifically regarding SIEM, EDR, or SOAR infrastructure. - AI/ML Infrastructure: Experience supporting infrastructure for large-scale AI/ML workloads (e.g., GPU scheduling, LLM serving optimization). - Startup Mentality: Background driving high-impact engineering initiatives in high-growth startups or enterprise SaaS. - Strong familiarity with Agentic Workflows such as Agno, Temporal, etc.. EDUCATION & CERTIFICATIONS - Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field. - Relevant certifications (CKA/CKAD, AWS/GCP Professional Cloud Architect, etc.) are a plus. WHY JOIN US? - Opportunity to work with cutting-edge AI-driven cybersecurity technologies and Google SecOps solutions. - Collaborate with a talented and innovative team focused on continuously improving security operations and system reliability. - Competitive salary and benefits package. - A culture of growth and development, with opportunities to expand your knowledge in AI, cybersecurity, and emerging technologies. If you're passionate about building resilient infrastructure, scaling AI systems, and working at the intersection of reliability and security, we encourage you to apply!
Salary Context
Similar Engineering roles on LokerDollar pay around $215k/yr (range $19.575k–2745.996k/yr, n=438 active listings).
Hiring at TENEX.AI
TENEX.AI has 20 other active roles on LokerDollar and has been hiring here since May 3, 2026 — across Engineering, Data & Analytics.
View all TENEX.AI openings →Free account · no credit card · Log in
Pro $9/mo · unlimited applies + AI resume
Frequently asked questions
- Is Staff Site Reliability Engineer at TENEX.AI a remote job?
- Yes. Staff Site Reliability Engineer at TENEX.AI is a fully remote role open to candidates worldwide.
- What type of employment is Staff Site Reliability Engineer at TENEX.AI?
- This is a full time position.
- How do I apply?
- Click the "Apply" button on this page to go to the official application at TENEX.AI.
Explore related
Market data & reports
Salary & skill-demand research built from our own listings data.
- Indonesia IT Jobs vs Global Remote (2026)Primary analysis of 2,049 listings: methodology, classification rules, downloadable datasets.
- AI-Skill Demand: Indonesia vs Global Remote (2026)10,000+ postings, taxonomy-first classifier, Wilson CIs, pre-registered before analysis.
- Indonesia Hiring Report: Tech vs Non-TechJob demand by field from aggregate open-job counts — never individual listings.
- Indonesia Salary BenchmarkAggregate salary ranges across roles, with open methodology and dataset.
- Indonesian Remote Work Salary & Demand IndexHow much of the global remote job corpus is open to Indonesia, and what it pays (USD) by role.
- Indonesia Quarterly Labor Market ReportLayoffs, funding, salaries & skills per quarter — open aggregates.
- Remote Market Reports by RoleAuto-generated per role family — skills, seniority, companies, salary.
- Global Remote Salary BenchmarkAnnual salary by role & currency, plus the share of listings open worldwide.
From the blog
- AI Job Trends in June 2026Analysis of seven latest AI jobs on Loker Dollar: OpenAI, Mistral AI, Spotify, and others with competition and application strategies.
- Top Remote USD Jobs: July 2026 Global EditionDiscover the latest USD-paying remote jobs for engineers, designers, and operators worldwide in July 2026. Salary insights and application tips included!
- 10 Remote Jobs Nobody Wants (Yet!)Are there remote openings with few applicants? We'll show you 10 high-demand, low-supply positions – and why your skills might be a perfect fit.
Free account · no credit card · Log in
Pro $9/mo · unlimited applies + AI resume