- Location
- United States
- Workplace
- Remote
- Employment
- Full time
- Level
- Staff
- Posted
- 2 months ago
About this role
Company Overview
TENEX is an AI-native, automation-first, built-for-scale Managed Detection and Response (MDR) provider. We are a force multiplier for defenders, helping organizations enhance their cybersecurity posture through advanced threat detection, rapid response, and continuous protection. Our team is composed of industry experts with deep experience in cybersecurity, automation, and AI-driven solutions. Backed by leading investors, we are rapidly growing and seeking top talent to join our mission of revolutionizing the AI-Native MDR landscape.
We’re a fast-growing startup backed by industry experts and top-tier investors led by Crosspoint Capital Partners and also backed by Shield Capital, DTCP (formerly Deutsche Telekom Capital Partners), Deepwork Capital, and the Florida Opportunity Fund. Seed round led by Andreessen Horowitz (a16z). As an early employee, you’ll play a meaningful role in defining and building our culture. Get in on the ground floor. We’re a small but well-funded team that just raised a substantial round - joining now comes with limited risk and unlimited upside.
As a Staff Site Reliability Engineer at TENEX, you will be a key technical driver responsible for ensuring the scalability, reliability, and performance of our AI-driven cybersecurity platform. You will play a crucial role in designing resilient infrastructure, automating operational workflows, and shaping the future of our production environments while collaborating across engineering teams to drive technical excellence.
Culture is one of the most important things at TENEX.AI - explore our culture deck at culture.tenex.ai to witness how we embody it, prioritizing the irreplaceable collaboration and community of in-person work.
Location: This role will require Monday - Thursday onsite in any of our locations. WFH Friday.
Job Responsibilities
System Resilience: Design, build, and maintain highly available, scalable, and secure infrastructure to support our AI-native cybersecurity platform.
Automation & Tooling: Develop internal tooling and automation to streamline deployment processes, incident response, and capacity planning.
Performance Engineering: Monitor system performance and proactively identify bottlenecks, optimizing infrastructure for low-latency, high-throughput AI workloads.
Incident Management: Lead incident response efforts, conduct post-mortems, and implement long-term solutions to prevent recurring reliability issues.
Infrastructure as Code (IaC): Manage infrastructure via code, driving consistency, auditability, and scalability across our cloud environments (e.g., AWS, GCP).
Cross-Functional Collaboration: Partner with sibling Engineering teams, Product, and Security teams to ensure reliability is baked into our development lifecycle from concept to production.
Required Skills & Qualifications
SRE & Infrastructure Expertise
Core Engineering: 10+ years of experience in SRE, DevOps, or Software/Systems Engineering, particularly in managing production systems at scale.
Cloud Infrastructure: Deep expertise in public cloud environments (AWS, GCP, or Azure) and managing services such as Kubernetes (EKS/GKE), networking, and storage.
Infrastructure as Code: Extensive experience with tools like Terraform, Pulumi, or similar technologies to manage complex infrastructure deployments.
Observability: Hands-on experience with monitoring, logging, and tracing stacks (e.g., Prometheus, Grafana, ELK, Datadog) to drive data-informed reliability decisions.
Distributed Systems: Solid understanding of microservices architecture, distributed databases, and event-driven systems.
Soft Skills
Communication: Clear, concise communication skills and a bias for collaborative problem-solving.
Leadership Alignment: Proven track record of guiding multi-stakeholder initiatives and influencing engineering practices across teams.
Analytical Rigor: Strong problem-solving, debugging, and analytical skills, especially in high-pressure environments.
Nice-to-have
Domain Background: Prior work in cybersecurity, specifically regarding SIEM, EDR, or SOAR infrastructure.
AI/ML Infrastructure: Experience supporting infrastructure for large-scale AI/ML workloads (e.g., GPU scheduling, LLM serving optimization).
Startup Mentality: Background driving high-impact engineering initiatives in high-growth startups or enterprise SaaS.
Strong familiarity with Agentic Workflows such as Agno, Temporal, etc..
Education & Certifications
Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field.
Relevant certifications (CKA/CKAD, AWS/GCP Professional Cloud Architect, etc.) are a plus.
Why Join Us?
Opportunity to work with cutting-edge AI-driven cybersecurity technologies and Google SecOps solutions.
Collaborate with a talented and innovative team focused on continuously improving security operations and system reliability.
Competitive salary and benefits package.
A culture of growth and development, with opportunities to expand your knowledge in AI, cybersecurity, and emerging technologies.
If you're passionate about building resilient infrastructure, scaling AI systems, and working at the intersection of reliability and security, we encourage you to apply!
As published by TENEX.AI. Applications are handled on their site.
Skills this posting mentions
About TENEX.AI
TENEX is a cybersecurity company leveraging advanced artificial intelligence and human expertise to transform enterprise security. Backed by Andreessen Horowitz (a16z) and Shield Capital, TENEX’s flagship offering is a next-generation Managed Detection and Response (MDR) service, transforming how organizations detect and respond to threats. With deep expertise in Google and Microsoft security ecosystems and state-of-the-art AI capabilities, TENEX empowers enterprises to enhance threat detection, agility, and resilience while maximizing the value of their security investments.
All 61 openings at TENEX.AIOne click, then it is written
Apply to TENEX.AI with a resume written for this role.
Queue Staff Site Reliability Engineer and I read the posting, rewrite your resume against it, draft the cover letter, and score the fit. Then you press send, or press one button and I fill in TENEX.AI’s form for you.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- 25 sent a week, free
- No card
- Nothing sent until you say so
More roles at TENEX.AI
See all- 6 weeks ago
- 7 weeks ago
- 7 weeks ago
- 8 weeks ago
- 8 weeks ago
- 8 weeks ago
Similar roles elsewhere
See more- Today
Senior Frontend Engineer, Ads Creative
RedditRemote - United StatesRemote
$191k - $267k/yrSeniorEngineering - Today
Senior Software Engineer, Agentic Ads Experience
RedditRemote - United StatesRemote
$217k - $304k/yrSeniorEngineering - Today
- Today
Staff Technical Product Manager, Ads ML Platform
RedditRemote - United StatesRemote
$217k - $304k/yrStaffEngineering
Put this to work
Paste your career in once. Every application after that is written for you.
Drop a resume or a LinkedIn URL. I rank the live openings against it, rewrite the resume and write a cover letter for the best of them, and fill in the employer's form when you press the button. You read, you decide what goes out.
01Drop your resume
A PDF or a LinkedIn URL. About a minute, once.
02I rank the openings
Every weekday morning, the live catalog scored against your history. Up to 20 worth your time, not two hundred links.
03Each one is written up
Resume rewritten for the posting, a cover letter, a fit score. Press send, or let me fill in the form.
- New matches ranked and written before you are up.
- Every bullet stays inside what your history supports. Nothing invented.
- Queued, submitted, interviewing, offer: one screen, not a spreadsheet.
500 free credits on sign-up. No card. Nothing is sent until you say so.
Listed from the job board TENEX.AI publishes. Refolk is not the employer and does not handle their hiring. Applications go to TENEX.AI directly.