Principal DevOps Engineer

Tenex
San Jose +4 more
Hybrid

Who this role is best for

Aimed at senior DevOps engineers with Azure platform expertise who can lead SRE and DevSecOps initiatives.

Best fit for

  • Candidates with 8+ years of Azure platform engineering experience and a track record in enterprise SaaS systems
    — “8+ years of progressive experience in DevOps, Site Reliability Engineering (SRE), or Platform Engineering roles.
  • Individuals who have led technical initiatives in high-growth startups and can drive automation and operational excellence
    — “Background leading technical initiatives in high-growth startups.
  • Candidates with hands-on experience in containerization, orchestration, and event-driven systems
    — “Automate deployment, scaling, and management of microservices and event-driven systems using containerization and orchestration technologies.

Things to consider

  • Onsite presence required in Kansas City, MO for four days a week, with limited remote flexibility.
    — “Monday through Thursday onsite in our Kansas City office (preferred), with San Jose or Sarasota, FL also considered. WFH Friday.
  • The role demands a deep understanding of Azure-specific tools and practices, not general cloud knowledge.
    — “Deep, hands-on expertise building production Azure platforms, including AKS, Entra ID and workload identity federation, VNet design and Private Link, Key Vault, Azure Policy, and subscription or landing zone architecture.

How to stand out

  • Emphasize experience with Azure governance, network topology, and identity management in your resume.
    — “Own our Azure governance and environment model, including subscription and management group structure, Azure Policy, network topology, and identity.
  • Highlight your ability to define and implement SLOs and SLIs in your interview responses.
    — “Lead Site Reliability Engineering (SRE) initiatives, defining and driving adherence to critical Service Level Objectives (SLOs) and Service Level Indicators (SLIs).
  • Showcase your work with infrastructure-as-code tools like Terraform or Bicep in your portfolio or resume.
    — “Extensive experience with Infrastructure-as-Code tools (e.g., Terraform, Bicep) and CI/CD best practices.
  • Demonstrate your ability to handle large-scale data pipelines and event-driven architectures.
    — “Experience with real-time data pipelines and stream processing (e.g., Kafka, Event Hubs, Service Bus, Pub/Sub).
  • Frame your experience in building secure and compliant environments as a core strength.
    — “Experience building secure and compliant (e.g., SOC 2, ISO 27001) environments.
Pace · SteadyCollaboration · HighAutonomy · HighDecision Impact · Company

Derived from job-description analysis by Serendipath's career intelligence engine.

What success looks like

  • high availability
  • secure and performant platform
  • automation and operational excellence
Typical background
devopssite reliability engineeringcloud platform engineering

Skills & requirements

Required

Azure PlatformSite Reliability EngineeringCi/cd PipelinesDevSecOpsMicroservicesContainerizationOrchestration

Preferred

Ai-native MdrCybersecurityAutomation-first

Stack & domain

AzureDevOpsKubernetesDockerCI/CDSREAIMLPythonLeadershipProblem-solvingTeamworkCommunicationAzure Solutions Architect ExpertCybersecurityCloudAutomation

About the role

Original posting from Tenex via Ashby

COMPANY OVERVIEW

TENEX is an AI-native, automation-first, built-for-scale Managed Detection and Response (MDR) provider. We are a force multiplier for defenders, helping organizations enhance their cybersecurity posture through advanced threat detection, rapid response, and continuous protection. Our team is composed of industry experts with deep experience in cybersecurity, automation, and AI-driven solutions. Backed by leading investors, we are rapidly growing and seeking top talent to join our mission of revolutionizing the AI-Native MDR landscape.

We’re a fast-growing startup backed by industry experts and top-tier investors led by Crosspoint Capital Partners and also backed by Shield Capital, DTCP (formerly Deutsche Telekom Capital Partners), Deepwork Capital, and the Florida Opportunity Fund. Seed round led by Andreessen Horowitz (a16z). As an early employee, you’ll play a meaningful role in defining and building our culture. Get in on the ground floor. We’re a small but well-funded team that just raised a substantial round – joining now comes with limited risk and unlimited upside.

As a Principal DevOps Engineer, you will be a key technical leader responsible for the architecture, evolution, and operation of our Azure infrastructure, CI/CD pipelines, and Site Reliability Engineering (SRE) practices. You will keep the platform highly available, secure, and performant as it scales to handle petabytes of security data and billions of daily events.

You'll work closely with Software Engineering, AI/ML, and Security Operations teams to define the technical vision and architecture for our production systems, driving automation and operational excellence to minimize toil and accelerate product delivery. This role requires deep, hands-on Azure platform expertise, a strong software engineering foundation, and fluency in DevSecOps principles.

Culture is one of the most important things at http://tenex.aiTENEX.AI http://TENEX.AI. Explore our culture deck at culture.tenex.ai http://culture.tenex.ai to witness how we embody it, prioritizing the irreplaceable collaboration and community of in-person work.

Location: This role will require Monday through Thursday onsite in our Kansas City office (preferred), with San Jose or Sarasota, FL also considered. WFH Friday. Candidates must live in or be willing to relocate to one of these three cities.

JOB RESPONSIBILITIES

  • Own the architecture of our Azure platform as it scales to petabytes of security data and billions of daily events.
  • Own our Azure governance and environment model, including subscription and management group structure, Azure Policy, network topology, and identity.
  • Lead Site Reliability Engineering (SRE) initiatives, defining and driving adherence to critical Service Level Objectives (SLOs) and Service Level Indicators (SLIs), and managing on-call rotations.
  • Drive operational excellence by implementing advanced monitoring, observability (logs, metrics, tracing), automated provisioning, and disaster recovery strategies.
  • Establish and enforce DevSecOps practices, standardizing CI/CD pipelines, infrastructure-as-code (IaC), security testing, and deployment mechanisms for rapid, secure, and reliable software delivery.
  • Automate deployment, scaling, and management of microservices and event-driven systems using containerization and orchestration technologies (Docker, Kubernetes, AKS).
  • Maintain workload portability across the platform so that infrastructure decisions remain reversible.
  • Partner with engineering teams to optimize application performance, resource utilization, and cloud cost efficiency.
  • Mentor and influence engineering teams on best practices in Azure architecture, reliability, and security-first development.
  • Collaborate with Product Management and Security Operations to translate new product requirements and operational needs into scalable and cost-effective platform solutions.
  • Evaluate and drive the adoption of new infrastructure technologies and engineering methodologies to maintain a competitive advantage.

REQUIRED SKILLS & QUALIFICATIONS

  • 8+ years of progressive experience in DevOps, Site Reliability Engineering (SRE), or Platform Engineering roles.
  • Deep, hands-on expertise building production Azure platforms, including AKS, Entra ID and workload identity federation, VNet design and Private Link, Key Vault, Azure Policy, and subscription or landing zone architecture. This is a platform engineering role rather than a Microsoft 365, Intune, or Windows administration role.
  • Experience standing up or moving production workloads across cloud environments, with the ability to describe the design, the data path, the cutover, and what broke.
  • Experience building secure and compliant (e.g., SOC 2, ISO 27001) environments.
  • Deep understanding of microservices architecture, containerization (Docker, Kubernetes), and event-driven systems.
  • Extensive experience with Infrastructure-as-Code tools (e.g., Terraform, Bicep) and CI/CD best practices.
  • Production experience writing and shipping software in Go or Python, beyond scripting and configuration.
  • Experience with monitoring and observability tools (Prometheus, Grafana, Azure Monitor, ELK stack, or similar).
  • Familiarity with real-time data pipelines and stream processing (e.g., Kafka, Event Hubs, Service Bus, Pub/Sub).
  • Proven track record of architecting, building, and operating highly scalable, distributed, and secure enterprise-grade SaaS platforms.

-

Nice-to-have

  • Working knowledge of more than one major cloud provider, deep enough to judge where the provider models differ rather than assume they match.
  • Prior experience in cybersecurity (SIEM, EDR, SOAR, or MDR) or an MSSP environment.
  • Experience with large-scale data warehousing/lakehouse technologies (e.g., Azure Data Explorer, Microsoft Fabric, Snowflake, BigQuery).
  • Background leading technical initiatives in high-growth startups or enterprise SaaS.
  • Familiarity with the underlying infrastructure to support AI/ML model deployment and monitoring (MLOps).

EDUCATION & CERTIFICATIONS

  • 10-12 years of experience, Bachelor's or Master's degree in Computer Science, Engineering, and or years of relative experience
  • Relevant certifications (Azure Solutions Architect Expert, Kubernetes, or security-related credentials) are a plus. Certifications complement production depth and do not substitute for it.

Source: Tenex careers (Ashby)

Similar roles