Own the architecture, governance, environment model, networking, identity, and operation of the Azure platform at large scale.
Lead SRE initiatives, including SLOs, SLIs, on-call rotations, monitoring, observability, automated provisioning, and disaster recovery.
Establish DevSecOps practices covering infrastructure as code, CI/CD pipelines, security testing, and deployment mechanisms.
Automate deployment, scaling, and management of containerized microservices and event-driven systems using Docker, Kubernetes, and AKS.
Optimize application performance, resource utilization, reliability, and cloud cost efficiency with engineering teams.
Mentor and influence engineers on Azure architecture, reliability, security-first development, and operational best practices.
Collaborate with Product Management and Security Operations to design scalable, cost-effective platform solutions.
Evaluate and adopt infrastructure technologies and engineering methodologies.
Requirements
8+ years of progressive experience in DevOps, Site Reliability Engineering, or Platform Engineering roles.
Deep hands-on production Azure platform expertise, including AKS, Entra ID, workload identity federation, VNet design, Private Link, Key Vault, Azure Policy, and subscription or landing-zone architecture.
Experience standing up or migrating production workloads across cloud environments, including explaining architecture, data paths, cutovers, and failure handling.
Experience building secure and compliant environments, such as SOC 2 or ISO 27001 environments.
Deep understanding of microservices, containerization, Docker, Kubernetes, and event-driven systems.
Extensive experience with infrastructure-as-code tools such as Terraform or Bicep and CI/CD practices.
Production experience writing and shipping software in Go or Python beyond scripting and configuration.
Experience with monitoring and observability tools such as Prometheus, Grafana, Azure Monitor, or ELK Stack.
Familiarity with real-time data pipelines and stream-processing technologies such as Kafka, Event Hubs, Service Bus, or Pub/Sub.
Bachelor’s or Master’s degree in Computer Science or Engineering, or relevant equivalent experience.
Relevant Azure, Kubernetes, or security certifications are a plus.
Experience with multiple major cloud providers, cybersecurity or MSSP environments, large-scale data warehousing or lakehouse technologies, high-growth startups or enterprise SaaS, or MLOps is preferred.
Benefits
Monday through Thursday onsite with Friday work-from-home; Kansas City is preferred, with San Jose or Sarasota, Florida also considered.
Candidates must live in or be willing to relocate to Kansas City, San Jose, or Sarasota.
Early-stage startup opportunity with meaningful influence on culture, architecture, and company growth.
TENEX.AI
TENEX is the first AI-native, human-led MDR powered by AI SOC. Backed by 24/7 U.S.based expert analysts with 8+ years avg. experience. Our human-led, AI-driven platform delivers 10x faster detection,