Senior/Staff DevOps Engineer, Platform Infrastructure
Redcellpartners · Seattle, WA · 2026-09-25
About this role
About Us
Red Cell Partners is an incubation firm building and investing in rapidly scalable technology-led companies that are bringing revolutionary advancements to market in three distinct practice areas: healthcare, cyber, and national security. United by a shared sense of duty and deep belief in the power of innovation, Red Cell is developing powerful tools and solutions to address our Nation’s most pressing problems.
About Trase
Co-founded in 2023 by Joe Laws and Grant Verstandig, Trase Systems is AI, Uncomplicated. Trase empowers enterprise leaders to harness the full potential of AI without the associated complexity and risks. We are an end-to-end solution for deploying, managing, and optimizing AI in the enterprise. Our platform specializes in bridging the “last mile” of AI adoption, unlocking AI's full potential while driving efficiency and significant cost savings. Trase is at the forefront of AI Agent innovation, topping the Hugging Face GAIA Leaderboard for Generalized AI Assistants, ahead of industry giants such as Google, Meta, Microsoft, and OpenAI. We are leveraging our cutting-edge technologies to develop mission-critical agentic applications in complex industries such as Healthcare, Oil & Gas, and National Security.
About the Role
As a Senior or Staff DevOps Engineer on Platform Infrastructure, you will design and operate the infrastructure that supports Trase OS across Trase-hosted and customer-controlled environments. You will build secure, repeatable deployment patterns spanning AWS, Microsoft Azure, Google Cloud Platform, private cloud, hybrid cloud, and on-premises infrastructure.
This is a hands-on engineering role with broad ownership. You will work across application packaging, Kubernetes, infrastructure as code (IaC), networking, security, observability, release engineering, and production reliability. You will also partner directly with engineering and customer-facing teams to turn deployment requirements into systems that can be installed, upgraded, operated, and supported consistently.
The level will reflect your experience and demonstrated scope. Staff-level candidates will be expected to lead architecture across teams, establish engineering standards, and mentor other engineers.
Why this Role is Needed
Trase OS supports mission-critical, long-running workflows in environments with different cloud services, network controls, security requirements, and operating models. Our deployment architecture must remain portable without sacrificing reliability, security, or operational clarity.
This role will reduce one-off deployment work, remove avoidable dependencies on a single cloud provider, and establish reusable infrastructure that internal teams and customers can operate with confidence.
What You'll Do
• Architect, build, and operate secure infrastructure across AWS, Microsoft Azure, and Google Cloud Platform, as well as private-cloud, hybrid-cloud, on-premises, and customer-controlled environments.
• Containerize and package Trase OS application services using Docker, Kubernetes, Helm, Kustomize, or equivalent tools.
• Create reusable infrastructure-as-code (IaC) modules and deployment workflows using Terraform, Pulumi, or comparable tooling.
• Design deployment patterns that account for customer-specific requirements such as restricted networks, limited or no egress, approved registries, data residency, and cloud account ownership.
• Identify, replace, or abstract hard dependencies on managed cloud services when they prevent portability across deployment environments.
• Troubleshoot complex issues across applications, Kubernetes clusters, cloud services, networks, and infrastructure rather than treating platform work as CI/CD scripting alone.
• Build and maintain CI/CD and GitOps workflows, release orchestration, environment promotion, upgrade paths, rollback procedures, and version compatibility controls.
• Build and operate observability for Trase OS across metrics, logs, traces, dashboards, alerting, and SLOs so teams can diagnose failures and support customer deployments.
Technical Leadership
• Lead infrastructure and reliability decisions that affect multiple product and customer teams.
• Set practical standards for production readiness, cloud portability, security, observability, and operational support.
• Make clear tradeoffs between speed, maintainability, cost, security, and customer requirements, and drive decisions through implementation.
• Build reusable platform capabilities that replace one-off deployment solutions.
• Mentor engineers and raise the team's ability to design, ship, and operate distributed systems.
Qualifications
• 10+ years of software, platform, infrastructure, SRE, or DevOps engineering experience, including ownership of production systems.
• Deep hands-on experience with Linux, Docker, Kubernetes, and production cluster operations.
• Strong experience with Helm and infrastructure as code, such as Terraform or Pulumi, including reusable modules, state management, testing, and change review.
• Hands-on infrastructure experience with at least two of AWS, Azure, and GCP, with working knowledge of core compute, networking, storage, identity, and managed-service patterns across all three.
• Experience deploying and operating software across customer-controlled, private-cloud, hybrid-cloud, or on-premises environments, including adapting cloud-native SaaS products for these deployment models.
• Strong understanding of networking, DNS, ingress, load balancing, certificates, IAM, secrets management, persistent storage, and service-to-service security.
• Experience building and operating CI/CD or GitOps systems, including production releases, upgrades, and rollbacks, with metrics, logs, traces, SLOs, and alerts for incident response.
• Strong software engineering and automation skills in Python, Go, TypeScript, or a similar language, with the ability to work across application and infrastructure…
Skills asked for
- devops
- aws
- azure
- google cloud
- kubernetes
- docker
- helm
- terraform
Similar jobs
- Senior Staff Cloud Backend EngineerCoupanginternal · Seattle
- Senior Staff Applied ScientistCoupanginternal · Seattle
- Senior Staff Applied ScientistCoupang · Seattle
- Senior Staff Software Engineer, Stripe DashboardStripe · Seattle
- Senior Staff Cloud Backend EngineerCoupang · Seattle
- Senior/ Staff AI Research EngineerXairatherapeutics · Seattle
- Senior Staff Software EngineerPingidentity · Seattle
- Senior Staff Software Engineer, Recognition PlatformMetropolis · Seattle
Your next role is already in here.
Search live openings from thousands of employers, save the ones worth a second look, and let JobBob keep watch for the rest.