Site Reliability Engineering Manager
Onebrief · Colorado Springs, CO · 2026-09-18
About this role
CONSEQUENTIAL WORK. DEDICATED PEOPLE.
ABOUT ONEBRIEF
Onebrief builds collaboration and AI-powered workflow software for military planning and operational coordination.
Military planning is complex by nature, requiring teams to coordinate information, people, and decisions across systems and locations. Onebrief brings planning, collaboration, simulation, and AI into one connected environment, helping teams test strategies, adapt to changing conditions, and make decisions with greater clarity when the stakes are real.
We are a distributed team of builders from military, operational, and technology backgrounds who care deeply about improving how important work gets done. Some team members work remotely, while others work directly alongside customers in operational environments around the world.
Founded in 2019, Onebrief is backed by leading investors including General Catalyst, Battery Ventures, Insight Partners, Sapphire Ventures, and Human Capital. Valued at more than $2 billion, we continue to invest in product innovation, AI capabilities, and team growth.
SECURITY CLEARANCE, LOCATION, AND ONSITE NOTICE
This role requires regularly working on-site at customer locations.
If you are not currently within commuting distance, you must be willing to relocate. Onebrief provides relocation assistance.
Active Secret clearance required.
ABOUT THE ROLE
We're hiring a Site Reliability Engineering Manager to lead our SRE team within Infrastructure & Security. You'll work closely with platform engineering, application engineering, security, and customer success to ensure Onebrief's mission-critical deployments are reliable, secure, and well supported across on-prem DoD and AWS environments.
You'll lead a team whose work spans customer-facing operations, infrastructure, observability, automation, and application reliability. Most of the team focuses on deploying and operating Onebrief in demanding customer environments.
You'll own the team's priorities, planning, execution, and development. A significant part of this role is coordinating work: understanding demand, balancing capacity, sequencing tasks, managing dependencies, and keeping commitments realistic as customer needs change. You'll help the team deliver immediate operational support while making steady progress on improvements that reduce future support demands.
You'll bring the technical grounding to evaluate risks, ask useful questions, and guide decisions. Your engineers will own technical implementation and lead incident response. You'll provide direction, remove blockers, and create the conditions for them to succeed.
ABOUT YOU
You care deeply about reliability and understand the challenges of operating software in environments where connectivity, access, and deployment options can be constrained. You treat infrastructure and operability as products that deserve clear ownership, thoughtful design, and continuous improvement.
You're an effective people manager who sets clear expectations, gives useful feedback, and helps engineers grow. You build accountability through clear priorities and meaningful ownership, and you recognize when your team needs direction, support, or room to solve a problem.
You're comfortable managing a changing workload. You can turn competing requests into an actionable plan, account for operational interruptions, and explain what the team can commit to with its available capacity. You surface tradeoffs early and work with stakeholders to make deliberate decisions about scope and timing.
You bring calm and structure when priorities shift or incidents occur. You support engineers leading the response, help resolve escalations, and coordinate with customer-facing partners. You build a culture where engineers can surface risks early and examine failures honestly.
You have the technical judgment to help the team determine whether a recurring problem needs an infrastructure change, better automation, an application fix, or a clearer process. You bring the right people together to address it and ensure they have time to follow through.
WHAT YOU'LL DO
- Lead and develop the SRE team: Hire, coach, and support engineers across infrastructure, operations, and application reliability. Set expectations, manage performance, support career development, and build the skills and coverage the team needs.
- Own capacity and work planning: Maintain a clear view of incoming requests, ongoing support needs, and planned engineering work. Break initiatives into manageable tasks with the team, establish ownership, sequence work, and adjust commitments as priorities or capacity change.
- Coordinate delivery across teams: Manage dependencies with platform engineering, application engineering, security, and customer success. Identify blockers early, resolve competing priorities, and communicate progress, risks, and decisions to stakeholders.
- Set the reliability roadmap: Translate customer needs, production data, incident patterns, and operational risks into a prioritized improvement plan. Protect capacity for work that reduces recurring failures and makes deployments easier to operate.
- Establish operational ownership: Ensure production deployments have clear support responsibilities, escalation paths, and readiness criteria. Plan with partner teams for new deployments, releases, and ongoing customer support.
- Support team-led incident response: Establish sustainable on-call coverage and clear incident response expectations. Coach engineers who serve as incident commanders and lead blameless postmortems / After Action Reviews (AARs). Help the team assess corrective actions, assign ownership, and schedule follow-through.
- Guide technical priorities: Work with engineers and technical leads to evaluate approaches to infrastructure, automation, observability, and application reliability. Ensure plans account for operability, security requirements, and the constraints of on-prem and…
Skills asked for
- sre
- aws
- devops
- terraform
- ansible
- python
- go
- bash
Similar jobs
- Senior Site Reliability Engineer, Colorado Springs (Top Secret Clearance Required, Relocation Provided)Onebrief · Colorado Springs
Your next role is already in here.
Search live openings from thousands of employers, save the ones worth a second look, and let JobBob keep watch for the rest.