
1001 - 5000 employees
Founded 2014
đź Consulting
đŁ Marketing
đ¤ Artificial Intelligence
đ° Secondary Market on 2020-11
Consulting ⢠Marketing ⢠Artificial Intelligence
GitLab is the most comprehensive AI-powered DevSecOps platform, offering tools for automated software delivery, security, and compliance throughout the software development lifecycle. It provides solutions across areas such as AI-assisted development, continuous integration/continuous deployment (CI/CD), source code management, and vulnerability management. GitLab aims to simplify and accelerate software delivery by uniting development, security, and operations on a unified platform. It is particularly recognized for its AI code assistants and has been named a leader in the Gartner Magic Quadrant⢠for DevOps Platforms, making it a preferred choice for many enterprises.
đ September 25
đ United States, Canada, +1 more countries â Remote
đľ $250k - $349k / year
â° Full Time
đ Senior
đ´ Lead
â DevOps & Site Reliability Engineer (SRE)
đť Ghost score 0%
Improve your chances of getting an interview by checking your resume score before you apply.

1001 - 5000 employees
Founded 2014
đź Consulting
đŁ Marketing
đ¤ Artificial Intelligence
đ° Secondary Market on 2020-11
Consulting ⢠Marketing ⢠Artificial Intelligence
GitLab is the most comprehensive AI-powered DevSecOps platform, offering tools for automated software delivery, security, and compliance throughout the software development lifecycle. It provides solutions across areas such as AI-assisted development, continuous integration/continuous deployment (CI/CD), source code management, and vulnerability management. GitLab aims to simplify and accelerate software delivery by uniting development, security, and operations on a unified platform. It is particularly recognized for its AI code assistants and has been named a leader in the Gartner Magic Quadrant⢠for DevOps Platforms, making it a preferred choice for many enterprises.
⢠Set and continuously refine the go-forward technical direction for CI, CD, Plan, and source code experiences. ⢠Partner directly with the VP of Engineering on planning, technical investment, and engineering priorities. ⢠Translate company-wide productivity, quality, and AI-native engineering goals into technical direction and iterative roadmaps for Core DevOps. ⢠Define measurable quality standards for availability, p95 and p99 latency, pipeline success rate, test reliability, data correctness, and defect escape rate. ⢠Identify systemic technical risks and drive prioritized plans to retire them. ⢠Lead complex design decisions through code, prototypes, and proofs of concept. ⢠Drive convergence on shared patterns, libraries, and paved paths. ⢠Design incremental migration, dual-run, rollback, and deprecation strategies for long-lived systems. ⢠Adopt AI platform capabilities into Core DevOps and feed requirements back to the AI organization. ⢠Serve as escalation point for complex or contested technical decisions. ⢠Review critical-path designs and merge requests, using reviews to teach. ⢠Ensure designs work across GitLab.com, GitLab Dedicated, and Self-Managed deployments while respecting multi-tenant, compliance, and data-governance boundaries. ⢠Partner with Infrastructure, Security, and SRE on observability, debuggability, and graceful failure modes. ⢠Communicate architectural constraints and opportunities to Product through roadmap terms. ⢠Write design documents, architecture narratives, and decision records. ⢠Grow Principal and Staff Engineers and contribute to the senior technical hiring bar. ⢠Participate in the Incident Management on-call rotation and complete Interview Training for technical interviewing.
⢠10+ years of software engineering experience, including 4+ years in a Staff, Principal, or equivalent senior technical leadership role. ⢠Deep expertise in AI and ML systems, including large language models, agentic frameworks, and autonomous workflow design at production scale. ⢠Proven track record of leading hands-on technical experimentation, including defining evaluation frameworks, running benchmarks, and translating findings into scalable architecture decisions. ⢠Strong background in scalable, multi-tenant distributed systems, including service decomposition, fault tolerance, observability, and operational resilience. ⢠Experience designing and implementing human-in-the-loop controls, safety guardrails, and responsible AI practices for production systems. ⢠Experience mentoring senior engineers and influencing technical direction across multiple teams or divisions without direct authority. ⢠Ability to work effectively in a fully remote, globally distributed organization with excellent written and asynchronous communication skills.
⢠Benefits to support your health, finances, and well-being ⢠Flexible Paid Time Off ⢠Team Member Resource Groups ⢠Equity Compensation & Employee Stock Purchase Plan ⢠Growth and Development Fund ⢠Parental Leave
Apply Nowđ September 25
Senior SRE maintaining MySQL and PostgreSQL reliability for Vultrâs global cloud infrastructure. Owning monitoring, disaster recovery, incident response, security compliance, and automation.
đşđ¸ United States â Remote
đľ $125k - $135k / year
đ° $329M Debt Financing - Vultr on 2025-06
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đ September 25
Staff SRE owning reliability strategy, observability, and incident readiness for Horizon3âs autonomous cybersecurity platform. Leading cross-functional initiatives across production infrastructure and services.
đşđ¸ United States â Remote
đľ $199.8k - $270k / year
â° Full Time
đ´ Lead
â DevOps & Site Reliability Engineer (SRE)
đ September 25
DevSecOps Engineer securing cloud-native software delivery for PingWind, a federal government services provider. Building CI/CD security, automating controls, and supporting vulnerability remediation.
đşđ¸ United States â Remote
â° Full Time
đ Senior
đ´ Lead
â DevOps & Site Reliability Engineer (SRE)
đ September 25
DevOps/SRE improving AWS infrastructure, CI/CD, and application reliability. Supporting technology tools for student athletes, coaches, and event operators.
đşđ¸ United States â Remote
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đ September 25
DevOps/SRE engineer optimizing CI/CD, AWS infrastructure, and observability. Supporting technology tools for student athletes, coaches, and event operators.
đşđ¸ United States â Remote
đ° Private Equity Round - IMG Academy on 2023-11
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)