Site Reliability Engineer

Job not on LinkedIn

🕒 July 31

🌐 United States, Argentina, +6 more countries – Remote

infoinfo

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 16%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Outpost

Outpost

51 - 200 employees

Founded 2021

🚗 Transport

📦 Logistics

🏠 Real Estate

Transport • Logistics • Real Estate

Outpost is a technology company that provides strategic locations for truck parking, fleet positioning, and drop trailer relays, facilitating seamless freight movement across the United States. Their rapidly growing nationwide network of secure yards supports trucking companies, from enterprise fleets to owner-operators, by maximizing property utilization and improving operational efficiency. By partnering with yard owners, Outpost enables better returns on investment while enhancing the logistics infrastructure essential for the supply chain.

📋 Description

• Own reliability targets across our backend/API, worker services, applications and CV pipeline; MTD, MTM, MTR, and follow-through on root causes. • Level up our monitoring and alerting, and build out auto-remediation, so on-call load scales with automation, not headcount. • Partner with our agentic engineering work to build agents that triage alerts and handle routine remediation. • Harden and optimize our GCP infrastructure (Cloud Run, Cloud SQL, GCS) for cost and performance as load scales. • Own database scale and performance; connection pooling, query optimization and indexing, read replicas, and capacity planning, so Postgres doesn't become the bottleneck as data volume grows. • Improve the reliability of our ML training and monitoring infrastructure, in partnership with the CV/ML team. • Run blameless postmortems and drive fixes for root causes, not just symptoms. • Participate in on-call rotation.

🎯 Requirements

• 4+ years in an SRE, infrastructure, or backend engineering role with production on-call ownership. • Deep experience with a major cloud provider (GCP preferred); compute, managed databases, object storage, networking. • Experience building monitoring/alerting/observability stacks (Grafana, Prometheus, Zabbix, Datadog, or similar). • Strong scripting/automation skills (Python, Bash, or similar). • Comfortable with containerized workloads (Docker) and CI/CD pipelines. • Track record of reducing incident volume or improving reliability metrics — not just responding to incidents. • Strong communication skills, comfortable working with both technical and non-technical stakeholders, know when and how to escalate urgency, and build strong working relationships across teams. • Strong communication skills in English — you write clearly and engage well async.

🏖️ Benefits

• Health insurance • 401(k) matching • Flexible work hours • Paid time off

Apply Now

Similar Jobs

🕒 July 29

Staq.io

51 - 200

💳 Fintech

🏦 Banking

☁️ SaaS

Senior DevOps Engineer responsible for managing cloud infrastructure and optimizing deployment processes for a fintech company. Leading CI/CD improvements, logging, monitoring, and security practices.

🇺🇸 United States – Remote

💰 Seed on 2023-08

⏳ Contract/Temporary

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 28

Mento

51 - 200

☁️ SaaS

👥 HR Tech

🤝 B2B

Mento coach providing one-on-one support and guidance to members through coaching sessions. Leading ongoing training to enhance skills and improve member engagement in a remote setup.

🇺🇸 United States – Remote

💰 $6M Seed Round - Mento on 2023-02

⏳ Contract/Temporary

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 22

Arctiq

201 - 500

💼 Consulting

🏥 Healthcare

📦 Logistics

Senior Site Reliability Engineer architecting reliability strategies for large-scale government systems. Leading implementation of SRE framework and mentoring mid-level engineers toward system resilience.

🇺🇸 United States – Remote

⏳ Contract/Temporary

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 17

Leland

11 - 50

💼 Consulting

📣 Marketing

📚 Education

Site Reliability & Engineering Coach for a remote platform connecting people to career experts. Focusing on practical skill development in site reliability engineering and career support.

🕒 June 24

Black Pearl Consult

11 - 50

💼 Consulting

🎯 Recruiter

DevOps Engineer focusing on automating deployment pipelines and managing cloud infrastructure for technology company. Collaborating with engineering and cloud teams to improve deployment speed and reliability.

🇺🇸 United States – Remote

💵 $750 - $1.5k / month

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)