Site Reliability Engineer

🔥 14 hours ago

🇺🇸 United States – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 14%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of OXIO

OXIO

51 - 200 employees

📡 Telecommunications

☁️ SaaS

💳 Fintech

Telecommunications • SaaS • Fintech

OXIO is a telecom company offering a cloud-native, programmable Telecom-as-a-Service platform that empowers businesses to launch and manage their own mobile networks. The company's platform, BrandVNO, enables the creation of custom mobile connectivity services without telecom expertise, allowing organizations to integrate these services seamlessly into their existing operations. OXIO provides tools for business intelligence, subscriber management, and personalized customer experiences, targeting retail, fintech, and various enterprise sectors. The company's API-first approach facilitates rapid innovation, cost reduction, and enhanced customer engagement across multiple markets.

📋 Description

• Design and implement cloud platforms supporting OXIO backend services • Automate technical operations, including deployments, scaling, and recovery • Monitor and maintain mission-critical production infrastructure for maximum uptime • Participate in an on-call rotation and continuous improvement through blameless postmortems • Enable Engineering, Telecom, and Data Engineering teams with tools to operate their services • Contribute to OXIO’s Carrier-as-a-Service telecom platform and modern connectivity infrastructure

🎯 Requirements

• Understanding of Linux/Unix systems • Familiarity with Linux/Unix system internals, including process management, filesystems, memory management, and networking • Proficiency in at least one programming language: Python, Go, or Ruby • Strong scripting skills in Bash or Perl • Experience with infrastructure provisioning tools such as Terraform, CloudFormation, or Ansible • Familiarity with Docker and Kubernetes • Familiarity with Prometheus, Grafana, or Datadog • Knowledge of alerts, log analysis, dashboards, and observability • Familiarity with incident management practices, including runbooks and postmortems • Experience participating in an on-call rotation and handling incidents • Experience setting up and maintaining CI/CD pipelines such as Jenkins, GitLab CI, or CircleCI • Hands-on experience with AWS, Google Cloud, or Azure • Knowledge of VMware, KVM, and cloud-native architecture • Understanding of TCP/IP, DNS, HTTP/HTTPS, load balancing, and firewalls • Nice-to-have: deployment strategies, high availability and failover, IAM and zero trust, distributed systems, custom monitoring, SQL and NoSQL databases, distributed tracing, log aggregation, performance profiling, load testing, and SaltStack configuration management

Apply Now

Similar Jobs

🔥 18 hours ago

Sphera

1001 - 5000

💼 Consulting

🏥 Healthcare

📦 Logistics

Lead DevOps Engineer leading Azure infrastructure, CI/CD, and MLOps for Sphera’s environmental, health, safety, and sustainability software platform. Driving reliability, security, observability, and cloud cost optimization.

🔥 18 hours ago

CACI International Inc

10,000+ employees

🎖️ Defense

🏛️ Government

🔒 Cybersecurity

Cloud/DevOps specialist securing AWS assessment infrastructure for CACI’s federal cybersecurity program. Building IaC, CI/CD, and authorized cloud-exploitation capabilities for continuous federal cyber assessments.

🇺🇸 United States – Remote

💵 $90.3k - $189.6k / year

🔥 Funding within the last year

💰 $500M Post-IPO Debt on 2026-02

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🔥 20 hours ago

MaintainX

501 - 1000

☁️ SaaS

🏭 Manufacturing

🏢 Enterprise

Site Reliability Engineer improving reliability, observability, and developer autonomy for MaintainX’s industrial work execution platform. Building tooling and standards for resilient, self-service operations.

🇺🇸 United States – Remote

💵 $120k - $249.3k / year

💰 $150M Series D on 2025-08

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🔥 20 hours ago

Rimutee

11 - 50

🤝 B2B

👥 HR Tech

Senior DevOps Engineer designing AWS cloud infrastructure, IaC, CI/CD, and disaster recovery for ReKluti’s international software clients. Remote, full-time role focused on secure, highly available platforms.

🗣️🇪🇸 Spanish Required

🕒 Yesterday

IREN

201 - 500

🤖 Artificial Intelligence

🤝 B2B

⚡ Energy

Azure DevOps Lead Engineer owning Azure architecture, governance, and reliability for IREN’s renewable-powered AI cloud and data centers. Leading DevOps teams and compliance automation.