Site Reliability Engineer

Job not on LinkedIn

🔥 3 minutes ago

🌏 Anywhere in the World

⏰ Full Time

🟠 Senior

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Yuno

Yuno

11 - 50 employees

💳 Fintech

🏢 Enterprise

☁️ SaaS

Fintech • Enterprise • SaaS

Yuno is a company that provides payment orchestration and infrastructure solutions on a global scale. Their technology empowers businesses to integrate over 300 payment methods, boosting acceptance rates and enabling seamless scaling of payment operations across multiple regions. Yuno focuses on providing a simplified payment process with features like smart routing, unified payment insights, auto reconciliation, and custom checkout options. They emphasize security and fraud management, ensuring safe transactions. Yuno facilitates global payouts and subscription management, making them an ideal partner for businesses looking to optimize their payment systems and increase revenue.

📋 Description

• Set the technical direction for reliability across Yuno’s infrastructure, beginning with the AWS platform that provisions, deploys, and manages AI agents at scale • Own the platform reliability strategy, including architectural decisions, reliability measurement, and engineering standards • Define SLO culture, error-budget policy, and incident practices across engineering teams • Design and own durable, reliable asynchronous messaging for inter-service communication • Own cloud infrastructure and automate provisioning with Infrastructure as Code • Ensure the platform scales reliably as transaction volume grows • Build monitoring, tracing, and alerting systems for platform health • Serve as senior escalation point for difficult production incidents • Run blameless postmortems and root-cause analyses that produce permanent fixes • Conduct continuous fault injection and resilience experiments • Mentor senior and mid-level engineers and raise organization-wide reliability standards

🎯 Requirements

• 7+ years of experience • Designed and owned event-driven systems using message queues such as Kafka, NATS, or RabbitMQ • Understanding of at-least-once delivery, consumer groups, dead letters, and backpressure • Experience migrating systems from synchronous to asynchronous communication • Deep AWS experience with EC2, VPC, IAM, S3, and RDS • Strong networking fundamentals • Infrastructure as Code experience with Terraform or Pulumi • Kubernetes and Docker production experience, including container lifecycle, resource limits, health checks, and orchestration at scale • Datadog fluency or equivalent experience with dashboards, monitors, APM, and distributed tracing • Track record defining and operating SLOs, SLIs, and error budgets across services • Hands-on fault injection, game day, or chaos experiment experience using Gremlin, Chaos Mesh, AWS FIS, or similar • Distributed systems debugging experience • Comfortable coding automation and tooling in Go, Python, or similar • Solid SQL and PostgreSQL knowledge • NoSQL experience with MongoDB and Redis, including indexing, replication, and performance tuning • Proven technical leadership, architecture influence across teams, and engineering mentorship • Advanced written and spoken English proficiency

🏖️ Benefits

• Competitive Compensation • Remote Work — you can work from everywhere • Home Office Bonus — a one-time allowance to set up your ideal home office • Work Equipment • Stock Options • Health Plan wherever you are • Flexible Days Off • Language, Professional, and Personal Growth courses

Apply Now

Similar Jobs

🕒 July 27

Empowers Staffing Inc

11 - 50

💼 Consulting

🎯 Recruiter

🤖 Artificial Intelligence

Infrastructure Automation Engineer at LAK Technology Inc managing cloud infrastructure with Terraform and Ansible. Focused on automation and CI/CD processes across AWS, Azure, or GCP environments.

🌏 Anywhere in the World

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 27

Supabase

51 - 200

☁️ SaaS

🔌 API

🤖 Artificial Intelligence

Release Engineer at Supabase, ensuring safe and observable deployments and operational reliability across systems. Engage in incident management, monitoring, and process documentation for improved deployment efficiency.

🌏 Anywhere in the World

💰 $80M Series B on 2022-05

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 20

Social Discovery Group

1001 - 5000

🌍 Social Impact

📱 Media

DevOps Engineer developing internal services and tools in Go for social discovery products. Collaborating on deployment automation in a remote working environment with a global team.

🌏 Anywhere in the World

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🗣️🇷🇺 Russian Required

🕒 July 11

GitLab

1001 - 5000

💼 Consulting

📣 Marketing

🤖 Artificial Intelligence

Site Reliability Engineer ensuring reliability of GitLab's user-facing services. Supporting operational excellence through engineering principles and automation.

🌏 Anywhere in the World

💵 $126.4k - $314.4k / year

💰 Secondary Market on 2020-11

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 8

Onchain

51 - 200

💸 Finance

💳 Fintech

₿ Crypto

DevOps Engineer evolving cloud infrastructure behind Lisk's financial platform as we scale. Join a remote-first team to improve security and reliability of money-handling systems.

🌏 Anywhere in the World

💰 Series B on 2018-07

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)