Site Reliability Engineer – SRE

Job not on LinkedIn

🕒 November 25, 2025

🌐 Russia, Serbia, +4 more countries – Remote

infoinfo

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 48%

infoinfo

🗣️🇷🇺 Russian Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of NOVACARD

NOVACARD

51 - 200 employees

Founded 2023

💳 Fintech

🏦 Banking

👥 B2C

Fintech • Banking • B2C

NOVACARD is a Mexican consumer credit-card provider and fintech offering a transparent, no-hidden-fee credit card with flexible credit limits (MXN 1,000–200,000), 0 annual fee, cashback rewards (5% at supermarkets, 0. 5% on other purchases), transfers, and a mobile app with a 28-day interest-free period. Operated by Unitron Technology S. A. P. I. de C. V. , NOVACARD targets consumers seeking simple, low-cost credit solutions and transparent terms.

📋 Description

• Ensure the stability, performance, and fault tolerance of production systems • Develop and maintain infrastructure automation and observability tools • Monitor system health, respond to incidents, and perform root cause analysis (RCA) • Collaborate with development teams to improve scalability and reliability of services • Define and manage SLIs, SLOs, and Error Budgets • Lead incident response by organizing recovery, documenting RCA, and running blameless post-mortems • Configure and administer Grafana and Zabbix, design dashboards, and fine-tune alerting • Integrate and monitor external vendor systems and collaborate with vendor technical support

🎯 Requirements

• Fluent Russian • English B1+ (comfortable with technical documentation) • 3+ years of experience as an SRE, DevOps, or Infrastructure Engineer • Strong understanding of observability principles (metrics, logs, traces) • Hands-on experience with Grafana and Zabbix (administration, configuration, alert optimization) • Experience working with AWS and CI/CD tools • Practical knowledge of SLI/SLO/Error Budget frameworks • Experience leading and documenting incidents and post-mortems • Scripting skills for automation (Python, Bash, or Go) • Solid understanding of distributed systems and networking fundamentals • Experience monitoring and supporting mobile applications • Familiarity with Terraform, Prometheus, Loki, ELK, or similar tools • Experience working with Kubernetes and containerized environments

🏖️ Benefits

• Fully remote work format • Official employment under the Russian Labor Code for residents of Russia • Contractor collaboration available for candidates from other countries • Opportunity to work in an international team on a new digital product for the Mexican market • Data-driven environment where contributions have a real impact

Apply Now