Senior/Staff DevOps Engineer

🔥 4 minutes ago

🇷🇸 Serbia – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 21%

infoinfo

🗣️🇷🇺 Russian Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of MEDvidi

MEDvidi

201 - 500 employees

Founded 2019

🏥 Healthcare

⚕️ Healthcare Insurance

💰 $2.8M Seed Round on 2022-09

Healthcare • Healthcare Insurance • Mental Health

MEDvidi is an online mental health treatment center aiming to make professional care accessible and affordable for everyone. The healthcare experts at MEDvidi provide personalized treatment plans for a variety of mental health conditions, including ADHD, anxiety, depression, insomnia, and OCD. With services that include initial assessments, ongoing support, and medication management through virtual consultations, MEDvidi seeks to enhance mental wellness with compassionate care tailored to individual needs.

📋 Description

• Set and drive the technical vision and quarterly roadmap for infrastructure with clear trade-offs and measurable goals • Run and evolve AWS and Kubernetes (EKS) infrastructure, including cluster management, autoscaling with Karpenter, policy enforcement with Kyverno, and zero-downtime operations • Own Infrastructure as Code end to end using Terraform and AWS CDK in TypeScript • Own GitLab CI/CD, including reusable/shared templates, OIDC, and self-managed GitLab • Build and own practical observability using Prometheus, Grafana, OpenTelemetry, OpenSearch, CloudWatch, log pipelines, and APM • Maintain fast blue-green deployments and health-gated automated rollback • Own zero-downtime PostgreSQL schema migrations using expand/contract and CI migration gating • Own security engineering in a HIPAA environment, including secrets hygiene, credential rotation, leak scanning, PHI-aware log and data handling, and Vault managed as code • Partner directly with product engineering teams to remove infrastructure friction and improve developer experience • Use agentic AI as a core workflow by integrating autonomous-agent output into production • Set technical direction, ship infrastructure changes hands-on, and own reliability, cost, security, performance, deployment health, and developer velocity outcomes

🎯 Requirements

• 6+ years in DevOps/infrastructure engineering • Strong systems fundamentals • Solid Linux administration and troubleshooting, including performance analysis, resource management, and process debugging • Hands-on AWS experience with EC2, EKS, RDS, ElastiCache, Lambda, SQS, EventBridge, API Gateway, ALB, and S3 • Production Kubernetes/EKS experience, including cluster management, node scaling, and policy enforcement; Karpenter, Kyverno, or similar • Strong Infrastructure as Code experience with Terraform and AWS CDK in TypeScript • CI/CD ownership with GitLab CI/CD, reusable/shared templates, OIDC id_tokens, and self-managed GitLab • Practical monitoring and observability experience with Prometheus, Grafana, OpenTelemetry, OpenSearch, CloudWatch, log-shipping, and error tracking/APM • Practical security engineering experience, including secrets rotation, short-lived credentials, leak scanning, and PHI-aware logging • HashiCorp Vault as code experience, including KV, JWT/OIDC authentication for CI, and policy design • Blue-green deployments with automated, health-gated rollback • PostgreSQL zero-downtime schema migrations using expand/contract and migration gating in CI • Containers experience with Docker, ECR, immutable tags, and image lifecycle management • Network and protocol fundamentals, including load balancing, TLS, and DNS • Hands-on agentic AI workflows, such as Claude Code or similar • Developer-focused mindset and strong problem-solving for complex system issues • Strong technical writing, including docs-as-code, ADRs, and design documents via MRs • Fluent Russian and English (B1) • Experience working effectively in remote, distributed teams • Experience in a regulated/compliance-heavy environment such as HIPAA or SOC 2 would be a plus • Ansible configuration management would be a plus • Node.js application operations with pm2 and npm would be a plus • GitOps tooling such as ArgoCD or Flux and deeper PostgreSQL administration would be a plus • AWS certifications would be a plus

🏖️ Benefits

• Competitive compensation package • Fully remote long-term collaboration under a B2B model • Health insurance after the probation period • Sports & wellness compensation • Personalized English lessons via Preply • 19 paid vacation days annually • 4 additional wellness days each year • Paid sick leave for the first 5 working days • Thoughtful gifts for key life events • Offline corporate events

Apply Now

Similar Jobs

🕒 August 31

Smartcat

51 - 200

🤖 Artificial Intelligence

☁️ SaaS

Senior DevOps engineer leading infrastructure, reliability, and CI/CD for Smartcat’s AI-powered localization platform. Building distributed engineering teams and scalable cloud systems.

AWS

Azure

Cloud

ElasticSearch

Google Cloud Platform

Kafka

Kubernetes

MongoDB

Prometheus

Vue.js

.NET

🕒 July 27

Cognativ Inc

51 - 200

🤝 B2B

💼 Consulting

🤖 Artificial Intelligence

Senior Site Reliability Engineer overseeing operational health of AI alerting platform and infrastructure. Responsible for reliability metrics, incident response, and disaster recovery planning.

Apache

AWS

Cloud

Grafana

IoT

Java

Kafka

Linux

Postgres

Prometheus

Python

Redis

Terraform

Go

🕒 July 21

Fundraise Up

51 - 200

🤲 Charity

💳 Fintech

☁️ SaaS

Join Fundraise Up as a Senior DevOps Engineer managing CI/CD and observability platforms. Support a global fundraising platform for nonprofits, focusing on reliability and mentorship.

🗣️🇷🇺 Russian Required

Ansible

Docker

Grafana

Jenkins

Kubernetes

Linux

Prometheus

Python

🕒 May 20

Collectly

51 - 200

🏥 Healthcare

⚕️ Healthcare Insurance

💳 Fintech

DevOps Engineer at Collectly managing infrastructure automation and CI/CD processes for healthcare tech. Join a dynamic team transforming revenue cycle management with AI-driven solutions.

Ansible

AWS

Azure

Cloud

Google Cloud Platform

Grafana

Jenkins

Kubernetes

Prometheus

Terraform