Senior/Staff DevOps Engineer

🔥 15 hours ago

🇵🇹 Portugal – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 20%

infoinfo

🗣️🇷🇺 Russian Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of MEDvidi

MEDvidi

201 - 500 employees

Founded 2019

🏥 Healthcare

⚕️ Healthcare Insurance

💰 $2.8M Seed Round on 2022-09

Healthcare • Healthcare Insurance • Mental Health

MEDvidi is an online mental health treatment center aiming to make professional care accessible and affordable for everyone. The healthcare experts at MEDvidi provide personalized treatment plans for a variety of mental health conditions, including ADHD, anxiety, depression, insomnia, and OCD. With services that include initial assessments, ongoing support, and medication management through virtual consultations, MEDvidi seeks to enhance mental wellness with compassionate care tailored to individual needs.

📋 Description

• Set and drive the technical vision and quarterly roadmap for infrastructure with clear trade-offs and measurable goals • Run and evolve AWS and Kubernetes (EKS) infrastructure, including cluster management, autoscaling with Karpenter, policy enforcement with Kyverno, and zero-downtime operations • Own Infrastructure as Code end to end using Terraform and AWS CDK in TypeScript • Own GitLab CI/CD, including reusable/shared templates, OIDC, and self-managed GitLab • Build and own practical observability using Prometheus, Grafana, OpenTelemetry, OpenSearch, CloudWatch, log pipelines, and APM • Maintain fast blue-green deployments and health-gated automated rollback • Own zero-downtime PostgreSQL schema migrations using expand/contract and CI migration gating • Own security engineering in a HIPAA environment, including secrets hygiene, credential rotation, short-lived credentials, leak scanning, PHI-aware log and data handling, and Vault managed as code • Partner directly with product teams to remove infrastructure friction and improve developer experience • Use agentic AI as a core part of the workflow and integrate autonomous-agent output into production • Set technical direction, ship infrastructure hands-on, and own outcomes for reliability, cost, security, performance, deployment health, and developer experience

🎯 Requirements

• 6+ years in DevOps/infrastructure engineering • Strong systems fundamentals • Solid Linux administration and troubleshooting, including performance analysis, resource management, and process debugging • Hands-on AWS experience with EC2, EKS, RDS, ElastiCache, Lambda, SQS, EventBridge, API Gateway, ALB, and S3 • Production Kubernetes/EKS experience, including cluster management, node scaling, and policy enforcement; Karpenter, Kyverno, or similar • Strong Infrastructure as Code experience with Terraform and AWS CDK in TypeScript • CI/CD ownership with GitLab CI/CD, including reusable/shared templates, OIDC id_tokens, and self-managed GitLab • Practical monitoring and observability experience with Prometheus, Grafana, OpenTelemetry, OpenSearch, CloudWatch, log-shipping, and error tracking/APM • Practical security engineering experience with secrets rotation, short-lived credentials, leak scanning, and PHI-aware logging • HashiCorp Vault as code experience, including KV, JWT/OIDC authentication for CI, and policy design • Blue-green deployments with automated, health-gated rollback • PostgreSQL zero-downtime schema migrations using expand/contract and migration gating in CI • Containers experience with Docker, ECR, immutable tags, and image lifecycle • Network/protocol fundamentals including load balancing, TLS, and DNS • Hands-on agentic AI workflows, such as Claude Code or similar • Developer-focused mindset and strong problem-solving for complex system issues • Strong technical writing, including docs-as-code, ADRs, and design docs via MRs • Fluent Russian and English (B1) • Experience working effectively in remote, distributed teams • Preferred: experience in regulated/compliance-heavy environments such as HIPAA or SOC 2 • Preferred: Ansible for VM fleet management • Preferred: Node.js application operations with pm2 and npm • Preferred: GitOps tooling such as ArgoCD or Flux and deeper PostgreSQL database administration • Preferred: AWS certifications

🏖️ Benefits

• Competitive compensation package • Health insurance after the probation period • Sports & wellness compensation • Personalized English lessons via Preply • 19 paid vacation days annually • 4 additional wellness days each year • Paid sick leave for the first 5 working days • Thoughtful gifts for key life events • Offline corporate events • Fully remote long-term collaboration under a B2B model

Apply Now

Similar Jobs

🕒 4 days ago

Intermedia Cloud Communications

1001 - 5000

💼 Consulting

🏥 Healthcare

⚖️ Legal

Site Reliability Engineer improving reliability, observability, and resilience for Intermedia’s cloud communications and AI-powered services. Automating production infrastructure and Voice/UC integrations.

Cloud

Distributed Systems

HDFS

Kubernetes

Linux

NFS

🕒 4 days ago

Keyrus

1001 - 5000

Cloud & DevOps Engineer building secure AWS platforms for Keyrus’s industrialized AI consulting. Automating infrastructure, serverless services, CI/CD, and enterprise cloud governance.

AWS

Cloud

Docker

Kubernetes

Terraform

🕒 4 days ago

Expleo Group

10,000+ employees

💼 Consulting

🎖️ Defense

📦 Logistics

Mid DevOps Engineer managing cloud infrastructure, CI/CD, and reliability for Expleo, a global engineering and technology services provider. Automating operations with Terraform, Kubernetes, and scripting.

AWS

Azure

Cloud

Docker

Google Cloud Platform

Grafana

Kubernetes

Prometheus

Python

Terraform

🕒 September 3

Tether.to

11 - 50

₿ Crypto

💳 Fintech

💸 Finance

DevOps Engineer building CI/CD, IaC, and release pipelines for Tether’s global digital finance platform. Automating secure web, desktop, and mobile deployments.

Android

Ansible

AWS

Distributed Systems

Docker

Firewalls

Grafana

iOS

JavaScript

Linux

Microservices

Prometheus

PyTorch

Shell Scripting

Tensorflow

Terraform

TypeScript

C++

🕒 August 25

Inetum

10,000+ employees

💼 Consulting

🏥 Healthcare

🛡️ Insurance

Arquitecto de microservicios gobernando entornos multi-cliente OpenShift y DevOps. DiseĂąando seguridad, escalabilidad, CI/CD y recuperaciĂłn ante desastres para Inetum.

🗣️🇪🇸 Spanish Required

Angular

Azure

Hibernate

Java

JUnit

Kafka

Kubernetes

OpenShift

Oracle

RabbitMQ

Spring

Spring Boot

SpringBoot

SQL