Search Remote Jobs

Senior Site Reliability Engineer

đŸ”„ 9 minutes ago

đŸ‡ȘđŸ‡ș Europe – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

đŸ‘» Ghost score 12%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Parasail

Parasail

11 - 50 employees

Founded 2023

☁ SaaS

💰 $10M Seed on 2025-05

SaaS

Parasail is a cloud platform that provides inference infrastructure for AI-native startups, focusing on running open-source and specialized models in production. It offers a global fleet of data centers and current-generation hardware, optimized endpoints, and a single API for deploying any model (including fine-tunes and Hugging Face models). Parasail emphasizes reliability, performance, flexibility to scale, cost efficiency (up to 30× cheaper than legacy clouds), day-0 access to frontier models, lossless defaults, customizable speed/quality tradeoffs, flexible drawdown billing, and dedicated engineering support with fast response. Customers use Parasail for high-throughput inference, batch processing, and production deployment of LLMs and multimodal models (vision, voice, OCR, reranking, retrieval).

📋 Description

‱ Build and improve Kubernetes infrastructure for provisioning, networking, storage, and service deployment across providers and regions ‱ Design isolation, failover, and recovery mechanisms to reduce the impact of hardware and infrastructure failures ‱ Automate capacity expansion, deployments, and maintenance ‱ Develop observability and diagnostics to reveal bottlenecks and surface failures ‱ Respond to incidents, identify root causes, and improve systems based on findings ‱ Improve platform performance, utilization, security, and reliability as inference demand grows ‱ Collaborate directly with infrastructure, platform, and inference engineers in a flat organization

🎯 Requirements

‱ Experience building and operating production infrastructure or distributed systems, with real ownership of reliability ‱ Strong Linux fundamentals ‱ Practical knowledge of networking, storage, and containers ‱ Hands-on experience running Kubernetes in production ‱ Ability to write maintainable software and automation to solve infrastructure problems ‱ Systematic approach to debugging issues across application, cluster, network, and hardware boundaries ‱ Good judgment about when to move quickly, simplify, and prioritize reliability ‱ Initiative to take problems from investigation through implementation and collaborate with teammates ‱ Nice to have: experience with multi-region, multi-provider, or bare-metal infrastructure ‱ Nice to have: familiarity with GPUs, model serving, vLLM, or SGLang ‱ Nice to have: experience with infrastructure as code, CI/CD, observability, or automated recovery ‱ Nice to have: experience building highly available services, multi-tenant platforms, or distributed data systems

đŸ–ïž Benefits

‱ Ownership and reach to shape how Parasail scales its inference cloud ‱ Meaningful architecture decisions and direct production impact ‱ Work close to hardware, deep in distributed systems, and alongside inference engineers ‱ Opportunity to help build the foundation for the next stage of AI infrastructure

Apply Now

Similar Jobs

🕒 September 14

MEDvidi

201 - 500

đŸ„ Healthcare

⚕ Healthcare Insurance

Senior/Staff DevOps Engineer owning AWS, Kubernetes, security, and reliability for MEDvidi’s AI-powered mental healthcare platform. Driving infrastructure strategy, observability, CI/CD, and developer experience remotely across EU.

đŸ‡ȘđŸ‡ș Europe – Remote

💰 $2.8M Seed Round on 2022-09

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

đŸ—ŁïžđŸ‡·đŸ‡ș Russian Required

🕒 September 10

Bet On Talent

1 - 10

đŸ’Œ Consulting

📣 Marketing

🎯 Recruiter

Senior DevOps Engineer scaling AWS and Cloudflare infrastructure for a real-time, real-money B2B platform. Owning Terraform, CI/CD, observability, security, and cloud architecture.

đŸ‡ȘđŸ‡ș Europe – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 27

Lisk

11 - 50

₿ Crypto

🌐 Web 3

💳 Fintech

DevOps Engineer managing AWS infrastructure for Lisk's fintech platform. Evolving cloud infrastructure and ensuring security and reliability in a remote-first environment.

đŸ‡ȘđŸ‡ș Europe – Remote

💰 Seed on 2024-02

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 24

Bet On Talent

1 - 10

đŸ’Œ Consulting

📣 Marketing

🎯 Recruiter

Senior DevSecOps Engineer responsible for implementing and supporting technological solutions. Building CI/CD pipelines and monitoring systems in a fully remote environment, predominantly from Europe.

đŸ‡ȘđŸ‡ș Europe – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 7

JobLeads

51 - 200

🎯 Recruiter

đŸ‘„ HR Tech

Senior DevOps Engineer optimizing infrastructure and supporting a worldwide user base. Working with a remote international team to improve deployment processes and performance in a tech-heavy environment.

đŸ‡ȘđŸ‡ș Europe – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)