
11 - 50 funcionários
🔧 Hardware
🏢 Corporativo
🤖 Inteligência Artificial
💰 $10.000.000 Seed Round em 2022-04
Hardware • Enterprise • Artificial Intelligence
Hydra Host é um fornecedor de soluções de computação de alto desempenho, oferecendo acesso a servidores dedicados de GPU bare metal otimizados para cargas de trabalho de IA e HPC. Sua plataforma permite que usuários acessem e aluguem GPUs de ponta globalmente, proporcionando desempenho, segurança e personalização incomparáveis. A infraestrutura da Hydra Host inclui um marketplace, conhecido como Brokkr, que oferece uma ampla gama de configurações e soluções de GPU adaptadas para aplicações críticas, como IA, big data e aprendizado de máquina. Através de suas soluções robustas, seguras e escaláveis, a Hydra Host garante que os clientes desfrutem de controle total sobre seus ambientes de servidor, com opções para escalabilidade e prontidão para o futuro. As ofertas da empresa são confiáveis por empresas líderes que buscam soluções de computação eficientes e inovadoras.
🕒 Outubro 29, 2025
🐊 Florida – Remoto
💵 $140.000 - $200.000 / ano
⏰ Tempo Integral
🟡 Pleno
🟠 Sênior
⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)
🗣️🇺🇸🇬🇧 Inglês obrigatório
Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

11 - 50 funcionários
🔧 Hardware
🏢 Corporativo
🤖 Inteligência Artificial
💰 $10.000.000 Seed Round em 2022-04
Hardware • Enterprise • Artificial Intelligence
Hydra Host é um fornecedor de soluções de computação de alto desempenho, oferecendo acesso a servidores dedicados de GPU bare metal otimizados para cargas de trabalho de IA e HPC. Sua plataforma permite que usuários acessem e aluguem GPUs de ponta globalmente, proporcionando desempenho, segurança e personalização incomparáveis. A infraestrutura da Hydra Host inclui um marketplace, conhecido como Brokkr, que oferece uma ampla gama de configurações e soluções de GPU adaptadas para aplicações críticas, como IA, big data e aprendizado de máquina. Através de suas soluções robustas, seguras e escaláveis, a Hydra Host garante que os clientes desfrutem de controle total sobre seus ambientes de servidor, com opções para escalabilidade e prontidão para o futuro. As ofertas da empresa são confiáveis por empresas líderes que buscam soluções de computação eficientes e inovadoras.
• Design, deploy, and maintain QA systems used by our development teams to test integration and live system responses across full-stack deployments in local, live, and ephemeral environments • Evaluate and integrate monitoring and QA tools to find the right tools for the job • Create a unified monitoring platform and processes that datacenter and device teams will integrate to monitor their components (live servers, lifecycle, networks, power, etc.) • Maintain monitoring processes and dashboards to provide complete visibility into the health, performance, and reliability of our CI systems, software deployments, and testing platforms • Create and maintain a systems test suite, in collaboration with our product managers, to validate marketplace changes against all business functions in live and ephemeral QA environments • Integrate all fore-mentioned systems to create holistic platform health statistics reporting • Design disaster-recovery processes in collaboration with devops • Ensure we are meeting uptime SLAs across all platform deployments • Work with datacenter and device teams to define service-level indicators (SLIs), service-level objectives (SLOs), and SLAs • Establish observability standards across the stack: logs, metrics, traces, and alerts, and actionable on-call playbooks • Automate everything from monitoring setups to incident responses to eliminate manual toil and increase reliability • Drive incident response, root cause analysis, and post‑mortems • Guide incident turn-around into tooling and process improvements • Establish the monitoring infrastructure and dashboards that enable everyone — from engineers to execs — to know what’s going on • Act as the reliability partner to engineering teams: review systems for reliability concerns, help design QA requirements and testing, and help teams meet reliability targets.
• 5–8+ years of experience in Reliability Engineering, DevOps, or infrastructure roles focused on large-scale, high-uptime production environments • Deep familiarity with monitoring and observability tooling: you've implemented and managed systems, esp. Prometheus, Grafana, and Zabbix • Strong experience with service orchestration in mutli-region environment (Nomad, Kubernetes, cloud VMs, distributed databases) • Track record of managing production system uptime and SLAs and building tools to support it • Experience writing and reviewing post-mortems and using those findings to drive improvements in tools and process • Proficient with scripting and programming languages (Python, Go, BASH, etc.) for automating operational tasks • Strong proficiency with infrastructure as code and devops workflows • Experience with distributed tracing, log aggregation, and alert tuning • Passion for building systems that fail gracefully, alert correctly, and empower others to operate confidently • Excellent communication skills: you can write clear documentation, drive incident reviews, and communicate reliability risks to technical and non-technical stakeholders.
• Competitive compensation: base salary + performance bonus + equity • Exposure to high-performance computing and state-of-the-art GPU environments • A core role in ensuring our systems are reliable, observable, and meet customer SLAs • Remote work environment with a strong culture of ownership and autonomy • No red tape: find the right solution, work with the team, get feedback, and get the job done.
Candidatar-se🕒 Outubro 24, 2025
DevOps Engineer designing and implementing automation processes at CARET, enhancing efficiency for legal and accounting firms. Collaborates with different teams leveraging cloud technologies for optimal service delivery.
🇺🇸 Estados Unidos – Remoto (EUA)
💵 $90.000 - $110.000 / ano
⏰ Tempo Integral
🟡 Pleno
🟠 Sênior
⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)
🗣️🇺🇸🇬🇧 Inglês obrigatório
🕒 Outubro 22, 2025
Senior DevOps Engineer responsible for application deployments and cloud infrastructure management. Collaborating with teams to automate processes and ensure high performance of applications.
🇺🇸 Estados Unidos – Remoto (EUA)
💵 $10.000 / mês
⏰ Tempo Integral
🟠 Sênior
⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)
🗣️🇺🇸🇬🇧 Inglês obrigatório
🕒 Outubro 18, 2025
Senior Full Stack TypeScript Developer designing a loan origination system for a financial services enterprise. Collaborating with tech leads and product teams using cutting-edge technologies.
🇺🇸 Estados Unidos – Remoto (EUA)
💵 $161.000 / ano
⏰ Tempo Integral
🟠 Sênior
⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)
🗣️🇺🇸🇬🇧 Inglês obrigatório
🕒 Outubro 16, 2025
CloudOps & DevOps Engineer enhancing secure data flows between Kafka, PostgreSQL, and S3. Collaborating with teams to ensure reliable infrastructure for client projects.
🇺🇸 Estados Unidos – Remoto (EUA)
⏰ Tempo Integral
🟡 Pleno
🟠 Sênior
⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)
🦅 Patrocina Visto H1B
🗣️🇺🇸🇬🇧 Inglês obrigatório
🕒 Outubro 15, 2025
Backend Developer creating scalable APIs and backend logic for an IT Services company. Collaborating with teams and ensuring performance, security, and code quality.
🇺🇸 Estados Unidos – Remoto (EUA)
💰 $1.000.000 Venture Round em 2006-04
⏰ Tempo Integral
🟡 Pleno
🟠 Sênior
⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)
🦅 Patrocina Visto H1B
🗣️🇺🇸🇬🇧 Inglês obrigatório