
10,000+ employees
Founded 1993
đĽ Healthcare
đ Manufacturing
đ¤ Artificial Intelligence
Healthcare ⢠Manufacturing ⢠Artificial Intelligence
NVIDIA is a leading technology company specializing in accelerated computing and artificial intelligence. NVIDIA pioneers advancements in graphical processing units (GPUs), cloud computing, data centers, and virtual reality, with a focus on gaming, automotive, healthcare, and robotics industries. The company's innovations, such as NVIDIA Omniverse, transform traditional digital processes by enabling high-fidelity simulations and rendering tasks. Their applications span various industries, from autonomous vehicles using NVIDIA DRIVE to healthcare solutions with NVIDIA Clara, and AI-driven analytics and workflows.
đ July 13
Improve your chances of getting an interview by checking your resume score before you apply.

10,000+ employees
Founded 1993
đĽ Healthcare
đ Manufacturing
đ¤ Artificial Intelligence
Healthcare ⢠Manufacturing ⢠Artificial Intelligence
NVIDIA is a leading technology company specializing in accelerated computing and artificial intelligence. NVIDIA pioneers advancements in graphical processing units (GPUs), cloud computing, data centers, and virtual reality, with a focus on gaming, automotive, healthcare, and robotics industries. The company's innovations, such as NVIDIA Omniverse, transform traditional digital processes by enabling high-fidelity simulations and rendering tasks. Their applications span various industries, from autonomous vehicles using NVIDIA DRIVE to healthcare solutions with NVIDIA Clara, and AI-driven analytics and workflows.
⢠Design, implement, and maintain large-scale HPC/AI clusters with state-of-the-art monitoring, logging, and alerting systems. ⢠Utilize and develop tools to manage infrastructure as code, ensuring scalable and repeatable deployments. ⢠Develop and maintain continuous integration and continuous delivery (CI/CD) pipelines to automate deployment processes. ⢠Develop automation scripts and tools to automate deployment, configuration management, and operational monitoring. ⢠Perform comprehensive troubleshooting from bare metal to application level, ensuring system reliability and efficiency. ⢠Serve as a technical resource, developing and sharing best practices with internal teams. ⢠Support R&D activities and engage in proof of concepts (POCs) and proof of values (POVs) for future improvements.
⢠B.Sc. in Computer Science, Engineering, or a related field with 5+ years of experience. ⢠Deep knowledge of HPC and AI solution technologies, including CPUs, GPUs, high-speed interconnects, and supporting software. ⢠Advanced proficiency in programming and scripting languages, with a solid understanding of object-oriented programming principles. ⢠Familiarity with Jenkins, Ansible, Puppet/Chef. ⢠Excellent knowledge of Windows and Linux (Redhat/CentOS and Ubuntu), networking and OS-level security. ⢠Deep understanding of networking protocols such as InfiniBand and Ethernet. ⢠Experience with job scheduling workloads and orchestration tools such as Slurm and Kubernetes. ⢠Background with multiple storage solutions like Lustre, GPFS, ZFS, and XFS. ⢠Expertise with virtual systems (VMware, Hyper-V, KVM, Citrix). ⢠Familiarity with cloud platforms (AWS, Azure, Google Cloud).
⢠Health insurance ⢠Professional development opportunities ⢠Flexible work arrangements
Apply Nowđ July 13
Senior Site Reliability Engineer at ClickHouse responsible for maintaining reliability and performance of cloud infrastructure. Collaborating with engineering teams to design scalable systems for real-time analytics.
Ansible
AWS
Azure
Cloud
Docker
Google Cloud Platform
Kubernetes
Puppet
Python
SQL
Terraform
Go
đ July 13
DevOps Lead overseeing infrastructure strategy for a rapidly scaling AI legal tech platform. Collaborating with engineering teams to modernize infrastructure and drive best practices.
đŠđŞ Germany â Remote
đ° $7.3M Seed Round - Jupus on 2025-05
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
AWS
Cloud
Grafana
Kubernetes
Terraform
đ July 11
DevOps Automation Specialist developing automation logic with VMware vRealize Automation. Working remotely for Dedalus, a global leader in healthcare technology.
đŠđŞ Germany â Remote
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đŁď¸đŠđŞ German Required
Ansible
Linux
Oracle
Postgres
Puppet
Python
SaltStack
VMware
đ July 9
Join XTEL as a Senior DevOps Engineer to build Azure infrastructure and automate deployments. Collaborate across teams in a remote, inclusive, growth-focused environment.
Azure
Cloud
DNS
Kubernetes
TCP/IP
Terraform
đ July 7
DevOps Generalist automating infrastructure tasks for a fast-growing tech company. Collaborate with international teams to enhance system reliability and performance.
đŠđŞ Germany â Remote
đľ âŹ60k - âŹ80k / year
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đŁď¸đŠđŞ German Required
Cloud
Kubernetes