
51 - 200 employees
Founded 2022
🤖 Artificial Intelligence
☁️ SaaS
🤝 B2B
💰 $20M Seed on 2024-06
Artificial Intelligence • SaaS • B2B
Runpod is a cloud platform that provides on-demand GPU compute and managed infrastructure tailored for AI development and deployment. It offers GPU "Pods" across 31 global regions, serverless GPU endpoints for low-latency inference, multi-node GPU clusters for distributed training, and a hub for deploying open-source models and templates. Runpod emphasizes fast startup (sub-200ms cold starts), autoscaling from zero to thousands of workers, support for 30+ GPU SKUs, and tooling for the full AI lifecycle from experiment to production, targeting developers and enterprise AI teams.
🔥 1 hour ago
🇺🇸 United States – Remote
💵 $152k - $175k / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
👷 Infrastructure Engineer
Improve your chances of getting an interview by checking your resume score before you apply.

51 - 200 employees
Founded 2022
🤖 Artificial Intelligence
☁️ SaaS
🤝 B2B
💰 $20M Seed on 2024-06
Artificial Intelligence • SaaS • B2B
Runpod is a cloud platform that provides on-demand GPU compute and managed infrastructure tailored for AI development and deployment. It offers GPU "Pods" across 31 global regions, serverless GPU endpoints for low-latency inference, multi-node GPU clusters for distributed training, and a hub for deploying open-source models and templates. Runpod emphasizes fast startup (sub-200ms cold starts), autoscaling from zero to thousands of workers, support for 30+ GPU SKUs, and tooling for the full AI lifecycle from experiment to production, targeting developers and enterprise AI teams.
• Design and implement robust workload and network isolation architectures for RunPod's multitenant GPU bare-metal and virtualized environments. • Harden Linux kernel configurations, container runtimes (e.g., Docker, containerd), and orchestration layers (e.g., Kubernetes) against breakouts and privilege escalation. • Conduct deep-dive security assessments and penetration testing specifically targeting our hypervisor, network stack, and hardware interfaces. • Write code (primarily C, Go, or Rust) to implement custom security controls, telemetry, and fixes at the OS and infrastructure level. • Evaluate and mitigate security considerations specific to GPU architecture, PCIe pass-through, and shared memory spaces. • Serve as the technical escalation point for infrastructure-level security incidents, developing forensic capabilities for ephemeral container environments.
• 5+ years of experience in infrastructure or systems-level security engineering. • Extensive knowledge of Linux kernel internals (cgroups, namespaces, eBPF, SELinux/AppArmor). • Deep understanding of virtualization technologies (KVM, QEMU) and workload/network isolation techniques in multitenant environments. • Strong systems-level programming skills in C, Go, Rust, or Python. • Familiarity with GPU architecture and hardware-level security considerations. • Experience in securing bare-metal cloud infrastructure and mitigating lower-level CVEs.
• Meaningful equity in a fast-growing company- everyone on the team receives stock options — your impact drives our growth, and you share in the upside. • Generous medical, dental & vision plans • Flexible PTO- take the time you need to recharge • Most roles are remote work first with an inclusive, collaborative teams utilizing slack as the main form of internal communication • $1,200 Home Office & Equipment Stipend- We set you up for success from day one with gear and support to create your ideal workspace
Apply Now🔥 1 hour ago
Senior Infrastructure Engineer developing internal tools and processes at Angi. Leveraging AWS and various programming languages to enhance website infrastructure and performance.
🇺🇸 United States – Remote
💵 $165k - $216k / year
💰 $30M Debt Financing - Angi on 2011-10
⏰ Full Time
🟠 Senior
👷 Infrastructure Engineer
AWS
Chef
Cloud
Docker
EC2
Kafka
Kubernetes
Node.js
Postgres
Python
Ruby
Terraform
Zookeeper
Go
🔥 1 hour ago
Infrastructure Engineer responsible for ownership and operation of cloud accelerator platform. Building control plane, managing provisioning, and ensuring reliability across compute fleet.
Ansible
Cloud
Distributed Systems
Kubernetes
Linux
Node.js
Python
Rust
Terraform
Go
🔥 2 hours ago
Senior Software Engineer developing high-performance databases and optimizing data pipelines for TRM Labs. Collaborating with cross-functional teams to build a safer financial system.
Airflow
MySQL
Postgres
RDBMS
SQL
🔥 2 hours ago
Senior or Staff AI Infrastructure Engineer at TRM Labs providing scalable solutions for AI/ML systems. Focused on building and enhancing infrastructure for next-gen AI applications.
Docker
Kubernetes
Prometheus
Python
Terraform
🔥 2 hours ago
Senior Infrastructure Engineer designing scalable infrastructure solutions for TRM Labs. Collaborate with data scientists and engineers to enhance operational capabilities.
Airflow
AWS
BigQuery
Google Cloud Platform
Kubernetes
Linux
React
Terraform
Unix