
51 - 200 Mitarbeiter
Gegründet 2022
🤖 Künstliche Intelligenz
☁️ SaaS
🤝 B2B
💰 €20.000.000 Seed im 2024-06
Artificial Intelligence • SaaS • B2B
Runpod ist eine Cloud-Plattform, die bedarfsgerechte GPU-Rechenleistung und verwaltete Infrastruktur bietet, die speziell für die Entwicklung und Bereitstellung von KI optimiert sind. Sie bietet GPU-"Pods" in 31 globalen Regionen, serverlose GPU-Endpunkte für latenzarme Inferenz, Multi-Node-GPU-Cluster für verteiltes Training und ein Hub zur Bereitstellung von Open-Source-Modellen und -Vorlagen. Runpod legt Wert auf schnellen Start (unter 200 ms Cold Starts), automatisches Skalieren von null auf tausende Arbeiter, Unterstützung für über 30 GPU-SKUs und Werkzeuge für den kompletten KI-Lebenszyklus von Experimenten bis hin zur Produktion, wobei sie Entwickler und KI-Teams in Unternehmen fokussiert.
🕒 vor 5 Tagen
🇺🇸 Vereinigte Staaten – Remote
💵 $150.000 - $240.000 / Jahr
⏰ Vollzeit
🟡 Mittelstufe
🟠 Senior
🔙 Backend-Entwickler
🗣️🇺🇸🇬🇧 Englisch erforderlich
Linux
NFS
Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

51 - 200 Mitarbeiter
Gegründet 2022
🤖 Künstliche Intelligenz
☁️ SaaS
🤝 B2B
💰 €20.000.000 Seed im 2024-06
Artificial Intelligence • SaaS • B2B
Runpod ist eine Cloud-Plattform, die bedarfsgerechte GPU-Rechenleistung und verwaltete Infrastruktur bietet, die speziell für die Entwicklung und Bereitstellung von KI optimiert sind. Sie bietet GPU-"Pods" in 31 globalen Regionen, serverlose GPU-Endpunkte für latenzarme Inferenz, Multi-Node-GPU-Cluster für verteiltes Training und ein Hub zur Bereitstellung von Open-Source-Modellen und -Vorlagen. Runpod legt Wert auf schnellen Start (unter 200 ms Cold Starts), automatisches Skalieren von null auf tausende Arbeiter, Unterstützung für über 30 GPU-SKUs und Werkzeuge für den kompletten KI-Lebenszyklus von Experimenten bis hin zur Produktion, wobei sie Entwickler und KI-Teams in Unternehmen fokussiert.
• Own Distributed Storage Architecture: Define, evolve, and operate Runpod’s global storage platforms, supporting training, inference, checkpointing, and dataset access at scale. • Build the Storage Engineering Team: Manage and grow a team of storage and systems engineers. Set clear ownership, technical direction, and operational standards across regions. • High-Performance Shared Filesystems: Design and operate large-scale SAN and NFS deployments, including performance-sensitive shared storage for GPU clusters. • Advanced Filesystems & Platforms: Lead deployments and operations of VAST Data and experience with Lustre or similar parallel filesystems used in HPC and AI environments. • End-to-End Performance Ownership: Drive performance optimization from NAND and NVMe media through controllers, networking, and client access patterns. • Next-Generation Storage Technologies: Evaluate and deploy cutting-edge capabilities such as NFS over RDMA, GPU Direct Storage (GDS), and low-latency data paths for accelerated workloads. • Reliability & Scale: Establish best practices for replication, data tiering, data protection, failure recovery, capacity planning, and lifecycle management. • Automation & Observability: Build automation for provisioning, expansion, upgrades, and monitoring. Ensure deep observability into throughput, latency, and error characteristics. • Cross-Functional Collaboration: Partner with Datacenter Networking, GPU Platform, SRE, and Product teams to ensure storage systems meet evolving workload and customer needs. • Vendor & Partner Management: Own technical relationships with storage vendors, hardware partners, and colocation providers; drive roadmap alignment and issue resolution.
• Engineering Leadership Experience: 3+ years managing storage, systems, or infrastructure engineering teams in production environments. • Distributed Storage Expertise: 8+ years designing and operating large-scale storage systems, including SAN and NFS architectures at multi-petabyte scale. • VAST Data Experience: Hands-on experience deploying, operating, or deeply integrating VAST Data in production environments is required. • Parallel Filesystems: Experience with Lustre or comparable HPC filesystems (e.g., GPFS, BeeGFS) supporting high-concurrency workloads. • Low-Level Storage Knowledge: Deep understanding of NAND, NVMe, PCIe, storage controllers, and performance characteristics across the stack. • High-Performance Data Paths: Proven experience with NFS over RDMA, RDMA-capable transports, or similar technologies. Familiarity with GPU Direct Storage strongly preferred. • Linux Systems Expertise: Strong Linux internals knowledge, including filesystems, I/O scheduling, memory management, and tuning for performance workloads. • Operational Excellence: Experience running 24/7 storage platforms with strong incident response, change management, and post-mortem discipline. • Communication & Leadership: Ability to clearly communicate complex technical tradeoffs and lead teams through high-stakes infrastructure decisions. • Successful completion of a background check.
• Meaningful equity in a fast-growing company- everyone on the team receives stock options — your impact drives our growth, and you share in the upside. • Generous medical, dental & vision plans • Flexible PTO- take the time you need to recharge • Most roles are remote work first with an inclusive, collaborative teams utilizing slack as the main form of internal communication • Join a passionate team on the cutting edge of AI infrastructure — where culture, learning, and ownership are at the heart of how we scale.
Jetzt Bewerben🕒 vor 5 Tagen
Backend Engineer owning core systems for Nebulock's threat hunting solutions. Focused on scalable data ingestion and processing for security telemetry.
🇺🇸 Vereinigte Staaten – Remote
🔥 Finanzierung im letzten Jahr
💰 €6.000.000 Seed im 2025-08
⏰ Vollzeit
🟡 Mittelstufe
🟠 Senior
🔙 Backend-Entwickler
🗣️🇺🇸🇬🇧 Englisch erforderlich
🕒 vor 5 Tagen
Software Engineer developing core database features for VillageSQL, a community-driven database technology. Join a passionate team building innovative solutions and engaging with open-source communities.
🇺🇸 Vereinigte Staaten – Remote
💵 $118.000 - $250.000 / Jahr
⏰ Vollzeit
🟡 Mittelstufe
🟠 Senior
🔙 Backend-Entwickler
🗣️🇺🇸🇬🇧 Englisch erforderlich
RDBMS
🕒 vor 5 Tagen
Senior Software Engineer developing scalable backend solutions for TRM's AI-powered crime investigation platform. Collaborate on building APIs and features for internal and external stakeholders in a remote role.
🗣️🇺🇸🇬🇧 Englisch erforderlich
🕒 vor 5 Tagen
Software Engineer focusing on building third-party API integrations for TRM’s AI-native products. Collaborate with teams to improve integration processes and systems.
🇺🇸 Vereinigte Staaten – Remote
💵 $180.000 - $240.000 / Jahr
⏰ Vollzeit
🟡 Mittelstufe
🟠 Senior
🔙 Backend-Entwickler
🗣️🇺🇸🇬🇧 Englisch erforderlich
🕒 vor 5 Tagen
Senior Python Engineer at eBay developing scalable backend services for AI-driven marketplaces. Collaborating with cross-functional teams to enhance platform capabilities and performance.
🇺🇸 Vereinigte Staaten – Remote
💵 $118.800 - $205.600 / Jahr
💰 €1.150.000.000 Post-IPO Debt - eBay im 2022-11
⏰ Vollzeit
🟠 Senior
🔙 Backend-Entwickler
🗣️🇺🇸🇬🇧 Englisch erforderlich