
11 - 50 employees
🤖 Artificial Intelligence
🏢 Enterprise
☁️ SaaS
Artificial Intelligence • Enterprise • SaaS
TensorWave is a cutting-edge AI compute company offering advanced solutions for enterprises focusing on training, fine-tuning, and inference. Powered by AMD's Instinct MI300X accelerators, TensorWave delivers a cost-effective and high-performance alternative to Nvidia's H100 chip. The company provides options for both bare-metal nodes and fully-managed Kubernetes clusters, ensuring ease of use with native support for PyTorch and TensorFlow. TensorWave is committed to data security with SOC2 Type II and HIPAA compliance, making it a reliable partner for businesses requiring dedicated and secure environments.
🔥 33 minutes ago
Improve your chances of getting an interview by checking your resume score before you apply.

11 - 50 employees
🤖 Artificial Intelligence
🏢 Enterprise
☁️ SaaS
Artificial Intelligence • Enterprise • SaaS
TensorWave is a cutting-edge AI compute company offering advanced solutions for enterprises focusing on training, fine-tuning, and inference. Powered by AMD's Instinct MI300X accelerators, TensorWave delivers a cost-effective and high-performance alternative to Nvidia's H100 chip. The company provides options for both bare-metal nodes and fully-managed Kubernetes clusters, ensuring ease of use with native support for PyTorch and TensorFlow. TensorWave is committed to data security with SOC2 Type II and HIPAA compliance, making it a reliable partner for businesses requiring dedicated and secure environments.
• Resolve Complex Escalations: Act as the final authority on issues exceeding GOC scope, utilizing code-level debugging and architectural investigation. • Direct Customer Engagement: Partner with customer technical leads to diagnose production issues, ensuring transparency and rapid resolution through active collaboration. • Iterative Problem Solving: Develop diagnostic scripts and workarounds to maintain customer operations while long-term patches are in development. • Drive Root Cause Analysis: Own end-to-end P1 resolution, partnering with TAMs to deliver clear, actionable post-incident analysis. • Bridge to Engineering: Convert recurring customer pain points into evidence-based feature requests, influencing product roadmap to resolve systemic failures. • Build Scalable Knowledge: Document non-obvious platform behaviors and refine GOC runbooks, ensuring institutional knowledge grows with every incident.
• 5–9 years in Infrastructure Engineering, Platform Engineering, or SRE, with a specific focus on high-performance computing or large-scale AI stacks. Proven track record of managing complex production environments where system reliability is mission-critical. • Kubernetes Expert: Deep experience in cluster administration and scheduler internals; comfortable reading/modifying controller code. • AI/GPU Infrastructure Specialist: Proficient in orchestrating GPU workloads and diagnosing training job failures using ROCm or CUDA. • Network Pathologist: Skilled in RDMA/RoCEv2, SRIOV, and BGP; capable of interpreting switch telemetry to identify silent packet drops. • Linux Power User: Expert in kernel networking, hugepages, and cgroups; able to debug at the OS layer when applications are silent. • Builder Mindset: Proficient in Python and Ansible; capable of writing custom diagnostic tools to automate remediation. • Executive Communicator: Strong technical rigor when presenting findings to VPs of Engineering, maintaining trust while delivering difficult updates. • Prior experience in a customer-facing engineering role (e.g., Solutions Engineering, Technical Support Engineering). • Experience in high-uptime environments where 24/7/365 availability is required.
• Stock Options • 100% paid Medical, Dental, and Vision insurance for Employees • Company Health Savings Account Contributions • 100% paid Short Term and Long Term Disability Insurance for Employees • Life and Voluntary Supplemental Insurance Options • Other Insurance Options, such as Pet & Legal Insurance • Various Supplementary Health Benefits, such as discounted Virtual Healthcare Appointments and Serious Illness Support • Flexible Spending Account • 401(k) • Employee Assistance Program • Flexible PTO • Paid Holidays • Parental Leave • Other In-Office Perks
Apply Now🔥 33 minutes ago
Manager leading sales operations for Comcast Business focusing on enterprise customers. Responsible for team management, operational objectives, and sales goals achievement.
🇺🇸 United States – Remote
💵 $110.4k - $184k / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
💻 Solutions Engineer
🦅 H1B Visa Sponsor
🔥 3 hours ago
Senior Partner Solutions Manager at Bizee supporting distribution partnerships and helping entrepreneurs through strategic solutions. Managing onboarding, growth, and operational excellence across partner channels.
🔥 4 hours ago
Senior Customer Solutions Engineer managing enterprise customer relationships in cloud security. Focused on strategic engagement and ensuring customer success with the Sysdig platform.
🇺🇸 United States – Remote
💵 $130k - $163k / year
💰 $350M Series G on 2021-12
⏰ Full Time
🟠 Senior
💻 Solutions Engineer
🦅 H1B Visa Sponsor
🔥 4 hours ago
Solution Architect responsible for designing and optimizing enterprise applications at a US manufacturing and distribution firm. Focus on core systems including ERP, CRM, and emerging AI-driven solutions.
🔥 5 hours ago
Solutions Excellence Engineer optimizing fulfillment operations and driving continuous improvement initiatives at Ocado Group. Collaborating with cross-functional teams to enhance quality and productivity in North America.