Fleet Reliability Engineer

🔥 0 minutes ago

IoT

Python

SaltStack

SQL

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Quartermaster

Quartermaster

1 - 10 employees

Founded 2023

📦 Logistics

🏭 Manufacturing

🎖️ Defense

Logistics • Manufacturing • Defense

Quartermaster is a company redefining maritime intelligence by providing unmatched precision and visibility in real-time to monitor the world's most remote waters. Their AI-powered SmartMast system transforms everyday vessels into a comprehensive ocean intelligence network, significantly enhancing the monitoring of maritime activities and addressing issues such as smuggling and illegal fishing. By leveraging onboard sensors and high-bandwidth satellite connectivity, Quartermaster delivers high-resolution detections and actionable insights to users, ensuring enhanced maritime awareness and decision-making.

📋 Description

• Own the health, uptime, and long-term reliability of the deployed SmartMast hardware fleet • Track fleet health, online status, degradation, failure causes, and corrective actions • Build and refine dashboards, alerting, and telemetry pipelines for power, thermal, connectivity, camera, radio, and compute systems • Define subsystem health criteria and proactive-intervention thresholds • Lead root-cause analysis from telemetry forensics through physical teardown of returned units • Maintain the fleet failure database and drive FMEA, reliability growth tracking, and CAPA to closure • Translate field failures into hardware, component-selection, manufacturing, and firmware improvements • Define preventive maintenance schedules, spares strategy, and RMA/repair-and-return workflows • Update installation, diagnostic, field-repair, and troubleshooting procedures • Support field deployments and complex repairs, including periodic travel to vessels, ports, and installation sites • Establish environmental and life-test protocols for vibration, corrosion, thermal, ingress, and power qualification • Feed reliability requirements and acceptance criteria into hardware revisions and supplier qualification • Design scalable reliability processes and tooling for a fleet growing into thousands of units

🎯 Requirements

• Bachelor's degree in Electrical, Mechanical, Systems, Reliability, or a related engineering discipline, or equivalent hands-on experience • 5+ years of engineering experience with deployed electro-mechanical hardware, including at least 2 years in reliability, sustaining/field engineering, or hardware operations for a fielded product • Direct ownership of hardware reliability outcomes for a fleet or installed base, including uptime, failure rates, or MTBF/MTTR • Hands-on proficiency with RCA methods including 8D, 5-Whys, and fishbone • Experience with FMEA, fault-tree analysis, and CAPA • Practical experience diagnosing electro-mechanical systems using telemetry/logs, bench instruments, and physical teardown • Ability to query, analyze, and visualize fleet telemetry using SQL and Python or equivalent • Working knowledge of electronics, power systems, mechanical enclosures, and harsh-environment hardware failure modes • Willingness and ability to travel periodically to field sites, including occasional international travel • Legally authorized to work in the United States and able to satisfy customer- or contract-driven eligibility requirements for government and maritime-security work • Preferred: experience with marine, maritime, offshore, automotive, aerospace/defense, satellite, telecom, or other harsh-environment fleets • Preferred: familiarity with IP-rated enclosures, corrosion and salt-fog effects, marine power systems, and IEC 60529, MIL-STD-810, or IEC 60068 • Preferred: experience with connected/IoT or edge devices, remote diagnostics, OTA firmware updates, and embedded-system/connectivity telemetry • Preferred: exposure to camera/optical systems, RF/software-defined radios, batteries, or edge-AI compute hardware • Preferred: experience building a reliability or sustaining-engineering function from scratch • Preferred: ASQ Certified Reliability Engineer or comparable credential

Apply Now

Similar Jobs

🔥 15 minutes ago

Blend360

501 - 1000

🏥 Healthcare

🏨 Hospitality

✈️ Travel

Lead DevOps/AIOps Engineer architecting GCP infrastructure, data platforms, and MLOps capabilities for Blend360’s enterprise consulting clients. Driving automation, reliability, security, and production AI delivery.

🔥 33 minutes ago

NVIDIA

10,000+ employees

🏥 Healthcare

🏭 Manufacturing

🤖 Artificial Intelligence

Senior Site Reliability Engineer operating NVIDIA’s Base Command Manager and large-scale GPU clusters. Handling incidents, deployments, and resilient Slurm/Kubernetes infrastructure for AI data centers.

🔥 1 hour ago

SAIC

10,000+ employees

☁️ SaaS

📣 Marketing

🏢 Enterprise

DevSecOps Engineer securing SAIC’s Windows, Linux, and cloud infrastructure. Managing Active Directory, automation, CI/CD pipelines, and compliance in controlled defense environments.

🇺🇸 United States – Remote

🔥 Funding within the last year

💰 $500M Post-IPO Debt - SAIC on 2025-09

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🔥 1 hour ago

Coupa Software

1001 - 5000

💼 Consulting

📦 Logistics

🏥 Healthcare

Lead Active Directory Site Reliability Engineer architecting secure identity infrastructure for Coupa’s AI-powered spend management platform. Driving automation, cloud integration, observability, and least-privilege administration across global environments.

🔥 4 hours ago

InnoData

2 - 10

🤝 B2B

💼 Consulting

🌍 Social Impact

Application Reliability Engineer supporting GCP and Google App Engine applications for Innodata, a global data engineering and AI services company. Managing incidents, deployments, microservices, and reliability improvements.