Platform Engineer – Incident Management

Job not on LinkedIn

🔥 6 minutes ago

🇺🇸 United States – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

🏗️ Platform Engineer

👻 Ghost score 15%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Bitso

Bitso

501 - 1000 employees

Founded 2014

₿ Crypto

💸 Finance

🛍️ eCommerce

Crypto • Finance • eCommerce

Bitso is a cryptocurrency exchange platform that allows users to buy, sell, and trade various digital assets, including Bitcoin (BTC), Ethereum (ETH), and stablecoins like Tether (USDT). Operating primarily in Latin America, specifically in countries such as Argentina, Brazil, Colombia, and Mexico, Bitso positions itself as a key player in bridging the gap between traditional finance and cryptocurrency. The platform also provides insights and updates on the latest trends in the crypto market through its blog, aiming to educate users on crypto technologies and investment opportunities.

📋 Description

• Own and execute on-call shifts end-to-end: acknowledge pages within SLA, declare incidents, assign roles, maintain communications cadence, and drive incidents to resolution • Build automation for the Sev1/Sev2 postmortem workflow, including scheduling, facilitation reminders, action-item assignment, ownership tracking, and due-date enforcement • Leverage AI to identify patterns across incidents and propose systemic fixes, including runbook improvements, alert tuning, platform hardening, and process changes • Build and extend internal automation and tooling, including AI-assisted incident-response workflows, to reduce manual toil and accelerate detection and resolution • Contribute to and improve observability: dashboards, alert configurations, and early-warning signals across Bitso’s platform • Participate in change and maintenance management processes, applying risk management to reduce deployment-related incidents • Collaborate with engineering squads across the company to surface platform risks and drive preventive actions • Keep incident tooling, runbooks, and severity criteria accurate, current, and useful for the broader engineering organization • Report to the Incident Management Manager

🎯 Requirements

• Proven ability to operate confidently in high-pressure incident scenarios, including communicating clearly with senior stakeholders and leadership while a production issue is live • Hands-on experience with Kubernetes — comfortable deploying, debugging, and navigating pod-level issues • Solid understanding of CI/CD pipelines and modern DevOps practices • Software development background in any language; ability to read, write, and debug code is essential (Python or Java experience is a plus) • Strong automation mindset: identify repetitive toil and eliminate it • Experience building or working with AI agents or LLM-based workflows is highly desirable • Strong interpersonal and written communication skills • Self-directed learner who doesn’t need a fully defined path to start contributing • Fintech or crypto industry background is a plus • Full incident-lifecycle ownership and incident-management experience implied by the role responsibilities

Apply Now

Similar Jobs

🔥 52 minutes ago

Peraton

10,000+ employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Senior Software Engineer integrating federal ESA species data into Peraton’s IPaC consultation platform. Designing APIs, architecture updates, and deployment workflows for government agencies.

🔥 6 hours ago

inKind

51 - 200

🍽️ Food & Beverage

📣 Marketing

🏨 Hospitality

Platform Engineer building and operating AWS infrastructure for inKind’s restaurant rewards platform. Improving reliability, security, observability, deployments, and cloud scalability.

🔥 7 hours ago

Danaher

10,000+ employees

🧬 Biotechnology

🏥 Healthcare

🔬 Science

Senior Data Platform Automation Engineer automating Danaher’s Snowflake, Azure, Matillion, dbt, and Airflow platform. Building IaC, CI/CD, observability, and AI-driven operations for life-sciences businesses.

🔥 9 hours ago

Humana

10,000+ employees

🏥 Healthcare

🛡️ Insurance

⚕️ Healthcare Insurance

Senior DevOps Platform Engineer securing Humana’s Azure/GCP cloud and CI/CD platforms. Owning JFrog, Kubernetes, GitOps, and software supply chain security for regulated healthcare delivery.

🔥 16 hours ago

Systems Planning & Analysis

1001 - 5000

📦 Logistics

🏥 Healthcare

💼 Consulting

Senior Platform Engineer supporting SPA’s platform-as-a-service solution for U.S. space and defense missions. Improving Kubernetes adoption, reliability, automation, and developer experience.