Staff Software Engineer, Platform Infrastructure

Job not on LinkedIn

🔥 5 minutes ago

🇮🇪 Ireland – Remote

⏰ Full Time

🔴 Lead

🏗️ Platform Engineer

👻 Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Pantheon Platform

Pantheon Platform

501 - 1000 employees

💼 Consulting

📣 Marketing

☁️ SaaS

Consulting • Marketing • SaaS

Pantheon Platform is the WebOps platform for managing websites built on Drupal and WordPress. The company focuses on empowering developers, IT professionals, and marketers to develop, test, and release website changes quickly and reliably, thereby enhancing the management of single or multiple websites. With its cloud-native software, Pantheon aims to help organizations create value through efficient website operations.

📋 Description

• Establish SRE as a discipline across the Platform Infrastructure Engineering team and broader Internal Platform Group • Define and drive adoption of SLO/SLI frameworks, reliability standards, and incident response practices • Own and evolve observability patterns using Prometheus metrics, OpenTelemetry tracing, and structured logging • Design and deliver service templates, Terraform modules, and infrastructure patterns for Go APIs, CLIs, Cloud Run services, and GKE workloads • Contribute to migrations and standardization involving GCP Secret Manager, GitHub Actions, Cloud Build, and Cloud SQL • Promote GCP-native managed and serverless services over self-hosted infrastructure • Apply security-first practices to IAP, IAM, VPC design, networking, and identity • Support compliance posture across SOC 2, ISO 27001, PCI DSS, and other frameworks • Mentor engineers and raise the technical bar across PIE and adjacent teams • Own development, testing, operations, and support for built systems in a full DevOps model • Join the on-call rotation after an initial ramp period of typically 3–6 months • Reduce on-call burden through automation, runbooks, and reliability improvements • Collaborate with teams across North America and Europe on cross-timezone standups and incident response

🎯 Requirements

• Deep understanding of SLOs, SLIs, error budgets, toil reduction, and operationalizing SRE practices • Hands-on expertise with GCP services, including Cloud Run, GKE, Cloud SQL, GCP Secret Manager, IAP/IAM, and networking • Practical experience implementing metrics, traces, and logs using Prometheus, OpenTelemetry, and structured logging • Practical experience with Grafana in production • Strong Terraform skills, including reusable module design • Experience setting technical direction for a platform or team • Experience translating ambiguous reliability goals into concrete architecture • Experience influencing engineers who do not report to you • Clear communication of reliability risk, architectural decisions, and incident postmortems • 8+ years building and operating production systems, with significant infrastructure, SRE, or platform engineering experience • Demonstrated SRE experience implementing SLO frameworks, incident management, on-call culture, and measurable reliability improvements • Strong hands-on Go proficiency, Go 1.21+ • Python as a secondary language • Deep practical GCP experience; comparable cloud experience accepted with clear willingness to ramp up on GCP • Hands-on experience building and maintaining Terraform modules • Solid operational experience with GKE or equivalent Kubernetes platforms • Production experience with Grafana, Prometheus, and OpenTelemetry • Practical experience with IAP, IAM, VPC design, and compliance frameworks such as SOC 2 or PCI DSS • Experience building reusable infrastructure patterns or developer platforms • Visa sponsorship is not available at this time

🏖️ Benefits

• Industry competitive compensation and equity plan • Robust vacation package with 28 days of holiday • Private medical and dental coverage • Life and critical illness insurance • Attractive workplace pension scheme • Access to Employee Resource Platform • Top-of-line equipment • Monthly allowance for wellness and reading • Access to LinkedIn Learning for continued development • Team-based and company-wide events and activities

Apply Now

Similar Jobs

🕒 July 27

Sky Betting & Gaming

1001 - 5000

🎲 Gambling

🎮 Gaming

👥 B2C

Staff Platform Engineer building cloud-native platforms at Fanatics Betting & Gaming. Collaborate with teams in UK and Ireland, drive technical initiatives in a remote role.

AWS

Cloud

Java

Kotlin

Kubernetes

Python

Terraform

Go