Senior Platform Engineer, Cloud Infrastructure

Job not on LinkedIn

🕒 August 10

🏈 North America – Remote

⏰ Full Time

🟠 Senior

☁️ Cloud Engineer

👻 Ghost score 44%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Virtasant

Virtasant

51 - 200 employees

💼 Consulting

🏢 Enterprise

🤝 B2B

Consulting • Enterprise • B2B

Virtasant is a cloud services and optimization company that helps businesses migrate to, manage, and build on public cloud platforms. They combine a proprietary automation platform with managed services and FinOps expertise to reduce cloud costs (claiming average savings of over 50%), support Cloud FinOps programs, and deliver outcome-based engagements rather than hourly or seat-based billing. Virtasant also provides enterprise AI guidance and partners with major cloud providers (AWS, Google Cloud, Azure) to deliver migration, optimization, and 24/7 managed operations.

📋 Description

• Design, build, and operate production Kubernetes clusters, including cluster networking, workload isolation, and multi-region topologies • Work with Kubernetes internals, including resource quota management, scheduling, cluster behaviour, NetworkPolicy enforcement, and custom controllers or operators • Implement and operate service mesh capabilities, including mTLS, service-account-level authentication and authorisation, and traffic management • Optimise containerised workloads for performance, cost, and resource efficiency • Write, refactor, and maintain production Go services, controllers, and middleware, with Python or Java also used • Read and contribute to complex existing codebases, including customised or extended open-source projects • Build HTTP, REST, and gRPC service interfaces for internal engineering teams • Write unit and integration tests • Lead incident response for platform-level issues, investigate root causes, and author postmortems • Troubleshoot production systems using logs, metrics, traces, and profiling tools • Diagnose and resolve performance and reliability problems across distributed systems • Define and drive SLOs and build actionable alerting • Own infrastructure as code, author reusable modules, and maintain them as the platform evolves • Build and improve CI/CD and GitOps delivery workflows • Balance developer velocity against reliability, security, and compliance requirements • Plan and execute cloud migration initiatives while maintaining reliability and minimising downtime • Build and maintain metrics, dashboards, alerting policies, and distributed tracing • Instrument services so failures are diagnosable without a code change • Partner with product, security, and infrastructure teams on requirements and architecture • Contribute to design reviews and help set technical direction • Mentor engineers and raise engineering practice standards

🎯 Requirements

• 6+ years of professional experience in software, platform, infrastructure, or site reliability engineering, including significant time operating production distributed systems • Demonstrated experience building and operating production Kubernetes platforms, not only deploying onto them • Production experience writing Go, Python or Java • Experience designing systems from an ambiguous starting point and carrying them to production • Experience migrating cloud services, including planning and executing moves of production workloads between providers or environments • Degree in Computer Science, Engineering, or a related field, or equivalent practical experience • Strong understanding of Kubernetes internals: networking (CNI), NetworkPolicy, resource management, and cluster behaviour under load • Hands-on experience with a service mesh (Istio, Envoy, Linkerd, or similar) and with mTLS and workload identity • Solid Linux fundamentals, including cgroups and resource management • Infrastructure as code at scale, Terraform or an equivalent • Production experience with at least one major cloud platform (GCP, AWS, or Azure); multi-cloud experience is a strong plus • Observability tooling: Prometheus, Grafana, OpenTelemetry, and query languages such as PromQL • Production experience with relational databases, including PostgreSQL or managed Postgres-compatible services, and an understanding of their replication and failover characteristics • Docker and container tooling as part of the delivery lifecycle • Strong debugging and performance profiling skills • Preferred: Experience with Go testing frameworks such as Ginkgo and Gomega • Preferred: Experience building Kubernetes controllers, operators, or other API-server extensions • Preferred: Additional strength in Python • Preferred: Experience with identity and access management: SSO, Keycloak, OIDC, SAML, or secret management with Vault or a cloud equivalent • Preferred: Experience designing for high availability and disaster recovery across regions or providers • Preferred: Experience working in a monorepo, and with build systems such as Bazel • Preferred: Ability to read Java • Preferred: Exposure to compliance frameworks such as SOC 2 or GDPR • Preferred: Interest in the reliability and safety of LLM-backed systems running in cloud-native environments • Preferred: Experience with Alibaba Cloud (AliCloud) • Preferred: Experience with large-scale Data Platform technologies such as Apache Spark and Apache Flink • Strong analytical and problem-solving ability • Clear written and verbal communication, particularly in design documents, postmortems, and cross-team requirements gathering • Able to work independently while contributing effectively within a distributed team • Comfortable in a fast-moving, highly technical environment

Apply Now

Similar Jobs

🕒 August 2

opinov8

201 - 500

💼 Consulting

🏥 Healthcare

📦 Logistics

Azure Cloud Architect responsible for designing scalable cloud-native solutions at Opinov8. Collaborating with cross-functional teams to ensure security and compliance in an Agile environment.

🏈 North America – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

☁️ Cloud Engineer

🕒 April 1

Talentxfactor.com

11 - 50

💼 Consulting

📣 Marketing

📦 Logistics

Community Manager for cloud-native services in a tech firm. Engaging users and improving community participation in Open Policy Agent project.

🏈 North America – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

☁️ Cloud Engineer

🕒 April 1

Canonical

501 - 1000

🤖 Artificial Intelligence

Cloud Engineering Manager at Canonical leading teams for private cloud infrastructure with open source software. Managing a global engineering team while driving project success and customer satisfaction.

🏈 North America – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

☁️ Cloud Engineer

🕒 December 24, 2025

HyTechPro

201 - 500

💼 Consulting

📦 Logistics

📣 Marketing

Senior D365/CRM Developer involved in building CRM solutions for clients across North America and Europe. Collaborating with teams and contributing to product development initiatives.

🏈 North America – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

☁️ Cloud Engineer

SQL

SSIS