Senior DevOps Engineer – Kubernetes

🕒 May 13

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of FICO

FICO

1001 - 5000 employees

Founded 1956

💸 Finance

🤖 Artificial Intelligence

☁️ SaaS

Finance • Artificial Intelligence • SaaS

FICO is a leading analytics and software company renowned for its FICO® Score, a tool widely used by lenders to assess credit risk. The company offers a comprehensive platform that leverages data, AI, and machine learning to power intelligent decision-making and customer engagement across various industries. FICO's solutions span fraud detection, credit scoring, and customer lifecycle management, making it vital to sectors such as finance and telecommunications. Its innovative products help businesses optimize outcomes through real-time analytics, business composability, and scenario management.

📋 Description

• Partner with product managers to translate platform priorities into scalable, forward-looking infrastructure strategies. • Architect and own CI/CD pipelines using ArgoCD, Tekton, or similar tools, ensuring high availability across environments. • Design and implement enterprise-grade observability solutions including monitoring, logging, and distributed tracing for Kubernetes-based applications. • Collaborate with architects and engineers to optimize deployments, performance, and scalability on Kubernetes. • Drive continuous improvement of engineering practices, establishing standards that elevate efficiency, reliability, and operational excellence. • Architect and automate infrastructure provisioning and configuration management leveraging AWS services. • Own security and compliance posture across cloud infrastructure, defining and enforcing policies organization-wide.

🎯 Requirements

• Extensive experience architecting and operating data-intensive, containerized applications on Kubernetes at scale, including debugging and troubleshooting complex production issues. • Advanced proficiency across AWS services including EKS, EC2, S3, IAM, Route 53, and ECR. • Deep expertise in observability using Prometheus and Grafana, with a strong grasp of security best practices across cloud and Kubernetes environments. • Strong scripting and automation capabilities (e.g., Python, Bash, GitHub Workflow). • Experience with infrastructure as code tools (Terraform, CrossPlane, AWS ACK). • Demonstrated expertise with Helm and ArgoCD, designing and managing complex CD pipelines across multi-environment setups.

🏖️ Benefits

• Highly competitive compensation, benefits and rewards programs that encourage you to bring your best every day and be recognized for doing so. • An inclusive culture strongly reflecting our core values: Act Like an Owner, Delight Our Customers and Earn the Respect of Others. • The opportunity to make an impact and develop professionally by leveraging your unique strengths and participating in valuable learning experiences. • An engaging, people-first work environment offering work/life balance, employee resource groups, and social events to promote interaction and camaraderie.

Apply Now

Similar Jobs

🕒 May 11

PlayOn! Sports

201 - 500

📱 Media

⚽ Sports

📚 Education

Senior Site Reliability Engineer focused on building tools and automation for system reliability at PlayOn. Collaborating with DevOps and engineering teams to enhance performance and scalability.

AWS

Azure

Cloud

Distributed Systems

Docker

Google Cloud Platform

Grafana

Java

Kubernetes

Linux

Prometheus

Python

Terraform

Go

🕒 May 9

Visionary Integration Professionals (VIP)

501 - 1000

🤝 B2B

🏛️ Government

Forward Deployment Engineer working on AI-enabled solutions for clients at Visionary Integration Professionals. Collaborating with customers to design, prototype, and support implementations across various sectors.

🇺🇸 United States – Remote

💵 $130k - $165k / year

💰 Debt Financing on 2018-11

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

AWS

Azure

ERP

JavaScript

Python

TypeScript

🕒 May 8

DroneDeploy

201 - 500

🚀 Aerospace

Senior DevOps Engineer optimizing DroneDeploy's cloud infrastructure and CI/CD pipelines. Collaborating with teams to enhance reliability and integrate AI tooling into workflows.

AWS

Azure

Cloud

Google Cloud Platform

Jenkins

Kubernetes

Linux

Python

Terraform

Go

🕒 May 8

Flock Safety

501 - 1000

🔐 Security

Site Reliability Engineer designing and building systems to enhance observability and scalability for aviation solutions at Flock. Empowering developers through optimized application stack management.

AWS

Grafana

Prometheus

Terraform

🕒 May 8

Zingtree

11 - 50

🤝 B2B

☁️ SaaS

🤖 Artificial Intelligence

Senior DevOps / Platform Reliability Engineer managing CI/CD, infrastructure for AI-driven platform. Collaborating across teams to automate processes and ensure reliability.

Ansible

AWS

Cloud

DNS

Grafana

Jenkins

Kafka

Kubernetes

Linux

Microservices

MySQL

Prometheus

Python

Redis

Terraform