Senior DevOps Engineer, Infrastructure & Reliability

Job not on LinkedIn

🔥 0 minutes ago

🐊 Florida – Remote

infoinfo

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 12%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Worth AI

Worth AI

11 - 50 employees

💼 Consulting

🛡️ Insurance

🤖 Artificial Intelligence

Consulting • Insurance • Artificial Intelligence

Worth AI is a company that specializes in enhancing financial security and risk management through advanced AI-driven solutions. Their platform offers a suite of tools designed for seamless onboarding, compliance, and credit risk assessment. They help banks, fintech companies, and credit unions streamline operations by automating processes such as KYC/KYB compliance, reputation monitoring, and automated credit underwriting. Worth AI's technology provides real-time risk monitoring, predictive analytics, and AI-generated insights to improve decision-making, safeguard institutions, and boost revenue. Key offerings include Worth Score™, an AI-driven credit score solution, and unique AI underwriting systems that enhance the accuracy and efficiency of financial assessments.

📋 Description

• Implement scalable Infrastructure-as-Code patterns using tools like Terraform to standardize cloud provisioning and reduce configuration drift • Own and evolve the Kubernetes platform (EKS or self-managed), ensuring workloads are secure, scalable, and resilient by default • Optimize CI/CD pipelines to improve deployment frequency, reduce lead time, and increase confidence in releases • Design and enforce secure networking, IAM, and secrets management strategies across environments • Improve observability by refining metrics, logs, and tracing using tools like DataDog • Optimize cloud cost efficiency through rightsizing, autoscaling strategies, and architectural improvements • Implement disaster recovery planning, backup strategies, and multi-region resilience initiatives • Refactor brittle or manually managed infrastructure into automated, testable, and reproducible systems • Introduce new infrastructure tooling or architectural shifts and drive adoption through documentation, workshops, and hands-on support • Partner with engineering teams to eliminate friction in CI/CD, deployments, and cloud environments • Communicate technical trade-offs clearly across engineering and product stakeholders, balancing speed with safety

🎯 Requirements

• 8+ years in DevOps, SRE, or infrastructure engineering • Proven experience designing and operating production Kubernetes environments at scale • Deep hands-on expertise with AWS infrastructure and cloud networking • Strong experience building and maintaining Terraform modules across large cloud environments • Demonstrated ownership of CI/CD systems and measurable improvement of DORA metrics • Experience leading incident response processes and driving meaningful postmortem outcomes • Strong understanding of distributed systems, event-driven architectures (Kafka), and database performance (PostgreSQL) • Proven ability to modernize legacy infrastructure and eliminate manual operational toil • Track record of taking a scoped infrastructure project from an ambiguous starting point to production without needing daily direction • Demonstrated ability to build trust across teams while raising the reliability bar • Bonus Points (Nice to Have): Experience coding applications; experience operating high-throughput Kafka clusters (MSK or self-managed); strong background in database performance tuning (PostgreSQL, Redis); experience implementing autoscaling strategies for high-traffic systems; familiarity with service mesh technologies; experience building internal developer platforms (IDP); background in security best practices (zero-trust networking, policy-as-code); experience with multi-region or globally distributed systems; experience introducing platform-wide reliability frameworks (SLOs, error budgets, chaos testing)

🏖️ Benefits

• Health Care Plan (Medical, Dental & Vision) • Retirement Plan (401k) • Life Insurance • Flexible Paid Time Off • 9 paid Holidays • Family Leave • Remote • Hybrid work (for Orlando Associates) • Free Food & Snacks (Orlando) • Wellness Resources

Apply Now

Similar Jobs

🔥 1 hour ago

SAIC

10,000+ employees

☁️ SaaS

📣 Marketing

🏢 Enterprise

Senior DevSecOps Engineer securing SAIC’s defense, space, intelligence, and civilian technology infrastructure. Automating systems, monitoring, and GitLab delivery pipelines across Linux, Windows, AWS, and air-gapped environments.

🇺🇸 United States – Remote

🔥 Funding within the last year

💰 $500M Post-IPO Debt - SAIC on 2025-09

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🔥 4 hours ago

The Home Depot

10,000+ employees

🏗️ Construction

📦 Logistics

🛒 Retail

Senior Principal Reliability Engineer designing resilient infrastructure for Home Depot store systems, payments, and COM platforms. Guiding multiple engineering teams on reliability, cloud costs, and technology strategy.

🔥 10 hours ago

Symbotic

501 - 1000

🔧 Hardware

📦 Logistics

🤖 Artificial Intelligence

Senior reliability manager scaling maintenance and asset performance across Exol’s automated warehouses. Driving uptime, safety, launches, vendor governance, and enterprise reliability standards.

🔥 13 hours ago

PingWind Inc. (SDVOSB)

51 - 200

💼 Consulting

📦 Logistics

🏥 Healthcare

DevSecOps Engineer building and deploying secure cloud-based IAM systems for federal government clients. Maintaining highly available architectures, automated delivery, compliance, and infrastructure upgrades.

🔥 13 hours ago

Hamsa

11 - 50

🤝 B2B

☁️ SaaS

Lead DevOps Engineer managing secure, scalable AWS infrastructure for Hamsa’s global finance operating system. Leading international DevOps delivery, CI/CD, Kubernetes, monitoring, and reliability practices.