Search Remote Jobs

Mid-Level SRE Analyst

Job not on LinkedIn

🔥 42 minutes ago

🤠 Texas – Remote

infoinfo

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

infoinfo

👻 Ghost score 10%

infoinfo

🗣️🇧🇷🇵🇹 Portuguese Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Experian

Experian

10,000+ employees

Founded 1996

💼 Consulting

📣 Marketing

📦 Logistics

Consulting • Marketing • Logistics

Experian is a global leader in digital experience, technology, and transformation. They partner with recognized brands to enhance customer understanding, innovate product strategies, and implement agile technology solutions. With a focus on delivering superior customer experiences through AI, cloud architecture, and project management, Experian helps businesses streamline their operations and achieve their objectives effectively.

📋 Description

• Design and operate highly available, scalable, and secure cloud platforms on AWS • Build and maintain Kubernetes-based infrastructure to support applications, data, and AI workloads • Improve platform reliability through automation, Infrastructure as Code (IaC), and self-service capabilities • Implement and enhance observability solutions using Datadog, including monitoring, logging, tracing, alerting, dashboards, and SLO management • Support and optimize large-scale data processing environments using Airflow, Amazon EMR, S3, and other AWS data services • Work with Data and AI teams to improve the reliability, scalability, and operational maturity of Machine Learning and Artificial Intelligence platforms • Lead incident response activities, root cause analyses, and post-incident reviews • Define and measure SLIs, SLOs, and error budgets • Improve deployment processes, CI/CD pipelines, and release reliability • Optimize cloud infrastructure utilization, performance, and costs • Mentor team members and promote SRE best practices across the engineering organization • Increase platform availability and reliability • Improve observability and reduce incident resolution time • Increase automation and reduce repetitive manual operational effort • Deliver reliable, scalable, and cost-efficient Data and AI platforms • Collaborate with Engineering teams to deliver resilient production systems

🎯 Requirements

• Currently pursuing or have completed a bachelor's degree • Solid experience in Site Reliability Engineering, Platform Engineering, Cloud Engineering, or DevOps roles • Strong hands-on experience with AWS services and cloud-native architectures • Deep knowledge of Kubernetes and containerized workloads in production environments • Experience managing and troubleshooting large-scale distributed systems • Solid experience with observability platforms, preferably Datadog • Experience supporting data platforms and pipelines using technologies such as Airflow, EMR, Spark, and S3 • Expertise in Infrastructure as Code (IaC) using Terraform or similar tools • Experience building and maintaining CI/CD pipelines and platform automation • Strong knowledge of Linux, networking, and system performance troubleshooting • Proficiency in scripting and automation using Python, Bash, or similar languages • Experience supporting large-scale cloud-native platforms in AWS environments • Experience operating Kubernetes platforms and managing cluster lifecycles • Knowledge of Site Reliability Engineering principles, including SLOs, SLIs, error budgets, and operational excellence practices • Experience implementing observability solutions using tools such as Datadog, Prometheus, Grafana, OpenTelemetry, or similar technologies • Familiarity with data processing and workflow orchestration platforms such as Airflow, Spark, or EMR • Experience with Infrastructure as Code (IaC) and platform automation practices • AWS, Kubernetes, Terraform, or Datadog certifications • Experience working in large-scale, highly available, or mission-critical enterprise environments • Intermediate technical English

🏖️ Benefits

• Remote work • Full-time Employee Status: Regular • Affirmative action position for women • Development opportunities related to gender equity and the Women in Experian group

Apply Now

Similar Jobs

🔥 50 minutes ago

Bitdeer Group

201 - 500

💼 Consulting

📦 Logistics

🏗️ Construction

Senior network SRE owning EVPN/BGP infrastructure and automation for Bitdeer's global AI GPU cloud. Connecting US, APAC, and Iceland data centers.

🔥 50 minutes ago

Bitdeer Group

201 - 500

💼 Consulting

📦 Logistics

🏗️ Construction

Senior Kubernetes SRE operating GPU cloud control planes for Bitdeer’s AI and Bitcoin mining infrastructure. Automating scheduling, multi-tenant isolation, BMaaS, and AIOps remediation.

🔥 50 minutes ago

Bitdeer Group

201 - 500

💼 Consulting

📦 Logistics

🏗️ Construction

Senior storage SRE operating high-performance systems for Bitdeer's AI GPU cloud. Designing resilient storage, telemetry, and automation for large-scale training clusters.

🔥 3 hours ago

Centex Technologies

51 - 200

💼 Consulting

📦 Logistics

📣 Marketing

DevOps Engineer 4 building secure CI/CD and cloud platforms for Centex Technologies’ programs and customers. Automating infrastructure, observability, security, and reliable software delivery.

🔥 3 hours ago

Paylocity

5001 - 10000

👥 HR Tech

☁️ SaaS

🤝 B2B

DevSecOps Engineer securing Paylocity’s cloud-based HR and payroll software platform. Developing security tooling, integrating build protections, and guiding vulnerability remediation across web and mobile applications.

🇺🇸 United States – Remote

💵 $96k - $130k / year

💰 $10M Venture Round - Paylocity on 2008-05

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)