
10,000+ employees
Founded 1996
đź Consulting
đŁ Marketing
đŚ Logistics
Consulting ⢠Marketing ⢠Logistics
Experian is a global leader in digital experience, technology, and transformation. They partner with recognized brands to enhance customer understanding, innovate product strategies, and implement agile technology solutions. With a focus on delivering superior customer experiences through AI, cloud architecture, and project management, Experian helps businesses streamline their operations and achieve their objectives effectively.
đĽ 42 minutes ago
đ¤ Texas â Remote
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đŚ H1B Visa Sponsor
đť Ghost score 10%
đŁď¸đ§đˇđľđš Portuguese Required
Improve your chances of getting an interview by checking your resume score before you apply.

10,000+ employees
Founded 1996
đź Consulting
đŁ Marketing
đŚ Logistics
Consulting ⢠Marketing ⢠Logistics
Experian is a global leader in digital experience, technology, and transformation. They partner with recognized brands to enhance customer understanding, innovate product strategies, and implement agile technology solutions. With a focus on delivering superior customer experiences through AI, cloud architecture, and project management, Experian helps businesses streamline their operations and achieve their objectives effectively.
⢠Design and operate highly available, scalable, and secure cloud platforms on AWS ⢠Build and maintain Kubernetes-based infrastructure to support applications, data, and AI workloads ⢠Improve platform reliability through automation, Infrastructure as Code (IaC), and self-service capabilities ⢠Implement and enhance observability solutions using Datadog, including monitoring, logging, tracing, alerting, dashboards, and SLO management ⢠Support and optimize large-scale data processing environments using Airflow, Amazon EMR, S3, and other AWS data services ⢠Work with Data and AI teams to improve the reliability, scalability, and operational maturity of Machine Learning and Artificial Intelligence platforms ⢠Lead incident response activities, root cause analyses, and post-incident reviews ⢠Define and measure SLIs, SLOs, and error budgets ⢠Improve deployment processes, CI/CD pipelines, and release reliability ⢠Optimize cloud infrastructure utilization, performance, and costs ⢠Mentor team members and promote SRE best practices across the engineering organization ⢠Increase platform availability and reliability ⢠Improve observability and reduce incident resolution time ⢠Increase automation and reduce repetitive manual operational effort ⢠Deliver reliable, scalable, and cost-efficient Data and AI platforms ⢠Collaborate with Engineering teams to deliver resilient production systems
⢠Currently pursuing or have completed a bachelor's degree ⢠Solid experience in Site Reliability Engineering, Platform Engineering, Cloud Engineering, or DevOps roles ⢠Strong hands-on experience with AWS services and cloud-native architectures ⢠Deep knowledge of Kubernetes and containerized workloads in production environments ⢠Experience managing and troubleshooting large-scale distributed systems ⢠Solid experience with observability platforms, preferably Datadog ⢠Experience supporting data platforms and pipelines using technologies such as Airflow, EMR, Spark, and S3 ⢠Expertise in Infrastructure as Code (IaC) using Terraform or similar tools ⢠Experience building and maintaining CI/CD pipelines and platform automation ⢠Strong knowledge of Linux, networking, and system performance troubleshooting ⢠Proficiency in scripting and automation using Python, Bash, or similar languages ⢠Experience supporting large-scale cloud-native platforms in AWS environments ⢠Experience operating Kubernetes platforms and managing cluster lifecycles ⢠Knowledge of Site Reliability Engineering principles, including SLOs, SLIs, error budgets, and operational excellence practices ⢠Experience implementing observability solutions using tools such as Datadog, Prometheus, Grafana, OpenTelemetry, or similar technologies ⢠Familiarity with data processing and workflow orchestration platforms such as Airflow, Spark, or EMR ⢠Experience with Infrastructure as Code (IaC) and platform automation practices ⢠AWS, Kubernetes, Terraform, or Datadog certifications ⢠Experience working in large-scale, highly available, or mission-critical enterprise environments ⢠Intermediate technical English
⢠Remote work ⢠Full-time Employee Status: Regular ⢠Affirmative action position for women ⢠Development opportunities related to gender equity and the Women in Experian group
Apply NowđĽ 50 minutes ago
Senior network SRE owning EVPN/BGP infrastructure and automation for Bitdeer's global AI GPU cloud. Connecting US, APAC, and Iceland data centers.
đşđ¸ United States â Remote
đľ $180k - $320k / year
đ° Post-IPO Equity on 2023-05
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đĽ 50 minutes ago
Senior Kubernetes SRE operating GPU cloud control planes for Bitdeerâs AI and Bitcoin mining infrastructure. Automating scheduling, multi-tenant isolation, BMaaS, and AIOps remediation.
đşđ¸ United States â Remote
đľ $180k - $260k / year
đ° Post-IPO Equity on 2023-05
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đĽ 50 minutes ago
Senior storage SRE operating high-performance systems for Bitdeer's AI GPU cloud. Designing resilient storage, telemetry, and automation for large-scale training clusters.
đşđ¸ United States â Remote
đľ $180k - $320k / year
đ° Post-IPO Equity on 2023-05
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
Linux
NFS
đĽ 3 hours ago
DevOps Engineer 4 building secure CI/CD and cloud platforms for Centex Technologiesâ programs and customers. Automating infrastructure, observability, security, and reliable software delivery.
đĽ 3 hours ago
DevSecOps Engineer securing Paylocityâs cloud-based HR and payroll software platform. Developing security tooling, integrating build protections, and guiding vulnerability remediation across web and mobile applications.
đşđ¸ United States â Remote
đľ $96k - $130k / year
đ° $10M Venture Round - Paylocity on 2008-05
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)