Data Site Reliability Engineer, SRE

Job not on LinkedIn

🔥 0 minutes ago

⚜️ Louisiana – Remote

infoinfo

💵 $111.2k - $150.4k / year

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 8%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of LouisianaNOW.Jobs

LouisianaNOW.Jobs

51 - 200 employees

🎯 Recruiter

🏪 Marketplace

Recruitment • Marketplace

LouisianaNOW. Jobs is a regional jobs portal and talent attraction site focused on connecting job seekers to employers and economic opportunities across Louisiana. The site highlights regional communities, industry sectors, current employers who are hiring, career advice articles, and job listings, and appears to be operated or supported by LED FastStart (Louisiana Economic Development). It’s positioned as a marketplace that promotes Louisiana’s growing industries and helps match talent with local openings and employer investment announcements.

📋 Description

• Provide technical leadership for day-to-day operational support, reliability, performance, and continuous improvement of CMM data platforms, pipelines, applications, and analytics services • Provide real-time monitoring, incident and event management, capacity planning, and operational reporting • Maintain and audit cloud user roles and responsibilities • Integrate SSO, MFA, and group identity management through JENIE to enforce least-privilege access • Assess and improve credential management processes • Provide disaster recovery, continuity of operations, high availability, fault tolerance, and automated failover designs • Integrate DevSecOps tools and processes with enterprise systems • Manage centralized secrets management with automated rotation, access logging, and policy enforcement • Integrate SAST, DAST, SCA, and CSPM security tools into pipelines • Implement continuous 24/7/365 monitoring for security, performance, and compliance • Provide supplemental monitoring of event response activities beyond normal business hours • Automate Software Bill of Materials generation and management • Provide diagnostics, metrics gathering, and performance tuning • Provide canary releases for end-user beta testing • Configure alerts for unusual behavior • Operate incident and event management integrated with enterprise SIEM solutions, automated alerting, escalation workflows, and root cause analysis • Detect, log, diagnose, escalate, and resolve incidents • Identify and eliminate recurring incident root causes • Recommend improvements to incident and problem management • Maintain a knowledge base of known issues, resolutions, and operational best practices • Perform automated full-stack health checks across operating systems, applications, databases, and PaaS services • Produce monthly issues management reports covering incidents, stability, performance, configuration issues, ticket volumes, and resolution times • Develop thresholds, rules, and response procedures • Monitor cloud resource utilization and manage threshold-breach resolution procedures • Improve reliability, observability, automation, scalability, and operational resilience • Monitor, maintain, and optimize cloud infrastructure, databases, and platform services • Act as FinOps Analyst and perform cost optimization

🎯 Requirements

• Bachelor's degree in Computer Science, Software Engineering, or related field, or equivalent experience • 5+ years’ experience in IT systems engineering, systems development, systems coding, and programming • Deep expertise with AWS services, including monitoring, logging, compute, storage, and networking • Proficiency in Infrastructure as Code tools such as Terraform, AWS CloudFormation, or Azure Bicep • Hands-on experience with monitoring and APM tools such as CloudWatch, Azure Monitor, Datadog, Prometheus, Grafana, or New Relic • Understanding of incident response, change management, and ITIL-based operational support • Familiarity with CI/CD toolchains and automation platforms including Jenkins, GitHub Actions, GitLab, or ArgoCD • Strong scripting skills in Python, PowerShell, or Bash • Advanced experience implementing DevSecOps using GitOps or similar tools • Experience developing, testing, and maintaining containerized applications • Expert knowledge of source version control, build/release tools, CI/CD pipelines, and software build processes • Experience building and maintaining CI/CD pipelines for large enterprises with complex applications • Experience with FinOps practices, cost modeling, forecasting, and cloud optimization tools • Understanding of federal compliance and security frameworks such as FedRAMP, NIST, and JISF Rev 5 • Ability to analyze logs and metrics and conduct performance tuning for cloud-based services and applications • Experience working across multiple product teams to assess overall product/program health • Must be able to pass a background check to obtain a position of Public Trust • Must be a US Person: Green Card Holder, US Permanent Resident Alien, Refugee, Asylee, or US Citizen • Excellent presentation and communication skills • Consultant mindset and ability to work with high-level customer stakeholders • Strong analytical and problem-solving skills • Experience with process design and documentation methodologies, quality deliverables, process and use-case modeling, and business-case development • Ability to work effectively independently and as part of a team

🏖️ Benefits

• Medical plan options, including plans with Health Savings Accounts • Dental plan options • Vision plan options • 401(k) plan with company match • Full-flex work weeks where possible • Paid vacation, sick, personal time, and holidays • Paid parental, military, bereavement, and jury duty leave • Typically 15 days of paid leave per calendar year • 10 paid holidays per year • Up to 160 hours of paid family leave in a rolling 12-month period for eligible employees • Short- and long-term disability benefits • Life insurance • Accidental death and dismemberment insurance • Personal accident insurance • Critical illness insurance • Business travel and accident insurance • AI-powered career tool identifying career steps and learning opportunities • Internal mobility team supporting career goals • Wellness packages • Competitive pay • Award-winning culture of innovation • Military-friendly workplace

Apply Now

Similar Jobs

🔥 29 minutes ago

General Dynamics Information Technology

10,000+ employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Senior DevOps Engineer automating Oracle-based healthcare deployments for GDIT’s CMS program. Advancing secure DevSecOps pipelines, release orchestration, and deployment governance.

🔥 4 hours ago

NVIDIA

10,000+ employees

🏥 Healthcare

🏭 Manufacturing

🤖 Artificial Intelligence

AI Tools Engineer building LLM-powered tools for NVIDIA’s global GeForce NOW service. Automating incident analysis and operational intelligence through AI/ML, data pipelines, Kubernetes, and AWS.

🔥 9 hours ago

Koniag Government Services

1001 - 5000

🏛️ Government

🎖️ Defense

💼 Consulting

Senior DevOps Engineer securing AWS/Azure cloud infrastructure for Koniag Government Services. Automating DevSecOps, CI/CD security, compliance, and incident response for federal customers.

🔥 11 hours ago

Guidehouse

10,000+ employees

🏥 Healthcare

🎖️ Defense

📦 Logistics

Senior DevOps Engineer automating cloud infrastructure, Kubernetes deployments, and CI/CD for Guidehouse government applications. Supporting secure, reliable delivery across development, QA, and operations.

🔥 14 hours ago

Virta Health

201 - 500

🏥 Healthcare

⚕️ Healthcare Insurance

🧘 Wellness

DevSecOps Engineer securing Virta Health’s cloud-native healthcare platform. Automating application security, IAM, vulnerability management, and compliance across GCP and Kubernetes.