Staff DevOps Engineer

Job not on LinkedIn

🔥 19 hours ago

🇨🇴 Colombia – Remote

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 14%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Abstra

Abstra

51 - 200 employees

Founded 2007

💼 Consulting

📦 Logistics

📣 Marketing

💰 $2.3M Seed Round - Abstra on 2022-01

Consulting • Logistics • Marketing

Abstra is a nearshore LATAM technology services company that provides staff augmentation, dedicated teams, software outsourcing, and managed IT services to U. S. and international clients. They deliver custom software development (web and mobile), UX/UI design, QA/testing, and platform & infrastructure work (DevOps/SRE), and offer Data & AI capabilities including data engineering, ML engineering, MLOps, and data science. Abstra emphasizes time-zone alignment, bilingual LATAM talent, cost-effective nearshore delivery, and long-term partnership models to help companies scale engineering capacity and accelerate roadmaps.

📋 Description

• Define repeatable service-delivery practices through modular, reusable automation and a developer platform enabling self-service service delivery • Participate in governance controls that reduce risk and promote organizational standardization • Refine enablement practices including design reviews, service launch coordination, production readiness assessments, service-level objective definition and review, incident management, and cost awareness • Support delivery teams throughout their service lifecycle and maturity • Collaborate with delivery teams to define product-specific metrics and remediations • Perform system analysis, testing, and fault troubleshooting • Respond to and triage production issue escalations during off-hours through an on-call rotation

🎯 Requirements

• 8+ years of engineering experience running high-availability systems and supporting infrastructure in customer-facing production environments • Proficiency in a high-level programming language, such as Python or Go • Proficiency in Bash scripting and Linux • Proficiency in modern technical operating practices • Experience in system architecture and design • Experience with continuous integration and continuous delivery using Jenkins, FluxCD, and GitHub Actions • Experience with Infrastructure as Code using Terraform, including Terraform modules • Experience with AWS cloud services and Kubernetes • Knowledge of SRE principles and practices • Experience with Kubernetes cluster concepts and design • Experience improving service observability through monitoring agents, metrics, logging, and dashboards • Knowledge of OpenTelemetry and Prometheus • Knowledge of observability platforms such as Datadog, Splunk, Dynatrace, or Observe • Comfort participating in an on-call rotation to respond to and triage production issue escalations during off-hours • Experience with AWS services and capabilities including ECS, EKS, ECR, EC2, S3, RDS, VPCs, IAM policy documents, policies, roles, instance profiles, and CloudWatch Logs • Experience with Docker containers and container orchestration using ECS and EKS • Successful completion of employment verification and background checks • Verification of job titles and employment dates with two previous employers

🏖️ Benefits

• Competitive compensation paid in USD • 20 days of paid time off (PTO) per year • Opportunities for professional growth and career development • Company-provided equipment • A collaborative, inclusive, and multicultural work environment • The opportunity to contribute to meaningful projects alongside a talented and supportive team

Apply Now