Senior Site Reliability Operations Engineer – Finance

Likely ghost job

🕒 July 28

💃 Latin America – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 67%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Truelogic Software

Truelogic Software

501 - 1000 employees

Founded 2004

☁️ SaaS

🤝 B2B

🏢 Enterprise

SaaS • B2B • Enterprise

Truelogic Software is a nearshore software development company specializing in agile staff augmentation services. They focus on providing custom outsourced software development with a team of highly skilled engineers from Latin America. Truelogic Software partners with both startups and Fortune 500 companies, offering solutions that align with their clients' time zones and ensuring high-quality outcomes through collaboration and responsiveness. With a presence in over 25 countries, Truelogic emphasizes remote work for better quality of life, and their engineers are experienced in various industries, delivering a wide range of successful projects globally.

📋 Description

• Ensure 24/7 stability of internal IT infrastructure and mission-critical backend systems • Monitor multi-platform IT infrastructure using AWS CloudWatch, New Relic, Nagios, and SumoLogic • Refine alert thresholds to minimize noise and enable proactive remediation • Serve as an escalation point for complex technical issues • Troubleshoot Linux/UNIX, Windows, virtual servers, and virtual desktop environments • Coordinate, automate, and execute code deployments using Jenkins, GitLab, or similar CI/CD tools • Drive non-disruptive releases and zero-downtime updates • Collaborate with Application Developers, third-party vendors, and internal Incident Management • Lead medium- to large-scale infrastructure projects, including migrations, cloud upgrades, and performance tuning • Maintain SOPs in team knowledge bases • Oversee enterprise backup operations using CommVault, Veeam, and AWS Backup • Work an Operations swing shift from 2 PM to 10:30 PM PST

🎯 Requirements

• 5+ years in an Operations Center (SRO/NOC) or cloud infrastructure environment • Hands-on experience in full-stack application deployments • Proficiency in Windows and UNIX/Linux administration • Experience with scripting, grepping logs, and analyzing performance metrics • Virtual server/desktop management experience • Practical experience with AWS cloud services, including Storage, VMs, and Networking • Experience with AWS CloudWatch, New Relic, Nagios, and SumoLogic • Scripting or programming capability in PowerShell, Python, or bash • Experience with Jenkins and GitLab CI/CD platforms • Experience with ServiceNow and Jira ITSM platforms • Experience with CommVault, Veeam, and AWS Backup • Good verbal and written communication skills • Advanced AWS Certifications are a plus • AI/ML infrastructure monitoring and predictive analytics exposure is a plus • ITIL-aligned or enterprise Change/Incident Management experience is a plus • Bachelor's Degree in Computer Science, Information Technology, or a related field is a plus • Must be currently based in Latin America

🏖️ Benefits

• 100% Remote Work • Highly Competitive USD Pay • Paid Time Off • Work with Autonomy • Work with Top American Companies • Engagement activities • Work-life balance support • Collaboration with a diverse, global network • Access to seasoned senior professionals

Apply Now

Similar Jobs

🕒 July 27

phData

201 - 500

💼 Consulting

🏥 Healthcare

🏭 Manufacturing

Join phData as a Senior DevOps Engineer for cloud-native data platforms. Deliver technical operations and platform reliability in a client-facing role across AWS and Azure.

AWS

Azure

Cloud

ETL

Linux

Python

SQL

Terraform

Unix

🕒 February 16

Agentero

11 - 50

🛡️ Insurance

☁️ SaaS

💳 Fintech

Site Reliability Engineer for a remote-first Latin American startup focused on innovative technology in insurance. Collaborate across teams and improve cloud infrastructure while owning incident response and monitoring solutions.

AWS

Cloud

Google Cloud Platform

Grafana

Linux

Prometheus

Python

Terraform

Go

🕒 February 5

Agentero

11 - 50

🛡️ Insurance

☁️ SaaS

💳 Fintech

Site Reliability Engineer working remote from Latin America to enhance reliability for a data-driven insurance technology platform. Collaborating with a distributed team aligned with US business hours.

AWS

Cloud

Google Cloud Platform

Grafana

Linux

Prometheus

Python

Terraform

Go