
11 - 50 employees
Founded 2024
đź Consulting
đŁ Marketing
đŚ Logistics
Consulting ⢠Marketing ⢠Logistics
Salve. Inno is a recruitment and consulting firm that connects exceptional talent with businesses through personalized hiring strategies and global remote sourcing. The company specializes in recruitment for roles across sectors such as marketing, forex, and iGaming, offering candidate sourcing, screening, and career-site driven hiring experiences while emphasizing DE&I, communication, and innovative process building. Founded in 2024 and headquartered in GdaĹsk, Poland, Salve. Inno operates with a small team and a global footprint via remote job listings and consulting services.
đĽ 0 minutes ago
đľđ Philippines â Remote
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đť Ghost score 25%
Improve your chances of getting an interview by checking your resume score before you apply.

11 - 50 employees
Founded 2024
đź Consulting
đŁ Marketing
đŚ Logistics
Consulting ⢠Marketing ⢠Logistics
Salve. Inno is a recruitment and consulting firm that connects exceptional talent with businesses through personalized hiring strategies and global remote sourcing. The company specializes in recruitment for roles across sectors such as marketing, forex, and iGaming, offering candidate sourcing, screening, and career-site driven hiring experiences while emphasizing DE&I, communication, and innovative process building. Founded in 2024 and headquartered in GdaĹsk, Poland, Salve. Inno operates with a small team and a global footprint via remote job listings and consulting services.
⢠Own the reliability, availability and operational health of production services running on AWS and Amazon EKS ⢠Operate and troubleshoot Kubernetes clusters in production, including cluster lifecycle, upgrades, node management, networking, scaling, capacity and workload reliability ⢠Participate actively in on-call and pager rotations and take ownership of production incidents ⢠Lead or play a key technical role during P1/P2 and Sev1/Sev2 incidents, including diagnosis, mitigation, recovery and communication ⢠Coordinate technical incident bridges and communicate directly with customers during production escalations ⢠Perform root cause analysis and lead blameless postmortems, ensuring incidents result in concrete engineering improvements ⢠Define, monitor and improve SLIs, SLOs and error budgets for production services ⢠Develop and maintain actionable alerts, operational runbooks and automated remediation ⢠Build and improve infrastructure using Terraform/Terragrunt and Infrastructure as Code practices ⢠Operate GitOps-based delivery environments using Argo CD or FluxCD ⢠Improve Kubernetes scaling and efficiency using Karpenter, KEDA and native Kubernetes autoscaling ⢠Build and improve observability using Prometheus, Grafana, OpenTelemetry, Datadog and/or ELK ⢠Support highly available distributed and event-driven systems, including Kafka/MSK environments ⢠Design, implement and validate disaster recovery and business continuity mechanisms against measurable RTO and RPO objectives ⢠Identify recurring operational problems and eliminate toil through automation and engineering ⢠Improve AWS performance, scalability, security and cost efficiency across production environments ⢠Work closely with software, platform and engineering teams to build reliability into systems throughout the development lifecycle ⢠Contribute to continuous improvement of incident management, operational readiness and SRE engineering practices
⢠Significant professional experience as a hands-on Site Reliability Engineer, Production Engineer or senior Platform Engineer with direct production ownership ⢠Several years of recent, hands-on experience operating production environments on AWS ⢠Strong, demonstrable experience operating Amazon EKS in production ⢠Deep Kubernetes operational knowledge, including cluster administration, upgrades, nodes, autoscaling, networking, troubleshooting and production failure scenarios ⢠Proven participation in a production on-call/pager rotation ⢠Demonstrable ownership of significant production incidents, including troubleshooting, mitigation, recovery, RCA and post-incident improvements ⢠Practical experience with SLIs, SLOs, error budgets, alerting and runbooks ⢠Strong Infrastructure as Code experience with Terraform and/or Terragrunt ⢠Production experience with Kubernetes delivery and GitOps practices; Argo CD or FluxCD strongly preferred ⢠Strong production observability experience with Prometheus, Grafana, OpenTelemetry, Datadog or ELK ⢠Experience operating highly available, distributed production systems ⢠Strong understanding of AWS networking, IAM, security, availability and resilience ⢠Experience implementing and testing disaster recovery strategies with measurable RTO/RPO objectives ⢠Proven external customer-facing technical experience, including technical discussions, production escalations, architecture/reliability conversations or incident communication ⢠Ability to explain complex technical problems clearly and make sound decisions during high-pressure production incidents ⢠Strong troubleshooting mindset and ability to work independently during complex production failures ⢠Strong professional English communication skills (minimum C1) for regular interaction with clients ⢠A coherent track record demonstrating sustained hands-on production engineering ownership
⢠Full-time permanent B2B cooperation ⢠Fully remote working environment ⢠Senior hands-on engineering position with meaningful ownership of business-critical production systems ⢠Opportunity to work on complex AWS, Kubernetes and distributed-system environments at scale ⢠Direct influence over reliability engineering, operational practices and platform improvements ⢠Modern engineering environment with strong emphasis on automation, observability and continuous improvement ⢠Collaboration with experienced engineering, platform and product teams ⢠Opportunity to introduce and use modern approaches, including AI-assisted engineering and operational automation ⢠Long-term opportunity for engineers who want to remain deeply technical and close to production ⢠Inclusive, respectful workplace regardless of gender, ethnicity, or background
Apply NowđĽ 7 hours ago
Power Platform engineer supporting, testing, and deploying Microsoft business solutions for A-TTC. Troubleshooting Power Apps, Automate, Pages, Dataverse, and release environments.
đľđ Philippines â Remote
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)
Azure
SQL
đ August 17
DevOps Engineer building and supporting scalable AWS and Nutanix infrastructure for DysrupITâs technology consulting clients. Automating deployments, monitoring platforms, and resolving incidents across enterprise environments.
đľđ Philippines â Remote
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)
Ansible
AWS
Cloud
DNS
Docker
Kubernetes
Linux
Microservices
MongoDB
MySQL
Python
TCP/IP
Terraform
đ August 11
DevOps Engineer automating AWS infrastructure and CI/CD pipelines for Satellite Officeâs Philippines-based offshore teams. Improving platform security, reliability, scalability, and deployment efficiency.
đľđ Philippines â Remote
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)
Ansible
AWS
Azure
Cloud
Docker
EC2
Grafana
Jenkins
Kubernetes
Linux
Prometheus
Python
Terraform
đ July 27
Lead Azure DevOps Engineer implementing cloud solutions in Azure for a client-facing role. Delivering and supporting technical implementations with a focus on client satisfaction.
Ansible
AWS
Azure
Cloud
Google Cloud Platform
Groovy
Jenkins
Linux
Python
Terraform
TFS
đ July 27
DevOps Team Manager at LegalMatch overseeing cloud infrastructure and driving best practices in DevOps team. Fostering team performance and operational excellence.
đľđ Philippines â Remote
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)
AWS
Cloud
Docker
Kubernetes
Linux