
51 - 200 employees
Founded 2010
💼 Consulting
🏥 Healthcare
🛡️ Insurance
Consulting • Healthcare • Insurance
<Verity Group> is a Brazil-based digital transformation and innovation consultancy that develops and accelerates products and services for enterprises using modern engineering, cloud, and AI-driven approaches. The company offers application modernization, digital experience design, technology outsourcing, and consulting services, and operates an AI orchestration platform called Verity Quantum to integrate and deploy AI agents. Verity Group serves B2B clients in banking, insurance, healthcare and other industries, has 200+ consultants and 1,000+ projects over 15+ years, and focuses on end-to-end strategy-to-execution delivery to generate measurable business value.
🔥 13 hours ago
🇧🇷 Brazil – Remote
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
👻 Ghost score 10%
🗣️🇧🇷🇵🇹 Portuguese Required
Ansible
AWS
Azure
Cloud
Docker
ElasticSearch
Google Cloud Platform
Grafana
Kubernetes
Linux
Prometheus
Terraform
Improve your chances of getting an interview by checking your resume score before you apply.

51 - 200 employees
Founded 2010
💼 Consulting
🏥 Healthcare
🛡️ Insurance
Consulting • Healthcare • Insurance
<Verity Group> is a Brazil-based digital transformation and innovation consultancy that develops and accelerates products and services for enterprises using modern engineering, cloud, and AI-driven approaches. The company offers application modernization, digital experience design, technology outsourcing, and consulting services, and operates an AI orchestration platform called Verity Quantum to integrate and deploy AI agents. Verity Group serves B2B clients in banking, insurance, healthcare and other industries, has 200+ consultants and 1,000+ projects over 15+ years, and focuses on end-to-end strategy-to-execution delivery to generate measurable business value.
• Define and track SLIs, SLOs, SLAs, MTTR, and MTTD • Implement observability, monitoring, alerting, and APM • Monitor latency, traffic, errors, saturation, availability, and performance • Work on incident prevention, identification, and resolution • Lead root cause analyses and define actions to prevent recurrence • Identify risks, bottlenecks, and single points of failure • Support the design of resilient, scalable, and highly available solutions • Automate operational activities and reduce manual tasks • Operate and evolve Kubernetes and Docker environments • Support capacity planning, business continuity, and disaster recovery strategies • Participate in deployments and support application stabilization • Collaborate with teams to improve reliability from the solution design stage • Create and maintain dashboards, alerts, procedures, and operational documentation • Promote a culture of reliability, observability, and continuous improvement
• Experience as a Site Reliability Engineer, SRE, or in an equivalent role • Hands-on experience with cloud environments using GCP, AWS, and/or Azure • Knowledge of Kubernetes and Docker • Experience with observability, monitoring, alerting, and APM • Knowledge of SRE metrics and practices, such as SLI, SLO, SLA, MTTR, and MTTD • Experience managing, investigating, and resolving incidents • Knowledge of application and infrastructure troubleshooting • Experience administering Linux environments • Knowledge of networking, security, performance, and high availability • Experience with automation and Infrastructure as Code • Experience with CI/CD pipelines • Strong communication skills and the ability to work with cross-functional teams • Analytical, proactive, collaborative, and prevention-oriented mindset • Nice to have: experience with GKE, EKS, or AKS • Nice to have: knowledge of Dynatrace, Datadog, Grafana, Prometheus, or similar tools • Nice to have: experience with the ELK Stack, Elasticsearch, and Kibana • Nice to have: knowledge of Terraform and Ansible • Nice to have: experience with mission-critical environments and distributed systems • Nice to have: experience in financial institutions or regulated environments • Nice to have: experience with capacity management and cloud cost optimization • Nice to have: knowledge of disaster recovery and business continuity • Nice to have: experience defining and managing error budgets • Nice to have: certifications in Cloud, Kubernetes, or SRE
• Meal voucher • Food allowance • Home office allowance • Health insurance • Dental insurance • Life insurance • Birthday Day Off • Total Pass / Wellhub app • Boon Saúde • Discount partnerships • Agreements with businesses and educational institutions • Welcome kit • Verity onboarding program • Verity Learning Interval • Great Place to Work certification and workplace improvement initiatives
Apply Now🕒 Yesterday
Arquiteto de Integração DevSecOps administrando plataformas Kubernetes e OpenShift críticas na Stefanini. Automatizando infraestrutura, segurança e operações para soluções tecnológicas globais.
🗣️🇧🇷🇵🇹 Portuguese Required
Kubernetes
Linux
OpenShift
🕒 2 days ago
DevOps Engineer automating CI/CD, cloud, containers, and infrastructure for MTP, a digital transformation consultancy serving enterprise clients.
🗣️🇧🇷🇵🇹 Portuguese Required
AWS
Azure
Cloud
Docker
Grafana
Jenkins
Kubernetes
Linux
Prometheus
Terraform
🕒 3 days ago
DevOps engineer automating CI/CD, cloud infrastructure, containers, and observability for DOMVS iT. Supporting critical Linux and Windows environments in São Paulo.
🗣️🇧🇷🇵🇹 Portuguese Required
AWS
Azure
Docker
Grafana
Jenkins
Kubernetes
Linux
NGINX
OpenShift
Terraform
🕒 5 days ago
Senior DevOps Engineer designing CI/CD, cloud, and IaC platforms for CI&T, a technology transformation company turning enterprise AI into business impact.
🇧🇷 Brazil – Remote
💰 $5.5M Venture Round on 2014-04
⏰ Full Time
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
AWS
Azure
Cloud
Docker
Jenkins
Kubernetes
Python
Terraform
🕒 September 19
SRE Engineer improving reliability, observability, and resilience for Verity, a digital transformation and engineering consultancy. Automating cloud, Kubernetes, and incident-management operations.
🗣️🇧🇷🇵🇹 Portuguese Required
Ansible
AWS
Azure
Cloud
Docker
ElasticSearch
Google Cloud Platform
Grafana
Kubernetes
Linux
Prometheus
Terraform