
1001 - 5000 employees
Founded 1999
🤖 Artificial Intelligence
☁️ SaaS
📡 Telecommunications
💰 Private Equity Round on 2020-12
Artificial Intelligence • SaaS • Telecommunications
Software Mind is a technology company that specializes in software development and digital transformation services. With a focus on AI and cloud solutions, the company offers a wide range of services including custom software development, mobile app development, and cloud consulting. Software Mind serves various industries such as financial services, telecom, biotech, and media, providing tailored solutions to accelerate digital transformations and business growth globally.
🔥 0 minutes ago
Improve your chances of getting an interview by checking your resume score before you apply.

1001 - 5000 employees
Founded 1999
🤖 Artificial Intelligence
☁️ SaaS
📡 Telecommunications
💰 Private Equity Round on 2020-12
Artificial Intelligence • SaaS • Telecommunications
Software Mind is a technology company that specializes in software development and digital transformation services. With a focus on AI and cloud solutions, the company offers a wide range of services including custom software development, mobile app development, and cloud consulting. Software Mind serves various industries such as financial services, telecom, biotech, and media, providing tailored solutions to accelerate digital transformations and business growth globally.
• Support the deployment, operation, and reliability of production services running on Kubernetes • Monitor service health and investigate production incidents across distributed applications • Participate in on-call support, incident response, root cause analysis, postmortems, and reliability improvements • Troubleshoot application runtime, networking, and service-to-service issues in collaboration with engineering teams • Support CI/CD, GitOps-based deployments, observability, and production monitoring • Work within a client-directed backlog and established priorities
• 5+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Production Engineering, or a closely related role • Strong recent hands-on experience supporting Kubernetes-based production services • 3+ years of hands-on production Kubernetes experience strongly preferred • Kubernetes production operations, including deployment, scaling, rollout/rollback, resource tuning, and service-to-service troubleshooting • Strong production incident response experience, including on-call, runbooks, postmortems, and paging hygiene • Splunk experience for log aggregation, search, and production troubleshooting • Prometheus and Grafana experience, specifically building alert rules and dashboards • CI/CD and infrastructure-as-code for containerized deployments, including Helm and GitOps tools such as ArgoCD or Flux • Strong Linux and networking fundamentals, including DNS, load balancing, TCP/HTTP, HTTP/2, and Kubernetes networking • Node.js production troubleshooting, including heap snapshots, CPU profiles, event-loop blocking, memory growth, worker/process isolation, and V8 isolates or similar runtime models • JVM/Java production troubleshooting, including GC log analysis, thread dump analysis, JVM tuning, and Java service latency investigation • In-memory cache experience with Redis/Valkey, including key design, TTL/eviction tuning, and cache invalidation • Service-to-service authentication experience, including mTLS, certificate rotation, certificate format conversion, and JWT-based service authentication • Web Components/Lit experience is nice to have • Server-side rendering or isomorphic runtime experience is nice to have • Canary rollout/multi-version production operations is nice to have • Distributed tracing and request-context correlation is nice to have • KEDA or event-driven autoscaling experience is nice to have • Experience with enterprise platform integration layers is nice to have
• Competitive salary and laptop • Professional development and training opportunities • Work with cutting-edge cloud and container technologies • Flexible work arrangements and collaborative team environment • Impact on organization-wide digital transformation initiatives
Apply Now🕒 Yesterday
Senior DevOps Developer building reliable AWS, Kubernetes, and MongoDB services for Autodesk Construction Solutions. Improving automation, observability, security, disaster recovery, and production reliability for construction software customers.
🇨🇦 Canada – Remote
💵 $107k - $157.3k / year
⏰ Full Time
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🕒 2 days ago
Cloud DevOps Engineer building AWS infrastructure, CI/CD pipelines, and automation for High Tech Genesis. Supporting containers, event-driven systems, service mesh, and observability.
🇨🇦 Canada – Remote
💵 CA$65 - CA$70 / hour
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🕒 3 days ago
DevOps Engineer building AWS infrastructure and automated systems for S&P Global’s financial data and technology solutions. Supporting resilient applications through Terraform, CI/CD, containerization, monitoring, and cloud operations.
🇨🇦 Canada – Remote
💵 $75k - $108k / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🕒 5 days ago
Site Reliability Engineer operating Yelp’s Kafka and Flink streaming platform across Canada. Automating cluster management, scaling, upgrades, migrations, and incident recovery for real-time data systems.
🇨🇦 Canada – Remote
💵 $135k - $185k / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🕒 August 6
Site Reliability Engineer operating Yelp’s Kafka and Flink streaming infrastructure across Canada. Automating cluster operations, upgrades, scaling, and incident recovery for real-time data systems.
🇨🇦 Canada – Remote
💵 $135k - $185k / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)