Site Reliability Engineer (SRE) – UI/UX

Job not on LinkedIn

🔥 4 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Software Mind

Software Mind

1001 - 5000 employees

Founded 1999

🤖 Artificial Intelligence

☁️ SaaS

📡 Telecommunications

💰 Private Equity Round on 2020-12

Artificial Intelligence • SaaS • Telecommunications

Software Mind is a technology company that specializes in software development and digital transformation services. With a focus on AI and cloud solutions, the company offers a wide range of services including custom software development, mobile app development, and cloud consulting. Software Mind serves various industries such as financial services, telecom, biotech, and media, providing tailored solutions to accelerate digital transformations and business growth globally.

📋 Description

• Support the deployment, operations, and ongoing maintenance of production services running on Kubernetes • Monitor service health, availability, and performance • Investigate and troubleshoot production incidents using logs, monitoring, and debugging tools • Perform log analysis and incident debugging using Splunk • Identify service issues and collaborate with engineering teams to support timely resolution • Participate in incident response and production support activities • Perform first-level debugging of UI-related issues involving Web Components • Support service reliability and continuous improvement initiatives • Assist with CI/CD pipelines and cloud-native application operations when needed • Work effectively within a client-directed backlog and established priorities

🎯 Requirements

• 4+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Production Support, or a related role • Hands-on experience supporting deployment, operations, and ongoing maintenance of production services running on Kubernetes • Experience monitoring service health, troubleshooting production issues, and supporting service reliability • Proficiency with Splunk for log analysis and incident debugging • Experience participating in production incident response and root-cause analysis • Working knowledge of Web Components and ability to perform first-level debugging of UI-related issues • Strong troubleshooting, analytical, and problem-solving skills • Experience collaborating with software engineering and cross-functional teams • Ability to work independently and effectively within a client-directed backlog • Excellent written and spoken English, at least B2 level • Experience supporting CI/CD pipelines • Familiarity with multi-tenant services • Experience with cloud-native application operations • Experience supporting high-availability enterprise or SaaS platforms • Familiarity with additional monitoring and observability tools • Experience with cloud platforms such as AWS, Azure, or GCP • Familiarity with container and deployment technologies such as Docker and Helm

🏖️ Benefits

• Competitive salary • Laptop • Professional development and training opportunities • Work with cutting-edge cloud and container technologies • Flexible work arrangements and collaborative team environment • Impact on organization-wide digital transformation initiatives

Apply Now

Similar Jobs

🕒 3 days ago

Mirantis

501 - 1000

💼 Consulting

🏥 Healthcare

📦 Logistics

Senior Site Reliability Engineer at Mirantis, contributing to cloud-based AI solutions using Kubernetes. Focused on deploying AI infrastructure and ensuring system reliability and performance.

🕒 4 days ago

Menlo Security Inc.

201 - 500

🔒 Cybersecurity

🏢 Enterprise

Platform Infrastructure Engineer enabling secure connectivity for customers while managing GCP and AWS infrastructure services. Join a global team with a mission-oriented approach to security and operations.

🇨🇦 Canada – Remote

💵 $112k - $168k / year

💰 $100M Series E on 2020-11

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 4 days ago

PerfectServe

201 - 500

🏥 Healthcare

⚕️ Healthcare Insurance

☁️ SaaS

Core role deploying AI products to new customers in healthcare setting. Collaborate with teams to optimize deployments and improve AI offerings.

🇨🇦 Canada – Remote

💵 $165k - $195k / year

💰 Private Equity Round on 2018-05

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 4 days ago

Mirantis

501 - 1000

💼 Consulting

🏥 Healthcare

📦 Logistics

Senior Site Reliability Engineer deploying AI infrastructure for cloud technologies. Collaborate with international teams on AI-driven automation and high-performance systems using Kubernetes.

🕒 4 days ago

Coinbase

1001 - 5000

💼 Consulting

₿ Crypto

💸 Finance

Senior Software Engineer improving reliability and security across Coinbase's services. Design and deliver projects for resilience and safe deployments within a fast-paced remote work environment.

🇨🇦 Canada – Remote

💵 $191.1k / year

💰 $21.4M Post-IPO Equity on 2022-11

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)