
51 - 200 employees
Founded 2006
Castillians is a company whose publicly available text is inaccessible without JavaScript; the provided content only shows a message asking the user to enable JavaScript. No information about the company's product, services, industry, or target customers can be determined from this text alone.
🔥 0 minutes ago
🇮🇪 Ireland – Remote
⏳ Contract/Temporary
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
Improve your chances of getting an interview by checking your resume score before you apply.

51 - 200 employees
Founded 2006
Castillians is a company whose publicly available text is inaccessible without JavaScript; the provided content only shows a message asking the user to enable JavaScript. No information about the company's product, services, industry, or target customers can be determined from this text alone.
• Monitor production systems, applications, and infrastructure using monitoring and alerting tools • Respond to alerts, investigate incidents, and take corrective action following established procedures • Perform routine health checks and scheduled maintenance activities • Review application, system, and server logs to troubleshoot issues • Restart services and carry out operational tasks using documented runbooks • Manage and resolve support tickets within agreed service levels (SLAs) • Escalate complex issues to second-level support or engineering teams as required • Track incidents through resolution and maintain accurate documentation • Conduct daily operational checks to ensure system stability and availability • Maintain and update operational documentation, runbooks, and troubleshooting guides • Participate in shift rotations, including weekends and public holidays • Support incident management activities and post-incident reviews • Identify recurring issues and recommend improvements to enhance reliability and efficiency
• 5+ years of experience in Technical Support, Operations, NOC, SOC, or Site Reliability Engineering roles • Strong troubleshooting, analytical, and problem-solving skills • Experience with monitoring and alerting tools, including Zabbix, Grafana, and Prometheus • Experience reviewing and analyzing application, system, and server logs • Understanding of incident management and escalation processes • Experience using ticketing tools such as Jira • Experience with cloud platforms, particularly Google Cloud • Familiarity with Linux and common command-line tools • Ability to follow structured procedures and operational runbooks • Strong attention to detail and commitment to service reliability • Good written and verbal communication skills • Ability to work independently and effectively within a shift-based team • Availability for shifts including weekends and public holidays • Ability to work within CET ± 2 hours • B2B contract outside IR35
• Flexible working arrangements • Opportunity for repeat engagements based on performance • Access to CX guidance and market insights through our professional network • 6-month duration with auto renew • Clear scope with no ambiguity over deliverables
Apply Now