Search Remote Jobs

Senior Automation, Observability Engineer

Job not on LinkedIn

🔥 0 minutes ago

🇺🇸 United States – Remote

đź’µ $113k - $147k / year

⏰ Full Time

đźź  Senior

👷🏻‍♀️ Engineer

🦅 H1B Visa Sponsor

infoinfo

đź‘» Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Ensono

Ensono

1001 - 5000 employees

đź’Ľ Consulting

Consulting • Cloud Services • IT Services

Ensono is a managed service provider and expert technology advisor focused on enabling clients to navigate complex IT environments. Ensono offers services including mainframe-as-a-service, cloud migration, and IT infrastructure management, with a strong emphasis on modernizing legacy systems and optimizing IT operations. The company is recognized for its flexibility and allyship, providing businesses with the ability to adapt and evolve through solutions like Ensono Flex®. It specializes in both mainframe and cloud solutions, holding competency in AWS mainframe modernization and is a global Azure Expert MSP. Ensono acts as an ally, helping companies achieve better business outcomes through IT innovation and strategic technology advisement.

đź“‹ Description

• Design, implement, and maintain enterprise monitoring and observability solutions • Develop and maintain Grafana dashboards, alerts, and visualizations • Monitor infrastructure, applications, middleware, IoT services, and enterprise telemetry using IBM Instana, Grafana, SolarWinds, and related tools • Configure and manage data collection with Telegraf, Prometheus, and monitoring agents • Analyze metrics, logs, traces, events, and telemetry to identify bottlenecks and service degradation • Support SLO, SLA, and operational health monitoring initiatives • Perform root cause analysis and troubleshoot infrastructure and application issues • Support onboarding, monitoring, and operational management of FOAK services and enterprise applications • Configure, validate, and troubleshoot ELT integrations and telemetry pipelines • Monitor log ingestion, event correlation, and telemetry data quality • Monitor Linux and Windows servers, VMware, Citrix VDI, DNS, proxy, middleware, integration services, enterprise applications, and IoT platforms • Investigate performance issues, recurring alerts, infrastructure anomalies, capacity, availability, and service health • Support platform upgrades, maintenance, and operational readiness reviews • Configure and maintain InfluxDB, including retention policies, performance tuning, and capacity planning • Develop operational dashboards and reports for performance insights • Acknowledge, investigate, troubleshoot, and resolve incidents • Coordinate incident resolution with Infrastructure, Network, Cloud, Security, Application, and Service Delivery teams • Participate in major incident bridges, disaster recovery exercises, and 24x7 operations support • Follow escalation procedures, SOPs, runbooks, and ITIL processes; support Problem Management and RCA documentation • Administer IBM Instana environments and support APM configuration, alerts, baselines, and thresholds • Develop Python, PowerShell, Bash, and VBScript automation for operational tasks, monitoring deployments, remediation, and event-driven workflows • Implement Ansible and Puppet infrastructure automation, server provisioning, configuration management, agent deployment, playbooks, and pipelines • Maintain SOPs, runbooks, monitoring procedures, escalation matrices, and observability documentation • Participate in knowledge-transfer sessions, service onboarding, service transition, migration, and continuous improvement initiatives

🎯 Requirements

• Required skills in Grafana, IBM Instana, SolarWinds, Telegraf, Prometheus, InfluxDB, OpenTelemetry, Grafana Alloy, APM monitoring, event management, alert management, observability concepts, and SLO/SLA monitoring • FOAK support and Enterprise Logging & Telemetry (ELT) experience • Log aggregation and correlation, telemetry data analysis, event correlation, application onboarding, and monitoring standards/observability frameworks • VMware, Linux administration, Windows Server, Citrix VDI, DNS services, proxy services, middleware technologies, and infrastructure performance monitoring • Python, PowerShell, Linux shell scripting, and VBScript • Ansible, Puppet, webhooks, and Infrastructure as Code (IaC) • ServiceNow, incident management, problem management, change management, ITIL Framework, and major incident management • Preferred experience: 7 to 10+ years in Monitoring, Observability, Infrastructure Operations, SRE, or Platform Engineering • Experience supporting large-scale enterprise environments and 24x7 operations • Hands-on experience with Grafana, Instana, SolarWinds, Telegraf, Prometheus, and InfluxDB • Experience supporting VMware, Citrix, middleware, enterprise applications, and cloud monitoring platforms • Experience with FOAK applications and ELT platforms • Strong troubleshooting, RCA, incident management, and operational support skills • Experience with automation frameworks and Infrastructure as Code (Ansible preferred) • Experience integrating observability platforms with enterprise automation solutions • Nice-to-have: Docker, Kubernetes, AWS, Microsoft Azure, Google Cloud Platform, Jenkins, GitHub Actions, GitLab CI/CD, REST APIs, microservices monitoring, and DevOps/SRE practices

🏖️ Benefits

• Unlimited Paid Days Off • Three health plan options • 401k with company match • Dental, vision, short-term disability, long-term disability, life and AD&D coverage • Flexible spending accounts • Family Forming Benefit including fertility coverage and adoption/surrogacy reimbursement • Paid childbearing and paternal leave • Education Reimbursement • Student Loan Assistance or 529 College Funding • Sabbatical leave • Wellness program • Flexible work schedule • Annual bonus plan based on company and individual performance • Equity grant under the Associate Equity Appreciation Program • Option to work from home or in Ensono offices when not required on a client site

Apply Now

Similar Jobs

🔥 1 hour ago

Higharc

11 - 50

🏗️ Construction

đź’Ľ Consulting

🏠 Real Estate

Forward Deployed Engineer building searchable, AI-powered plan data for Higharc, a company transforming how new homes are designed and built. Embedding with homebuilders to ship production data systems.

🔥 1 hour ago

CACI International Inc

10,000+ employees

🎖️ Defense

🏛️ Government

đź”’ Cybersecurity

Data integration engineer building Oracle EBS replication pipelines with GoldenGate and Qlik Replicate. Supporting Advana data analytics, dashboards, compliance, and DoD modernization for CACI’s federal clients.

🇺🇸 United States – Remote

đź’µ $75.2k - $158.1k / year

🔥 Funding within the last year

đź’° $500M Post-IPO Debt on 2026-02

⏰ Full Time

🟡 Mid-level

đźź  Senior

👷🏻‍♀️ Engineer

🔥 2 hours ago

FUJIFILM Corporation

10,000+ employees

🏭 Manufacturing

🏥 Healthcare

Senior Project Engineer implementing Synapse medical imaging solutions for FUJIFILM Healthcare Americas. Leading healthcare IT infrastructure, integrations, migrations, and customer go-live delivery.

Citrix

🔥 2 hours ago

Emsar

501 - 1000

Biomedical Equipment Engineer I repairing and calibrating medical equipment for EMSAR, a healthcare and life-science field-services company. Managing customer service responsibilities across a nationwide territory.

🔥 15 hours ago

Omega Technical Services

201 - 500

🏗️ Construction

đź’Ľ Consulting

🏥 Healthcare

Applied Mechanics Engineer analyzing seismic, structural, and thermomechanical behavior for Omega Technical Services’ advanced nuclear reactor projects. Supporting DOE and DoD mission-critical nuclear infrastructure.

Graphite