Senior Site Reliability Engineer, SRE

Stelle nicht auf LinkedIn

🕒 vor 9 Monaten

🗽 New York – Remote

infoinfo

⏳ Vertrag

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

👻 Geisterscore 45%

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of Gov Services Hub

Gov Services Hub

51 - 200 Mitarbeiter

Gegründet 2015

💼 Beratung

🎖️ Verteidigung

📦 Logistik

Consulting • Defense • Logistics

Gov Services Hub ist ein in den USA ansässiges IT-Dienstleistungs- und Regierungsvertragsunternehmen, das Personaldienstleistungen, Cybersicherheit, Anwendungs- und mobile Entwicklung, Infrastruktur- und Cloud-Dienste sowie Datenanalysen für Bundes-, Landes- und Kommunalbehörden einschließlich Verteidigungs- und Geheimdienstkunden anbietet. Das Unternehmen bietet auch Unterstützung bei Regierungsverträgen wie Hilfe bei Vorschlägen, Compliance-Überprüfungen und Anbieterregistrierung, um Unternehmen zu helfen, öffentliche Aufträge zu gewinnen und zu verwalten. Mit Hauptsitz in Manassas, VA, konzentriert sich Gov Services Hub auf sichere, den Vorschriften entsprechende und skalierbare Technologielösungen, die auf die Bedürfnisse des öffentlichen Sektors zugeschnitten sind.

Beschreibung

• Lead incident response and develop sustainable on-call practices, including runbooks, blameless postmortems, and continuous improvement to reduce MTTR • Build and maintain self-service observability tools (Datadog, Prometheus, ELK) for proactive monitoring and troubleshooting • Create and maintain Infrastructure as Code (IaC) using Terraform or CloudFormation for consistent, secure AWS environments • Partner with development teams to architect resilient, scalable infrastructure for critical components like databases, networking, async workflows, and data pipelines • Design and implement robust CI/CD pipelines (GitHub Actions) with advanced deployment strategies (blue/green, canary) • Drive best practices in reliability and performance early in the design phase to future-proof January’s systems

🎯 Anforderungen

• Proven experience leading incident response and postmortem processes for high-availability production systems • Deep expertise in designing highly available architectures (EC2, Fargate, auto-scaling, health checks, graceful degradation) • Strong experience with AWS cloud infrastructure and IaC tools (Terraform, CloudFormation) • Hands-on experience with CI/CD automation using GitHub Actions or equivalent tools • Proficiency in observability and monitoring stacks (Datadog, Prometheus, ELK) • Solid scripting/programming skills in Python (for automation, tooling, and debugging) • Excellent communication and documentation skills, with the ability to collaborate across engineering and platform teams

🏖️ Vorteile

• Remote role (NYC-based preferred for hybrid collaboration) • Opportunity to build and own the entire SRE practice for a growing FinTech startup • Fast-paced, innovative environment working on AI-forward consumer finance products

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 9 Monaten

Atmosera

51 - 200

☁️ SaaS

🔒 Cybersecurity

DevOps Engineer supporting GitHub Enterprise Cloud migration projects at Atmosera. Collaborating with client teams to ensure successful migrations from various platforms.

🇺🇸 Vereinigte Staaten – Remote

⏳ Vertrag

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 9 Monaten

Pierce Professional Resources

11 - 50

🏢 Unternehmen

🎯 Rekrutierung

💼 Beratung

DevOps Engineer focusing on building CI/CD pipelines and developing cloud infrastructure on AWS and Azure. Managing containerized applications and ensuring compliance and security in the cloud.

🇺🇸 Vereinigte Staaten – Remote

⏳ Vertrag

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich