Principal DevOps Engineer

🕒 vor 2 Monaten

🗽 New York – Remote

infoinfo

💵 $180.000 - $230.000 / Jahr

⏰ Vollzeit

🔴 Experte

⛑ DevOps- und Site Reliability Engineer (SRE)

🦅 H1B-Visum-Sponsor

infoinfo

👻 Geisterscore 24%

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of NBCUniversal

NBCUniversal

10.000+ Mitarbeiter

Gegründet 2004

📱 Medien

Media • Entertainment

NBCUniversal ist ein führendes globales Medien- und Unterhaltungsunternehmen, bekannt für die Erstellung und Verbreitung von Inhalten über verschiedene Plattformen. Mit über 100 Jahren Erfahrung ist es Teil von Comcast und umfasst Marken wie Peacock, NBC Sports und viele andere, um ein weltweites Publikum zu informieren, zu unterhalten und zu stärken. Das Unternehmen ist in den Bereichen Fernsehübertragung, Filmproduktion und Freizeitparks tätig und ist auch für seine Initiativen in den Bereichen Technologie und unternehmerische soziale Verantwortung anerkannt. NBCUniversal setzt sich für Innovation und sozialen Einfluss ein, was es zu einem dynamischen Arbeitsplatz für Medien- und Technologieprofis macht.

Beschreibung

• Architect a Kubernetes-native platform that models broadcast infrastructure as custom resources. • Lead the technical strategy leveraging Crossplane compositions and custom Go functions to automate provisioning across multi-account AWS environments and on-prem control rooms. • Design, build, and maintain production-grade Kubernetes operators, controllers, and internal platform APIs in Go. • Actively develop custom Crossplane providers to deeply integrate external enterprise platforms (such as NRCS, Venafi, and Infoblox) into our control plane, managing resource lifecycles and approval workflows. • Lead the design of cloud networking, DNS strategies, and cross-account connectivity across hybrid environments, automating VPC topology and dynamic network routing. • Partner closely with broadcast systems engineers, system integrators, and external vendors to bridge the gap between broadcast hardware and automated infrastructure. • Write RFCs, drive architectural decisions, mentor engineers, and establish high-confidence CI/CD pipelines, testing strategies, and GitHub Actions automation. • Own the platform's authorization model, designing hierarchical RBAC systems, resource identifier schemes, and identity integrations that enforce fine-grained access control. • Drive GitOps-based continuous delivery (Flux, Kustomize, Helm) and manage configuration-as-code for compute fleets using Puppet. • Ensure deep operational visibility by designing comprehensive observability and alerting stacks. • Oversee the integration of remote desktop/VDI connectivity solutions, focusing on session authentication, credential management, and gateway routing.

🎯 Anforderungen

• 10+ years of experience designing, building, and operating production infrastructure and cloud-native platforms at enterprise scale. • Strong proficiency in Go (systems-level programming, API servers). • Expert-level knowledge of the Kubernetes ecosystem, including CRD/XRD generation, operators, informers, admission webhooks, and RBAC. • Deep production experience with Crossplane, including composite resources, composition functions, and specifically developing custom Crossplane providers in Go to integrate external enterprise platforms. • Extensive production experience with AWS multi-account architectures, cross-account networking patterns, and identity federation. • Production experience with GitOps tooling, specifically Flux (HelmRelease, Kustomization) or ArgoCD for continuous delivery on Kubernetes. • Hands-on experience with Puppet, including module development, PuppetDB, Hiera, and r10k. • Experience designing REST APIs with middleware patterns and modern authentication (OAuth/JWT). • Keen eye for information security, including cross-account IAM trust chains, least-privilege policies, JWT token lifecycles, and secrets abstraction. • Strong background in designing telemetry platforms using Grafana, Prometheus/Mimir, Loki, OpenTelemetry, and metrics collection agents (Alloy, Prometheus Node Exporter). • Working knowledge of PostgreSQL, SQLite or similar relational databases, encompassing schema design, migrations, and query optimization. • Excellent problem-solving skills with a proven ability to present architectural decisions to executives, engage with vendors, and write clear technical documentation.

🏖️ Vorteile

• Health insurance • Dental insurance • Vision insurance • 401(k) • Paid leave • Tuition reimbursement • Variety of discounts and perks

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 2 Monaten

Assured

11 - 50

🛡️ Versicherung

☁️ SaaS

🤖 Künstliche Intelligenz

Staff Site Reliability Engineer optimizing database systems for tech-driven insurance provider. Leading design, automation, and performance initiatives for a modern claims processing platform.

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 2 Monaten

AlphaSense

1001 - 5000

💼 Beratung

🏥 Gesundheitswesen

📣 Marketing

Staff SRE architecting reliability platforms for AlphaSense’s AI-powered market intelligence platform. Leading AIOps, incident response, observability, and SRE culture across global engineering.

🇺🇸 Vereinigte Staaten – Remote

💵 $150.000 - $225.000 / Jahr

💰 Debt Financing im 2022-06

⏰ Vollzeit

🔴 Experte

⛑ DevOps- und Site Reliability Engineer (SRE)

🦅 H1B-Visum-Sponsor

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 2 Monaten

Lyric - Clarity in motion.

201 - 500

🏥 Gesundheitswesen

💼 Beratung

📦 Logistik

Staff Azure DevOps Engineer managing secure, scalable Azure infrastructure for Lyric’s healthcare decision intelligence platform. Automating deployments, improving reliability, and supporting mission-critical claims workflows.

🇺🇸 Vereinigte Staaten – Remote

💵 $150.289 - $225.434 / Jahr

⏰ Vollzeit

🔴 Experte

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 2 Monaten

Gorilla Logic

501 - 1000

💼 Beratung

📣 Marketing

📦 Logistik

Technical Engineering Manager leading high-performing cloud and DevOps teams. Guiding architecture and delivery of scalable, reliable, and secure cloud solutions for clients.

🇺🇸 Vereinigte Staaten – Remote

⏰ Vollzeit

🟠 Senior

🔴 Experte

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 2 Monaten

ClassWallet

11 - 50

💳 Fintech

📚 Bildung

🏛️ Regierung

DevOps Engineer optimizing AWS infrastructure, GitHub Actions, and observability for ClassWallet's public-funds digital wallet platform. Ensuring scalable, compliant, and reliable systems for government agencies.

🇺🇸 Vereinigte Staaten – Remote

💰 €500.000 Debt Financing im 2020-05

⏰ Vollzeit

🟠 Senior

🔴 Experte

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich