
5001 - 10000 employees
đ Cybersecurity
đ° Post-IPO Equity on 2001-07
Cloud Computing âą Cybersecurity âą Content Delivery
Akamai Technologies is a leading cloud services provider that specializes in delivering security, cloud computing, and content delivery solutions. It offers a range of services such as API security, DDoS protection, and performance optimization for web applications, ensuring secure and reliable user experiences. With a robust global infrastructure, Akamai empowers businesses to streamline their digital presence while safeguarding against various cyber threats and enhancing application performance.
đ 6 days ago
đ Massachusetts â Remote
đ” $121.4k - $218.6k / year
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đŠ H1B Visa Sponsor
Improve your chances of getting an interview by checking your resume score before you apply.

5001 - 10000 employees
đ Cybersecurity
đ° Post-IPO Equity on 2001-07
Cloud Computing âą Cybersecurity âą Content Delivery
Akamai Technologies is a leading cloud services provider that specializes in delivering security, cloud computing, and content delivery solutions. It offers a range of services such as API security, DDoS protection, and performance optimization for web applications, ensuring secure and reliable user experiences. With a robust global infrastructure, Akamai empowers businesses to streamline their digital presence while safeguarding against various cyber threats and enhancing application performance.
âą Oversee, scale, and optimize next-generation dedicated AI hardware infrastructure âą Ensure uptime and reliability of AI hardware infrastructure offerings âą Collaborate with product teams from early development through deployment to ensure reliability, scalability, and performance âą Define key performance indicators and defend them when breached âą Develop Python programmatic tooling and infrastructure-as-code utilities to automate fleet-wide provisioning and reduce operational toil âą Integrate automated workflows across corporate ticketing systems for hardware and network break-fix events âą Use AI utilities and LLM-assisted development to accelerate technical execution and system analysis âą Improve availability, latency, and systemic health of high-density hardware environments using private cloud and compute technologies âą Design telemetry pipelines, Prometheus/Grafana dashboards, and AI-based anomaly detection for bare-metal and virtualized environments âą Participate in 24x7x365 on-call rotations and lead real-time incident management âą Manage high-severity service disruption protocols through PagerDuty and Slack workflows âą Partner with third-party infrastructure vendors and coordinate on-site field technicians
âą 5 years of relevant experience âą Bachelor's degree in Computer Engineering, Computer Science or equivalent âą Tooling and coding ability in Python for scalable operational tools, API integrations, and automation frameworks âą Hands-on experience with Prometheus, Grafana, OpenTelemetry, and Loki âą Working understanding of advanced networking topologies and high-bandwidth routing/switching infrastructure âą Knowledge of BGP and dual-stack IPv4/IPv6 networks âą Experience designing new service rollouts, including operational readiness criteria, telemetry baselines, and alerting thresholds âą Extensive experience building technical runbooks, leading complex incident response bridges, and driving blameless post-mortems âą Ability to own ambiguous technical problems, coordinate cross-functional teams, and deliver production-grade solutions âą Participation in 24x7x365 on-call rotations
âą Flexible work through Akamai's FlexBase program: at home, in an office, or a combination of both âą Annual bonus or incentives âą Equity awards âą Employee Stock Purchase Plan (ESPP) âą Healthcare âą 401K savings plan âą Company holidays âą Vacation/PTO âą Sick time âą Parental leave âą Employee assistance program âą Mental and financial wellness support
Apply Nowđ 6 days ago
Senior SRE ensuring reliable, scalable infrastructure for PrizePicksâ daily fantasy sports platform. Leading incident response, observability, Kubernetes operations, and reliability improvements.
đșđž United States â Remote
đ” $120k - $175k / year
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đ 6 days ago
Senior SRE building reliability engineering for Veeam Data Cloudâs Government and Sovereign Cloud SaaS platform. Designing Azure infrastructure, observability, incident response, and compliance-ready delivery practices.
đșđž United States â Remote
đ” $158.4k - $294.1k / year
đ° $500M Private Equity Round on 2019-01
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đŠ H1B Visa Sponsor
đ 6 days ago
Site Reliability Engineer building reliability practices for Veeamâs Government and Sovereign Cloud SaaS platform. Designing Azure infrastructure, observability, automation, and incident-response systems in regulated environments.
đșđž United States â Remote
đ” $138.9k - $231.4k / year
đ° $500M Private Equity Round on 2019-01
â° Full Time
đ Senior
đŽ Lead
â DevOps & Site Reliability Engineer (SRE)
đŠ H1B Visa Sponsor
đ 6 days ago
Senior Deployment Engineer helping Karat, a technical interviewing company, implement and optimize enterprise interview frameworks. Advising clients, analyzing hiring performance, and delivering executive training.
đșđž United States â Remote
đ” $116.9k - $148.2k / year
đ° Funding Round on 2022-04
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đŠ H1B Visa Sponsor
đ 6 days ago
Site Reliability Engineer architecting Ciscoâs developer platform and infrastructure for cloud application delivery. Consolidating legacy tools and improving scalable, resilient engineering workflows.
đșđž United States â Remote
đ” $138.1k - $198.2k / year
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đŠ H1B Visa Sponsor