SRE Monitoring Platform Software Engineer, Entry Level

Job not on LinkedIn

🔥 4 minutes ago

🏄 California, Texas – Remote

infoinfo

💵 $105k - $155k / year

⏰ Full Time

⚪️ Entry-level

🧑‍💻 Full-stack Engineer

🚫👨‍🎓 No degree required

👻 Ghost score 4%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Bitdeer Group

Bitdeer Group

201 - 500 employees

💼 Consulting

📦 Logistics

🏗️ Construction

💰 Post-IPO Equity on 2023-05

Consulting • Logistics • Construction

Bitdeer Group is a leader in the blockchain and high-performance computing industry. It is one of the world’s largest holders and suppliers of hash rate, offering specialized mining infrastructure and high-quality hash rate sharing products. Founded by cryptocurrency pioneer Jihan Wu and led by CEO Matt Linghui Kong, the company is headquartered in Singapore with mining datacenters in the United States, Norway, and Bhutan. Bitdeer is committed to providing comprehensive computing solutions, including cloud services and AI capabilities, while emphasizing dedication, authenticity, and trustworthiness in its mission to be the most reliable provider in the industry.

📋 Description

• Contribute to Bitdeer's NeoCloud SRE platform for observing, protecting, and operating a multi-region GPU rental fleet • Build collection agents, metrics/logs/traces/profiles stores, enrichment services, and collection monitors • Write ingestion, query, and storage-path code • Contribute to alerting, correlation, and SLO frameworks; implement and tune default alert rules • Contribute to topology services, cluster-health rollups, and OSS-SRE-tool collection plugins for Kubernetes, Slurm, Ray, Volcano, Kueue, and KubeRay • Help build remediation actuators, orchestration/workflow components, inspection probes, and job schedulers • Instrument services with metrics, logs, and traces using OpenTelemetry • Build dashboards and write actionable on-call runbooks • Write unit, integration, and contract tests for shipped components • Participate in chaos and soak tests led by senior engineers • Develop components from design through production using GitOps and the CI/CD release pipeline • Meet declared SLOs and maintain drift-free systems • Operate what you build under senior-engineer guidance • Participate in on-call as a shadow before taking primary responsibility • Progress toward independently delivering components and owning a sub-context within 12 months

🎯 Requirements

• 0–2 years of software engineering experience; new graduates with strong projects or internships welcome • Solid fundamentals in Go, Python, Java, or Rust; Go preferred • Ability to write clean, tested, readable code and explain design choices • Data structures, algorithms, concurrency, TCP/HTTP networking, and operating-system fundamentals • Understanding of distributed-systems concepts including idempotency, retries, back-pressure, caching, and eventual consistency • Hands-on exposure to Prometheus, Grafana, Loki, or similar observability tools • Ability to write basic PromQL queries and instrument services • Familiarity with Linux, the shell, system logs, and standard debugging tools • Kubernetes basics, including Pods, Services, and Deployments; experience running something on Kubernetes • Git, branching, pull requests, and CI pipelines such as GitHub Actions or GitLab CI • Unit and integration testing discipline • Clear written and verbal English • Curiosity and willingness to learn GPU/AI infrastructure, AIOps, distributed systems, and observability • Nice-to-have: internship or project experience in monitoring, observability, telemetry pipelines, or platform/SRE tooling • Nice-to-have: exposure to GPU/AI infrastructure such as DCGM, InfiniBand/RoCE, Kubernetes GPU Operator, Slurm, or Ray • Nice-to-have: exposure to AIOps/ML-adjacent tooling • Nice-to-have: contributions to open-source observability or cloud-native projects

🏖️ Benefits

• Mentorship from senior and principal engineers • Opportunity to learn the full observability stack at production scale • Participation in on-call as a shadow before taking primary responsibility • Equal employment opportunities and non-discrimination

Apply Now

Similar Jobs

🕒 2 days ago

Cambium Learning Group

501 - 1000

📚 Education

🤖 Artificial Intelligence

Software Engineer Intern contributing to AI-powered educational assessment systems at Cambium Learning Group, an educational technology leader. Supporting LLM features, agent workflows, and generative AI platform integration.

🕒 September 2

McKesson

10,000+ employees

💼 Consulting

📦 Logistics

🏥 Healthcare

Software Engineer Intern building approval workflows in Macro Helix’s 340B Architect SaaS platform. Improving healthcare implementation transparency and speed at McKesson.

🕒 September 2

ShippyPro

51 - 200

📦 Logistics

☁️ SaaS

🛍️ eCommerce

Junior software engineer building ShippyPro’s high-volume shipment tracking and notification services. Developing Python, TypeScript, AWS, microservices, and event-driven systems for global merchants.

🇺🇸 United States – Remote

💵 €29k - €35k / year

💰 $15M Series B - ShippyPro on 2023-11

⏰ Full Time

⚪️ Entry-level

🧑‍💻 Full-stack Engineer

🚫👨‍🎓 No degree required

🕒 August 31

IGS Energy

1001 - 5000

⚡ Energy

🛍️ eCommerce

Software Engineering Intern developing web applications and APIs for IGS Energy. Contributing to agile software development, automated testing, and production releases.

🕒 August 21

VivSoft

51 - 200

💼 Consulting

🏥 Healthcare

📦 Logistics

Entry-level software engineer supporting secure personnel-vetting software for the U.S. Department of Defense. Contributing to coding, testing, debugging, and DevSecOps delivery under senior engineers.