
1001 - 5000 employees
đ Retail
đď¸ eCommerce
đ° Venture Round on 2019-01
Retail ⢠eCommerce ⢠AI
CSC Generation is a retail technology platform that enhances revenue growth and unit margin management through automation and AI. The company manages a diverse portfolio of 325,000 products online and receives over 10 million monthly page views across its brands. CSC Generation specializes in retail, ecommerce, and wholesale, with a commitment to expanding by acquiring successful brands. Founded in 2016 by Justin Yoshimura, CSC Generation has acquired several well-known brands including One Kings Lane and Sur La Table, and continues to seek out new brands to join its network. The company offers extensive career opportunities and focuses on creating an inspiring and challenging work environment.
đĽ 0 minutes ago
đ¨đˇ Costa Rica â Remote
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)
Ansible
AWS
Azure
Cloud
DNS
Google Cloud Platform
Grafana
JavaScript
Kubernetes
Linux
Node.js
Prometheus
Python
Terraform
TypeScript
Improve your chances of getting an interview by checking your resume score before you apply.

1001 - 5000 employees
đ Retail
đď¸ eCommerce
đ° Venture Round on 2019-01
Retail ⢠eCommerce ⢠AI
CSC Generation is a retail technology platform that enhances revenue growth and unit margin management through automation and AI. The company manages a diverse portfolio of 325,000 products online and receives over 10 million monthly page views across its brands. CSC Generation specializes in retail, ecommerce, and wholesale, with a commitment to expanding by acquiring successful brands. Founded in 2016 by Justin Yoshimura, CSC Generation has acquired several well-known brands including One Kings Lane and Sur La Table, and continues to seek out new brands to join its network. The company offers extensive career opportunities and focuses on creating an inspiring and challenging work environment.
⢠Work on service resiliency, performance tuning, and system design across Backcountry's platform ⢠Drive resolution of critical incidents and ensure fixes are methodically implemented through postmortems ⢠Leverage AI-assisted engineering tools (Claude Code, GitHub Copilot, MCP-based agents) to investigate, automate, and ship fixes across infrastructure and application repositories ⢠Reduce toil by designing and implementing automation ⢠Partner with other Site Reliability Engineers, developers, and architects to evaluate and implement best practices for current and future workloads ⢠Monitor system health and capacity, taking proactive action to fix problems before they occur ⢠Collaborate with engineering teams to build, deploy, and support features ⢠Build and maintain observability (metrics, logs, traces, profiles) and SLI/SLO instrumentation for Backcountry services ⢠Participate in FinOps initiatives across GCP and AWS, including capacity planning and committed-use discount strategy ⢠Participate in the on-call support rotation within the SRE team
⢠3+ years of experience supporting containerized production services, preferably running Kubernetes ⢠3+ years of experience with Infrastructure as Code (Terraform, AWS CDK, Ansible, etc.) ⢠3+ years of cloud experience operating in Google Cloud Platform and/or AWS (multi-cloud stack; Azure/Entra exposure is a plus) ⢠Comfortable diagnosing issues and shipping bug fixes directly to application code (not just infrastructure) to keep services reliable and stable ⢠Comfortable performing deep dives across both infrastructure and application/software git repositories to trace issues end-to-end ⢠Proficient with AI-assisted coding tools (e.g., Claude Code, GitHub Copilot) and MCP-based agents, used to accelerate investigation, code review, and automation ⢠Strong knowledge of scripting and programming languages (Bash, Python, and TypeScript/Node.js) ⢠Experience managing Linux (any major distribution) in production environments ⢠Excellent understanding of internet application protocols (DHCP, DNS, HTTPS, SSH, etc.) ⢠Understanding of how DevOps (CI/CD) and SRE practices (SLOs, SLIs) apply to daily work ⢠Hands-on experience with observability tooling (Grafana, Prometheus, Loki, OpenSearch, or equivalents) and SLI/SLO instrumentation ⢠Experience with GitOps and Kubernetes packaging (ArgoCD, Helm, Kustomize) ⢠Proactively track emerging technology trends and developments, evaluating which ones are worth bringing into engineering practice ⢠Bachelor's degree in computer science or similar, or equivalent experience ⢠Advanced-level English communication skills, both verbal and written.
⢠Competitive Benefits: We offer an attractive benefits package including primarily remote work, private medical and life insurance, additional paid time off, monthly allowances and reimbursements, employee discounts, and opportunities for professional growth.
Apply NowđĽ 17 hours ago
DevOps Support Engineer improving platform reliability and operational performance for TransUnion's credit products. Support 24/7 operational readiness and manage complex technical applications.
đ¨đˇ Costa Rica â Remote
đ° Post-IPO Debt on 2018-04
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)
Cloud
Google Cloud Platform
ITSM
Shell Scripting
đ 4 days ago
Site Reliability Engineer ensuring reliable operation of company services at Veeam Software. Responsible for proactive monitoring, incident response, and resilient observability practices.
đ¨đˇ Costa Rica â Remote
đ° $500M Private Equity Round on 2019-01
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)
AWS
Azure
Cloud
Grafana
Kubernetes
Linux
đ 4 days ago
Senior DevOps Engineer responsible for cloud infrastructure and CI/CD platforms at Experian. Leverage AI-driven tooling to enhance operational excellence and automation practices.
AWS
Cloud
Distributed Systems
Kubernetes
Linux
Python
Terraform
đ 4 days ago
DevOps Senior Engineer enhancing security boundaries and policies in AWS at Experian. Leading projects in cloud security while collaborating with DevOps and engineering teams.
AWS
Cloud
Firewalls
Linux
Python
SDLC
đ 4 days ago
đ¨đˇ Costa Rica â Remote
đ° Private Equity Round on 2021-10
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
Ansible
AWS
Azure
Chef
Cloud
Docker
Kubernetes
Linux
Prometheus
Puppet
Python
SQL