Customer Reliability Engineer at Andromeda
Confirmed open
Diagnose GPU-cluster incidents for Andromeda customers at the Linux and Kubernetes layers, coordinate fixes with infrastructure providers, and build monitoring and scripts that catch failures early. The posting lists global remote work or San Francisco.
- Location
- Remote
- Advertised region
- Global Remote / San Francisco, CA / San Francisco / United States
- Employment
- FullTime
- Work arrangement
- remote
- Technologies
- Python, Kubernetes, Linux
Role overview
Diagnose GPU-cluster incidents for Andromeda customers at the Linux and Kubernetes layers, coordinate fixes with infrastructure providers, and build monitoring and scripts that catch failures early. The posting lists global remote work or San Francisco.
How to apply
Review the original posting and apply through the employer’s careers page.
View source and applyApply directlySource checked: 2026-10-03
Show what you can do
Build a free profile around your CV and real work. Share its link when you are ready. Applications for this role still go directly to the employer.
Build a free profile