About this role
Site Reliability Engineer (SRE) at Thinking Machines Lab. Location: San Francisco, California, United States. Role: Define reliability, Own incident response, Design monitoring Requirements: Bachelor's degree or equivalent experience; distributed systems, cloud infrastructure or SRE; reliability tooling; incident response; strong cross-team communication. Category: Engineering Seniority: Mid Level Tools: Kubernetes, Docker, Cloud Platforms, Monitoring Tools, Automation Commitment: Full Time Workplace: Onsite Languages: English