Responsibilities
- Work closely with technical leads, architects, and engineering teams to design solutions for challenges in distributed systems and infrastructure operations.
- Develop automated responses to operational concerns including system monitoring, performance optimization, capacity planning, and incident recovery.
- Take part in on-call duties to maintain and support essential infrastructure services.
- Maintain high standards of code quality by implementing Infrastructure as Code best practices.
- Communicate clearly through the creation and evaluation of technical documentation such as design specs, operational runbooks, and system change logs.
- Use Agile principles to consistently deliver customer-focused improvements and features.
- Serve as the primary contact for existing and new components within the compute platform.
Compensation
Not specified
Work Arrangement
Remote-first with potential for ad-hoc in-person requirements
Team
Engineering team focused on cloud infrastructure and distributed systems
Other
- Occasional travel may be necessary for in-person project or team meetings.
- Although remote-first, employees might need to attend occasional on-site gatherings, functional offsites, or customer engagements.
- Hiring decisions are made by human team members; AI is used during parts of the hiring process to improve efficiency.
- The position is not available for candidates based in San Francisco, Oakland, San Jose, or surrounding areas in California.
Not specified