Responsibilities
- Lead the full lifecycle of a centralized automation framework, including strategic roadmap development and long-term vision alignment across network operations and engineering teams.
- Serve as the technical governance lead for automation architecture, defining standards for engineering practices, API design, data models, and cross-domain consistency to prevent siloed solutions.
- Design a modular extensibility model allowing domain teams to build custom workflows using standardized, reusable components without duplicating core logic.
- Evaluate build-versus-buy decisions for critical tooling—including orchestration, data storage, observability, and AI platforms—and communicate cost and operational impacts to leadership.
- Define the operational framework for automation, including ownership models, readiness criteria, promotion workflows, and cross-functional coordination to align engineering and operations.
- Design a reusable component library for health checks, diagnostics, remediation actions, and notifications, with versioned and governed interfaces for safe composition.
- Establish a unified API layer for integrating alerts, executing actions across systems, and reporting outcomes to downstream tools like ITSM and collaboration platforms.
- Enforce code quality and delivery standards through testing frameworks, CI/CD pipelines, security scanning, and GitLab-based deployment workflows, setting benchmarks via reference implementations.
- Define a DAG-based orchestration model using tools like Apache Airflow to enable auditable, composable workflows for diagnostics, decisions, remediation, and validation.
- Architect fully automated closed-loop responses for lower-severity incidents, from detection through root cause analysis, remediation, and resolution, with full traceability and no manual intervention.
- Manage a tiered automation model progressing from recommendation to human-approved actions to full autonomy, with defined confidence metrics and promotion rules based on incident performance.
- Champion AI-assisted development practices using tools like Claude and GitHub Copilot, demonstrating rapid prototyping capabilities to accelerate solution delivery.
- Design intelligent diagnostic agents and correlation engines that integrate signals from multiple domains to generate accurate root-cause hypotheses and reduce alert noise.
- Integrate AI into core operational processes with defined governance, human oversight points, and trust thresholds to ensure safe and reliable deployment.
- Build real-time data pipelines and an operational data lake to store telemetry, incidents, KPIs, and automation logs for analysis, forecasting, and performance measurement.
- Implement DataOps practices including version-controlled pipelines, data validation, schema management, and monitoring to ensure data reliability matching network standards.
- Define standardized dashboarding with Grafana using a single source of truth to deliver consistent operational visibility across teams and leadership levels.
- Enforce automation governance policies requiring rollback plans, audit logs, human validation for critical actions, and fallback mechanisms when confidence thresholds are breached.
- Embed security into automation workflows through least-privilege access, secrets management, audit trails, and controlled third-party access with explicit oversight.
- Assess downstream impact of automation on provisioning, inventory, billing, and reporting systems to prevent unintended operational or financial side effects.
- Act as the central intake and design authority for automation requests, transforming manual toil into prioritized, measurable automation initiatives with clear ownership and targets.
- Collaborate with platform, security, and operations teams to align on event schemas, API contracts, and policy enforcement before automation deployment.
- Mentor senior and staff-level developers, raising technical standards and establishing review processes for automation across all engineering locations.
- Drive continuous improvement by measuring automation coverage, impact on resolution times, false positive rates, override frequency, and adoption metrics for all workflows.
Benefits
- Competitive pay with equity through stock options
- Comprehensive health and welfare benefits
- Monthly stipends for wellness and educational expenses
- Generous paid time off and holiday allowances
- Eligibility for temporary international assignments
- Rare opportunity to help build and operate the world’s first live direct-to-device satellite network
- Exposure to elite talent across software, hardware, chipsets, telecom, satellite, and virtualization domains
Compensation
Competitive compensation packages including a stock option based equity program
Work Arrangement
Remote (Worldwide) — five continents
Team
Open, transparent, inclusive culture blending Silicon Valley, Nordic and South Asia characteristics
Other
- Equal opportunity employer with inclusive culture blending Silicon Valley, Nordic and South Asia characteristics.
- Flexible approach to work.
- Open, transparent, inclusive culture.
- Opportunity to temporarily work abroad.
Not specified