Responsibilities
- Accelerate researchers by taking on the heavy parts of large-scale ML pipelines and building robust tools.
- Interface cutting-edge research with production: integrate checkpoints, streamline evaluation, and expose APIs.
- Conduct experiments on the latest deep-learning techniques (sparsified 70 B + runs, distributed training on thousands of GPUs).
- Design, implement and benchmark ML algorithms; write clear, efficient code in Python.
- Deliver prototypes that become production-grade components for Le Chat and our enterprise API.
Requirements
- Master’s or PhD in Computer Science (or equivalent proven track record).
- 4 + years working on large-scale ML codebases.
- Hands-on with PyTorch, JAX or TensorFlow; comfortable with distributed training (DeepSpeed / FSDP / SLURM / K8s).
- Experience in deep learning, NLP or LLMs.
- Strong software-design instincts: testing, code review, CI/CD.
- Self-starter, low-ego, collaborative.
Nice to Have
- Bonus for CUDA or data-pipeline chops.
Benefits
- Comprehensive benefits package designed to support well-being, growth, and work-life balance.
- Benefits vary by country and may include healthcare coverage, parental leave, retirement plans, relocation support, wellness programs, meal and transportation allowances, and other location-specific perks.
Work Arrangement
Remote (Worldwide) — Europe, North America, Asia, Middle East
Additional Information
- Your privacy matters to us. You can learn more about how we handle your personal data in our Applicant