Responsibilities
- Keep current with advancements in code large language models, agent architectures, and related research areas, integrating innovative concepts into production systems.
- Develop and execute scalable training methodologies for code models and deploy agent frameworks for inference and sampling, working closely with pretraining teams, generating supervised fine-tuning trajectories, and advancing reinforcement learning algorithms.
- Improve performance on existing evaluation benchmarks and create new benchmarks aligned with enterprise customer requirements.
- Lead high-impact experiments on advanced compute platforms to expand the capabilities of state-of-the-art language models.
Benefits
- Weekly lunch allowance of $75 or £75, or equivalent in local currency.
- Comprehensive medical and dental coverage, including dedicated funding for mental health support.
- Retirement savings matching through RRSP, 401K, or pension plan contributions.
- Full salary top-up for parental leave up to six months, available to either parent.
- Annual enrichment fund covering arts and culture, fitness and wellness, personal time, workspace upgrades, and professional development such as courses, conferences, and coaching.
- Six weeks of paid vacation annually, equivalent to 30 working days.
- Budget for travel to other offices if remote, plus attendance at an annual company-wide offsite event.
- Co-working space benefit to facilitate local collaboration in your city.
- One-time $500 stipend for home office setup.
Work Arrangement
Hybrid — London, Toronto, New York, San Francisco
Team
Team operates across ET to CET time zones, requiring alignment with these hours for effective collaboration.
Other
- Candidates should be located in time zones compatible with ET to CET to ensure seamless team coordination.
- Applicant screening may include AI-powered tools to evaluate alignment with role criteria.
- Be cautious of fraudulent schemes: no payment or third-party service fees are ever required during the hiring process.