Responsibilities
- Create datasets and moderation models to assess large language models and end-to-end systems for content safety and machine learning fairness, including text-to-text and multimodal-to-text models.
- Build training datasets for large language models using supervised fine-tuning and reinforcement learning methods, targeting content safety, fairness, and security.
- Investigate and apply advanced methods to detect and reduce bias in language models and associated systems.
- Establish and monitor key performance indicators related to responsible use and behavior of large language models.
- Adhere to best practices in automation, system monitoring, scalability, and safety protocols.
- Contribute to shared code repositories and build safety-focused tools to enhance machine learning team productivity.
- Perform data preprocessing tasks by working with data scientists and engineers to gather, clean, and transform complex, large-scale datasets.
- Carry out exploratory data analysis to discover meaningful patterns and improve model effectiveness.
- Work with cross-functional teams including product engineers, data scientists, and analysts to interpret business needs and design appropriate machine learning solutions.
Benefits
- Comprehensive benefits package
- Paid time off
- Positive and supportive workplace culture
Compensation
Competitive pay ranging from $90 to $130 per hour, depending on experience, education, location, and other factors
Work Arrangement
Hybrid – Santa Clara, CA
Other
- This is a full-time W-2 contract position.
- Hourly rate is competitive and ranges from $90 to $130 based on experience, education, location, and related factors.
Not specified