Amazon's Artificial General Intelligence (AGI) Data Services organization is focused on developing diverse datasets to train and evaluate AI models. The Language Engineer will support the creation of complex datasets and collaborate with cross-functional teams to ensure the effectiveness of AI systems.
Responsibilities:
- Design complex data collections with human participants in response to science needs: author instructions, define and implement quality targets and mechanisms, provide day-to-day coordination of data collection efforts (including planning, scheduling, and reporting), and be responsible for the final deliverables
- Design and conduct complex data creation tasks using synthetic and model-based data generation methods, following state-of-the-art approaches
- Analyze and extract insights from large amounts of data
- Build tools or tool prototypes for data analysis or data creation, using Python or another scripting language
- Use modeling tools to bootstrap or test new AI functionalities
- Collaborate with scientists, software engineers, and other data creators to evaluate performance of AI models