About the job
Apple is where individual imaginations gather together, committing to the values that lead to great work. Every new product we build, service we create, or experience we deliver is the result of us making each other’s ideas stronger. The diversity of our people and their thinking inspires the innovation that runs through everything we do. When we bring everybody in, we can do the best work of our lives. Here, you’ll do more than join something — you’ll add something.
Responsibilities
Work closely with ML Engineers to understand data annotation needs
Design and manage data annotation processes, including the development of user instructions, annotation pipeline processing, and process improvement
Develop LLM auto-judges and judging criteria for generative AI model evaluation
Analyze collected data annotations to assess and refine LLM auto-judges
Qualifications
Minimum
BA or Master’s degree in Data Science, Statistics, or a quantitative social science field
2+ years of hands-on experience working in survey design and human data annotation
Proficiency in Python
Excellent communication skills
Preferred
PhD in Data Science, Statistics, or a quantitative social science field
Hands-on industry experience with product-focused statistical analysis
Experience working with large-scale multimodal data and data-annotation pipelines
Experience with LLM prompt engineering & prompt optimization
Experience with LLM auto-judges for generative AI model evaluation
A track record of publications or technical presentations in Data Science or a related field
Excellent at cross-functional collaboration