Apple Inc. is seeking an expert in evaluating machine learning and deep learning models, including foundation models and multimodal systems.
You will craft robust evaluation frameworks, using traditional statistics and LLMs as judges to assess tasks like summarization and multimodal generation. The role requires strong Python expertise and a deep understanding of statistical methods, data quality, and model robustness, collaborating with ML engineers, data scientists, and ML infrastructure teams
#J-18808-Ljbffr