← All jobs

Remote

AI Evaluation PhD Expert

About this role

Seeking a full-time STEM PhD Expert in AI Evaluation to remotely assess and enhance advanced AI models through evaluating AI-generated content, crafting relevant questions, and ranking responses, with a short-term engagement lasting until the end of June and potential for extension. Key Responsibilities Assessing the factuality and relevance of domain-specific text produced by AI models Crafting and answering questions related to Machine Learning and AI Evaluating and ranking domain-specific responses generated by AI models Required Qualifications A PhD (completed or in final stages) in Machine Learning/AI, Computer Science, Engineering, Statistics, or a closely related quantitative field Strong analytical and critical-thinking skills to identify subtle errors in technical reasoning Fluent written English with the ability to communicate complex ideas clearly

Source listing: virtualvocations_main