Use advanced physics and analytical reasoning to help improve large language models. You will probe model limitations and explain complex concepts through clear, step-by-step reasoning.
What You Will Do
- Design and solve challenging physics problems that test large language models.
- Create high-quality, step-by-step solutions with detailed reasoning.
- Work with AI researchers to align tasks with evaluation goals, including symbolic manipulation, abstraction and multi-step reasoning.
- Help define evaluation benchmarks across physics curricula from early undergraduate through PhD-level topics.
- Construct novel scenarios, evaluate reasoning pathways and provide structured feedback and detailed annotations.
What You Will Bring
- * A PhD in Physics, Applied Physics or an equivalent technical field.
- Strong physics knowledge spanning advanced engineering-entrance through graduate and PhD-level concepts.
- Strong analytical, research and problem-solving skills with excellent English comprehension.
- Exceptional structured written communication and constructive feedback skills.
- Lateral thinking and the ability to work independently in a fast-paced remote setting.
- A personal desktop or laptop and stable high-speed internet.
Experience in AI evaluation, data annotation, content review, quality assurance or a related analytical role is preferred, not required.
PROJECT DETAILS
Expected engagement length is 12 weeks, with an immediate start. A minimum of 20 hours per week, up to 40 hours per week, is required. Four hours of daily Pacific Time overlap is required. Selection includes an assessment and a delivery review.
PAYMENT
Turing pays $80 USD directly for each approved task.