Turing is a rapidly growing AI company focused on advancing and deploying powerful AI systems. The company collaborates with leading AI labs to push the boundaries of model capabilities in areas such as reasoning, coding, multimodality, multilinguality, STEM, and frontier knowledge. These advancements are then applied to build real-world AI solutions that address critical business challenges. In this role, you will engage in pioneering projects aimed at fine-tuning large language models by applying your strong analytical and mathematical expertise. The ideal candidate will have a solid foundation in mathematics, ranging from advanced engineering entrance-level concepts to graduate and PhD-level topics, and the ability to deconstruct complex problems into clear, logical steps. You will contribute directly to identifying model limitations while gaining experience with cutting-edge AI tools to enhance your career prospects.
Key Responsibilities:
- Problem Creation & Solutioning: Design and solve challenging physics problems that push the limits of large language models.
- Authoring Gold-Standard Data: Develop clear, high-quality, step-by-step solutions with detailed reasoning and explanations.
- Research Collaboration: Partner with LLM researchers to ensure task designs align with evaluation objectives, focusing on areas where models face difficulties such as symbolic manipulation, abstraction, and multi-step reasoning.
- Benchmark Definition: Assist in defining and constructing new evaluation benchmarks covering physics topics from early undergraduate to PhD levels.
Required Qualifications:
- Educational background or Doctorate in Mathematics or an equivalent technical discipline.
- Strong analytical, research, and problem-solving skills, combined with excellent English comprehension.
- Exceptional written communication skills, capable of providing structured, constructive feedback and detailed annotations in a remote work environment.
- Ability to think laterally and creatively to generate novel scenarios and assess complex reasoning pathways.
- Self-motivated and able to work independently within a fast-paced, remote-first setting.
- Access to a personal desktop or laptop with a stable, high-speed internet connection.
Preferred Qualifications and Benefits:
- Experience in AI evaluation, data annotation, content review, quality assurance, or related analytical roles is advantageous but not mandatory.
- Commitment of at least 4 hours per day, up to 40 hours per week, with a minimum of 4 hours overlapping with Pacific Standard Time (PST).
- Engagement as a contractor for a duration of 12 weeks.
The evaluation process includes sending a Job Interest Form to shortlisted candidates, followed by communication of next steps and onboarding details for those selected. This opportunity offers the chance to work at the forefront of AI research and development, contributing to projects that shape the future of intelligent systems.