Turing is a rapidly growing AI company focused on advancing and deploying powerful AI systems. The company collaborates with leading AI labs to push the boundaries of model capabilities in areas such as reasoning, coding, multimodality, multilinguality, STEM, and frontier knowledge. These advancements are then applied to develop real-world AI solutions that address critical business challenges.
In this role, you will engage in cutting-edge projects aimed at fine-tuning large language models by applying your strong analytical skills and deep expertise in mathematics. The ideal candidate will have a solid foundation in mathematics, ranging from advanced engineering entrance level concepts to graduate and PhD-level topics. You will be responsible for dissecting complex phenomena into clear, logical steps and play a key role in identifying model limitations. This position offers an excellent opportunity to work with frontier AI tools and future-proof your career in a fast-evolving field.
Key Responsibilities:
- Problem Creation & Solutioning: Design and solve challenging physics problems that push the limits of large language models.
- Authoring Gold-Standard Data: Develop clear, high-quality, step-by-step solutions with detailed reasoning and articulation.
- Research Collaboration: Partner with LLM researchers to ensure task designs align with evaluation goals, focusing on areas where models face difficulties such as symbolic manipulation, abstraction, and multi-step reasoning.
- Benchmark Definition: Assist in defining and building new evaluation benchmarks covering physics curricula from early undergraduate to PhD-level topics.
Required Qualifications:
- Strong analytical, research, and problem-solving skills coupled with excellent English comprehension.
- Exceptional structured written communication skills, with the ability to provide constructive feedback and detailed annotations in a remote work environment.
- Strong lateral thinking abilities to create novel scenarios and evaluate complex reasoning pathways.
- Self-motivated and capable of working independently in a fast-paced, remote-first setup.
- Access to a personal desktop or laptop with a stable, high-speed internet connection.
- Educational background or Doctorate in Mathematics or an equivalent technical field.
Preferred Qualifications and Benefits:
- Experience in AI evaluation, data annotation, content review, quality assurance, or related analytical roles is preferred but not mandatory.
- Flexible contractor engagement for a 12-week period, requiring a commitment of at least 4 hours per day and up to 40 hours per week, with a minimum of 4 hours overlapping with Pacific Standard Time (PST).
Candidates who are shortlisted will be invited to complete a Job Interest Form. Those selected for final consideration will be contacted with further steps, including onboarding instructions. This role offers a unique chance to contribute to setting new benchmarks in AI capabilities within advanced physical sciences while working remotely with a global team.