Top Daily Deal: Cometeer5% offShop Now
STEM & Research

Mathematical Reasoning & LLM Evaluation

Mathematics Expert (Master's/Ph.D.): Shaping Frontier AI at Turing

Turing··3 min read
Location
Remote — Global
Engagement
Contractor
Apply at Turing

Earns 25 points on this device — once per role per day

Applications are handled by Turing on their own site. Dealuxe is not the employer and does not screen applicants.

Mathematical Reasoning & LLM Evaluation
Advanced AI Research & Mathematics

Engage your advanced mathematical training to evaluate, train, and future-proof large language models with the world's leading AI research accelerator.

The rapid expansion of artificial intelligence has moved far beyond simple pattern matching. Today, frontier AI research laboratories require deep logical reasoning, complex mathematical problem-solving, and rigorous scientific validation to ensure advanced models operate reliably. Bridging this gap demands elite human intellect—scholars and researchers who can translate abstract theorems and rigorous proofs into structured evaluation benchmarks.

Based in San Francisco, California, Turing stands as the world's leading research accelerator for frontier AI labs and a trusted partner for global enterprises deploying advanced AI systems. Turing empowers customers by accelerating frontier research with high-quality data, advanced training pipelines, and elite STEM researchers, while helping enterprises transform AI proofs of concept into robust proprietary intelligence.

Role Focus
Mathematical Reasoning & LLM Evaluation
Location
Remote (Global)
Commitment Options
20, 30, or 40 Hours / Week

About the Role: You Are a Fit If...

In this specialized role, you will work on cutting-edge projects that improve and evaluate large language models by applying advanced mathematical reasoning, problem-solving, computational thinking, and clear written communication. The ideal candidate possesses a robust foundation in mathematics—particularly at the level expected in rigorous engineering entrance examinations as well as graduate or PhD-level coursework.

You are an ideal fit for this engagement if you excel at breaking down intricate mathematical concepts into simple, clear explanations while working with high efficiency. This role also features computational projects where you design precise, closed-ended prompts, write reliable Python code, verify numerical answers, and provide comprehensive rationales.

Key Required Skills & Experience:

  • Core Mathematical Domains: Deep proficiency in Algebra, Calculus & Analysis, Geometry & Topology.
  • Analytical & Research Skills: Strong academic research background with structured, logical problem-solving abilities.
  • Feedback & Annotation: Exceptional ability to provide constructive feedback, identify mathematical errors, and write detailed annotations.
  • Technical Proficiency: Familiarity with Python programming using approved scientific libraries, plus theorem-prover experience in Lean.
  • Communication & Independence: Self-motivated remote operator with superior written explanation skills and the ability to work autonomously.

What a Day-to-Day Look Like

Working as a Mathematics Expert with Turing is intellectually stimulating and dynamically varied. Your daily responsibilities include:

  • Problem Generation: Designing original, highly challenging mathematics problems that test the reasoning limits of large language models in multi-step, abstract, and proof-based settings.
  • Independent Solutions: Solving complex problems independently and drafting detailed, logically structured solutions with clear justifications.
  • Model Evaluation: Reviewing model-generated outputs, isolating mathematical errors or missing arguments, and delivering precise annotations and corrections.
  • Benchmark Development: Contributing to the creation of new evaluation benchmarks spanning early undergraduate to PhD-level mathematics curricula.
  • Computational & Theorem Work: Developing and validating Python-based solutions for computational tasks and translating mathematical problems into formal language using Lean.

Perks of Freelancing With Turing

Contracting with Turing offers elite professionals unmatched career flexibility and engagement quality:

  • Fully Remote Autonomy: Work from anywhere in the world with complete schedule flexibility.
  • Cutting-Edge Collaboration: Direct involvement with premier LLM research labs shaping the future of artificial intelligence.
  • Career Future-Proofing: Master how to leverage AI tools effectively, enhancing your analytical workflow in an AI-first economy.
  • Extension Potential: Opportunities for contract extensions based on consistent performance and ongoing project needs.

Engagement Details & Onboarding Excellence

This engagement operates on flexible contractor terms (no medical or paid leave included). Candidates can select from three weekly time commitments: 20 hours/week, 30 hours/week, or 40 hours/week, requiring at least 4 hours per day with a minimum of 4 hours of overlap with Pacific Standard Time (PST).

To ensure long-term stability and unlock access to advanced, higher-tier tasks, successful applicants must complete the thorough onboarding process and complete their initial 10 hours of work with precision and dedication.


Ready to Shape the Future of AI?

Submit your qualifications and begin your application process directly through the secure portal.
Join top-tier mathematicians advancing artificial intelligence at Turing.

Apply as Mathematics Expert
Secure application powered by Turing. Remote contractor position.

What the work is

  • Problem Generation: Designing original, highly challenging mathematics problems that test the reasoning limits of large language models in multi-step, abstract, and proof-based settings.
  • Independent Solutions: Solving complex problems independently and drafting detailed, logically structured solutions with clear justifications.
  • Model Evaluation: Reviewing model-generated outputs, isolating mathematical errors or missing arguments, and delivering precise annotations and corrections.
  • Benchmark Development: Contributing to the creation of new evaluation benchmarks spanning early undergraduate to PhD-level mathematics curricula.
  • Computational & Theorem Work: Developing and validating Python-based solutions for computational tasks and translating mathematical problems into formal language using Lean.

Ready to apply for Mathematical Reasoning & LLM Evaluation?

The application is on Turing's own site and takes a few minutes.

Apply at Turing

Earns 25 points on this device — once per role per day

Dealuxe is not the employer, does not set the pay or the hiring terms, and cannot guarantee a role is still open. If you complete a purchase or form, we may earn a small commission at no extra cost to you.

Following an offer here banks 10 points on this device — once per page, within the 500 points a day anything on the site can earn.

Ad Disclosure: the application link is a referral link.

While you job-hunt, save on the brands you already use

Browse 2,287 vetted brands with live commission offers and exclusive deals — every one open to everyone, no sign-in needed.

Explore all brand categories →
Reading reward0% · worth 5 pts

Scroll through the piece and stay a moment. Reading pays 5 points and sharing pays 50. Following the apply link pays 25, and buying coins pays back 25 points a dollar.

Get stories in your inbox

New brand drops, deal breakdowns and the best of the Journal, straight from The Storefront Blog. Free forever — and subscribing pays you 25 points.

Go paid, earn 75

Drop your email in the box above, then bank the bonus. A paid plan pays 75 — three times the free tier — plus every paid-only post.

Copies the link with your caption. Grab a username to bank points across devices.

Boost this listing

See what's trending

Trade the points you've earned to push this up Trending and the homepage, where more readers will find it.

Be the first to comment

Attach a gift:
Comments earn points once per article per day.

Loading comments…

Similar roles

AI Prompt & Policy Specialist
STEM & Research

AI Prompt & Policy Specialist

Turing

An exhaustive guide to working with Turing, navigating the AI Prompt & Policy Specialist position, required technical stacks, day-to-day workflows, and the onboarding pipeline.

Location
Remote — Global
Posted
Mathematics Research Specialist
STEM & Research

Mathematics Research Specialist

Turing

Discover how advanced algebraic geometry, formal proof development, and high-level mathematical frameworks are driving the future of artificial intelligence with Turing.

Location
Remote — Global
Posted
Small Business Owners (AI Response Evaluation) - Korean Business Document
STEM & Research

Small Business Owners (AI Response Evaluation) - Korean Business Document

Turing

Shape the future of enterprise artificial intelligence by evaluating chatbot performance for real-world small business operations.

Location
Remote — Global
Posted
Small Business Owners (AI Response Evaluation)
STEM & Research

Small Business Owners (AI Response Evaluation)

Turing

Shape the future of enterprise artificial intelligence by leveraging your operational business expertise and Spanish documentation skills.

Location
Remote — Global
Posted
Physics Expert
STEM & Research

Physics Expert

Turing

Discover how graduate-level physicists and researchers are training next-generation large language models through rigorous scientific benchmarking and problem architecture.

Location
Remote — Global
Posted
Chemistry Expert (PhD / Master's)
STEM & Research

Chemistry Expert (PhD / Master's)

Turing

Discover how PhDs and Master's-level chemists are leveraging their expertise to train next-generation large language models in advanced STEM reasoning.

Location
Remote — Global
Posted

More stem & research listings

All stem & research roles →

Other roles at Turing

All Turing roles →

Recommended For You

Explore curated deals from top brands across every category.