Scientific Computing AI Evaluation Expert

Posted yesterday

opentrain aiNew York (NY)

SENIORITY

Senior

Apply

About the role

The work You will turn mathematical and research workflows into self-contained, terminal-based tasks that test AI agents. Your work will cover numerical analysis, optimization, statistics, mathematical modeling, probability, dynamical systems, numerical integration, differential equations, matrix computation, stochastic modeling, and algorithm analysis. Build tasks with datasets, equations, model definitions, constraints, and expected outputs. Implement expert solutions in Python, R, Julia, C/C++, Bash, or another relevant language. Create grading criteria covering accuracy, convergence, complexity, feasibility, and mathematical correctness. Define tolerances, stopping criteria, stability requirements, and reproducibility controls. Write automated tests for edge cases and alternative valid implementations. Debug floating-point precision, solver, conditioning, convergence, and performance issues. Document assumptions, mathematical formulations, expected outputs, and known limitations. What it pays and takes This is a remote contractor assignment for five weeks. The role requires advanced technical experience in mathematics, statistics, or a closely related field, along with the ability to independently implement, test, and validate computational algorithms.
Pay: $300 per approved task.
Schedule: At least 6 hours per day and 40 hours per week. Time overlap: At least 4 hours overlapping with Pacific Standard Time.
Location: Remote and open worldwide.
Language: English. Background: Ph. D., postdoctoral experience, or equivalent advanced technical experience in mathematics, statistics, or a closely related discipline. Programming: Strong ability in Python, R, Julia, C/C++, Bash, or another relevant language. Environment: Experience working in Linux or terminal-based environments. Expertise: Practical experience in numerical methods, optimization, statistics, mathematical modeling, or scientific computation. About AI training work AI training is the human work behind systems that learn from examples, including writing, testing, and evaluating model outputs. This role uses specialized mathematical and programming knowledge to create reliable evaluations that show whether AI agents can solve computational problems correctly.

Before you apply

Applying takes about a minute. These four things decide how fast it moves after that.

Your profile is current

It's what we read first. Occupations, seniority and locations matter more than a long history.

Two examples you can talk through

Not a portfolio — just two pieces of work where you can explain the decisions and what you'd change.

A number in mind

What you're on now and what would make you move. We negotiate better when we know both.

Your notice period

Employers plan around it, and it's the question that stalls offers most often.

Once you apply, someone reads it and calls you before anything reaches the employer — usually within two working days.

More like this