Meet Nikita: ML Systems Researcher

Verified expert · 2 min read
Portrait of Nikita Annenkov, ML Systems Researcher

Nikita Annenkov

ML Systems Researcher · United States

Turned precision and judgment into a track record.

Read the full story

Where precision is the point

Nikita spends his weeks in the space between ML systems and the people trying to evaluate them. Coding-agent benchmarks, scientific-computing debugging, evaluations for software that has to operate other software. It is the kind of work where the difference between a good task and a great one is not difficulty, it is rigor.

When he found Terac, the pitch resonated immediately: work where careful reasoning matters more than throughput. He started with a few tasks and kept going.

What makes a task meaningful

"The most interesting part so far has been how much the work rewards precision," he says. "A good task is not just hard, it has to be fair, deterministic, reviewable, and calibrated so that an agent's failure actually means something."

The word he keeps coming back to is calibrated. A task that is too easy tells you nothing. One that is impossible tells you nothing different. The signal lives somewhere in the middle, and getting there requires far more judgment than most people expect.

Inside Computer Control

Nikita contributes across Terac's Computer Control track: ML systems tasks where agents operate real software, evaluations for coding agents, and scientific-computing workflows where correctness often depends on expert judgment rather than pattern matching.

The work asks for two things at once: constructing tasks an agent might fail in instructive ways, and reviewing outputs with the kind of rigor usually reserved for research environments.

Multiple accepted evaluations later

Since joining the platform, Nikita has contributed accepted evaluations across machine learning, data science, and model-training categories. He says the review process itself became part of the appeal: accepted work feels earned, and rejected work usually comes back with actionable feedback.

The platform, he says, stays out of the way in the right places. Tasks are self-contained, expectations are clear, and the asynchronous workflow fits naturally around his existing engineering and research schedule.

For ML researchers and technical evaluators considering the work, his advice is simple: bring the same standards you would bring to a paper or benchmark. The projects reward depth, rigor, and clear reasoning far more than speed.

Join the panel

Have a story worth telling?

Terac runs on verified experts. Apply once, get verified, and start matching to research and AI work in your field.

  1. 01Apply
  2. 02Get verified
  3. 03Start matching

Share your story

Tell us what you have worked on and why it mattered. A few lines is enough to start a conversation.

Email stories@terac.comOr start the full application
© 2026 All Rights Reserved by Terac