Mindrift.ai: Be the “I” in AI

Mindrift.ai: Be the “I” in AI

Join 10,000+ experts earning $15-50/hr training AI models remotely. Flexible freelance work, weekly payments. No AI experience required. Apply in 5 minutes.

Internet Software & Services

Description

  • Build realistic virtual developer environments with codebases, infrastructure, tickets, documentation, and conversation history.
  • Design tasks from intermediate states of those environments, including prompts and clear definitions of what counts as solved.
  • Write tests that validate agent solutions by accepting all correct approaches and rejecting incorrect ones.
  • Review agent outputs, analyze failures, and iterate on tasks and tests based on QA feedback.
  • Refine evaluation criteria to ensure tasks are fair, robust, and solvable by an AI agent.
  • Support the assessment of how well AI coding agents handle realistic development tasks.
  • Collaborate through the project workflow from qualification to task completion and submission by deadline.

Requirements

  • 5+ years of software development experience.
  • Experience with Python and FastAPI.
  • Experience with JavaScript or TypeScript, including React.
  • Experience with Docker, Postgres, Kafka, and Redis.
  • Experience writing functional and integration tests.
  • English proficiency at B2 level or higher.
  • Ability to work on a project-based basis rather than as a permanent employee.
  • Ability to complete tasks estimated at around 20 hours each and submit them by the deadline.

Benefits

  • Up to $50/hr equivalent compensation, depending on level and pace.
  • Flexible schedule: choose when and how to work.
  • Project-based work with estimated 20-hour tasks.
  • Opportunity to work on AI evaluation tasks for leading tech companies.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

QA Contractor

FUTURESEARCH Internet Software & Services

FUTURESEARCH is hiring a meticulous quality checker to verify research evaluation tasks used for training and benchmarking frontier AI models.

13 hours, 48 minutes ago

Expert Rater Project - Dutch (Netherlands)

Welo Global Professional Services

Welo Data is hiring a Dutch-speaking Expert Rater in the Netherlands to support rater quality, client test rounds, and AI advertising data evaluation on a short-term freelance project.

1 day, 13 hours ago

Hydrus -Voice Contributor — English (US)

Welo Global Professional Services

Welo Data is hiring a freelance native-level English (US) speaker to record conversational customer-service audio for a remote AI voice project.

1 day, 13 hours ago

AI / ML Consultant

Crosslake 251-1K Capital Markets

Crosslake is hiring a US-based remote consultant to help technology and private equity clients evaluate, improve, and implement AI and software-related initiatives across the investment lifecycle.

Agile Computer Vision Generative AI LLM Machine Learning MLOps NLP Statistics
1 day, 14 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers