Mindrift.ai: Be the “I” in AI

Mindrift.ai: Be the “I” in AI

Join 10,000+ experts earning $15-50/hr training AI models remotely. Flexible freelance work, weekly payments. No AI experience required. Apply in 5 minutes.

Internet Software & Services

Description

  • Build realistic developer environments with codebases, infrastructure, tickets, documentation, and conversation history.
  • Design tasks from intermediate states of simulated environments and define what counts as a solved solution.
  • Write tests that verify agent solutions and accurately distinguish correct from incorrect approaches.
  • Review agent outputs, analyze failures, and refine tasks and tests based on QA feedback.
  • Iterate on evaluation criteria to make tasks fair, robust, and challenging for AI coding agents.
  • Ensure tasks remain solvable by an AI agent while still testing real-world developer judgment.
  • Submit completed tasks by the deadline and meet the listed acceptance criteria.

Requirements

  • 5+ years of software development experience.
  • Experience with the core stack: Python (FastAPI), JavaScript/TypeScript (React), Docker, Postgres, Kafka, and Redis.
  • Experience writing functional and integration tests.
  • English proficiency at B2 level or higher.
  • Ability to work on a project-based basis rather than permanent employment.
  • Ability to complete tasks that are estimated at around 20 hours each, with flexibility to choose when and how to work.

Benefits

  • Up to $50/hr equivalent compensation, depending on level and pace.
  • Project-based flexible schedule with no fixed working hours.
  • Estimated 20 hours per task, allowing candidates to choose when and how to work.
  • Opportunity to work on AI evaluation projects for leading tech companies.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

QA Contractor

FUTURESEARCH Internet Software & Services

FUTURESEARCH is hiring a meticulous quality checker to verify research evaluation tasks used for training and benchmarking frontier AI models.

13 hours, 49 minutes ago

Expert Rater Project - Dutch (Netherlands)

Welo Global Professional Services

Welo Data is hiring a Dutch-speaking Expert Rater in the Netherlands to support rater quality, client test rounds, and AI advertising data evaluation on a short-term freelance project.

1 day, 13 hours ago

Hydrus -Voice Contributor — English (US)

Welo Global Professional Services

Welo Data is hiring a freelance native-level English (US) speaker to record conversational customer-service audio for a remote AI voice project.

1 day, 13 hours ago

AI / ML Consultant

Crosslake 251-1K Capital Markets

Crosslake is hiring a US-based remote consultant to help technology and private equity clients evaluate, improve, and implement AI and software-related initiatives across the investment lifecycle.

Agile Computer Vision Generative AI LLM Machine Learning MLOps NLP Statistics
1 day, 14 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers