Prolific

Prolific

Prolific is a platform that enables researchers to quickly find trustworthy research participants. With a pool of over 120,000 active and verified participants, Prolific ensures high-quality responses through continuous monitoring and engagement. The p...

Professional Services
51-250
Founded 1997
$0M raised

Description

  • Evaluate AI-generated explanations of model architectures, loss functions, and backpropagation for technical accuracy.
  • Audit ML code and notebooks, including training loops, data preprocessing scripts, and model evaluations, for correctness and efficiency.
  • Provide human feedback to refine RLHF frameworks and improve model alignment, safety, and helpfulness.
  • Analyze model reasoning on complex chain-of-thought prompts and identify where reasoning breaks down.
  • Benchmark and compare model outputs using technical taxonomies and performance metrics.
  • Review technical content for hallucinations, biased outputs, and logical failures.
  • Help train and evaluate next-generation LLMs through paid expert tasks.

Requirements

  • BS, MS, or PhD in Computer Science, Artificial Intelligence, Robotics, or a related quantitative field with a focus on Machine Learning.
  • Experience building, deploying, or fine-tuning ML models in a production environment.
  • Professional-level understanding of neural network architectures such as Transformers, CNNs, and RNNs, plus optimization techniques.
  • Hands-on experience with Prompt Engineering, RLHF, or RAG workflows.
  • Ability to audit complex model logic, identify training data contamination, and evaluate mathematical proofs behind ML algorithms.
  • High attention to detail in spotting hallucinations, biased outputs, and logical failures in AI-generated technical content.
  • Expert proficiency in PyTorch or TensorFlow/Keras.
  • Advanced Python experience, including NumPy, Pandas, Scikit-learn, and Hugging Face Transformers.
  • Experience with AWS SageMaker, Google Cloud Vertex AI, Weights & Biases, or LangChain.
  • Familiarity with Pinecone, Milvus, or Weaviate for RAG evaluation.

Benefits

  • Pay rates up to $80 per hour for researchers seeking your skills.
  • Flexible hours.
  • Ability to work from home.
  • Quick 10- to 15-minute assessment and fast onboarding if successful.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

AI Engineer (Freelance)

Multiplica 251-1K Internet Software & Services

Multiplica busca un AI Engineer freelance para diseñar e implementar soluciones de inteligencia artificial y automatización en una plataforma propia conectada con sistemas internos y externos.

GitHub GitLab JavaScript Python React
1 day, 13 hours ago

Chatbot Developer (WhatsApp, Telegram, Discord) - Freelance

Mindrift.ai: Be the “I” in AI Internet Software & Services

Mindrift is hiring a part-time freelance Bot Developer for the Tendem project to build and refine conversational bots and messaging-platform integrations in a hybrid AI + human workflow.

Docker Node.js OAuth Python REST API Serverless
2 days, 13 hours ago

Chatbot Developer (WhatsApp, Telegram, Discord) - Freelance

Mindrift.ai: Be the “I” in AI Internet Software & Services

Mindrift is hiring a freelance, part-time remote Bot Developer for the Tendem project to build and refine conversational bots and messaging-platform integrations in a hybrid AI + human workflow.

Docker Node.js OAuth Python REST API Serverless
2 days, 13 hours ago

Desarrollador Full stack Senior

Coderio 51-250 Internet Software & Services

Coderio busca un Desarrollador Fullstack Senior para un equipo internacional que desarrolla aplicaciones web de արտադրcción con componentes de IA y orquestación de LLMs para múltiples mercados.

Next.js PostgreSQL Tailwind CSS TypeScript
3 days, 13 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers