Prolific

Prolific

Prolific is a platform that enables researchers to quickly find trustworthy research participants. With a pool of over 120,000 active and verified participants, Prolific ensures high-quality responses through continuous monitoring and engagement. The p...

Professional Services
51-250
Founded 1997
$0M raised

Description

  • Evaluate AI-generated explanations of model architectures, loss functions, and backpropagation for technical accuracy.
  • Audit ML code and notebooks, including training loops, data preprocessing scripts, and model evaluations, for correctness and efficiency.
  • Provide human feedback to refine RLHF frameworks and improve model alignment, safety, and helpfulness.
  • Analyze model reasoning on complex chain-of-thought prompts and identify where reasoning breaks down.
  • Benchmark and compare model outputs using technical taxonomies and performance metrics.
  • Review technical content for hallucinations, biased outputs, and logical failures.
  • Help train and evaluate next-generation LLMs through paid expert tasks.

Requirements

  • BS, MS, or PhD in Computer Science, Artificial Intelligence, Robotics, or a related quantitative field with a focus on Machine Learning.
  • Experience building, deploying, or fine-tuning ML models in a production environment.
  • Professional-level understanding of neural network architectures such as Transformers, CNNs, and RNNs, plus optimization techniques.
  • Hands-on experience with Prompt Engineering, RLHF, or RAG workflows.
  • Ability to audit complex model logic, identify training data contamination, and evaluate mathematical proofs behind ML algorithms.
  • High attention to detail in spotting hallucinations, biased outputs, and logical failures in AI-generated technical content.
  • Expert proficiency in PyTorch or TensorFlow/Keras.
  • Advanced Python experience, including NumPy, Pandas, Scikit-learn, and Hugging Face Transformers.
  • Experience with AWS SageMaker, Google Cloud Vertex AI, Weights & Biases, or LangChain.
  • Familiarity with Pinecone, Milvus, or Weaviate for RAG evaluation.

Benefits

  • Pay rates up to $80 per hour for researchers seeking your skills.
  • Flexible hours.
  • Ability to work from home.
  • Quick 10- to 15-minute assessment and fast onboarding if successful.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Agentic AI Engineer

Resilient Co 11-50 Professional Services

Aon is seeking a senior Agentic AI Engineer to build an orchestration layer that modernizes internal automation through Google Agentic Orchestration, MCP integrations, and LLM-driven agent-to-system interactions.

Azure LLM
1 day, 23 hours ago

Computer Vision Expert

Weekday 11-50 Construction & Engineering

An independent contractor role with a client-focused AI research initiative where a Computer Vision Expert helps design vision tasks, evaluate model performance, and improve frontier AI systems.

Computer Vision Deep Learning Generative AI Python PyTorch TensorFlow
1 week ago

LLM Engineer - Freelancer

Monterail 251-1K Internet Software & Services

Monterail is building a freelance network of LLM Engineers to add practical AI features into existing production products.

Azure LLM Node.js Python Ruby
1 week, 1 day ago

Founding Director of Engineering - AI-Native Venture Studio

AI Acquisition 51-200 Business Consulting and Services

An AI-native venture studio is hiring a player-coach to lead product delivery across multiple software products and help run a small team of humans plus AI coding agents.

Docker GraphQL JavaScript Kubernetes LLM Microservices Python REST API Terraform
1 week, 6 days ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers