Prolific

Prolific

Prolific is a platform that enables researchers to quickly find trustworthy research participants. With a pool of over 120,000 active and verified participants, Prolific ensures high-quality responses through continuous monitoring and engagement. The p...

Professional Services
51-250
Founded 1997
$0M raised

Description

  • Evaluate AI-generated explanations of model architectures, loss functions, and backpropagation for technical accuracy.
  • Audit ML code and notebooks, including training loops, data preprocessing scripts, and model evaluations, for correctness and efficiency.
  • Provide human feedback to refine RLHF frameworks and improve model alignment, safety, and helpfulness.
  • Analyze model reasoning on complex chain-of-thought prompts and identify where reasoning breaks down.
  • Benchmark and compare model outputs using technical taxonomies and performance metrics.
  • Review technical content for hallucinations, biased outputs, and logical failures.
  • Help train and evaluate next-generation LLMs through paid expert tasks.

Requirements

  • BS, MS, or PhD in Computer Science, Artificial Intelligence, Robotics, or a related quantitative field with a focus on Machine Learning.
  • Experience building, deploying, or fine-tuning ML models in a production environment.
  • Professional-level understanding of neural network architectures such as Transformers, CNNs, and RNNs, plus optimization techniques.
  • Hands-on experience with Prompt Engineering, RLHF, or RAG workflows.
  • Ability to audit complex model logic, identify training data contamination, and evaluate mathematical proofs behind ML algorithms.
  • High attention to detail in spotting hallucinations, biased outputs, and logical failures in AI-generated technical content.
  • Expert proficiency in PyTorch or TensorFlow/Keras.
  • Advanced Python experience, including NumPy, Pandas, Scikit-learn, and Hugging Face Transformers.
  • Experience with AWS SageMaker, Google Cloud Vertex AI, Weights & Biases, or LangChain.
  • Familiarity with Pinecone, Milvus, or Weaviate for RAG evaluation.

Benefits

  • Pay rates up to $80 per hour for researchers seeking your skills.
  • Flexible hours.
  • Ability to work from home.
  • Quick 10- to 15-minute assessment and fast onboarding if successful.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

AI Developer

Capco 5K-10K Internet Software & Services

Capco Poland is hiring a Poland-based Cloud and Platform Engineering consultant to help design and implement cloud, AI, and enterprise integration solutions for financial services clients.

AWS Azure Envoy Generative AI Grafana Kubernetes LLM Splunk Terraform
15 hours, 38 minutes ago

Agentic AI Engineer

Resilient Co 11-50 Professional Services

Aon is seeking a senior Agentic AI Engineer to build an orchestration layer that modernizes internal automation through Google Agentic Orchestration, MCP integrations, and LLM-driven agent-to-system interactions.

Azure LLM
3 days, 16 hours ago

Computer Vision Expert

Weekday 11-50 Construction & Engineering

An independent contractor role with a client-focused AI research initiative where a Computer Vision Expert helps design vision tasks, evaluate model performance, and improve frontier AI systems.

Computer Vision Deep Learning Generative AI Python PyTorch TensorFlow
1 week, 2 days ago

LLM Engineer - Freelancer

Monterail 251-1K Internet Software & Services

Monterail is building a freelance network of LLM Engineers to add practical AI features into existing production products.

Azure LLM Node.js Python Ruby
1 week, 3 days ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers