Prolific

Prolific

Prolific is a platform that enables researchers to quickly find trustworthy research participants. With a pool of over 120,000 active and verified participants, Prolific ensures high-quality responses through continuous monitoring and engagement. The p...

Professional Services
51-250
Founded 1997
$0M raised

Description

  • Evaluate AI-generated explanations of model architectures, loss functions, and backpropagation for technical accuracy.
  • Audit ML code and notebooks, including training loops, data preprocessing scripts, and model evaluations, for correctness and efficiency.
  • Provide human feedback to refine RLHF frameworks and improve model alignment, safety, and helpfulness.
  • Analyze model reasoning on complex chain-of-thought prompts and identify where reasoning breaks down.
  • Benchmark and compare model outputs using technical taxonomies and performance metrics.
  • Review technical content for hallucinations, biased outputs, and logical failures.
  • Help train and evaluate next-generation LLMs through paid expert tasks.

Requirements

  • BS, MS, or PhD in Computer Science, Artificial Intelligence, Robotics, or a related quantitative field with a focus on Machine Learning.
  • Experience building, deploying, or fine-tuning ML models in a production environment.
  • Professional-level understanding of neural network architectures such as Transformers, CNNs, and RNNs, plus optimization techniques.
  • Hands-on experience with Prompt Engineering, RLHF, or RAG workflows.
  • Ability to audit complex model logic, identify training data contamination, and evaluate mathematical proofs behind ML algorithms.
  • High attention to detail in spotting hallucinations, biased outputs, and logical failures in AI-generated technical content.
  • Expert proficiency in PyTorch or TensorFlow/Keras.
  • Advanced Python experience, including NumPy, Pandas, Scikit-learn, and Hugging Face Transformers.
  • Experience with AWS SageMaker, Google Cloud Vertex AI, Weights & Biases, or LangChain.
  • Familiarity with Pinecone, Milvus, or Weaviate for RAG evaluation.

Benefits

  • Pay rates up to $80 per hour for researchers seeking your skills.
  • Flexible hours.
  • Ability to work from home.
  • Quick 10- to 15-minute assessment and fast onboarding if successful.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

AI Software Engineer

Zone & Co 251-1K Diversified Financial Services

Zone & Co is seeking an AI Software Engineer to build and deploy production-ready AI solutions embedded in its Oracle NetSuite financial operations platform.

AWS Azure JavaScript Machine Learning Python SQL
1 week, 1 day ago

AI Software Engineer

Zone & Co 251-1K Diversified Financial Services

Zone & Co is seeking an AI Software Engineer to build and deploy production-ready artificial intelligence solutions that enhance its ERP-native financial operations platform for Oracle NetSuite customers.

AWS Azure JavaScript Machine Learning Python SQL
1 week, 4 days ago

Senior AI Engineer (Agentic AI / AWS) - REMOTE

Gramian Consultancy Group Professional Services

Senior AI Engineer contractor for a Big4 consultancy serving financial institutions, building secure production-grade autonomous agents, multi-agent systems, and LLM-powered enterprise applications.

Generative AI MLOps Python
2 weeks ago

Senior Python Developer - Freelance

Netguru 251-1K Internet Software & Services

Netguru is seeking a Senior Python Developer for a three-month, full-time remote freelance engagement in Poland to build a production AI agent for an image-processing system.

AWS Azure CI/CD Docker GCP Node.js OpenTelemetry PostgreSQL Python TypeScript Vertex AI
4 weeks, 1 day ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers