Prolific

Prolific

Prolific is a platform that enables researchers to quickly find trustworthy research participants. With a pool of over 120,000 active and verified participants, Prolific ensures high-quality responses through continuous monitoring and engagement. The p...

Professional Services
51-250
Founded 1997
$0M raised

Description

  • Evaluate AI-generated explanations of model architectures, loss functions, and backpropagation for technical accuracy.
  • Audit ML code and notebooks, including training loops, data preprocessing scripts, and model evaluations, for correctness and efficiency.
  • Provide human feedback to refine RLHF frameworks and improve model alignment, safety, and helpfulness.
  • Analyze model reasoning on complex chain-of-thought prompts and identify where reasoning breaks down.
  • Benchmark and compare model outputs using technical taxonomies and performance metrics.
  • Review technical content for hallucinations, biased outputs, and logical failures.
  • Help train and evaluate next-generation LLMs through paid expert tasks.

Requirements

  • BS, MS, or PhD in Computer Science, Artificial Intelligence, Robotics, or a related quantitative field with a focus on Machine Learning.
  • Experience building, deploying, or fine-tuning ML models in a production environment.
  • Professional-level understanding of neural network architectures such as Transformers, CNNs, and RNNs, plus optimization techniques.
  • Hands-on experience with Prompt Engineering, RLHF, or RAG workflows.
  • Ability to audit complex model logic, identify training data contamination, and evaluate mathematical proofs behind ML algorithms.
  • High attention to detail in spotting hallucinations, biased outputs, and logical failures in AI-generated technical content.
  • Expert proficiency in PyTorch or TensorFlow/Keras.
  • Advanced Python experience, including NumPy, Pandas, Scikit-learn, and Hugging Face Transformers.
  • Experience with AWS SageMaker, Google Cloud Vertex AI, Weights & Biases, or LangChain.
  • Familiarity with Pinecone, Milvus, or Weaviate for RAG evaluation.

Benefits

  • Pay rates up to $80 per hour for researchers seeking your skills.
  • Flexible hours.
  • Ability to work from home.
  • Quick 10- to 15-minute assessment and fast onboarding if successful.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

AI Developer

AI Developer (Contract) at a remote company, responsible for building AI-powered workflows and internal tools that improve business efficiency and create new capabilities.

JavaScript LLM Python
15 hours, 5 minutes ago

AI Agent Implementations Specialist

Extreme Networks 1K-5K IT Services

Extreme Networks is hiring an AI Agent Developer & Marketo Administrator to support marketing automation, AI agent implementation, and MarTech operations across its global marketing organization.

CRM JavaScript Machine Learning Node.js Python REST API
1 day, 15 hours ago

Entrepreneur in Residence - Technical Co-founder (CTO)

FutureSight 11-50 Internet Software & Services

FutureSight is seeking a Co-Founder & CTO to lead the technical direction of a new B2B AI venture from inception to launch and early growth.

LLM System Design
1 day, 15 hours ago

Technical Co-founder (CTO) - AI Collections Agent

FutureSight 11-50 Internet Software & Services

Argentl is seeking a Co-Founder & CTO to lead product and engineering for an AI collections platform that automates debtor follow-up for finance teams at B2B companies.

LLM
1 day, 15 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers