Prolific

Prolific

Prolific is a platform that enables researchers to quickly find trustworthy research participants. With a pool of over 120,000 active and verified participants, Prolific ensures high-quality responses through continuous monitoring and engagement. The p...

Professional Services
51-250
Founded 1997
$0M raised

Description

  • Evaluate AI-generated explanations of model architectures, loss functions, and backpropagation for technical accuracy.
  • Audit ML code and notebooks, including training loops, data preprocessing scripts, and model evaluations, for correctness and efficiency.
  • Provide human feedback to refine RLHF frameworks and improve model alignment, safety, and helpfulness.
  • Analyze model reasoning on complex chain-of-thought prompts and identify where reasoning breaks down.
  • Benchmark and compare model outputs using technical taxonomies and performance metrics.
  • Review technical content for hallucinations, biased outputs, and logical failures.
  • Help train and evaluate next-generation LLMs through paid expert tasks.

Requirements

  • BS, MS, or PhD in Computer Science, Artificial Intelligence, Robotics, or a related quantitative field with a focus on Machine Learning.
  • Experience building, deploying, or fine-tuning ML models in a production environment.
  • Professional-level understanding of neural network architectures such as Transformers, CNNs, and RNNs, plus optimization techniques.
  • Hands-on experience with Prompt Engineering, RLHF, or RAG workflows.
  • Ability to audit complex model logic, identify training data contamination, and evaluate mathematical proofs behind ML algorithms.
  • High attention to detail in spotting hallucinations, biased outputs, and logical failures in AI-generated technical content.
  • Expert proficiency in PyTorch or TensorFlow/Keras.
  • Advanced Python experience, including NumPy, Pandas, Scikit-learn, and Hugging Face Transformers.
  • Experience with AWS SageMaker, Google Cloud Vertex AI, Weights & Biases, or LangChain.
  • Familiarity with Pinecone, Milvus, or Weaviate for RAG evaluation.

Benefits

  • Pay rates up to $80 per hour for researchers seeking your skills.
  • Flexible hours.
  • Ability to work from home.
  • Quick 10- to 15-minute assessment and fast onboarding if successful.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Fullstack Data Scientist - freelancer

Lingaro 5K-10K IT Services

Lead GenAI engineering initiatives by designing, building, evaluating, and governing production-ready AI applications and architectures for enterprise use cases.

Apache Spark AWS Azure CI/CD Databricks Docker Generative AI GitHub Actions GitLab Kubernetes Machine Learning Microservices Python
2 hours, 38 minutes ago

[Job-00023] Senior AI Engineer, Brazil

CI&T 5K-10K Internet Software & Services

CI&T is seeking a Senior AI Engineer to build reliable AI-powered information extraction solutions and enhance a web application using LLMs, document intelligence, and full-stack engineering.

Next.js Python React
2 hours, 38 minutes ago

[Job 32221] SR. AI Engineer -Foco em Criação e Orquestração de Agentes (Agentic SDLC)

CI&T 5K-10K Internet Software & Services

A CI&T busca um(a) Software Engineer Mid-Senior para arquitetar, construir e entregar soluções de software para clientes, usando agentic SDLC e IA aplicada da arquitetura ao código em produção.

CI/CD Microservices NestJS Node.js TypeScript
2 hours, 53 minutes ago

Senior Data & AI Engineer

CareDx 251-1K Pharmaceuticals

CareDx is seeking a Senior Data and AI Engineer to build production data engineering and clinical NLP solutions that convert structured and unstructured EMR data into standardized, reliable transplant-focused clinical data for patient care, research, and business operations.

CI/CD Deep Learning Git HIPAA Machine Learning NLP Pandas Python PyTorch SQL TensorFlow
3 hours, 23 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers