Data annotation job
Applied Philosophy Benchmark Specialist
Mercor · Posted 2 weeks ago
- AI evaluation
- STEM
About this role
Role Overview
We are seeking expert philosophers to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core philosophy domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities.
You will be assigned one of two task types
- Question Authoring, Create original, challenging multiple-choice questions in your area of philosophy expertise, rate their difficulty, and submit them for review.
- Question Verification, Review pre-written questions for accuracy, clarity, and rigor. Edit where needed, rate difficulty, and document any changes made.
Philosophy Domains Covered
Formal Ontology & Knowledge Representation, AI Ethics, Applied Epistemology, Philosophy of Technology & Robotics, Philosophy of Science.
Key Responsibilities
- Author original philosophy questions that test deep conceptual understanding, not surface-level recall
- Ensure questions are unambiguous, self-contained, and precisely defined, all necessary information must be in the problem statement
- Rate each question's difficulty: Medium (intro undergraduate), Hard (advanced undergraduate), or Expert (post-graduate and above)
- Provide 1 correct answer and 9 plausible but subtly incorrect alternatives that challenge expert-level solvers
- Write step-by-step Chain-of-Thought solutions with clear, concise reasoning in markdown format
- Supply 1 to 5 academic references per question from reputable sources (peer-reviewed journals, university repositories)
- For verification tasks: flag issues with clarity, completeness, precision, or solvability and justify any edits made
Ideal Qualifications
- PhD or doctoral candidate in Philosophy or a closely related field
- Master's degree considered for candidates with exceptional depth in a specific subdomain
- Strong command of philosophical argumentation, formal logic, and canonical texts across traditions
- Research publications or teaching experience in philosophy is a strong plus
- Excellent written English and ability to express complex ideas clearly and concisely
More About the Opportunity
- Expected commitment: 10+ hours/week
- Asynchronous, fully remote work
This listing is aggregated by DataAnnotationJobs.org. Applications are processed by Mercor, not by this site.
Mercor is an AI-native talent marketplace connecting vetted domain experts such as doctors, lawyers, and engineers directly with frontier AI labs for RLHF, evaluation, and specialized training data at massive scale. See all Mercor jobs.
Apply for this role
Your details first, then a short screening test. Both are read as part of your application and go to Mercor, and we will tell you when similar roles appear.