An older role: first posted 64d ago, and Turing still listed it when we checked today. Newer roles tend to fill faster. See the jobs hiring now.
Before you applyHow to pass the Turing assessment and interviewsMost people who don't get in fail the screening, not the CV. Five minutes here first.What the work is
Role Overview
We are seeking a highly qualified Art Domain Expert to evaluate and improve Large Language Models (LLMs). You will design challenging prompts, assess factual accuracy and reasoning, and identify knowledge gaps within fine arts, visual arts, architecture, design, art history, museums, and artists.
Key Responsibilities
- Create advanced prompts covering fine arts, visual arts, architecture, design, art history, museums, and artists.
- Evaluate AI-generated responses for factual accuracy, reasoning quality, completeness, and nuance.
- Identify hallucinations, logical inconsistencies, outdated information, and edge cases.
- Develop benchmark datasets and adversarial test cases.
- Provide evidence-based feedback with reliable references.
- Collaborate with AI researchers to improve model performance.
- Maintain high annotation quality and documentation.
Day - to - day task
- Write domain-specific prompts of varying difficulty.
- Compare multiple AI responses and rank them.
- Explain why a response is correct or incorrect using authoritative references.
- Identify ambiguous questions and propose improved prompts.
Minimum Qualifications
- Master’s degree or higher in any field; degrees in Art History, Fine Arts, Visual Arts, Museum Studies, Cultural Studies, Design, Humanities, or other art-related disciplines are preferred.
- Strong knowledge of art, including familiarity with major artists, artworks, art movements, styles, periods, techniques, museums, cultural traditions, and developments in visual arts.
- 3+ years of relevant professional experience, preferably in art research, academia, museums, galleries, curation, art criticism or journalism, education, publishing, arts and cultural content, or a related field.
- Excellent written English, research, and analytical skills.
- Strong attention to detail and ability to assess artistic, historical, and cultural information accurately and within the appropriate context.
Preferred Qualifications
- Experience with LLMs, Generative AI, prompt engineering, or AI evaluation.
- Published research, industry recognition, or teaching experience is a plus.
- Ability to work independently and provide objective, evidence-backed reviews.
Offer Details
- Commitments Required: 40 hours per week with at least 4 hours PST overlap
- Employment type: Contractor assignment (no medical/paid leave)
- Duration of contract: 8 weeks.
Application Process
- Complete the assessment shared with you.
- Our team will evaluate your submission.
- Selected candidates will be contacted regarding the next steps.
Pay
See listing, fully remote. How payouts and tax work.
Check history
How we score →today–Open