Weekday AI (YC W21)

AI Evaluation Specialist

Weekday AI (YC W21)

Remote · Part Time

Be the first to apply

Experience
Any
Salary
USD 70 – USD 70 / hour
Openings
1
Posted
3 সপ্তাহ আগে
Work mode
Work from home
Education
Bachelor's degree
Resume
Required to apply

Job description

About the Role

This position supports a leading AI research project dedicated to enhancing the quality, precision, and reasoning capabilities of cutting-edge artificial intelligence systems. It invites skilled analytical professionals with strong critical thinking and communication abilities to meticulously evaluate AI-generated responses across diverse topics.

Key Responsibilities

  • Assess AI-generated content for correctness, clarity, logical structure, and comprehensiveness.
  • Detect factual inaccuracies, gaps in reasoning, contradictions, and unsupported assertions.
  • Utilize structured evaluation frameworks and thorough quality standards to appraise responses.
  • Compose clear, succinct, and evidence-based explanations for evaluation decisions.
  • Point out both positive aspects and areas needing enhancement within AI outputs.
  • Apply uniform evaluation practices consistently across multiple tasks.
  • Adhere strictly to detailed project guidelines and assessment criteria.
  • Ensure objectivity, precision, and reproducibility in all evaluations.
  • Work independently to deliver high-quality outcomes.

Qualifications

  • Bachelor's degree from a globally recognized university, preferably from top-ranked institutions.
  • Outstanding analytical and problem-solving skills.
  • Excellent written communication abilities, capable of clearly articulating complex reasoning.
  • Advanced critical reading skills, including identifying nuanced arguments, implicit meanings, logical inconsistencies, missing context, and weak or unsupported reasoning.
  • Meticulous attention to detail and consistent application of structured evaluation guidelines.
  • Ability to manage responsibilities autonomously and efficiently.
  • Fluency in English at a native level.

Preferred:

  • Experience in content evaluation, research, quality assurance, editing, or analytical review.
  • Knowledge of AI, large language models, or AI evaluation techniques.
  • Experience with structured annotation or assessment frameworks.
  • Capability to deliver thoughtful, impartial, and well-substantiated written evaluations meeting defined quality standards.

Engagement Details

  • Independent contractor role.
  • Fully remote with flexible scheduling allowing work at your convenience.
  • Project duration is flexible and can be adjusted according to business requirements and individual performance.
  • Compensation is $70 per hour, with weekly payments processed upon submission of approved work.
  • No visa sponsorship is available for this role.

Why Participate

  • Make meaningful contributions to the advancement of next-generation AI technologies.
  • Help enhance the reasoning, accuracy, and dependability of sophisticated AI systems.
  • Engage in stimulating intellectual projects with practical impact.
  • Collaborate indirectly with top-tier AI researchers by providing high-caliber evaluation outputs.

Equal Opportunity

Applicants will be evaluated fairly without discrimination based on legally protected characteristics. Reasonable accommodations are provided upon request.

Contract Specifics

  • Contractor status with independent working arrangements.
  • Work is completely remote and scheduled by the contractor.
  • Payments are made weekly based on completed and approved assignments.
  • This role does not involve access to confidential or proprietary information.

Work styles they’re looking for

Critical thinking Attention to Detail Written Communication Research Skills English fluency Independent Work Analytical Reasoning

Leave it if you'd like a reply — we won't use it for anything else.

Click to browse, drag & drop, or paste a screenshot

PNG, JPG, GIF, MP4, WebM, MOV · Max 20MB each · Up to 5 files

🤖
Online · instant AI help