evaluator remote

10,000 evaluator remote job listings in United States. Find daily updated positions from leading job boards.

  • Senior FP&A

    3 days ago


    Washington, DC, United States Volga Partners Full-time

    Volga Partners is seeking a highly experienced Quantitative Finance Subject Matter Expert for remote, project-based work evaluating AI-generated financial analyses and improving AI model reasoning. You will validate technical accuracy, identify inaccuracies, and help refine datasets and prompts used in LLMs within the finance sector. The role demands deep...


  • Miami, FL, United States Associated Builders and Contractors (ABC) Full-time

    Alignerr is seeking an Electrical Engineering Expert to design and evaluate AI training problems across circuits, power systems, and signal processing. You will author gold-standard solutions and audit AI outputs for accuracy and safety, ensuring alignment with industry standards.The role is fully remote, with flexible hours and 10–40 hours per week. A...


  • Fairfax, VA, United States Delphi Technologies Full-time

    Empower Mental Health is seeking a Licensed Clinical Psychologist to perform VA compensation and pension evaluations. This independent contractor role offers flexibility with telehealth and in-person opportunities, and a focus on objective, fact-based assessments for veterans.You will operate under VA guidelines and deliver concise, properly formatted...


  • Northern, KY, United States Associated Builders and Contractors (ABC) Full-time

    Alignerr is seeking an Electrical Engineering Expert to design and evaluate AI training problems across circuits, power systems, and signal processing. You will author gold-standard solutions and audit AI outputs for accuracy and safety, ensuring alignment with industry standards.The role is fully remote, with flexible hours and 10–40 hours per week. A...


  • Dallas, TX, United States YO AI Labs Full-time

    YO AI Labs in the United States is seeking a Turkish Bilingual Expert for a remote, contractor role to support language and AI training projects. You will evaluate Turkish audio for nativeness, fluency, pronunciation, and intonation, and provide feedback to improve AI systems, all in English. No prior AI experience required. Applicants should have native...


  • Moultrie, GA, United States Delphi Technologies Full-time

    Empower Mental Health is seeking a Psychiatrist to conduct VA Compensation & Pension evaluations as a non-treatment, independent contractor. Evaluations may be conducted via telehealth or in-person, with per-evaluation compensation.The ideal candidate holds an MD or DO license, is adept at objective, structured assessments, and can complete reports in...


  • United States RWS Full-time

    RWS in Germany is seeking a Speech AI Evaluation Specialist to support the improvement of AI-generated content in German. This freelance, part-time role is remote with a flexible 10+ hour weekly schedule and starting immediately. You will conduct short voice conversations with AI models, follow prompts, evaluate responses for relevance and clarity, and...

  • Electrical Engineer

    5 days ago


    Northern, KY, United States OpenTrain AI Full-time

    OpenTrain AI seeks an Electrical Engineering AI Response Evaluator to assess AI-generated answers, prompts, explanations, and calculations. You will judge technical accuracy, logical soundness, completeness, clarity, safety, and practical feasibility across a wide range of electrical engineering topics.The work may cover circuit design and analysis, power...


  • Austin, Texas, United States YO AI Labs Full-time

    YO AI Labs seeks experienced Developer & Infrastructure Experts to evaluate AI-powered workflows across software development, cloud infrastructure, DevOps, SRE, and platform engineering. You will test AI-generated commands, configurations, and deployment workflows, and provide practical feedback based on real-world engineering standards.This remote...


  • New York, NY, United States YO AI Labs Full-time

    YO AI Labs invites a Tamil Bilingual Expert to contribute to a language and AI training project focused on Tamil understanding and generation. You will evaluate Tamil audio content, assess linguistic quality, provide feedback, and ensure alignment with guidelines. No prior AI experience required; strong Tamil proficiency and English communication are...


  • Washington, DC, United States YO AI Labs Full-time

    YO AI Labs is seeking Tamil bilingual experts to contribute to language and AI training projects focused on evaluating Tamil audio quality and linguistic authenticity. You will assess nativeness, fluency, pronunciation, and overall accuracy, providing clear feedback in English. No prior AI experience is required; strong Tamil and English communication, plus...


  • United States Airbnb Full-time

    Airbnb is seeking a Senior Staff Machine Learning Engineer to drive evaluation and the data flywheel for Airbnb Assistance Engineering. You will set the direction for GenAI systems, align offline metrics with online outcomes, and partner with product, engineering, and design to build scalable evaluation platforms. In this role you will design data...


  • California City, CA, United States YO AI Labs Full-time

    YO AI Labs seeks PhD and academic experts to support AI research projects. You will evaluate and improve AI model responses across technical and humanities disciplines using your subject-matter expertise and research skills. This remote contractor role focuses on identifying errors, gaps, and weak reasoning, and on creating expert prompts, reference answers,...


  • New York, NY, United States YO AI Labs Full-time

    YO AI Labs is seeking Italian Bilingual Experts to contribute to a language and AI training project focused on improving Italian-language understanding and generation. You will evaluate Italian audio content, assess linguistic quality, and provide detailed feedback based on guidelines. No prior AI experience is required. Strong Italian proficiency, English...


  • Seattle, Washington, United States YO AI Labs Full-time

    YO AI Labs is seeking an experienced Developer & Infrastructure Expert to evaluate AI-powered workflows across software development, cloud infrastructure, DevOps, SRE, and platform engineering. You will test AI-generated commands, configurations, and deployment workflows, and provide practical feedback.This remote contractor role emphasizes hands-on...


  • Illinois City, IL, United States YO AI Labs Full-time

    YO AI Labs is seeking a PhD and academic expert to support AI research projects remotely. You will evaluate model responses across technical and humanities disciplines, leveraging your subject-matter expertise and research skills to improve AI outputs. The role involves creating expert prompts, reference answers, and critiques, identifying errors and gaps,...


  • United States Dorado Part-time

    RWS is seeking Generative Audio Evaluation Specialists to join a remote freelance project focused on testing and assessing a cutting-edge generative audio model. You will evaluate AI-produced audio for short video and text-to-audio tasks, ensuring quality, naturalness, and alignment with captions. Strong Arabic (Kuwait) proficiency and meticulous attention...


  • Ann Arbor, MI, United States Mathematica, Inc. Full-time

    Sr. Researcher, Healthcare Evaluation (Remote Eligible) About Mathematica: Mathematica applies expertise at the intersection of data, methods, policy, and practice to improve well-being around the world. We collaborate closely with public- and private-sector partners to translate big questions into deep insights that improve programs, refine strategies, and...


  • Chicago, IL, United States Mathematica, Inc. Full-time

    Sr. Researcher, Healthcare Evaluation (Remote Eligible) About Mathematica: Mathematica applies expertise at the intersection of data, methods, policy, and practice to improve well-being around the world. We collaborate closely with public- and private-sector partners to translate big questions into deep insights that improve programs, refine strategies, and...


  • New York, NY, United States Obsidian Full-time

    Obsidian is seeking expert Evaluators in Government / public administration to review AI-generated documents, spreadsheets, and slide decks for accuracy and quality. You will apply deep subject-matter expertise to grade outputs in a remote, hourly engagement. Candidates should have 5+ years in Government/public administration, native or professional English...