AI Agent Evaluation Analyst for Autonomous Agents (No coding required)

Remote Full-time
We’re hiring detail-oriented, analytical contributors to help test and improve autonomous AI agent evaluations. This is part-time, fully remote work with flexible hours, ideal for people who enjoy finding edge cases, questioning assumptions, and strengthening complex systems. What you’ll do • Review and refine agent evaluation tasks and scenarios for logic, completeness, and realism • Identify inconsistencies, ambiguities, and missing assumptions • Define gold-standard expected behaviors for agents • Annotate reasoning paths, cause-effect relationships, and plausible alternatives • Collaborate with QA, writers, and developers to suggest refinements and expand edge case coverage • Ensure autonomous agents are tested thoroughly and realistically What we’re looking for • Strong analytical thinking and excellent attention to detail • Fluent written English with clear documentation skills • Comfort reading structured formats such as JSON or YAML (no need to write code) • Ability to reason about complex systems and spot what could break or be misinterpreted Nice to have Prior exposure to QA/test-case thinking, logic puzzles, or evaluation frameworks Apply tot his job
Apply Now

Similar Opportunities

Experienced Registered Behavior Technician for In-Home ABA Therapy - Atlanta, GA

Remote Full-time

Immediate Hiring: Experienced Registered Behavioral Technician (RBT) for Clinic-Based ABA Therapy Services

Remote Full-time

Experienced Registered Behavioral Technician (RBT) - ABA Therapy for Children with Autism Spectrum Disorder

Remote Full-time

Experienced Registered Nurse - Telehealth: Providing Remote Care Coordination and Patient Support

Remote Full-time

Experienced Substitute Teacher for Riverside County Schools - Join Scoot Education's Innovative Team

Remote Full-time

Experienced Substitute Teacher for San Bernardino County - Flexible Schedules & Competitive Pay

Remote Full-time

Experienced School Year Instructional Coach for High-Dosage Tutoring Programs in Edgewater Park, NJ

Remote Full-time

Experienced School Year Tutor for K-8 Students in Math and Literacy - Mickleton, NJ

Remote Full-time

Experienced Secondary Social Studies Teacher for Kansas - Flexible Hybrid Remote Arrangement

Remote Full-time

USPS Office Helper

Remote Full-time

**Experienced Part-Time Virtual Customer Care Representative - American Express Remote Jobs: Delivering Exceptional Service from the Comfort of Your Home**

Remote Full-time

Experienced Flight Paramedic - AirMed Professional | Part-Time Opportunity in Emergency Medical Services with Sanford Health

Remote Full-time

Experienced Remote Live Chat Customer Support Agent – Social Media and Website Chat Expertise for Dynamic Customer Engagement

Remote Full-time

Biostatistician – Clinical Trials (Drug/Biologic)

Remote Full-time

Market Risk Analyst - DRG

Remote Full-time

**Experienced Data Entry Clerk – Remote Opportunity at arenaflex**

Remote Full-time

**Experienced Remote Customer Service Representative – Deliver Exceptional Client Experiences and Unlock Career Growth Opportunities at arenaflex**

Remote Full-time

League of Conservation Voters Education Fund: Climate Equity Policy Fellow, Chispa AZ

Remote Full-time

Senior Director of Product Marketing, Sustainability Solutions

Remote Full-time

**Experienced Full Stack Customer Service Representative – Remote Customer Experience Specialist**

Remote Full-time
← Back to Home