[Remote] AI Agent Evaluation Analyst for Autonomous Agents (No coding required)

Remote Full-time
Note: The job is a remote job and is open to candidates in USA. OpenTrain AI is hiring detail-oriented, analytical contributors to help test and improve autonomous AI agent evaluations. The role involves reviewing evaluation tasks, identifying inconsistencies, and collaborating with teams to ensure thorough testing of agents. Responsibilities • Review and refine agent evaluation tasks and scenarios for logic, completeness, and realism • Identify inconsistencies, ambiguities, and missing assumptions • Define gold-standard expected behaviors for agents • Annotate reasoning paths, cause-effect relationships, and plausible alternatives • Collaborate with QA, writers, and developers to suggest refinements and expand edge case coverage • Ensure autonomous agents are tested thoroughly and realistically Skills • Strong analytical thinking and excellent attention to detail • Fluent written English with clear documentation skills • Comfort reading structured formats such as JSON or YAML (no need to write code) • Ability to reason about complex systems and spot what could break or be misinterpreted • Prior exposure to QA/test-case thinking, logic puzzles, or evaluation frameworks Company Overview • OpenTrain AI connects companies with vetted data labeling experts, supports any annotation tool, and manages escrow payments. It was founded in 2022, and is headquartered in Seattle, Washington, USA, with a workforce of 2-10 employees. Its website is Apply tot his job
Apply Now

Similar Opportunities

Experienced Registered Behavior Technician for In-Home ABA Therapy - Atlanta, GA

Remote Full-time

Immediate Hiring: Experienced Registered Behavioral Technician (RBT) for Clinic-Based ABA Therapy Services

Remote Full-time

Experienced Registered Behavioral Technician (RBT) - ABA Therapy for Children with Autism Spectrum Disorder

Remote Full-time

Experienced Registered Nurse - Telehealth: Providing Remote Care Coordination and Patient Support

Remote Full-time

Experienced Substitute Teacher for Riverside County Schools - Join Scoot Education's Innovative Team

Remote Full-time

Experienced Substitute Teacher for San Bernardino County - Flexible Schedules & Competitive Pay

Remote Full-time

Experienced School Year Instructional Coach for High-Dosage Tutoring Programs in Edgewater Park, NJ

Remote Full-time

Experienced School Year Tutor for K-8 Students in Math and Literacy - Mickleton, NJ

Remote Full-time

Experienced Secondary Social Studies Teacher for Kansas - Flexible Hybrid Remote Arrangement

Remote Full-time

USPS Office Helper

Remote Full-time

Customer Support Associate (Entry Level) - Join blithequark's Remote Team for a Rewarding Career in Customer Experience

Remote Full-time

Case Mgr Behavioral Health (Crisis Triage) - remote (PA/NJ/DE)

Remote Full-time

Experienced Full-Time Work-from-Home Live Chat Support Agent – Customer Experience & Product Information Expert

Remote Full-time

Director/ Contract Manufacturing

Remote Full-time

Customer Service Representative - Remote Opportunity with blithequark: Flexible Schedule, Career Growth, and Supportive Community

Remote Full-time

Airbnb Customer Support (Remote Jobs ? Part Time)

Remote Full-time

**Experienced Full Stack Data Entry Specialist – Remote Opportunity with arenaflex**

Remote Full-time

**Experienced Customer Service Representative – Delivering Exceptional Patient Experiences in a Dynamic Remote Work Environment**

Remote Full-time

Experienced Customer Support Specialist – Delivering Exceptional Pet Parent Experiences in a Fully Remote Environment at blithequark

Remote Full-time

Registered Nurse job at Trinity Health in Fayetteville, NY, Syracuse, NY

Remote Full-time
← Back to Home