[Remote] AI Agent Evaluation Analyst for Autonomous Agents (No coding required)

Remote Full-time

Note: The job is a remote job and is open to candidates in USA. OpenTrain AI is hiring detail-oriented, analytical contributors to help test and improve autonomous AI agent evaluations. The role involves reviewing evaluation tasks, identifying inconsistencies, and collaborating with teams to ensure thorough testing of agents. Responsibilities • Review and refine agent evaluation tasks and scenarios for logic, completeness, and realism • Identify inconsistencies, ambiguities, and missing assumptions • Define gold-standard expected behaviors for agents • Annotate reasoning paths, cause-effect relationships, and plausible alternatives • Collaborate with QA, writers, and developers to suggest refinements and expand edge case coverage • Ensure autonomous agents are tested thoroughly and realistically Skills • Strong analytical thinking and excellent attention to detail • Fluent written English with clear documentation skills • Comfort reading structured formats such as JSON or YAML (no need to write code) • Ability to reason about complex systems and spot what could break or be misinterpreted • Prior exposure to QA/test-case thinking, logic puzzles, or evaluation frameworks Company Overview • OpenTrain AI connects companies with vetted data labeling experts, supports any annotation tool, and manages escrow payments. It was founded in 2022, and is headquartered in Seattle, Washington, USA, with a workforce of 2-10 employees. Its website is Apply tot his job

Apply Now

Experienced Structural Concrete Carpenters Wanted for Immediate Hire – Join Our Team of Skilled Professionals in the Construction Industry

Remote Full-time

Assistant/Associate Professor, Educator - Department of Mechanical and Materials Engineering - Tenure Track Faculty Position

Remote Full-time

Experienced Pharmacy Technician and Data Entry Specialist - Remote Job Opportunity with Walgreens, $75,000/Yearly

Remote Full-time

Experienced Part-Time Remote Data Entry Specialist – Entry-Level Opportunity at blithequark

Remote Full-time

Experienced Customer Support Representative – Remote Position for Delivering Exceptional Service and Driving Customer Satisfaction

Remote Full-time

[Remote] AI Agent Evaluation Analyst for Autonomous Agents (No coding required)

Similar Opportunities

AI Automation Developer for Gemini API & Workflow Automation

Experienced AI Developer Needed for Automations & CRM Integrations (Python, n8n, Make) - Contract to Hire

Zoho One Automation - AI Systems Specialist

Automation and AI Specialist

[Remote] Senior Data Engineer, Data Platform

Data Engineer and AI Engineer Needed for Innovative Project

[Remote] Senior Data Engineer- AI/ML (Remote)

Agentic Artificial Intelligence (AI) Engineer - Talent Pipeline (Remote - US)

Senior AI Data Engineer, Risk Engineering

Principal Product Manager, AI

Experienced Structural Concrete Carpenters Wanted for Immediate Hire – Join Our Team of Skilled Professionals in the Construction Industry

Assistant/Associate Professor, Educator - Department of Mechanical and Materials Engineering - Tenure Track Faculty Position

Experienced Pharmacy Technician and Data Entry Specialist - Remote Job Opportunity with Walgreens, $75,000/Yearly

Experienced Part-Time Remote Data Entry Specialist – Entry-Level Opportunity at blithequark

Experienced Customer Support Representative – Remote Position for Delivering Exceptional Service and Driving Customer Satisfaction

Associate Consultant, US

Chief Information Officer for the Coalition for the Common Good (Remote)

Experienced Entry-Level Chat Support Agent – Deliver Exceptional Customer Experiences in a Dynamic Remote Environment

Experienced Customer Service Associate - Nights and Weekends (Full Time) at arenaflex

[Work From Home] Data Modeler with strong mongo DB

[Remote] AI Agent Evaluation Analyst for Autonomous Agents (No coding required)

Similar Opportunities

AI Automation Developer for Gemini API & Workflow Automation

Experienced AI Developer Needed for Automations & CRM Integrations (Python, n8n, Make) - Contract to Hire

Zoho One Automation - AI Systems Specialist

Automation and AI Specialist

[Remote] Senior Data Engineer, Data Platform

Data Engineer and AI Engineer Needed for Innovative Project

[Remote] Senior Data Engineer- AI/ML (Remote)

Agentic Artificial Intelligence (AI) Engineer - Talent Pipeline (Remote - US)

Senior AI Data Engineer, Risk Engineering

Principal Product Manager, AI

Experienced Structural Concrete Carpenters Wanted for Immediate Hire – Join Our Team of Skilled Professionals in the Construction Industry

Assistant/Associate Professor, Educator - Department of Mechanical and Materials Engineering - Tenure Track Faculty Position

Experienced Pharmacy Technician and Data Entry Specialist - Remote Job Opportunity with Walgreens, $75,000/Yearly

**Experienced Part-Time Remote Data Entry Specialist – Entry-Level Opportunity at blithequark**

Experienced Customer Support Representative – Remote Position for Delivering Exceptional Service and Driving Customer Satisfaction

Associate Consultant, US

Chief Information Officer for the Coalition for the Common Good (Remote)

**Experienced Entry-Level Chat Support Agent – Deliver Exceptional Customer Experiences in a Dynamic Remote Environment**

**Experienced Customer Service Associate - Nights and Weekends (Full Time) at arenaflex**

[Work From Home] Data Modeler with strong mongo DB

Experienced Part-Time Remote Data Entry Specialist – Entry-Level Opportunity at blithequark

Experienced Entry-Level Chat Support Agent – Deliver Exceptional Customer Experiences in a Dynamic Remote Environment

Experienced Customer Service Associate - Nights and Weekends (Full Time) at arenaflex