Remote Senior AI Backend Engineer - Agent Evaluation & Quality

Posted 19 hours ago

Share:

Please let Salla know you found this job on RemoteYeah. This helps us get more companies to post jobs here for you.

Description:

  • Own the evaluation systems for multi-agent systems, ensuring performance and quality.
  • Build judges, test harnesses, and simulators to evaluate agent performance and identify failures.

Requirements:

  • Strong software engineering fundamentals with experience in Python or Typescript.
  • Hands-on experience with LLMs and agent systems, including orchestration frameworks.
  • A measurement mindset focused on metrics and quantifying performance.
  • Production experience with LLM systems, addressing reliability and observability.
  • 5+ years of software engineering experience, particularly with LLM/agent work.

Benefits:

  • Opportunity to contribute to agent development alongside evaluation systems.
  • Work in a dynamic environment with a focus on quality and performance.

Job title

Job type

Experience level

Required experience

5 years

Salary

-

Degree requirement

No degree required

Location requirements

Benefits

-

Report this job

Job expired or something else is wrong with this job?

Report job