Skip to main content
Tallo logoTallo logo

Find Jobs

Find Jobs Near You – Available Work in Your Location

Skip to job details
Apply for this opportunity

To apply for this job, you'll continue to an external website or email application.

MaxIT Consulting - Max Corporate Group

Agent Evaluation Infrastructure Engineer

Career Insights for Talent / Sports Agent

See where this job fits in the broader career landscape. Knowing your career path helps you see what's possible from here.

Scorecard

Based on California data

Review key factors to help you decide if this role fits your goals. How is this calculated?

Were these scores useful?

What they do

A Talent or Sports Agent represents and promotes artists, performers, and athletes in dealings with current or prospective employers. May handle contract negotiation and other business matters for clients.

$81,015 / year median in California

+9% projected growth

Explore Career

Job Description

Agent Evaluation Infrastructure Engineer MaxIT Consulting - Max Corporate Group San Francisco, CA Job Details Full-time 4 hours ago Qualifications Software engineering Testing and evaluation Research Infrastructure architecture design Full Job Description San Francisco, California | Primarily On-site We are seeking an Agent Evaluation Infrastructure Engineer to build the environments, evaluation systems, and supporting infrastructure used to train and assess long-horizon enterprise AI agents. The Opportunity You will work on the engineering and research problems behind realistic agent environments, post-training systems, and reliable evaluation of complex multi-step workflows. Key Responsibilities Design evaluation environments for long-horizon enterprise agent workflows. Define tasks, state, tools, graders, and reward signals used to evaluate and improve agents. Build high-fidelity representations of complex enterprise software environments. Develop infrastructure for rollouts, orchestration, trajectory inspection, and grader pipelines. Measure both correctness and efficiency across multi-step agent behavior. Investigate evaluation failures, reward-quality issues, and agent behavior. Build production-quality systems rather than notebook-only research prototypes. Required Qualifications Hands-on experience with AI environments, evaluations, reinforcement learning infrastructure, or related agent-training systems. Strong software engineering fundamentals. Demonstrated ability to build and ship technical infrastructure. Understanding of evaluation methodology, reward design, graders, and agent trajectories. Ability to work across languages and technology stacks based on system requirements. Candidate Profile A PhD is not required. Strong engineering and shipped environment or evaluation systems are more important than academic credentials or publication history. Seniority The opportunity is open to exceptional new graduates, early-career engineers, and experienced senior candidates. Selection is based primarily on engineering strength and relevant technical work. Work Arrangement The role is anchored in San Francisco with a strong preference for in-person collaboration. Limited flexibility may be considered case by case for exceptional candidates.