Mercor is building high-fidelity simulated work environments to evaluate AI agents on real marketing tasks. You will design organic growth tasks, specify the required dashboards, personas and tool states, and write reference answers that reveal robust reasoning.
You will review AI agent attempts, judge performance against a clear rubric, and ensure analyses cover diagnostics, prioritization, planning and reporting across multiple tools.
#J-18808-Ljbffr…
