Raycaster/ Eval

APEX-Agents

gpt-5.4-mini on World224-HS-10

0/3Fail
Domain
Investment Banking
Category
AI Agents for Take-Private Analysis
Harness
dual

Grader rubric

Criteria verdict

  1. States Net Debt at Exit is $647.15M

  2. States Sponsor Equity Value at Exit is $28,382.93M

  3. States IRR% is 20.28%

Prompt excerpt

Task context

Please create a new scenario within the LBO model to assess if the sponsor can still meet the 20% IRR target at the Year 5 exit, given higher Net Working Capital needs. 1. Increase accounts receivable as % of sale by 2 percentage points value in Year 1, then keep it constant in the remaining projection years 2. Increase prepaid expenses and other current assets as % of sale by 2 percentage points in Year 1, then keep it constant in the remaining projection years. Output: Add a new worksheet to the LBO model titled “NWC Scenario”. Add the Net Debt at Exit, Sponsor Equity Value at Exit, IRR %. All monetary results must be displayed in USD millions, rounded to two decimal places, and all percentages must also be rounded to two decimal points.

Response trace

Agent response, tools, files, and edits

Open full trace

On a phone, the interactive viewer works best full-screen — pick the narrative report or the files & trajectory workspace.

Why Eval exists · why Workspace exists

Public evidence and cloud agents are the same harness.

Eval exists so scores are inspectable—tasks, trajectories, artifacts, and rubric verdicts anyone can open.Workspace exists so people can automate real file work with that harness, and so Raycaster never evaluates work it cannot perform.