Agents Last Exam — Task Input Data Input files (the materials each task hands to the agent at run start) for the Agents Last Exam (ALE) benchmark. Browsable per task directory layout. The Agents Last Exam dataset family ALE is published as three companion HuggingFace datasets: Dataset Contents Access Task Card Metadata One row per task: titles, prompts, taxonomy, input file descriptors Open Task Input Data The input/ files each task hands the agent at run start Open Reference (Ground Truth) Data The reference/ outputs used to score runs ⚠️ Gated (manual approval) You are viewing: Task Input Data. Source repository: https://github.com/rdi berkeley/agents last exam Official site: https://agents last exam.org Layout What's Included input/ — the materials the agent is given when a run begins. software/ — task software fixtures, for the tasks that ship them. What's Excluded reference/ — ground truth outputs. Published separately in the gated reference repo (see the family table above) to protect benchmark integrity. VM images and any internal artifacts.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy