futureofagents.org
RESEARCH DOSSIER / 2026.10INITIAL EDITION · OPEN FOR REVIEW
← Back to research index

Observable deployments, not miracle demos

Agent Field Reports.

A credible field report includes the environment, task, limitations and failures.

EVIDENCE STATUS / LAUNCH

This is a scoped research dossier, not a completed systematic review or an independently tested result. It identifies methods, questions and source trails for future reporting.

01

Report what ran

Date the run, model, software versions, tools, prompts or relevant constraints, and permitted environments.

02

Record bad outcomes

A failure case with a reproducible trace can teach more than a polished highlight. It deserves equal editorial attention.

03

Keep revisions visible

If a report is later corrected or replicated, publish the new evidence with an edition note rather than silently replacing the earlier account.

SUGGESTED VERIFICATION METHOD

What would count as evidence?

Publish a redacted trace, access matrix, verified outcome and reproducibility limitations.

STARTING SOURCE TRAIL

Documents to examine

  • NIST AI RMF
  • OpenAI Agents SDK — Running Agents

These are starting points, not claims that every document has been independently reproduced.

Read our cited field note →
EDITORIAL / VERSION RECORD

Edition 1.0 · 09 October 2026

Initial research brief published. No earlier revisions or submitted public corrections are claimed.

Suggest a documented correction ↗