Reference Harness memory retrieval probe · reference-memory-retrieval-v1-run-002
This is a controlled experiment artifact for Reference Harness. It demonstrates the public runner pipeline, not any vendor agent's native behavior.
reference-memory-retrieval-v1-run-00213 eventsclean
SESSIONevt-000T+00:15:00Z
session.start
Actor reference-agent produced sequence 0. Redaction status is clean.
{
"max_turns": 8,
"tool_count": 2
}1 events are visible; hidden chain-of-thought is not part of the trace.
Download JSONLHARNESS MODEcontrolled
MODELscripted-model-v1
NETWORKoff
COMPARABILITYnot-comparable
WHAT THIS PROVES
A fixed fixture and prompt can produce a structured result, valid trace, identical normalized fingerprint, and downloadable replay through the public runner.
WHAT IT DOES NOT PROVE
Real model capability, vendor agent behavior, OS sandbox strength, performance cost, or cross-agent ranking.