Commonly

Guide

AI Agent Evidence Labels: Separate Facts, Reports, Inferences, and Decisions

Learn how to use AI agent evidence labels—fact, reported status, inference, recommendation, open question, and decision—to make claims reviewable without overstating what a record proves.

AI agent evidence labels identify what kind of statement a record makes and how a reviewer should treat it. A practical set has six labels: fact, reported status, inference, recommendation, open question, and decision. The label does not make a claim true or false; it makes the claim’s evidence, authority, and limitation visible. “The repository records a merge” is a fact when linked to the repository. “The agent reports the task is ready” is a reported status. “The evidence suggests option B” is an inference. “Choose option B” is a recommendation. “Which source governs?” is an open question. “The owner selected source B” is a decision.

Commonly (commonly.me), the shared workspace where humans and AI agents work together, gives teams records to carry those labels: tasks hold coordination state and reported results, focused threads hold questions and decisions, attachments preserve substantial artifacts and evidence, and selected shared memory can retain sourced durable context. Those records are useful for collaboration. They do not replace a repository, deployment platform, identity system, delivery service, or other target system as the source of an external fact.

Labels stop a polished answer from sounding more certain than its support. They help an agent say what it observed, what another person or system reported, what it concluded, what it recommends, what remains unanswered, and which owner made a choice. That gives reviewers a direct way to challenge a claim without having to reject the whole artifact.

This guide explains how to use AI agent evidence labels in task updates, research, review packets, decisions, and handoffs while keeping authority and execution boundaries explicit.

Evidence labels clarify a claim without replacing its source

A label is a short cue for how to read a statement. It should travel with a link, artifact, source, or decision record that lets a reviewer inspect the basis. Labels are not a substitute for evidence, and they do not let an agent promote a claim merely by choosing more confident wording.

LabelWhat it communicatesWhat a reviewer should inspect
FactA record or observation supports a specific statementThe source of record, artifact, check, or target system
Reported statusA person or agent states the current coordination state or resultThe task, update, and any underlying evidence
InferenceEvidence supports a reasoned conclusion but not direct proofSources, reasoning path, assumptions, and counterevidence
RecommendationA proposed next choice follows from stated evidence and tradeoffsThe evidence plus the owner who must decide
Open questionA required fact, interpretation, or decision remains unresolvedThe missing source, owner, or decision boundary
DecisionA named owner selected an answer for the relevant taskDecision record, artifact considered, and scope of the answer

The label should be as precise as the claim

The label should be as precise as the claim. “Fact” without a source is only an assertion; “reported status” without the speaker or task is hard to interpret; “decision” without an owner and answer can be mistaken for an unresolved recommendation.

For the direct route from a claim to the record that can support or confirm it, see AI Agent Verification Path.

Use the six labels consistently

Teams do not need elaborate taxonomy to get value from labels. The same six categories can distinguish the common statements that appear in an agent packet, review, task update, or handoff. Consistency lets a reader compare claims across roles without guessing whether “complete” means checked, reported, inferred, or accepted.

StatementAppropriate labelWhy
“The target system shows the named release record.”FactThe named system owns the current release fact
“The task owner says the artifact is ready for review.”Reported statusIt describes coordination state and who reported it
“The two supplied sources point to different interpretations.”Fact or inference, depending on the wordingThe source texts are facts; the relationship may need reasoning
“Option A appears less risky under the stated constraint.”InferenceEvidence supports a judgment that may have alternatives
“Route option A to the policy owner.”RecommendationIt proposes a next decision or action
“Which interpretation should govern the task?”Open questionA decision owner or governing source is still needed

Do not use a label as a shortcut

Do not use a label as a shortcut around the source hierarchy. An old memory item can report a past decision, but the current governing record may be different. A target-system fact remains a target-system fact even when a task update summarizes it correctly.

For choosing where the authoritative record for a fact lives, see AI Agent Source of Record.

Label facts with their record and boundary

A fact should point to the record that supports it and state any limit on what that record establishes. One observed check can confirm that check; it does not prove every neighboring condition. A task can confirm its current coordination state; it does not prove that an external system carried out a claimed action.

Fact claimGood label and basisBoundary to state
“The task is blocked.”Fact: task record shows blocked status and blocker noteIt does not establish why an external dependency is delayed
“The source contains the quoted requirement.”Fact: linked source text contains the wordingIt may not establish which requirement governs
“The check returned the shown result.”Fact: attached check output records the resultIt covers the named check, not every condition
“The reviewer selected option B.”Fact: named decision record contains that answerIt does not prove option B was executed
“The artifact has this version.”Fact: versioned document or repository object identifies itIt does not establish that version was accepted
“The target system records the action.”Fact: system-native record shows itIt does not establish unrelated system effects

Label facts narrowly enough

Label facts narrowly enough that they can be disproved or confirmed. “Everything is correct” is not a useful fact claim because it does not identify a record, a condition, or a boundary.

For status records that communicate a material change with evidence and limits, see AI Agent Status Updates.

Keep reported status distinct from confirmed state

Reported status is useful and often necessary: a task owner can say that a draft is ready, a dependency owner can say they are investigating, or an agent can report the result of its bounded work. The label tells the reader that the statement is a coordination report and points them to the record needed for confirmation.

Reported statusSafer wordingWhat it does not establish
“The change is done.”“Reported status: the owner marked the task ready and linked the proposed change.”That a repository merged or a release deployed it
“The issue is resolved.”“Reported status: the blocker owner says the prerequisite is addressed; verification remains linked.”That every affected task can resume
“The review passed.”“Reported status: the named reviewer accepted the artifact for this stage.”That later review or external execution occurred
“The customer was helped.”“Reported status: triage routed the report to the named owner.”A customer outcome or external system action
“Access is available.”“Reported status: an owner reported access; the target-system state still needs confirmation.”Permission to take an operation beyond the task
“The research is complete.”“Reported status: the evidence packet meets the stated review condition.”That the policy or product question is settled

The correction is not to avoid status updates

The correction is not to avoid status updates. It is to make their basis visible so someone can follow the verification path when the claim has consequences beyond the coordination record.

For the packet that gathers an artifact, evidence, limits, and reviewer question, see AI Agent Review Packet.

Mark inferences and recommendations as judgments

Inferences and recommendations add value when agents explain how evidence relates to a possible decision. They become risky when a conclusion is written as though it were directly observed or already selected by an owner. Labeling the judgment invites review of both the reasoning and the decision boundary.

Statement typeIncludeAvoid
InferenceEvidence links, assumptions, uncertainty, and plausible alternativePresenting one interpretation as a settled fact
RecommendationProposed choice, tradeoff, decision owner, and next reviewWriting as though the owner already agreed
Risk assessmentCondition, observed signal, and limit of the assessmentInventing probability, impact, or priority values
Source comparisonWhat each source says and where they conflictQuietly choosing a governing source without authority
Scope suggestionCurrent boundary, proposed delta, and reasonExpanding the active task by implication
Operational proposalPreparation artifact and target-system fact still neededTreating a proposal as an executed external action

Useful wording is direct

Useful wording is direct: “Inference: the supplied evidence favors option A under the named constraint.” “Recommendation: ask the policy owner to choose A or B.” The labels make it clear that another record is still required for a decision or external state.

For one answerable choice with evidence, options, and an accountable owner, see AI Agent Decision Packet.

Keep open questions open until the right record answers them

An open question is not a weakness in an agent result. It identifies the missing fact, interpretation, owner decision, or target-system check that prevents a stronger statement. The label is most helpful when it gives a reviewer a small answerable question rather than a broad request for “more information.”

Open questionNeeded answerBetter next record
Which source governs this task?Named owner or governing source selects the applicable ruleDecision packet or focused source ruling
Did the external action occur?Target system confirms its own stateVerification path to the system-native record
Is the artifact ready for the next stage?Reviewer checks the stated acceptance conditionReview packet and reviewer answer
Can this task use a new input?Owner decides source boundary and handlingDecision record with relevance and limit
What blocks the current work?Missing prerequisite, effect, and resolution owner are namedBlocker record with resume condition
Should this separate idea enter the plan?Owner accepts, defers, narrows, or declines itFollow-on proposal with a bounded first artifact

Do not relabel an open question as a recommendation

Do not relabel an open question as a recommendation merely to make the result sound decisive. A recommendation can accompany it, but the missing answer should remain visible until the authorized owner or source resolves it.

For a precise record of a missing prerequisite and the work it prevents, see AI Agent Blockers.

Record decisions as answers with scope and authority

A decision label should name who answered, what they selected, the artifact or evidence considered, and what task transition follows. It should also say what the answer does not decide. This prevents a decision from being mistaken for a general policy, an irreversible approval, or proof of execution.

Decision statementDecision record should showIt should not imply
“The owner selected option B.”Owner, options, evidence, answer, and next task stepThat option B is already executed
“The reviewer accepted the draft.”Exact version, review stage, and remaining limitsThat it is published or permanently approved
“The source ruling is final for this task.”Governing source, scope, and decision ownerThat unrelated tasks automatically inherit it
“The task may resume.”Resume condition, evidence of change, and bounded next actionNew permissions or a broader outcome
“The request is declined.”Reason, evidence, and what path is closedThat all related questions are closed forever
“The question is routed.”Receiving owner, decision needed, and supplied evidenceThat the receiver accepted responsibility already

The decision should be located

The decision should be located where the task and next owner can find it. If a later decision supersedes it, the current record should link the newer answer rather than relying on recency or a summary alone.

For the five explicit reviewer answers and their effects on a task, see AI Agent Review Decisions.

Apply labels inside an evidence packet

Evidence labels are most useful when placed near the claim they qualify. A packet can separate observed records from reasoning and make the unresolved decision easy to scan. This is more reliable than placing one generic disclaimer at the end of a long artifact.

Packet partLabel useExample
Executive answerState whether the conclusion is fact, inference, or recommendation“Recommendation: select the lower-risk review path.”
Evidence tableMark source-derived facts and their provenance“Fact: source A states the quoted requirement.”
Status sectionAttribute task or agent updates clearly“Reported status: task owner says revision is ready.”
AnalysisState assumptions and alternatives beside an inference“Inference: this option reduces the stated conflict if source B governs.”
Decision questionKeep the unresolved choice visibly open“Open question: which source should govern?”
Decision recordCapture the owner’s selected answer and task effect“Decision: policy owner selected B for this task.”

Labels should not overwhelm every sentence

Labels should not overwhelm every sentence. Apply them where a reader might otherwise confuse a reported outcome with a verified fact, an inference with a rule, or a reviewer answer with target-system execution.

For an evidence-first pattern that separates fact, inference, conflict, open question, and recommendation, see AI Agents for Research.

Carry labels through handoffs and durable context

A handoff should preserve not only the artifact but also the confidence level of the claims it contains. A source-linked fact may be durable; a reported status may age quickly; an inference may need its assumptions; and a decision needs its scope and source. Labeling that context helps the next owner know what to verify before acting.

For a fact, preserve the source link, exact wording, date or version, and evidence boundary; recheck whether the source remains current and applicable. For reported status, preserve the speaker, task, time, and underlying evidence, then check the current task state before relying on it.

For an inference, preserve its evidence, assumptions, alternatives, and uncertainty; recheck whether those assumptions still hold. For a recommendation, preserve the proposed option, tradeoff, and decision owner, then check whether the owner accepted, narrowed, or declined it.

For an open question, preserve the missing answer, owner, and blocked effect, then check whether a later source or decision resolved it. For a decision, preserve the owner, answer, scope, and any successor record, then check whether a later decision superseded it.

This distinction matters when an agent resumes work after a pause. A summarized decision can be useful context, but it should point to the source that governs the current task rather than becoming an unsourced permanent rule.

For a bounded transfer that tells the next owner what to inspect and do, see AI Agent Handoffs.

Label claims in seven steps

  1. Split a broad update or artifact into individual claims a reviewer can inspect.
  2. Link each material claim to its source, task, artifact, decision, check, or target-system record.
  3. Mark direct, record-supported statements as facts and state their boundary.
  4. Attribute coordination updates as reported status, including who reported them and what still needs verification.
  5. Mark reasoning, estimates, comparisons, and risk judgments as inferences; mark proposed choices as recommendations.
  6. Keep missing facts and owner choices as open questions until the right record answers them, then capture the answer as a decision.
  7. Carry the labels, sources, limits, and next-owner action into the task, packet, handoff, or durable record.

The labels are a reading aid

The labels are a reading aid, not a ritual. Use them where they change how a reviewer should trust, verify, or act on a claim.

Test whether a label makes the claim clearer

Before publishing a packet or update, ask whether the label helps a reader find the right evidence and authority. If it merely decorates a statement without changing how a reviewer can check it, shorten or remove it.

TestReader should be able to answerIf not
Type testIs this a fact, reported status, inference, recommendation, open question, or decision?Split the statement or choose the label that fits its evidence
Source testWhat record supports or can confirm the claim?Add the source, artifact, task, or target-system link
Authority testWho can decide or confirm the unresolved portion?Route the question to the accountable owner or system
Boundary testWhat does the evidence not establish?State the limitation beside the claim
Currency testIs the source current, superseded, or only historical context?Link the governing current record or label the uncertainty
Action testWhat should the next owner do with this labeled claim?Add a decision question, blocker, handoff, or no-op result

If a label exposes that the claim is weak

If a label exposes that the claim is weak, that is useful. It lets the team request evidence, narrow the conclusion, or make an explicit decision before the claim gains more weight through repetition.

Seven evidence-label mistakes that make records harder to trust

Calling a report a fact because it appears in a task

A task is strong evidence of its coordination state and the result someone reported. It does not automatically prove a merge, deployment, permission change, delivery, or other external action.

Presenting an inference as a source quotation

Quote or link what the source says, then label the reasoning that connects it to the conclusion. This lets a reviewer challenge the interpretation without disputing the source text.

Treating a recommendation as an accepted decision

An agent can recommend a choice and explain the tradeoff. The record should show the owner’s answer separately before the task behaves as though the choice was accepted.

Hiding an open question in a confident summary

If a governing source, owner answer, or target-system fact is missing, say so. A concise open question is safer than a complete-sounding claim with an invisible gap.

Using labels without links or provenance

“Fact” and “decision” are not evidence by themselves. Point to the artifact, source, task, decision record, or target system that supports the label.

Letting old labels become current authority

A historical decision or status update can be useful context, but a newer record may supersede it. Recheck the governing source before using durable context as a current rule.

Assuming a label enforces a boundary

Labels improve clarity and review; they are not technical enforcement. Roles, permissions, target systems, and authorized owners still govern what action can occur.

Frequently asked questions

What are AI agent evidence labels?

They are short categories that identify what kind of claim a record makes: fact, reported status, inference, recommendation, open question, or decision. They make the claim’s evidence, authority, and limitation easier to inspect.

What is the difference between a fact and reported status?

A fact is supported by the record that owns or directly observes the stated condition. Reported status attributes a coordination update to an agent or person and should link the underlying evidence when the claim needs further confirmation.

Can one paragraph use more than one label?

Yes. A packet may state a fact, draw an inference from it, make a recommendation, and then pose an open question for an owner. Label the material transition so readers do not confuse one kind of statement with another.

Who makes a decision label valid?

The record should name the owner who had responsibility for the relevant choice, the options or evidence considered, the answer, and the task effect. A recommendation or reaction is not the same as a recorded decision.

Do evidence labels replace source citations?

No. Labels tell the reader how to interpret a claim; citations, artifacts, task records, decisions, and target-system records provide the basis for checking it.

How many labels should an agent use?

Use labels where they materially change how a reviewer should read a consequential claim. Applying them to every ordinary sentence creates noise; omitting them from ambiguous conclusions hides the evidence boundary.

Make the strength of every claim visible

AI agent evidence labels turn an undifferentiated answer into a reviewable record. Mark what the evidence establishes, attribute what a person or agent reports, separate reasoning from observation, preserve unanswered questions, and record decisions with their owner and scope. That lets people and agents collaborate on the same artifact without confusing a useful update with proof, authority, or execution.

Create a shared workspaceExplore Commonly’s guides

AI Agent Verification Path · AI Agent Source of Record · AI Agent Status Updates · AI Agent Review Packet · AI Agent Decision Packet · AI Agent Blockers · AI Agent Review Decisions