Cited briefing
A source may be current but use a different population or method.
Coordinate source gathering, synthesis, and review for evidence-based briefs.
Planning guidance, not a claim of installed capabilities or customer results. Validate scope and feasibility before deployment.
Research teams need traceable evidence rather than fluent but unsupported summaries. A bounded agent workflow can separate source collection, comparison, and editorial review while keeping uncertain findings visible.
A source may be current but use a different population or method.
Two estimates may disagree because their dates or definitions differ.
A broken source link makes a polished summary difficult to audit.
Plan separate retrieval and synthesis tasks, use permission-scoped document or browser tools, and select models against citation tests. Retrieved text is untrusted input, not authority to change the task or access policy.
Keep source version, tool inputs, and execution status with each draft. Missing evidence remains visible.
A draft whose material claims can be traced to source passages.
Define acceptanceProposed capabilities must be tested against representative inputs. Unsupported or low-confidence results belong in a review queue, not an automatic decision.
Hypothetical: two reports disagree on a market estimate. A collection agent records dates and methods, a synthesis agent highlights the conflict, and an analyst decides whether either figure belongs in the brief.
A source may be current but use a different population or method.
A draft or exception needs source verification before it can leave the workspace.
Prepare a cited briefing from an approved source list.
A draft whose material claims can be traced to source passages.
Receive a bounded task and permission-limited documents. Record source versions and reject access outside the agreed scope.
Route retrieval, calculation, or drafting to selected models and scoped tools. Preserve failures, evidence, and approval boundaries in the run record.
An assigned person checks evidence, records the outcome, and follows the existing operational procedure. Rejected signals feed back into evaluation.
Measure source coverage, citation accuracy, and time to human approval.
Accepted observations / reviewed observations. Count missed cases separately against the manual reference.
Target: agree before pilotRecord minutes per reviewed item, including rework and escalations. Compare the same task with the manual baseline.
Target: agree before pilotRecord whether stale inputs, denied access, and unavailable sources stop or visibly degrade the workflow.
Target: agree before pilotEstablish a manual baseline, agree acceptance thresholds with the operational owner, and compare review effort as well as accuracy. Any benefit must be measured in the pilot; no savings or ROI are promised here.
Pilot one research question with a fixed source budget and approved corpus. Test missing citations and conflicting sources, assign an editorial approver, and stop runs that exceed access or cost limits.
Review document formats, search access, browser restrictions, model hosting, and licensed-source terms. CCTV is not required; the prerequisite is authorized access to the research material.
| Check | Required evidence |
|---|---|
| Source access | Approved formats, source versions, licenses, and read permissions. |
| Tool boundaries | Sandbox, denied-command tests, credential isolation, and context limits. |
| Compute & recovery | Model endpoint, per-run budget, timeout behavior, and resumable evidence. |
A workflow is not a universal connector. Verify file formats, authentication, tool permissions, model context limits, and failure recovery in the actual environment before committing to a setup.
Example: export a cited draft and evidence table into an approved document workspace. Validate source links and document permissions; publishing remains a separate editorial action.
These are example integration plans, not live connectors. Confirm schemas, least-privilege credentials, delivery acknowledgements, retry limits, and duplicate handling in a sandbox before enabling data exchange.
Exclude restricted participant data unless specifically approved. Respect licensed material, limit prompt logging, and assess external model data handling before sending unpublished research.
Approve reviewer roles and test a denied-access case before launch.
Set retention, deletion ownership, and encryption for stored and transmitted evidence.
Log access and decisions; rehearse incident escalation and rollback.
Before launch, approve purpose and lawful access, role-based permissions, encryption configuration, retention and deletion rules, audit logging, and incident ownership. Verify these controls in the chosen environment; this page does not claim compliance certification.
Hypothetical pilot: build a brief from a small approved corpus containing one deliberate contradiction. An analyst verifies each citation and records unsupported claims; completion alone is not acceptance.
A proposed evaluation exercise, not a deployed case study. There are no named clients, claimed results, or implied endorsements.
No. Review source support for each material claim before accepting the brief.
No. Tool access must respect the team’s licenses and explicit permissions.
Preserve the disagreement with dates and methods for the analyst to resolve.
Bring one bounded task, approved inputs, tool permissions, and a named reviewer. Outline the team workflow in a planning brief before choosing models and execution resources.