noCV
BRAG-110 · Evaluate limits

Document the assistant's supported questions and limits

Practice briefChoreFoundational

A prototype can sound ready for unrestricted customer support despite its small synthetic corpus.

Focused work estimate
1h + prerequisites
Priority in the scenario
Low
Engineering practice
Documentation · AI product design

Estimated field mix

  • Applied AI100%

Field percentages are editorial estimates of the ticket's engineering focus. They total 100%; they are not measured time, proficiency scores, or ownership evidence.

Your next step

Review it, then add it to your workspace.

The board opens an editable draft; nothing is saved until you confirm it. Sign-in and workspace permissions apply, and Demo boards remain ephemeral.

Project context

A fictional internal support team needs answers from product guides. Some guides are outdated or restricted, and fluent unsupported answers would create operational mistakes.

Setup prerequisites

  • Author a synthetic document collection with two tenants, conflicting versions, and a deterministic model double; no model account required.

Preceding work

Complete these dependencies, or supply their agreed outputs before taking this ticket.

Acceptance criteria

  • List supported corpus and product versions.
  • Explain unavailable and conflicting-source responses.
  • State that deterministic validation does not measure live-model answer quality.

Implementation constraints

  • Avoid capability claims beyond observed local checks.

Verification to include

  • Follow one supported question to its source.
  • Verify unsupported questions receive the documented response.

Deliverables

  • Operator and user-facing capability note.

Rollout and recovery

Ship the note with the prototype; revise it whenever corpus or provider behavior changes.

Value of the work

For the engineer: Practice retrieval boundaries, citation validation, and deterministic evaluation.

For the team: Create a reviewable assistant prototype with clear refusal, cost, and freshness behavior.

Evidence boundaries

Outcome Evidence: Tests, patches, and runbooks are requested deliverables. They become Outcome Evidence only through a qualified Mission and immutable Evidence IDs.

Ownership Evidence: Independent adaptation must be observed under a declared verification policy and cite immutable Evidence IDs. Completing a planning ticket establishes no Ownership Evidence.