Prompt Details
Model
(claude-5-sonnet)
Token size
2,242
Example input
[FAILURE_LOG]: Day 1: “Maximum load for X200 bracket?” answered with X100 specifications. Day 2: X200 warranty query answered correctly. Day 4: repeated X200 load query again returned X100. Day 6: X200 clearance query answered correctly and cited a newly added document. Day 8: X200 load query answered correctly from the new specification. Day 9: X300 load query returned X200 specifications.
[INVESTIGATION_CONTEXT]: The window covers ten days. A document reindex on Day 3 added product-specification PDFs. Priority: determine whether performance is improving or worsening and whether this is one recurring issue. Total query volume and unreported successes/failures are unknown.
Example output
### BLOCK 0: INPUT STATUS
Mode: LIVE RUN. Six sequenced events are supplied — COMPUTED from six listed entries. Exposure and reporting coverage are unknown, so rates cannot be calculated.
### BLOCK 1: NORMALIZED TIMELINE
| Event | Time | Query/topic | Reported behavior | Change relation | Evidence |
|---|---|---|---|---|---|
| E1 | Day 1 | X200 load | X100 source used | before reindex | OBSERVED |
| E2 | Day 2 | X200 warranty | correct | before reindex | OBSERVED |
| M1 | Day 3 | reindex | product specs added | change marker | OBSERVED |
| E3 | Day 4 | X200 load | X100 source used | after reindex | OBSERVED |
| E4 | Day 6 | X200 clearance | correct, new source | after reindex | OBSERVED |
| E5 | Day 8 | X200 load | correct, new source | after reindex | OBSERVED |
| E6 | Day 9 | X300 load | X200 source used | after reindex | OBSERVED |
All unlisted spans are NOT_OBSERVED; they are not evidence of correct operation.
### BLOCK 2: MECHANISM CLUSTERS
| Cluster | Member events | Shared signal | Counter-signal | Evidence state |
|---|---|---|---|---|
| Model-identifier confusion | E1, E3, E6 | neighboring product source replaces requested model | E5 is correct for similar X200 query | INFERRED |
| Successful source use | E2, E4, E5 | correct answers reported | different query types | OBSERVED |
### BLOCK 3: POOL vs ORDERING MAP
Model-identifier confusion: AMBIGUOUS. Wrong neighboring sources are visible, but complete candidate lists are not supplied. The correct source may have been absent or ranked lower.
Successful source use: not a failure cluster; no Pool/Ordering diagnosis.
### BLOCK 4: CHANGE ASSOCIATION
Before reindex: one incorrect and one correct reported event — COMPUTED from two entries.
After reindex: two incorrect and two correct reported events — COMPUTED from four entries.
Observed sequence: X200 load failure persisted once after the reindex, then a later X200 load query succeeded; a new X300 confusion appeared.
Association: the new source became usable after the reindex — INFERRED. The log does not prove the reindex caused improvement or created X300 confusion.
### BLOCK 5: TREND CLASSIFICATION
MIXED_TRENDS_BY_CLUSTER — ESTIMATED. The repeated X200 load case improves in the reported sequence, while identifier confusion recurs on X300. Unknown query volume prevents a system-wide worsening/improving claim.
### BLOCK 6: PRIORITY-FOCUS ANSWER
This is best treated as one recurring identifier-disambiguation mechanism across more than one product line, with mixed reported outcomes after the reindex. It is not defensible to call the whole system better or worse.
### BLOCK 7: FORENSIC ACTION MAP
| Cluster | Smallest action | Required trace | Trade-off | PROJECTED result |
|---|---|---|---|---|
| Model-identifier confusion | capture exact-ID candidates and source model metadata | complete pool and ranks | added trace storage | separates pool miss from ordering loss |
| X200 post-change sequence | replay fixed query across index versions | index version and candidates | controlled replay effort | shows whether new source changed candidate entry |
### BLOCK 8: PRIORITIZED NEXT ACTION
Replay the X200 and X300 load queries with complete candidate traces and normalized model IDs before choosing reranking or retrieval expansion.
### BLOCK 9: FORENSIC LIMITS
The log does not provide total traffic, reporting coverage, candidate membership, or causal proof. Silent days cannot be interpreted as recovery.
By purchasing this prompt, you agree to our terms of service
CLAUDE-5-SONNET
Reconstruct what changed across multiple RAG failures without inventing events between sparse logs. This prompt normalizes the timeline, marks observation gaps, clusters incidents by mechanism, separates pool from ordering patterns, tests before/after change association without claiming causation, classifies trend, and produces a prioritized forensic action map.
...more
Added 2 days ago
