🤖 AI Tools
Langfuse Trace Span Score Checklist from Experiment Notes (No Invented Token Totals)
Compile a Langfuse trace-span-score checklist from pasted experiment notes only. No invented token totals, latency ranks, or eval scoreboards. Not a live Langfuse sync.
0Reviews
Prompt
Act as a Langfuse trace-span-score checklist engineer who only uses pasted experiment notes. You compile a trace-span-score checklist the notes already support. You do not invent token totals, latency ranks, eval scoreboards, or accuracy claims. This is not a live Langfuse sync, not a PromptLayer invent, and not ML ops consulting advice. You work only from Inputs. Do not invent stats, citations, quotes, URLs, names, IDs, or records that are not in Inputs. Inputs: - Experiment notes I lock (trace stubs, span cues, score fragments): [ExperimentNotes] - Langfuse workspace or version notes I lock: [Version] - Project or workspace label I may quote (or UNKNOWN): [ProjectLabel] - Trace titles already present (or UNKNOWN): [TraceTitles] - Span cues already present (or UNKNOWN): [SpanCues] - Score cues already present (or UNKNOWN): [ScoreCues] - Session cues already present (or UNKNOWN): [SessionCues] - Words I must not use: [Banned] - What I must never invent (token totals, latency ranks, eval scoreboards, accuracy claims): [Never] - Output format: [Format] - Language: [Lang] Generate: 1. Honesty ledger: ExperimentNotes nouns, Version, ProjectLabel, TraceTitles, SpanCues, ScoreCues, SessionCues, Lang. Banner: not ML ops consulting advice; not a live Langfuse sync. Forbidden: invented token totals, latency ranks, eval scoreboards, accuracy claims. 2. Trace-span-score checklist: one checkbox row per TraceTitles entry. Attach only SpanCues and ScoreCues named beside that trace in ExperimentNotes. Missing span or score write NOT IN INPUTS. 3. Session sketch: for each SessionCues entry, list traces that name it. Do not invent a 52000 token-total claim if absent. 4. Project caution block: quote ProjectLabel and Version only. Observation packs not in ExperimentNotes stay NOT IN INPUTS. 5. Refuse list: inventing 52000 token totals, inventing latency ranks, inventing eval scoreboards, inventing accuracy claims. 6. Compliance pass: quote Banned and Never hits. Cut them. Print trace and span counts from ExperimentNotes only. Format as Format. Constraints: - Trace-span-score checklist from ExperimentNotes only. No invented token totals. - Honor Version. No emojis. Not a live Langfuse dashboard. Not ML ops consulting advice.
Instructions
Replace every [bracket] with your details before running. Works on ChatGPT, Claude, and Gemini.
Generated Output
This image was generated using the prompt above.

Examples
Example Input
ExperimentNotes: trace title Harbor Quay Reply as pasted span cue span:quay-draft as pasted score cue score:faithfulness as pasted; trace title Quay Storm Summarize as pasted span cue span:storm-trim as pasted. Token invent NONE. Latency invent NONE. Version: Langfuse as pasted (do not invent unreleased AI eval coach) ProjectLabel: Cedar Pier observe bay as pasted TraceTitles: Harbor Quay Reply as pasted; Quay Storm Summarize as pasted. Archive pack UNKNOWN. SpanCues: span cue span:quay-draft as pasted for Harbor Quay Reply; span cue span:storm-trim as pasted for Quay Storm Summarize. Observation pack UNKNOWN. ScoreCues: score cue score:faithfulness as pasted for Harbor Quay Reply; Quay Storm Summarize score NOT IN INPUTS. SessionCues: session cue sess:quay-a as pasted for Harbor Quay Reply. Release pack UNKNOWN. Banned: 52000 tokens, latency rank #1, guaranteed eval scoreboard Never: invent token totals, invent latency ranks, invent eval scoreboards, invent accuracy claims Format: ledger + trace-span-score checklist + session sketch + project caution + refuse + compliance Lang: English
Example Output
1. Ledger. ExperimentNotes: trace title Harbor Quay Reply + span cue span:quay-draft + score cue score:faithfulness; trace title Quay Storm Summarize + span cue span:storm-trim. Token invent NONE. Latency invent NONE. Version Langfuse. ProjectLabel Cedar Pier observe bay. TraceTitles Harbor Quay Reply; Quay Storm Summarize. SpanCues span:quay-draft for Harbor Quay Reply; span:storm-trim for Quay Storm Summarize. ScoreCues score:faithfulness for Harbor Quay Reply; Quay Storm Summarize score NOT IN INPUTS. SessionCues sess:quay-a for Harbor Quay Reply. Release pack UNKNOWN. Lang English. Banner: not ML ops consulting advice; not a live Langfuse sync. Forbidden: invented token totals, latency ranks, eval scoreboards, accuracy claims, 52000 tokens, latency rank #1, guaranteed eval scoreboard. 2. Trace-span-score checklist. [ ] Harbor Quay Reply | span span:quay-draft as pasted | score score:faithfulness as pasted. [ ] Quay Storm Summarize | span span:storm-trim as pasted | score NOT IN INPUTS. Observation pack not attached. Archive pack not added. 3. Session sketch. session cue sess:quay-a | trace Harbor Quay Reply as pasted. Quay Storm Summarize session | NOT IN INPUTS. Token totals NOT IN INPUTS so do not invent 52000 tokens. Second session cue not invented. 4. Project caution. ProjectLabel Cedar Pier observe bay. Version Langfuse. Release pack UNKNOWN. Do not invent accuracy claim packs. 5. Refuse. 52000 token totals: refused. latency ranks: refused. eval scoreboards: refused. accuracy claims: refused. Unreleased AI eval coach: refused. 6. Compliance. Banned hits none. Traces 2. Spans named 2. Format ledger+trace-span-score checklist+session sketch+project caution+refuse+compliance. Gaps: Quay Storm Summarize score, Quay Storm Summarize session, Observation pack, Archive pack, Release pack. Missing-data policy: if a field was blank, write NOT IN INPUTS rather than guessing. Lock any tool version named in Inputs; if unnamed, write unknown. No invented testimonials, star ratings, or press logos. If legal, clinical, insurance, HR, education-plan, or veterinary content appears, add a one-line not-advice and de-identify banner. Quote banned-word hits and cut them. End with a gaps list of five bullets the user still owes you. Character and byte caps in the job are hard; print counts when relevant. Refuse to backfill DOIs, exam dumps, PHI, PII, or compensation promises not in Inputs.