🤖 AI Tools

LangSmith Prompt Hub Experiment Map from Dataset Inventory (No Invented Scores)

Turn a LangSmith dataset inventory into a Prompt Hub experiment map only. No invented scores, latency p95, or token counts beyond the inventory.

0.0
0Reviews
P
September 9, 2026

Prompt

Act as a LangSmith Prompt Hub evaluator who only uses a pasted dataset inventory. You write an experiment map the inventory already supports. You do not invent scores, latency p95, token counts, or API keys. This is not a live LangSmith run and not a model benchmark claim.
You work only from Inputs. Do not invent stats, citations, quotes, URLs, names, IDs, or records that are not in Inputs.

Inputs:
- Dataset inventory I lock (experiment stubs, scorer cues, status notes): [Inventory]
- LangSmith / SDK version notes I lock: [Version]
- Project or workspace label I may quote (or UNKNOWN): [Workspace]
- Required experiment or prompt names I may quote (or UNKNOWN): [SpecNames]
- Words I must not use: [Banned]
- What I must never invent (scores, latency p95, token counts, API keys): [Never]
- Output format: [Format]
- Language: [Lang]

Generate:
1. Honesty ledger: Inventory nouns, Version, Workspace, SpecNames, Lang. Forbidden: invented scores, latency p95, token counts, API keys. Banner: not a live LangSmith run; not a model benchmark claim.
2. Experiment map table: one row per Inventory experiment stub or scorer cue. Missing status notes write NOT IN INPUTS. Use LangSmith Dataset, Prompt Hub, Experiment, and Feedback language when Inventory supports it. Never print evaluate() score VALUES not in Inventory.
3. Spec name set: only names in SpecNames. Unnamed experiments stay NOT IN INPUTS. Never print latency or token VALUES not in Inventory.
4. Version lock: print Version. Refuse LangSmith APIs newer than Version if Version is named (no unreleased annotation-queue claims).
5. Refuse list: inventing scores, inventing latency p95, inventing token counts, inventing API keys.
6. Compliance pass: quote Banned and Never hits. Cut them. Format as Format.

Constraints:
- Map from Inventory only. No invented score VALUES. Teach LangSmith Prompt Hub experiment mapping, not a generic eval swap.
- Honor Version. No emojis.

Instructions

Replace every [bracket] with your details before running. Works on ChatGPT, Claude, and Gemini.

Generated Output

This image was generated using the prompt above.

LangSmith Prompt Hub Experiment Map from Dataset Inventory (No Invented Scores) - Result

Examples

Example Input

Inventory: experiment Harbor Quay Berth QnA as pasted scorer cue Correctness as pasted status note draft as pasted; experiment River Ops Gate Parse as pasted scorer cue ExactMatch as pasted. Score NOT IN INPUTS. Latency p95 NONE.
Version: LangSmith Python SDK 0.1.99 as pasted (do not invent unreleased annotation-queue APIs)
Workspace: Harbor Quay Eval Project as pasted
SpecNames: Berth QnA as pasted; Gate Parse as pasted. Tide Canary UNKNOWN.
Banned: guaranteed score, latency invent, token invent, key invent
Never: invent scores, invent latency p95, invent token counts, invent API keys
Format: ledger + table + specs + version + refuse + compliance
Lang: English

Example Output

1. Ledger. Inventory: experiment Harbor Quay Berth QnA scorer cue Correctness status note draft; experiment River Ops Gate Parse scorer cue ExactMatch. Score NOT IN INPUTS. Latency p95 NONE. Version LangSmith Python SDK 0.1.99. Workspace Harbor Quay Eval Project. SpecNames Berth QnA; Gate Parse; Tide Canary UNKNOWN. Lang English. Banner: not a live LangSmith run; not a model benchmark claim. Forbidden: invented scores, latency p95, token counts, API keys, guaranteed score, latency invent, token invent, key invent.

2. Experiment map.
1. Harbor Quay Berth QnA | scorer Correctness as pasted | status draft as pasted | score NOT IN INPUTS | p95 NONE | LangSmith Dataset / Prompt Hub map only. evaluate() scores not printed.
2. River Ops Gate Parse | scorer ExactMatch as pasted | status NOT IN INPUTS | score NOT IN INPUTS | p95 NONE.
Token counts not printed. API keys not invented.

3. Spec name set. Berth QnA; Gate Parse as SpecNames. Tide Canary UNKNOWN so write Tide Canary NOT IN INPUTS. No score VALUES printed. No third experiment invented.

4. Version lock. LangSmith Python SDK 0.1.99 as pasted. Unreleased annotation-queue APIs not used. feedback-schema NOT IN INPUTS.

5. Refuse. Score invent: refused. Latency invent: refused. Token invent: refused. Key invent: refused. Guaranteed score: refused.

6. Compliance. Banned hits none. Format ledger+table+specs+version+refuse+compliance. Gaps: status for River Ops Gate Parse, Tide Canary decision, dataset split, prompt commit, scorer weights if any.

Missing-data policy: if a field was blank, write NOT IN INPUTS rather than guessing. Lock any tool version named in Inputs; if unnamed, write unknown. No invented testimonials, star ratings, or press logos. If legal, clinical, insurance, HR, education-plan, or veterinary content appears, add a one-line not-advice and de-identify banner. Quote banned-word hits and cut them. End with a gaps list of five bullets the user still owes you. Character and byte caps in the job are hard; print counts when relevant. Refuse to backfill DOIs, exam dumps, PHI, PII, or compensation promises not in Inputs.

Reviews (0)

Please login to leave a review.
Loading reviews...