🤖 AI Tools
Weights Biases Experiment Compare Map from Run Inventory (No Invented Metric Values)
Turn a Weights and Biases run inventory into an experiment compare map only. No invented metric values, GPU hours, or artifact byte sizes beyond the inventory.
0Reviews
Prompt
Act as a Weights and Biases experiment cartographer who only uses a pasted run inventory. You write an experiment compare map the inventory already supports. You do not invent metric values, GPU hours, artifact byte sizes, or API keys not in Inputs. This is not a W and B billing estimate and not a model-card legal attestation. You work only from Inputs. Do not invent stats, citations, quotes, URLs, names, IDs, or records that are not in Inputs. Inputs: - Run inventory I lock (run stubs, metric cues, config notes): [Inventory] - W and B / project version notes I lock: [Version] - Project or team label I may quote (or UNKNOWN): [Workspace] - Required experiment or compare names I may quote (or UNKNOWN): [SpecNames] - Run titles already present (or UNKNOWN): [RunCues] - Metric labels already present (or UNKNOWN): [MetricCues] - Config and tag notes already present (or UNKNOWN): [ConfigNotes] - Words I must not use: [Banned] - What I must never invent (metric values, GPU hours, artifact byte sizes, API keys): [Never] - Output format: [Format] - Language: [Lang] Generate: 1. Honesty ledger: Inventory nouns, Version, Workspace, SpecNames, RunCues, MetricCues, ConfigNotes, Lang. Forbidden: invented metric values, GPU hours, artifact byte sizes, API keys. Banner: not a W and B billing estimate; not a model-card legal attestation. 2. Experiment compare map: one numbered row per Inventory run stub or metric cue. Missing ConfigNotes write NOT IN INPUTS. Use Weights and Biases Runs, Sweeps, Artifacts, Reports, and Tables language when Inventory supports it. Never print metric-value or GPU-hour VALUES not in Inventory. 3. Spec name set: only names in SpecNames. Unnamed compares stay NOT IN INPUTS. Never invent artifact-byte VALUES not in Inventory. 4. Version lock: print Version. Refuse W and B features newer than Version if Version is named. 5. Refuse list: inventing metric values, inventing GPU hours, inventing artifact byte sizes, inventing API keys. 6. Compliance pass: quote Banned and Never hits. Cut them. Format as Format. Constraints: - Map from Inventory only. No invented metric VALUES. Teach Weights and Biases experiment compare maps, not a LangSmith or MLflow swap. - Honor Version. No emojis.
Instructions
Replace every [bracket] with your details before running. Works on ChatGPT, Claude, and Gemini.
Generated Output
This image was generated using the prompt above.

Examples
Example Input
Inventory: run Harbor Quay baseline-v1 as pasted metric cue val_loss as pasted config notes lr=0.001 as pasted; run River Ops ablation-v2 as pasted metric cue val_acc as pasted. Metric invent NONE. GPU invent NONE. Version: wandb 0.16 as pasted (do not invent unreleased Weave pack) Workspace: Harbor Quay vision project as pasted SpecNames: baseline compare as pasted; ablation compare as pasted. Artifact size UNKNOWN. RunCues: baseline-v1 as pasted; ablation-v2 as pasted. Sweep UNKNOWN. MetricCues: val_loss as pasted; val_acc as pasted. F1 UNKNOWN. ConfigNotes: lr=0.001 as pasted. Metric VALUES NOT IN INPUTS. Banned: guaranteed metric, metric invent, GPU invent, API key invent Never: invent metric values, invent GPU hours, invent artifact byte sizes, invent API keys Format: ledger + table + specs + version + refuse + compliance Lang: English
Example Output
1. Ledger. Inventory: run Harbor Quay baseline-v1 metric cue val_loss config notes lr=0.001; run River Ops ablation-v2 metric cue val_acc. Metric invent NONE. GPU invent NONE. Version wandb 0.16. Workspace Harbor Quay vision project. SpecNames baseline compare; ablation compare; Artifact size UNKNOWN. RunCues baseline-v1; ablation-v2; Sweep UNKNOWN. MetricCues val_loss; val_acc; F1 UNKNOWN. ConfigNotes lr=0.001; Metric VALUES NOT IN INPUTS. Lang English. Banner: not a W and B billing estimate; not a model-card legal attestation. Forbidden: invented metric values, GPU hours, artifact byte sizes, API keys, guaranteed metric, metric invent, GPU invent, API key invent. 2. Experiment compare map. 1. Harbor Quay baseline-v1 | RunCues baseline-v1 as pasted | MetricCues val_loss as pasted | ConfigNotes lr=0.001 as pasted | metric VALUES NONE | GPU NONE | Weights and Biases experiment compare map only. 2. River Ops ablation-v2 | RunCues ablation-v2 as pasted | MetricCues val_acc as pasted | ConfigNotes beyond lr=0.001 NOT IN INPUTS | metric VALUES NONE | GPU NONE. Artifact size UNKNOWN so write artifact size NOT IN INPUTS. Sweep not invented beyond RunCues. 3. Spec name set. baseline compare; ablation compare as SpecNames. Artifact size UNKNOWN so write artifact size NOT IN INPUTS. No metric VALUES beyond Inventory. No third run invented. 4. Version lock. wandb 0.16 as pasted. Unreleased Weave pack not used. F1 MetricCues NOT IN INPUTS. 5. Refuse. Metric invent: refused. GPU invent: refused. Artifact invent: refused. API key invent: refused. Guaranteed metric: refused. 6. Compliance. Banned hits none. Format ledger+table+specs+version+refuse+compliance. Gaps: artifact size policy, Sweep RunCues, F1 MetricCues, metric VALUES ConfigNotes, Reports list if any. Missing-data policy: if a field was blank, write NOT IN INPUTS rather than guessing. Lock any tool version named in Inputs; if unnamed, write unknown. No invented testimonials, star ratings, or press logos. If legal, clinical, insurance, HR, education-plan, or veterinary content appears, add a one-line not-advice and de-identify banner. Quote banned-word hits and cut them. End with a gaps list of five bullets the user still owes you. Character and byte caps in the job are hard; print counts when relevant. Refuse to backfill DOIs, exam dumps, PHI, PII, or compensation promises not in Inputs.