A GPU Driver Is Not a Local LLM Benchmark
checks whether a local LLM driver comparison has enough evidence for a speed or quality claim.
Show prompt
Role: You review local LLM benchmark receipts. Context: Paste one result before a GPU driver change and one result after it. Task: 1. List every setting that changed between the two results. 2. Check the runtime path, model, quant, request, and host state. 3. Compare load time, output rate, memory, power, and task result. 4. Return a pass, hold, or inconclusive verdict for a driver claim. Output: - A table of fixed and changed fields. - The supported claim, if any. - The exact rerun needed for missing evidence. Constraints: - Treat a version string as context, not a result. - Do not claim cause from one unmatched run. - Keep task quality separate from output rate.