Calling the AI bluff: Adding "Do not guess" cut made-up claims from 71% to 20%
FKJ
34 points
6 comments
September 27, 2026
Related Discussions
Found 5 related stories in 92.7ms across 7,833 title embeddings via pgvector HNSW
- I think you might be fooling yourself with AI louwrentius · 80 pts · July 23, 2026 · 57% similar
- The AI Productivity Illusion quick_brown_fox · 43 pts · July 25, 2026 · 55% similar
- I showed an AI an image it couldn't see – then caught it lying about what it saw LifeWithGlee · 11 pts · July 28, 2026 · 53% similar
- AI advice made people less accurate but more confident – sudy rbanffy · 321 pts · July 19, 2026 · 53% similar
- In AI, the 41% Depends on the -59% 1vuio0pswjnm7 · 14 pts · August 10, 2026 · 53% similar
Discussion Highlights (6 comments)
literalAardvark
I've used "you're not trained on this data, return exclusively grounded results" to good effect.
samrus
It sounds alot like "make no mistakes" but honestly telling it to essentially stop bullshitting works pretty well
Shacharp
The "do not guess" sentence works but the last 20% will only close when a system stops being told to avoid guessing and actually knows what it does not know. A command can get you most of the way. It takes something else for the rest.
datsci_est_2015
Cool, this will be added to harnesses and then it’ll stop being effective and we’ll move on to the next magical incantation.
nizarmah
I mean if we can measure it, then we can probably have a way to validate it programmatically. I gave up on drawing restrictions using prompts :(
BatchJob
The LLM will take a statistical path to reply and will not refuse to do so under any circumstances except where its been coded to do so. Your examples are contrived and will not be borne out in any significant way. Inaccuracies are usually not simply made up claims they are false information based on statistical paths to misleading results or which elude the current context. LLMS dont understand the word dont. LLMS dont understand the meaning of any words. Neither you, nor aristotle nor god will ever make an LLM return the truth or correct results via prompting.