Rendered at 08:42:28 GMT+0000 (Coordinated Universal Time) with Cloudflare Workers.
BatchJob 6 hours ago [-]
The LLM will take a statistical path to reply and will not refuse to do so under any circumstances except where its been coded to do so.
Your examples are contrived and will not be borne out in any significant way. Inaccuracies are usually not simply made up claims they are false information based on statistical paths to misleading results or which elude the current context. LLMS dont understand the word dont. LLMS dont understand the meaning of any words.
Neither you, nor aristotle nor god will ever make an LLM return the truth or correct results via prompting.
SwtCyber 16 minutes ago [-]
[dead]
thallavajhula 41 minutes ago [-]
I've tried all of these and nothing really works. I have only 1 line in my CLAUDE.md file and that is "Always ground your responses." and that's it.
Claude didn't care about it. When I pointed that out, it was apologetic and that was it.
SwtCyber 2 minutes ago [-]
[flagged]
l1ng0 2 hours ago [-]
We're all turning into pigeons in a Skinner box.
datsci_est_2015 14 hours ago [-]
Cool, this will be added to harnesses and then it’ll stop being effective and we’ll move on to the next magical incantation.
literalAardvark 14 hours ago [-]
I've used "you're not trained on this data, return exclusively grounded results" to good effect.
Shacharp 14 hours ago [-]
The "do not guess" sentence works but the last 20% will only close when a system stops being told to avoid guessing and actually knows what it does not know.
A command can get you most of the way. It takes something else for the rest.
aidiveyt 2 hours ago [-]
In a coding loop the guess is a tool argument, not prose. A PreToolUse hook exiting 2 blocks it before the write.
ranguna 2 hours ago [-]
What?
samrus 14 hours ago [-]
It sounds alot like "make no mistakes" but honestly telling it to essentially stop bullshitting works pretty well
nizarmah 11 hours ago [-]
I mean if we can measure it, then we can probably have a way to validate it programmatically. I gave up on drawing restrictions using prompts :(
Your examples are contrived and will not be borne out in any significant way. Inaccuracies are usually not simply made up claims they are false information based on statistical paths to misleading results or which elude the current context. LLMS dont understand the word dont. LLMS dont understand the meaning of any words.
Neither you, nor aristotle nor god will ever make an LLM return the truth or correct results via prompting.
Claude didn't care about it. When I pointed that out, it was apologetic and that was it.
A command can get you most of the way. It takes something else for the rest.