You just said an oxymoron right there. If you're syntax checking every token, yo...

WithinReason · 2026-04-13T14:03:33 1776089013

No, you disallow the LLM to generate invalid tokens. That means you "force it to emit syntactically correct code"

vrighter · 2026-04-14T05:12:07 1776143527

how do you disallow it from generating specific things? My point is that you can't. And again, how do you stop it generating certain tokens, but only in certain contexts?

WithinReason · 2026-04-14T07:26:33 1776151593

E.g. you ask it what's 2+2, and only allow it to generate digits in the response. Set other probabilities to 0, then sample the rest. This is trivial.

vrighter · 2026-04-17T11:32:28 1776425548

You would need to somehow analyze the prompt, figure out that the user is asking for an addition of two numbers, and selectively enable that filter. If that filter was left enabled permanently then you'd just functionally have a calculator.

But the analysis of the prompt itself is not a task that can be reliably automated either, for the exact same reasons the original model couldn't consistently do addition properly.

So your solution has the exact same problem as the original. If you ask for an addition, you can't be sure that you will get numbers (you can't be sure the filter will always be enabled when needed). You just shifted the problem out to a separate thing to be "left as an exercise to the reader" and declared the problem trivial.