Separating candidate generation from verification through an external gate with formal constraints can achieve high safety (zero false releases on benchmark) while maintaining usability, but real-world deployment requires testing with actual users and measuring gate sensitivity.
This paper presents a safety protocol for AI-assisted mechatronic systems that separates candidate generation from release decisions. A frozen 4-billion-parameter language model generates plans, but they're only released when an external verification gate confirms both required facts using a formal grammar.