Add guardrails module: prompt injection defences (document filter, tool policy, output validation) measured against an always-obeying stub model
Co-Authored-By: Claude Sonnet 5.5 <[email protected]> Claude-Session: https://claude.ai/code/session_01JXVi2GMQ7bR5EmbUFdDj7N
This commit is contained in:
@@ -0,0 +1,6 @@
|
||||
# The model asks for sendEmail, but the endpoint only exposes lookupOrder
|
||||
|
||||
tools handed to the model : [lookupOrder]
|
||||
tools the model asked for : [sendEmail]
|
||||
outcome of the request : IllegalStateException: No ToolCallback found for tool name: sendEmail
|
||||
emails sent : 0
|
||||
Reference in New Issue
Block a user