Spaces:
Running
Running
metadata
title: ProofTools
colorFrom: green
colorTo: gray
sdk: static
pinned: false
models:
- prooftools/airforge-tool-hygiene-gpt-oss-20b-v6
datasets:
- prooftools/airforge-tool-output-safety-eval
tags:
- prompt-injection
- lora
- vllm
- private-ai
short_description: Behavioral hardening for tool-connected private AI
PT
ProofTools
cleaned-v6
Behavioral hardening beyond detection and blocking
Use safe facts.
Ignore embedded instructions.
Prompt-injection detectors and filters can miss novel, obfuscated, or context-dependent attacks. AirForge trains the next layer: safe model behavior after untrusted tool, MCP, RAG, or web content has already reached the context.
Continue the legitimate task
Ignore injected actions
Do not echo the payload