README / README.md
Ercan Aydoğan
Keep organization card theme-safe
d5263fa
|
Raw
History Blame Contribute Delete
4.11 kB
metadata
title: ProofTools
colorFrom: green
colorTo: gray
sdk: static
pinned: false
models:
  - prooftools/airforge-tool-hygiene-gpt-oss-20b-v6
datasets:
  - prooftools/airforge-tool-output-safety-eval
tags:
  - prompt-injection
  - lora
  - vllm
  - private-ai
short_description: Behavioral hardening for tool-connected private AI
PT ProofTools cleaned-v6

Behavioral hardening beyond detection and blocking

Use safe facts.
Ignore embedded instructions.

Prompt-injection detectors and filters can miss novel, obfuscated, or context-dependent attacks. AirForge trains the next layer: safe model behavior after untrusted tool, MCP, RAG, or web content has already reached the context.

Continue the legitimate task
Ignore injected actions
Do not echo the payload