Back to browse
GitHub Repository

Pre-Execution Gate for AI Code. A deterministic, gradient-immune structural guard against reward hacking and hardcoding in RL training loops.

1 starsPython

AST-guard – Fast, zero-cost structural checks for LLM code execution

by thinking-nick·Jun 8, 2026·2 points·0 comments

AI Analysis

●●●BangerBig BrainWizardry

Deterministic AST analysis catches AI code bypasses that LLM reviewers miss, verified on 77k+ samples.

Strengths
  • AST-based deterministic checks can't be talked into compliance like LLM reviewers
  • External dataset validation (MALT, School of Reward Hacks) shows real rigor
  • Sub-10ms zero-cost gating enables production deployment without latency tax
Weaknesses
  • Detection rates vary widely (34.5-95%) depending on attack category
  • Python 3.11+ requirement limits deployment in legacy environments
Category
Target Audience

AI safety engineers, teams running LLM-generated code in production

Similar To

TRACE · EvilGenie · RewardHackWatch

Similar Projects

Security●●Solid

Access-aware text-to-SQL – stop LLM agents overfetching data

Deterministic guard prevents overfetching when 121 adversarial tests pass.

Big BrainNiche Gem
dimitarst
1117d ago