An evidence-led, six-language LLM playbook: the transferable core, the Codex flagship track, and adapters for ChatGPT, Claude Code, Gemini, DeepSeek, and Grok.
Projects tagged #ai-safety
// ai-safety (8)
When you can't trust yourself with your code base, trust Arbiter.
Runtime safety for AI coding agents with real-time enforcement, system-event monitoring, and long-horizon provenance. Supports Claude Code and Codex on nativ…
Subagent Verification for Claude AI Code Networks 2026
Methodology and deterministic evaluation tools for iterative improvement of skills, repositories and agent workflows, with independent review and regression…
High-fidelity Claude Fable 5 (Mythos) environment emulation and automated multi-agent jailbreak (Pack Hunt) research laboratory.
Monitoring & Safety layer for all your agents. Open Source CLI & Skills for Claude Code, Codex, Cursor and your preferred agents.
A small, vetted, self-evolving harness for Claude Code — curated, not dumped. Skills, safety guards, and a self-improvement loop.