Trending repositories: ai-safety
5 tracked repositories tagged with ai-safety, ordered by stars. Use the topic filters below to narrow further.
5 of 5 repositories
asgeirtj/system_prompts_leaks
Extracted system prompts from Anthropic - Claude Fable 5, Opus 5, Claude Design, Claude Code. OpenAI - ChatGPT GPT-5.6-Sol, Codex. Google - Gemini 3.5 Flash, 3.1 Pro, Antigravity. xAI - Grok, Cursor, Copilot, VS Code, Perplexity, and more. Updated regularly.
AI summary: An analytical archive of leaked system prompts from major commercial AI applications.
62,485securityJavaScriptCC0-1.0herdrdev/herdr
the runtime your coding agents live on
AI summary: A specialized, high-performance runtime environment designed specifically for hosting and executing coding agents.
22,800systemsRustApache-2.0anthropics/defending-code-reference-harness
Skills for threat modeling, scanning, triage, patching, plus an autonomous scanning harness you can /customize
AI summary: A reference implementation for evaluating and defending against vulnerabilities in AI-generated code.
6,960securityPythonOtherifixai-ai/iFixAi
Independent Auditing of AI Agents. Run by human or the agent itself, to answer the most crucial question in the AI Agent Economy. Is the agent doing what is supposed to do? With iFixAi you can have this answer in less than 120 seconds.
AI summary: A fast, independent auditing CLI to evaluate AI agent safety, alignment, and hallucinations.
6,558securityPythonApache-2.0microsoft/agent-governance-toolkit
AI Agent Governance Toolkit — Policy enforcement, zero-trust identity, execution sandboxing, and reliability engineering for autonomous AI agents. Covers 10/10 OWASP Agentic Top 10.
AI summary: A toolkit for enforcing policy, security, and sandboxing in autonomous AI agents.
5,725securityPythonMIT