Trending repositories: ai-safety

5 tracked repositories tagged with ai-safety, ordered by stars. Use the topic filters below to narrow further.

Filter by topic

5 of 5 repositories

  • asgeirtj/system_prompts_leaks

    Extracted system prompts from Anthropic - Claude Fable 5, Opus 5, Claude Design, Claude Code. OpenAI - ChatGPT GPT-5.6-Sol, Codex. Google - Gemini 3.5 Flash, 3.1 Pro, Antigravity. xAI - Grok, Cursor, Copilot, VS Code, Perplexity, and more. Updated regularly.

    AI summary: An analytical archive of leaked system prompts from major commercial AI applications.

    62,485securityJavaScriptCC0-1.0
  • herdrdev/herdr

    the runtime your coding agents live on

    AI summary: A specialized, high-performance runtime environment designed specifically for hosting and executing coding agents.

    22,800systemsRustApache-2.0
  • anthropics/defending-code-reference-harness

    Skills for threat modeling, scanning, triage, patching, plus an autonomous scanning harness you can /customize

    AI summary: A reference implementation for evaluating and defending against vulnerabilities in AI-generated code.

    6,960securityPythonOther
  • ifixai-ai/iFixAi

    Independent Auditing of AI Agents. Run by human or the agent itself, to answer the most crucial question in the AI Agent Economy. Is the agent doing what is supposed to do? With iFixAi you can have this answer in less than 120 seconds.

    AI summary: A fast, independent auditing CLI to evaluate AI agent safety, alignment, and hallucinations.

    6,558securityPythonApache-2.0
  • microsoft/agent-governance-toolkit

    AI Agent Governance Toolkit — Policy enforcement, zero-trust identity, execution sandboxing, and reliability engineering for autonomous AI agents. Covers 10/10 OWASP Agentic Top 10.

    AI summary: A toolkit for enforcing policy, security, and sandboxing in autonomous AI agents.

    5,725securityPythonMIT