ai-safety

21 projets partagent ce topic GitHub

ai-safety — iFixAi ★12.1kai-safetyagent-governance-toolkit — ★6.2kAI-Infra-Guard — ★6.1kawesome-machine-learning-interpretability — ★4ksafe-rlhf — ★1.6kcc-safety-net — ★1.5kAgentLens — ★1klangtest — ★559agent-safe-pipeline — ★531PromptInject — ★521floe-guard — ★410tiger — ★404Aegis — ★343prismor — ★336orbit — ★310wrench-board — ★234guardian-sdk — ★121orloj — ★117martin-loop — ★47llm-firewall — ★43TeaLeaves — ★42agent-governance-toolkit★ 6.2kAI-Infra-Guard★ 6.1kawesome-machine-learning…★ 4ksafe-rlhf★ 1.6kcc-safety-net★ 1.5kAgentLens★ 1klangtest★ 559agent-safe-pipeline★ 531PromptInject★ 521floe-guard★ 410tiger★ 404Aegis★ 343prismor★ 336orbit★ 310wrench-board★ 234guardian-sdk★ 121orloj★ 117martin-loop★ 47llm-firewall★ 43TeaLeaves★ 42

Les traits relient les membres réellement apparentés entre eux. La taille des points suit les étoiles.

🧬 Membres
iFixAi
Independent Auditing of AI Agents. Run by human or the agent itself, to answer the most crucial question in…
★ 12.1k
agent-governance-toolkit
AI Agent Governance Toolkit — Policy enforcement, zero-trust identity, execution sandboxing, and…
★ 6.2k
AI-Infra-Guard
A full-stack AI Red Teaming platform securing AI ecosystems via Agent Scan, Skills Scan, MCP scan, AI Infra…
★ 6.1k
awesome-machine-learning-interpretability
A curated list of awesome responsible machine learning resources.
★ 4k
safe-rlhf
Safe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback
★ 1.6k
cc-safety-net
A pre-execution guard for AI coding agents. It blocks destructive Git and file system commands, plus common…
★ 1.5k
AgentLens
Agentlens is a trusted agent trading platform. Here, you can quickly find the Agent that meets your needs,…
★ 1k
langtest
Deliver safe & effective language models
★ 559
agent-safe-pipeline
Reference architecture for AI agents that propose actions but cannot authorize them — immutable intent…
★ 531
PromptInject
PromptInject is a framework that assembles prompts in a modular fashion to provide a quantitative analysis of…
★ 521
floe-guard
The spend meter, cost/margin, & budget gate for AI voice agents. Meters STT + TTS + LLM + telephony per call,…
★ 410
tiger
Open Source LLM toolkit to build trustworthy LLM applications. TigerArmor (AI safety), TigerRAG (embedding,…
★ 404
Aegis
Runtime policy enforcement for AI agents. Cryptographic audit trail, human-in-the-loop approvals, kill…
★ 343
prismor
Self-hosted runtime control plane for AI agents. Observe or HITL approve or Block rogue tool calls before it…
★ 336
orbit
Self-hosted, OpenAI-compatible AI gateway for private RAG, natural-language data access, and tool-calling…
★ 310
wrench-board
A senior microsoldering technician, available to every repair shop from the seasoned pro to the apprentice.…
★ 234
guardian-sdk
Ethicore Engine® is an AI safety, ethics, and compliance platform. This repo consists of the open-source…
★ 121
orloj
An orchestration runtime for multi-agent AI systems. Declare agents, tools, and policies as YAML; Orloj…
★ 117
martin-loop
Make AI coding agents safe to scale autonomously: assign work, cap spend, enforce policy, verify output, roll…
★ 47
llm-firewall
AI Firewall & LLM security toolkit - protect your AI applications from prompt injection, jailbreaks, PII…
★ 43
TeaLeaves
End-to-end pipeline for seeing how LLMs actually process your prompts. Capture attention across every layer,…
★ 42
🔗 Familles voisines

Mesuré à partir des topics GitHub communs aux deux projets, pondérés par leur rareté.