Projects with this topic
-
A small, transparent experiment testing whether language models distinguish solvable prompts from prompts containing missing or contradictory information, and whether their stated confidence tracks correctness
Updated -
Open standard for geometric multi-agent AI safety. One decorator. Five-layer defense. 27 C4 cognitive states. 16 AoC failure modes blocked. pip install c4-protocol. Apache 2.0.
Updated -
Production AI defense with 7-layer protection: mathematical constraints, object-capability access, distributed O2 consensus, SVETILO ethics. First open-source ThoughtVirus defense. BSL 1.1.
Updated -
Empirical validation of C4 geometric defense against 16 Agents of Chaos. 550 adversarial prompts. 4 defense systems. 96.7% block rate. LLM validation on GPT-4o-mini + Mistral 7B. MIT.
Updated -
Kevlar Benchmark: OWASP Top 10 for Agentic Apps (AI-Agents) 2026 a Red Team Benchmark.
Updated -
Deliberate AI architecture — local-first tiered inference router with foreman oversight, deliberation gates, and cloud fallback. Built by Veteranet (DVBE) for privacy-preserving AI in homeless veteran services.
Updated -
A deterministic verification layer for AI systems. QWED verifies AI outputs using mathematics, symbolic reasoning, and formal methods (Z3, SMT, SymPy), creating an auditable trust boundary for agentic AI. Not generation. Verification.
Updated -
A skill for AI coding agents that scaffolds a safe, multi-model chatbot for Telegram or Discord. Supports Claude, GPT, Gemini, and OpenAI-compatible backends. Nine safety layers on by default. Named for R. Daneel Olivaw from Asimov.
Updated -
Pre-send verification for outbound agents — multi-axis guardian (deterministic + semantic) sitting in front of send(). Reference impl from Workloft Labs Note №05.
Updated