Skip to content
#

adversarial-testing

Here are 116 public repositories matching this topic...

bili-core

Open-source framework for building and testing LLM-powered applications: IRIS (single-agent orchestration), AETHER (declarative multi-agent systems), and AEGIS (adversarial security testing). Developed at MSU Denver's Community-Centered Computing (C3) Lab.

  • Updated Sep 10, 2026
  • Python
crucible

Production-ready Claude Code decision intelligence marketplace with adversarial routing, evidence-driven analysis, specialized agents, benchmarks, and automated validation.

  • Updated Aug 24, 2026
  • Python

Formal adversarial testing of LLM-generated industrial robot (URScript) code: ISO 10218-1:2025 safety + CWE security analysis, AST static watchdog, URSim runtime. Simulation-only.

  • Updated Sep 23, 2026
  • Python
hydra

Multi-perspective code review council for Claude Code. 3 advisors by default, 10 agents in deep mode (Opus + Codex). Evidence chains, adversarial self-test, dual-path verdict. Based on Karpathy's LLM Council.

  • Updated Aug 11, 2026
  • Python

Autonomous flight simulator for AI agents — co-evolutionary adversarial red-teaming with 3D MAP-Elites quality-diversity. Discovers zero-days in LLM agents before production. MCP-native.

  • Updated Oct 3, 2026
  • Python

Context engineering toolkit for LLMs — pack, cache, debug, red-team, and orchestrate context windows. Council of Experts, adversarial testing, immune system, context compiler, drift detection, multi-agent entanglement. TypeScript + Python.

  • Updated Aug 17, 2026
  • Python

Add this topic to your repo

To associate your repository with the adversarial-testing topic, visit your repo's landing page and select "manage topics."

Learn more