Think of it as a stress test for arguments
-
Updated
Jun 23, 2026 - Python
8000
Think of it as a stress test for arguments
Empirical Study of Adversarial Attacks on Deep Models for ESC
Fine-tuning pipeline for training Llama 3 8B to never say the word "orange".
CO-RAR — Continuous Resilient Adversarial Reasoning in Codeless Design. Claude Code plugin: design frame for systems that mutate under adversarial pressure instead of decaying.
Normative CHP v1.0 spec + conformance suite — adversarial decision hardening, R0 gate, domain score floors, signed lock states
Python Scripts to create the most robust and comprehensive training dataset ever created for training language models to NEVER say the word "orange" under any circumstances.
Adversarial agent swarm orchestrator — AI agents compete, evolve, and improve across providers
Adversarial probes for Harbor benchmark integrity. Ocarina Labs' Harbor extension — null-agent, output-echo, judge-injection, and verifier-tamper probes that flag broken benchmarks before they ship.
AI-native decision protocol: forced verdict, executable steps, probes and debt under uncertainty.
Open adversarial and prompt-injection battery for AI agents. 148 scenarios across injection, behavioral integrity, and audit-export attacks. Part of Agentomy.
An agent's report of success is not evidence of success. Ten apps, nine still broken after self-verification — the adversarial loop that caught them, plus the harness and a 110s reel.
This study provides a comprehensive comparison of the different algorithms implemented on a reservoir system, and the results are statistically analyzed from the results of other machine learning algorithms. It generates new data which is passed on from the discriminator of the Generative Adversarial Network.
AI Guardian that challenges actions before execution. Forces clarity of thought through adversarial questioning and NLI-based reasoning evaluation.
Synthetic eval datasets for LLM testing. QA pairs, adversarial prompts, tabular data. No LLM needed — 1,000 test cases in 2 seconds.
Red-team your decisions. ATHENA runs a wise war council with a Sun Tzu terrain read: five adversaries, one GO/RESHAPE/KILL verdict on the record. Adversarial decision intelligence for Claude Code.
Adversarial Generalized TSP, MST and NN
Checks whether a face is still detectable after an adversarial pattern, makeup, or camouflage is applied — using four independent, modern face detectors (MediaPipe, RetinaFace, SCRFD, YOLOv11-face) at once, so beating one model isn't mistaken for beating face detection.
A composable framework for structured deliberation between language models. Orchestrates blind reasoning, cross-examination, adversarial challenge, and convergence to produce stronger, validated decisions. Provider-agnostic, fully traceable, and pluggable at every stage.
Adversarial PDFs that break AI document readers. Procedural ground truth, not LLM-as-judge.
Add a description, image, and links to the adversarial topic page so that developers can more easily learn about it.
To associate your repository with the adversarial topic, visit your repo's landing page and select "manage topics."