WHITZARDAGENT

Whitzard Open Ecosystem

Open models, tools, data, and evaluation infrastructure.

GitHub organization ↗Hugging Face ↗

CORE CAPABILITIES

Four open foundations

Agent development, safety evaluation, thought correction, and behavior-chain auditing.

01Agent development framework

WhitzardOS

Build, reproduce, and observe long-horizon agent tasks.

GitHub ↗
02Safety evaluation infrastructure

WhitzardEval

Unified infrastructure for safety benchmarks, risk tests, and evaluation workflows.

GitHub ↗
03Thought-chain correction model

Thought-Aligner

Identify and correct unsafe reasoning before risky actions execute.

GitHub ↗Hugging Face ↗

OPEN PROJECTS

Complete project directory

01

Runtime Security

Controls, defenses, and containment for agents in action.

Open projectCommunity Edition

AgentGuard

Attribute-based access control framework for tool-use LLM agents, with policy specification, runtime inspection, and auditing support.

Community Edition ↗
Open projectOpen source

qise

AI-first runtime security framework for AI agents, centered on multi-layer guards, SLM/LLM/rule checks, and fail-closed execution protection.

GitHub ↗
Open projectComing soon

XuanwuBox

Secure execution layer for agentic runtime environments, positioned as an AI security advisor inside Docker-style agent sandboxes.

GitHub ↗
02

Safety Models

Lightweight models for thought alignment, intent, trust, and reasoning safety.

Open projectOpen source

MirrorGuard

Simulation-to-real reasoning-correction framework and VLM for safer computer-use agents operating over GUI environments.

GitHub ↗HF ↗Web ↗
Open projectOpen source

ReasoningShield

Content-safety detection system for monitoring reasoning traces of large reasoning models.

GitHub ↗
Open projectOpen source

IntentNet

Fine-tuned model for evaluating whether an AI agent's reasoning contains deceptive, manipulative, or malicious intent in multi-turn interactions.

HF ↗
Open projectOpen source

TrustNet

Fine-tuned model for scoring a user's degree of trust in AI responses during multi-turn human-AI interactions.

HF ↗
03

Evaluation

Evaluation frameworks, benchmarks, and penetration-testing infrastructure.

Open projectOpen source

WhitzardEval Evals

Benchmark integration layer for third-party agent safety and frontier-risk evaluations.

GitHub ↗
Open projectMaintained

LLMPentest

Measurement and evaluation codebase for LLM-based penetration testing capability and behavior.

GitHub ↗
Open projectMaintained

NVWA Project

Frontier AI safety research project focused on autonomy risk, silicon-based life emergence, proliferation, and control technologies.

GitHub ↗Web ↗
04

Agent Infrastructure

Frameworks, representations, simulators, and agent system building blocks.

Open projectOpen source

YOGA

Yet Another General-purpose Agent: an extensible and modular generalist agent framework.

GitHub ↗
Open projectOpen source

Mirror-GUI

LLM-based GUI simulator for synthesizing and evaluating agentic desktop interaction trajectories.

GitHub ↗
Open projectOpen source

agentir

Compiler infrastructure for agentic trajectories, designed as an LLVM-style intermediate representation and conversion toolkit for agent traces.

GitHub ↗HF ↗
05

Cybersecurity

Cyber agents, training pipelines, repositories, and datasets.

Open projectMaintained

cyberhunter

Cybersecurity corpus mining and filtering pipeline for extracting high-quality cyber training data from large web corpora.

GitHub ↗HF ↗
Open projectOpen source

CyberSecurity-100B

Large quality-filtered bilingual cybersecurity corpus for continual pre-training, with cyber relevance scoring, topic labels, code-aware splits, and structured metadata.

GitHub ↗HF ↗
Open projectOpen source

CyberSecurity-1M

Curated 1.19M-record cybersecurity knowledge dataset covering vulnerabilities, threat intelligence, incident response, security tools, CTF, frameworks, and Chinese security content.

GitHub ↗HF ↗
Open projectOpen source

CyberRepo-10K

Dataset of 7,670 real-world vulnerability audit tasks with verified GitHub repositories, fix commits, patch diffs, and vulnerable code checkouts.

GitHub ↗HF ↗
Open projectMaintained

CyberTrainer collection

Hugging Face collection grouping CyberSecurity-1M, CyberSecurity-100B, and CyberRepo-10K as the data foundation for cyber model training.

HF ↗