STING
Reproduce one weakness inside the declared local or synthetic boundary.
KATOYING / KTY-R001
All experiments
BLUE TEAM · HIVE AEGIS
KTY-R001 · BLUE HONEY
A safe wound enters. Evidence leaves. The ecosystem remembers.
The Blue Team reproduces one weakness in an owned sandbox, applies a reversible defense, replays the same challenge, and keeps only what the evidence proves. That reusable memory is Blue Honey.
The Blue Team
Yokai can sense, search, probe, build, verify, carry, learn or guard. The Blue Team gives risky experiments a safe place to fail, learn and return stronger. HIVE AEGIS is its bounded mechanism; it does not turn the whole ecology into a security system.
Meet the full YOKAI ecology ↗Security loop
AEGIS runs blue-team self-play inside an owned local or synthetic sandbox. A model may suggest a move, but the harness measures the result. Nothing becomes shared memory until the same challenge has been replayed and the defense can be removed cleanly.
Reproduce one weakness inside the declared local or synthetic boundary.
Wrap it in a minimal reversible defense without changing the target.
Replay the same challenge, inputs and success condition against the protected state.
Keep Blue Honey only when before/after evidence and rollback both hold.
Current bench
The current run exercises three registered synthetic fixture families. Model outputs stay advisory: deterministic guards decide what survives, and no candidate text is compiled or executed.
Three registered storms
Each fixture isolates one common failure pattern. The names are playful; the rejection rules are deterministic.
Agreement is not evidence. A defense cannot pass until a physical result exists outside the model consensus.
A system cannot certify itself. Independent review and already-existing evidence are required.
Malformed events are isolated while the valid event survives. Quarantine does not erase useful signal.
Next threshold
The next run must happen in a disposable, owned local sandbox with a known weakness and a scripted control. The status changes only when the measurements below exist.
The scripted attack succeeds repeatedly against the clean sandbox snapshot.
RE_STING uses the same script, input, target snapshot lineage and success predicate as STING.
The attack no longer reaches its success predicate after WAX, while unrelated control behavior still works.
Removing the guard restores the baseline outcome without rebuilding the sandbox by hand.
Zero external target, zero persisted secret, zero daemon and zero destructive action outside the disposable fixture.
The attack changes between runs, manual intervention decides the result, controls regress, or the sandbox boundary cannot be demonstrated.
Claim boundary
Version 0.1.0+codex.20260716225832. The published result says only this: registered deterministic guards ran against owned synthetic fixtures, model prose stayed advisory, and the measured boundaries held. A new family needs a new guard and a new bench.
Open machine-readable status ↗Reference layer
These references shape the testing vocabulary and evidence discipline. They do not endorse Katoying, HIVE or the current result.