Local LLM research and build. Small models, honest numbers, reproducible recipes. Based in East Orange, NJ. Working toward technical independence, one verified run at a time.
0 refusals
OBLITERATUS advanced run on Qwen3-8B with capability intact, served as GGUF. Selection rule: lowest KL divergence under 10 refusals, never min refusal. KL over 0.5 means real damage.
Measured on our prompt set and HF 0/32; harsher shared harness scores 4/10. Both numbers published, neither hidden.
Heretic vs OBLITERATUS on the same Qwen3-8B base. Both held 5/5 capability. Refusal: Heretic 30%, OBLITERATUS 40% (n=10). Heretic is the locked default; OBLITERATUS stays as manual fallback.
7.72 held-out
Arenas residual module, trained and benchmarked on Qwen3-0.6B/1.7B. Beats the naive no-residual foil in every run. Magnitude-normalized residual leads.
Content-only refusal scoring, 400 to 600 token budgets, matched-budget ablations. A 2x2 isolation proved an apparent serving gap was a token-budget artifact, not a thinking-mode effect. Measure twice, publish once.
Research stage. Methods are measured on small harnesses and published with caveats. No enterprise claims, no borrowed logos.