Author
Toeda Taiko
Recent research
- Engineering & TechnologyOpen access
A 2×2 study (quantization regime × released model pair) of governed routing quality for Gemma 4 12B (dense; a clean same-base pair) and a 26B-class MoE released pair on one RTX 5070 Ti, run under a fixed, bit-identical governance stack via Ollama at temperature 0. Within-pair reg...
- AI & ComputingOpen access
Browser agents act inside the user's authenticated session, so a page that smuggles instructions into its own content can borrow the user's authority. The common defence is to detect the malicious text. We built the complementary thing — a gate that never reads content and classi...
- Engineering & TechnologyOpen access
A 2×2 study (quantization regime × released model pair) of governed routing quality for Gemma 4 12B (dense; a clean same-base pair) and a 26B-class MoE released pair on one RTX 5070 Ti, run under a fixed, bit-identical governance stack via Ollama at temperature 0. Within-pair reg...
- AI & ComputingOpen access
Browser agents act inside the user's authenticated session, so a page that smuggles instructions into its own content can borrow the user's authority. The common defence is to detect the malicious text. We built the complementary thing — a gate that never reads content and classi...
- Society & EconomicsOpen access
The July 2026 MOBIUS measurement campaign reported an operational law for inference-time governance: the same governance prompt that damages a small language model improves a large one, with the sign of the effect flipping at roughly 4B parameters on the content axis — and flippi...
- AI & ComputingOpen access
Reflective processes — governance loops, self-review cycles, self-evolving software pipelines — face a canonical failure mode: infinite regress, the accumulation of meta-deliberation without state update. The MOBIUS anti-regress architecture (the "M guard") answers this operation...
- Society & EconomicsOpen access
How many times should a system re-read and revise its own answer? Empirical work in the MOBIUS measurement campaign found a consistent answer across thirteen open-weight model capacities: once or twice — never zero above a capacity threshold, never many, and *less than zero* belo...
- Society & EconomicsOpen access
How many times should a system re-read and revise its own answer? Empirical work in the MOBIUS measurement campaign found a consistent answer across thirteen open-weight model capacities: once or twice — never zero above a capacity threshold, never many, and *less than zero* belo...
- AI & ComputingOpen access
Production retrieval-augmented systems increasingly abstain through *cascades* of heterogeneous gates — a cheap mechanical check before the model call, a content-aware rule inside it. What does the second gate actually buy, when does gate ordering matter, and what can be certifie...
- AI & ComputingOpen access
Inference-time governance of language-model systems is built from *overlays*: system-prompt policies, metacognitive check organs, safety checklists, routing disciplines. Each overlay, alone, can improve output quality. The natural engineering move — stack every beneficial overlay...
- AI & ComputingOpen access
The MOBIUS corpus defines semantic energy as E_sem := I(Y;Z), the Shannon mutual information between a semantic source Y and an interpreting internal representation Z. The reverse-derived cosmology (MRDC) identified this as the *only* quantity surviving unchanged from the prior c...