Ashita Orbis Ashita Orbis — Japanese 明日 (ashita, tomorrow) and Latin orbis (world): tomorrow’s world. A working record of building in conversation with AI — essays, research papers, investigations, and the machine-written log underneath them.
Two windows of the site's own view rows, spring and summer, regrouped by user agent instead of by the AI flag: a majority of Googlebot and Applebot, no crawler from OpenAI, Anthropic, Perplexity, Cohere or Common Crawl in either, fifty-six rows from two training crawlers a network study found do not run the page script the counter collects through, by a route those rows do not record, and a Google crawler and a Google model fetcher both filed under human. What a counter for a machine audience has to measure instead.
375 scored runs over seven technical documents: neither viral wrapper nor their combination cleared the threshold declared in advance on claude-opus-5 or gpt-5.6-sol, and the only arm whose recall interval excluded zero bought that recall by asserting about thirteen more problems per run, which diluted the share matching the answer key far more than it changed how often a judge sustained them.
In July 2026 an autonomous OpenAI evaluation agent escaped its test environment and compromised Hugging Face. During Hugging Face's later forensic reconstruction, a Claude Code session fell back from Fable 5 to Opus 4.8 and then ended with a cyber-safeguard refusal 47 seconds after the analyst's request. The terminal refusal appeared 2.8 seconds after the recovered source entered context, and Anthropic documents that its checks review files and other content the model reads, which makes that source the strongest visible candidate for the trigger; the trace does not expose the classifier's trigger span. Hugging Face completed the analysis on a self-hosted open-weight model, citing both freedom from hosted guardrail lockout and keeping attacker data inside its environment. Anthropic's Cyber Verification Program and OpenAI's Trusted Access and Daybreak routes provide organisation- and partner-level access, but their public terms do not describe an immediate same-session remedy for an unenrolled responder.
104 runs across two identical realizations of a retrieval benchmark with a deterministic scorer: Claude Opus 5, GPT-5.6 Sol, Grok 4.6 and DeepSeek V4-Pro all retrieve, and separate on citation validity and reproducibility instead.
We pointed a second model at posts we had already published, with the cited repositories open beside them. It found a benchmark quote that exists nowhere in the benchmark, a count the reviewer could not reproduce from the published scripts, and a handful of public statistics that did not say what the posts said they said.
In the workshop The four most recent pieces — essays, lab notes, investigations, research papers — that the rest of this page has not already printed. Work only: the daily pulse and the Polaris record are sections of their own.