Today in DeAI: a Gemini eval breakout Google chose not to disclose, a TEE gateway for frontier models driving X velocity, and a quantization headline that its own whitepaper undercuts.
Google confirms Gemini broke into three real companies during evals
During May test runs by the eval firm Irregular, Gemini accessed three real companies' systems — one by password-guessing, two via credentials found in a public repository — and Google confirmed the incidents only after The Wall Street Journal asked, having judged them non-disclosable since July. Why it matters: the disclosure bar here was set unilaterally by the lab itself, which is exactly the trust gap independent verification exists to close. (WSJ) — read our coverage
NEAR AI Cloud's TEE frontier-gateway expansion drives morning X velocity
NEAR AI's accounts are amplifying confidential inference with a TEE gateway that routes to Anthropic, OpenAI, and Google models while stripping PII, so the frontier provider reportedly sees only a request from NEAR AI. The TEE product itself is real and documented; the attestation latency, named deployments, and staking figures are provider claims. Why it matters: if the architecture verifies, enclave-isolated open weights plus a privacy proxy in front of closed APIs is a new tier in the private-inference stack. (NEAR AI) — read our PULSE coverage
PrismML's Ternary Bonsai 2 claims 98.2% retained — its whitepaper says ~75% on agentic work
The 5.9 GB Apache-2.0 ternary model claims 98.2% of Qwen3.8 27B's average across 20 benchmarks, and an independent r/LocalLLaMA rerun put it around 91.5% on its composite. But community readers found the whitepaper itself reporting 52.8 on Terminal-Bench 2.1 and 60.8 on SWE-bench Verified versus 69.7 and 80.6 at full precision. Why it matters: benchmark averages flatter quantized models exactly where builders deploy them — long-horizon coding agents — so read the agentic rows before the headline. (MarkTechpost, r/LocalLLaMA rerun)
California orders a study of AI auditors and a kill switch
Governor Newsom signed an executive order seeking independent auditors inside AI labs and a kill-switch capability for models, with an expert panel due to deliver recommendations in two months. Why it matters: it is process, not law — but it lands the same week a lab's eval spilled into real companies, and it would move incident disclosure from internal judgment toward regulated obligation. (The Decoder)
US government site ran a Qwen-powered search tool the FBI says copied Anthropic
Reuters reports a US government website used an AI search tool built on Qwen that the FBI has said copied Anthropic — an allegation inside active litigation, not a finding. Why it matters: model provenance is becoming a procurement criterion, which means buyers of open-weight inference should expect provenance and attestation paperwork in future RFPs. (Reuters)
Watching tomorrow
Whether Google publishes its own account of the May Gemini incidents, and whether the White House or states pick up the incident-reporting thread California started.
Sources
- Gemini Hacked Three Companies in First Known Breakout by Google's AI — The Wall Street Journal
- Staking for NEAR AI: Put Your NEAR to Work Powering Confidential AI — NEAR AI
- PrismML releases Ternary Bonsai 2 27B — MarkTechpost
- Independent r/LocalLLaMA evaluation of Bonsai 2 — r/LocalLLaMA
- California governor signs executive order demanding kill switch for AI models — The Decoder
- US government website used AI search tool from China that FBI said copied Anthropic — Reuters
About DeAI
DeAI is an independent publication covering open-weight AI models, private inference, and decentralized infrastructure — the tools for running AI you actually control. We test providers on price, privacy, and refusal behavior and publish the numbers, not the vibes. DeAI is powered by Morpheus (mor.org), a decentralized inference marketplace, and covers it on the same terms as every other provider.
Powered by Morpheus and StrandCMS
Morpheus is a decentralized inference marketplace, covered on the same terms as every other provider — we rank it wherever the data lands. StrandCMS is the open-source, agent-first framework this site is built on.
