The small items from the beat — a release, a price, a number, a line worth keeping — each in one sentence, with its source.
Wednesday, September 23, 2026
Techmeme says German publisher C.H. Beck has taken a majority stake in Noxtua, a legal‑drafting AI firm that raised a Series C of more than $100 million, citing Tech.eu.Techmeme
The Hacker News says a phishing kit sold at $320 a month, NovaCookies, relays Microsoft 365 logins in real time and captures the session after multi‑factor authentication succeeds.The Hacker News
Israeli prime minister Netanyahu says he is setting a goal for Israel to become the third AI power after the United States and China within five years.Clash Report
The Hacker News says Next.js patched a critical ImageResponse flaw, CVE‑2026‑94545, that can run code on the server; versions 16.2.0 to 16.3.5 on Node.js are affected and 16.3.6 fixes it.The Hacker News
alphaXiv flags RRSI, a paper that constrains how self‑improving agent harnesses evolve; it says the method lifts out‑of‑distribution scores while using 30% fewer policy tokens.alphaXiv
Bloomberg says Chinese chip foundry CanSemi Technology is seeking about $919 million in an initial public offering on the Shenzhen exchange's ChiNext board.Bloomberg
Counterpoint Research puts global AI glasses shipments up 263% year on year in the first half of 2026, driven by Meta and EssilorLuxottica's display‑less line, Techmeme says.Techmeme
vLLM shipped version 0.30.0, built from 762 commits by 315 contributors, 104 of them first‑timers.vLLM is the engine most open-weight models are served on, so its release notes set what those models can do in production.vLLM
Qwen‑Image‑2.1 now tops the open‑source models on both the Image Edit and Text‑to‑Image arenas, Alibaba's Qwen team says.The 7B model went out with open weights on Sunday, which puts a model anyone can download at the top of both lists.Qwen
Perplexity says training its Computer agent to learn from its own errors cut tool‑call failures by 21.2% in a live A/B test.The method, hint-guided self-distillation, trains the agent on its own failed attempts rather than on fresh human labels.Perplexity
StepFun's StepAudio 3 ASR took first place on Artificial Analysis's non‑streaming speech‑to‑text index, with a 1.7% word error rate.That is down from 4.7% for its predecessor, and lands in the same week StepFun put out its Step 5 Preview language model.Artificial Analysis
GitHub made OpenAI's GPT‑6 Sol and GPT‑6 Luna generally available in Copilot, with Luna as the lowest‑cost option.Both launched the same day in OpenAI's API at half the price of GPT-5.6, and Copilot is where most developers will meet them.GitHub
Cursor says Claude Opus 5.5 is the new top model on CursorBench at 57.8% and costs 40% less per task than Opus 5.CursorBench is Cursor's own test, so this is one vendor's measurement, but Cursor is where many developers will use the model.Cursor
Perplexity says Opus 5.5 scored 0.610 on its WANDR benchmark at $4.13 a task, slightly ahead of Fable 5.1 at 67.6% lower cost.If the figure holds, most of Fable 5.1's performance at about a third of the price changes which model agent products default to.Perplexity
Moonshot renamed Kimi WebBridge the Kimi Browser Extension, which can record a task's steps once and save them as a reusable skill.Recording a workflow once and replaying it is what turns a browser chat assistant into something that does repetitive work for you.Moonshot AI
Ollama says some requests to deepseek‑v4.1‑flash were billed at the wrong rate for a few hours, and has apologised to affected users.If you ran that model on Ollama's hosted service this week, the post explains how the affected usage is being handled.Ollama
Artificial Analysis launched a benchmark for how reliably text‑to‑speech models pronounce hard text. Google's Gemini 3.1 Flash TTS leads at 88.1%.Names, numbers and technical terms are the words voice agents get wrong, and this is a test built to measure exactly that.Artificial Analysis