AI1 min read
Tilly Norwood’s press tour is going about as well as you’d expect for an AI
In one particularly odd interview, Norwood seems to malfunction and begin speaking Chinese.
From TechCrunch AI
Blog
Daily notes on new models, LLM releases, agent frameworks and AI research, written from the sources we follow and delivered as a newsletter every day.
Get the daily issue
Every new post of the day, in one email. Confirmation required.
AI1 min read
In one particularly odd interview, Norwood seems to malfunction and begin speaking Chinese.
From TechCrunch AI
AI1 min read
Chinese startup Manus is raising $500 million at a $4 billion valuation. It resumed independent operations after breaking off its merger with Meta.
From TechCrunch AI
How this blog is made
Each feed entry becomes one request to OpenSmartRoute: the router picks a model with a cost-weighted objective, the editorial-writer skill is layered on the prompt, and the outcome trains the learners - the same pipeline available to every workspace.
Open any post to see which target answered, its confidence, the alternatives and what the request cost. Run the same pipeline yourself: register feeds in the operator console, map a small model under Providers, or call POST /api/v1/route with execute: true.
Garry Tan advocates for U.S. open-weight AI labs to utilize distillation techniques on frontier models, arguing against regulatory overreach and aiming for a more diverse open-weight AI ecosystem.
From TechCrunch AI
LLMs1 min read
arXiv:2609.10722v1 Announce Type: new Abstract: Structured extraction from Chinese military news supports intelligence analysis, decision-making, and knowledge base construction. However, existing resources provide limited support for jo...
From arXiv cs.CL
LLMs1 min read
SinoGlyphBench, a new diagnostic benchmark, highlights the vulnerability of LLMs and MLLMs to Chinese glyph-level obfuscation. Evaluations revealed a significant increase in false negatives and positives, alongside reduced accuracy, demonstrating the need for robust moderation strategies.
From arXiv cs.CL
LLMs1 min read
A new benchmark evaluates whether large language models can interpret social meaning in Chinese online comments, focusing on indirect and playful language. The strongest model achieves 81.42% accuracy.
From arXiv cs.CL
LLMs1 min read
Introducing Hy4 Preview New open weight text input (no vision) LLM from Chinese company Tencent today: 770B total parameters, 49B active parameters, 1M token context window, 1.56TB on Hugging Face. This is a big size increase from their ...
From Simon Willison on LLMs
Agents1 min read
Z.ai released GLM-5.3-Flash, a natively multimodal model with a 1M-token context window and 320B parameters, achieving strong performance benchmarks and competitive pricing, sparking significant community interest.
From Latent Space
Posts are drafted from public feeds by models OpenSmartRoute routes to - the same router, skill and metering customers use - and always link to the original source. Corrections: support.
Archive (39)