AI1 min read
Anthropic merges Claude chat and Cowork into one interface
Anthropic combines its chat and Cowork tools into a single window. This removes the need to switch tabs for different tasks.
From TechCrunch AI
Blog
Daily notes on new models, LLM releases, agent frameworks and AI research, written from the sources we follow and delivered as a newsletter every day.
Get the daily issue
Every new post of the day, in one email. Confirmation required.
AI1 min read
Anthropic combines its chat and Cowork tools into a single window. This removes the need to switch tabs for different tasks.
From TechCrunch AI
Research1 min read
arXiv:2609.13548v1 Announce Type: new Abstract: Web agents can utilize reusable tools to reduce the cost and latency of low-level browser interaction, but automatically discovered tool collections can be large, redundant, and poorly alig...
From arXiv cs.AI
How this blog is made
Each feed entry becomes one request to OpenSmartRoute: the router picks a model with a cost-weighted objective, the editorial-writer skill is layered on the prompt, and the outcome trains the learners - the same pipeline available to every workspace.
Open any post to see which target answered, its confidence, the alternatives and what the request cost. Run the same pipeline yourself: register feeds in the operator console, map a small model under Providers, or call POST /api/v1/route with execute: true.
Research1 min read
Research demonstrates a new approach, Debate-to-Skill, for annotating query-to-agent interactions by focusing on executable capability rather than simple topical relevance. This method achieves better results compared to existing techniques on an industrial benchmark, particularly in complex scenarios.
From arXiv cs.AI
Research1 min read
Research identified seven distinct sources contributing to Physical AI capabilities: Recorded-Experience, Predictive-Modeling, Evaluative-Interaction, Surrogate-Environment, Mechanism-Grounded, Embodied-Coupling, and Evolution-Driven. This framework allows analysis of how capabilities are formed, supporting applications like explanation and transfer.
From arXiv cs.AI
Research1 min read
Introduces Harbor Adapters for evaluating agents across over 80 benchmarks and presents Harbor-Index, a curated set of challenging tasks for comprehensive assessment.
From arXiv cs.AI
Research1 min read
Enhanced capabilities in large language models may lead to more correlated behaviors, increasing systemic risk, especially when models share reasoning or misinformation environments.
From arXiv cs.AI
Models1 min read
GPT-6 Astra is the most capable broadly deployed model from OpenAI and has reached the Critical cybersecurity level under the Preparedness Framework.
From OpenAI news
Models1 min read
Astra is the first OpenAI model to meet the Critical cybersecurity capability threshold under the Preparedness Framework, with enhanced safeguards for release.
From OpenAI news
LLMs1 min read
A capability threshold I've been carefully monitoring.
From Interconnects
Posts are drafted from public feeds by models OpenSmartRoute routes to - the same router, skill and metering customers use - and always link to the original source. Corrections: support.
Archive (34)