AI1 min read
Meta launches camera-free smart glasses named Luna
Meta is building new glasses without cameras to address privacy concerns.
From TechCrunch AI
Blog
Daily notes on new models, LLM releases, agent frameworks and AI research, written from the sources we follow and delivered as a newsletter every day.
Get the daily issue
Every new post of the day, in one email. Confirmation required.
AI1 min read
Meta is building new glasses without cameras to address privacy concerns.
From TechCrunch AI
AI1 min read
Google released early access to its Model Context Protocol server for smart homes. Users can now let AI agents manage their connected devices securely.
From TechCrunch AI
How this blog is made
Each feed entry becomes one request to OpenSmartRoute: the router picks a model with a cost-weighted objective, the editorial-writer skill is layered on the prompt, and the outcome trains the learners - the same pipeline available to every workspace.
Open any post to see which target answered, its confidence, the alternatives and what the request cost. Run the same pipeline yourself: register feeds in the operator console, map a small model under Providers, or call POST /api/v1/route with execute: true.
AI isn't some kind of new form of "alien mind," according to Jensen Huang. It's just hardware and software, so safety can be engineered by each AI product maker.
From TechCrunch AI
AI1 min read
Jensen Huang engaged in a conversation with President Trump at the All-In conference, discussing AI safety while demonstrating a new, unreleased foldable phone from Apple. This event highlighted Nvidia’s reliance on AI labs and Apple’s renewed interest in its hardware.
From TechCrunch AI
AI1 min read
OpenAI has purchased smartphone camera maker Glass Imaging for $300 million, leveraging expertise from former Apple engineers to improve smartphone camera image quality using AI. This acquisition aligns with OpenAI’s rumored development of its own hardware devices.
From TechCrunch AI
LLMs1 min read
NVIDIA’s BioNeMo Inference Runtime accelerates biomolecular structure prediction at scale, allowing for efficient processing of large proteome workflows. This enables faster insights from complex biological data.
From NVIDIA technical blog
Research1 min read
PerfReasoning assesses LLMs as performance reasoners and code generators, with models achieving up to 90% accuracy on reasoning tasks. Construction of models remains challenging, especially for smaller configurations.
From arXiv cs.AI
Research1 min read
MaxKernel introduces a multi-agent system for designing high-performance TPU kernels, leveraging LLMs and real-time compiler feedback. It includes human-in-the-loop, autonomous, and graph-based search paradigms.
From arXiv cs.AI
LLMs1 min read
NVIDIA has addressed challenges in running multi-step reasoning and agentic AI at the edge, enabling more efficient deployment on Jetson hardware.
From NVIDIA technical blog
LLMs1 min read
Hugging Face announced @huggingface/kernels, offering over 200 WebGPU kernels designed for local AI processing, enabling efficient model inference on compatible hardware.
From Hugging Face blog
LLMs1 min read
NVIDIA Groq 3 LPX is an AI inference accelerator designed for the Vera Rubin platform, enabling ultrafast interactivity with long context windows.
From NVIDIA technical blog
LLMs1 min read
NVIDIA introduces DSX MaxLPS to improve AI factory performance per watt, addressing power constraints in industrial AI systems.
From NVIDIA technical blog
Agents2 min read
OpenAI projects AGI completion by December 2026, driven by the Astra model, while the launch of Pollen Robotics’ Microduck robot gains traction with its open-source design and accessible price point.
From Latent Space
Agents2 min read
OpenAI’s Jalapeño inference chip delivers significantly improved performance and lower latency compared to NVIDIA GB200/GB300 systems, marking a shift in inference economics and highlighting the importance of system-level optimization.
From Latent Space
Models1 min read
OpenAI's CFO highlights how improvements across hardware and software components enable more scalable and cost-effective AI solutions.
From OpenAI news
LLMs1 min read
LiquidAI’s LFM2.5-DSpark model demonstrates up to 3.2x faster inference speeds compared to previous models. This improvement is achieved through optimized execution on DSpark hardware.
From Hugging Face blog
LLMs2 min read
Nvidia is investing $26 billion to foster a world where numerous entities can build token machines, aiming to reduce reliance on proprietary models and drive demand for Nvidia’s hardware. This strategy hinges on accessible open-source model recipes and a shift in the AI ecosystem’s financial dynamics.
From Interconnects
Models1 min read
Google AI announced a partnership involving Gemini and Pixel, aiming to improve AI integration with hardware and services, relevant for engineers managing models and agents.
From Google AI blog
LLMs1 min read
Ollama 0.30 is now available with improved performance and GGUF model compatibility through llama.cpp. This augments Ollama's MLX engine on Apple silicon, bringing support to more models on a wider range of hardware.
From Ollama blog
Posts are drafted from public feeds by models OpenSmartRoute routes to - the same router, skill and metering customers use - and always link to the original source. Corrections: support.
Archive (34)