Robot manipulation datasets have grown far slower than the models trained on them, mostly because collection stays closed and centralized. […]
Category: Applications
IFM Releases K2 Horizon: Six Apache 2.0 Models From 0.9B to 375B
Most open model launches release one checkpoint and a benchmark table. The Institute of Foundation Models (IFM) released something wider […]
H Company Releases NeoMME: A Family of 260M and 800M Single-Tower Multimodal Encoders That Drop the Vision Tower and Causal Decoder
Most visual document retrievers in production today are hand-me-downs. ColPali and the models that followed it take a generative vision-language […]
Meta FAIR Introduces AI Research Preference Models (RPMs): Ranking ML Experiments Before Spending GPU Hours
AI research agents can already propose, implement and score their own machine learning experiments. Idea generation is cheap; verification is […]
UC Berkeley Researchers Release CUA-Lite, an Open Platform Unifying Sandboxes, Data, Evaluation and RL for Computer-Use Agents
A team of researchers from UC Berkeley have released CUA-Lite, an open platform for computer-use agents (CUAs). The argument behind […]
Perplexity Details Its GPU Embedding Stack: How Ivy, Tulip and ROSE Serve pplx-embed
Retrieval quality in an AI search product is bounded by two things: how good the embedding model is, and how […]
GitHub Introduces Project HydraFusion: Runtime Multi-Model Orchestration That Builds a Workflow Per Coding Task in Copilot CLI
GitHub has released Project HydraFusion, a research preview that stops treating model choice as a one-time setting. Instead of routing […]
Adaption Labs Introduces ‘Invent a Dataset’: Training Data Generated From a Task Description, Not a Seed Corpus
This week, Adaption Labs released Invent a Dataset. The feature generates a structured, training-ready dataset from a description of the […]
Google Launches Agentic Video Understanding for Gemini Flash Models, Cutting Video Tokens by Up to 88%
Video has been the most expensive modality to reason over. A Gemini model handed a 90-minute lecture has, until now, […]
NVIDIA Releases Personal AI Router (PAIR): An Open Source Virtual Inference Router that Distributes Local AI Requests Across RTX, DGX Spark, and Mac Nodes
Multi-agent workflows have changed the shape of local inference. A lead agent decomposes a task and spawns subagents. What looked […]
