Robot manipulation datasets have grown far slower than the models trained on them, mostly because collection stays closed and centralized. […]
Category: Staff
IFM Releases K2 Horizon: Six Apache 2.0 Models From 0.9B to 375B
Most open model launches release one checkpoint and a benchmark table. The Institute of Foundation Models (IFM) released something wider […]
H Company Releases NeoMME: A Family of 260M and 800M Single-Tower Multimodal Encoders That Drop the Vision Tower and Causal Decoder
Most visual document retrievers in production today are hand-me-downs. ColPali and the models that followed it take a generative vision-language […]
Meta FAIR Introduces AI Research Preference Models (RPMs): Ranking ML Experiments Before Spending GPU Hours
AI research agents can already propose, implement and score their own machine learning experiments. Idea generation is cheap; verification is […]
UC Berkeley Researchers Release CUA-Lite, an Open Platform Unifying Sandboxes, Data, Evaluation and RL for Computer-Use Agents
A team of researchers from UC Berkeley have released CUA-Lite, an open platform for computer-use agents (CUAs). The argument behind […]
Perplexity Details Its GPU Embedding Stack: How Ivy, Tulip and ROSE Serve pplx-embed
Retrieval quality in an AI search product is bounded by two things: how good the embedding model is, and how […]
GitHub Introduces Project HydraFusion: Runtime Multi-Model Orchestration That Builds a Workflow Per Coding Task in Copilot CLI
GitHub has released Project HydraFusion, a research preview that stops treating model choice as a one-time setting. Instead of routing […]
Nous Research Adds One-Click Local Model Setup to Hermes Desktop
The hard part of running an open-weights model locally was never the model. It was everything before it, reading VRAM […]
Adaption Labs Introduces ‘Invent a Dataset’: Training Data Generated From a Task Description, Not a Seed Corpus
This week, Adaption Labs released Invent a Dataset. The feature generates a structured, training-ready dataset from a description of the […]
Google Launches Agentic Video Understanding for Gemini Flash Models, Cutting Video Tokens by Up to 88%
Video has been the most expensive modality to reason over. A Gemini model handed a 90-minute lecture has, until now, […]
