Architect Financial Technologies has launched Liquid Inference, an LLM router that runs a live auction for every request. Liquid Inference […]
Category: Staff
NVIDIA PivotOPD Teaches Multi-Turn AI Agents to Recover From Pivotal Mistakes
NVIDIA researchers, with Princeton University and the University of Maryland, have introduced PivotOPD, an on-policy distillation method for multi-turn LLM […]
Perplexity AI Releases pplx-embed-v2-late: A 0.6B Edge Model and a 9B Model Scoring 92.4% on MADQA
Perplexity has released pplx-embed-v2-late, a pair of ColBERT-style multimodal embedding models. They come in 2 sizes: 0.6B for fast, cheap […]
- agentic AI
- AI
- AI Agents
- AI infrastructure
- AI Shorts
- Applications
- Artificial Intelligence
- Editors Pick
- Enterprise AI
- For Devs
- Guardrail
- Hardware
- harness
- Language Model
- Large Language Model
- Machine Learning
- New Releases
- Promote
- Security
- Small Language Model
- Software Engineering
- Sponsored
- Staff
- Tech News
- Technology
What Happens When a Trusted Model Repo Changes? Unsloth Studio Re-Checks Before It Runs
Beta Lessons Learned After over 500 million downloads, years of requests from the open-source community, and being a top product […]
Anthropic Releases Claude Haiku 5.5: A Small Model With 1M Context Priced at $0.10 per Million Input Tokens
Anthropic has released Claude Haiku 5.5, its cheapest and fastest small model to date. It targets high-volume work like summaries, […]
Liquid AI Releases Open-Weight d1-3B and d1-omni-600M: Multimodal Decision Models With Zero Output Tokens
Liquid AI has released Open d1, two open-weight multimodal models in its d1 decision model family. d1-3B reads text and […]
Meta AI Open-Sources Rebalancer: A C++ Assignment Solver That Runs About 40 Million Placement Problems a Day
Meta has open-sourced Rebalancer, a C++ library with a Python interface for solving assignment problems. It decides which objects go […]
A Developer’s Guide to Laya: Zero-Shot Decisions and Calibration
In this tutorial, we work with Laya, the open-source decision engine from Convai Innovations that became one of the most-starred […]
Google DeepMind Releases EmbeddingGemma 2, a 740M Open Multimodal Embedding Model Built on Gemma 4
Google DeepMind has released EmbeddingGemma 2, an open model that embeds text, code, images, video and audio into one 768-dimensional […]
Mistral AI Releases Mistral Large 4 (Le Chonk): A 1.05T Parameter Multimodal MoE Model
Mistral AI has just announced the release of Mistral Large 4 (ML4), internally nicknamed Le Chonk, as a public preview. […]
