Alibaba’s Qwen team has released Qwen-Image-2.1-Turbo, an accelerated checkpoint of its open-weight Qwen-Image-2.1 model. It generates and edits images in […]
Category: Staff
OpenAI Decisions API Hits Public Beta With 10x Faster Typed Answers
OpenAI has released the Decisions API in public beta. It turns text and images into typed answers your code can […]
Meet the Underdog Saluki 27B: A 2-bit Qwen3.8-27B That Beats the Original at Tool Calling
Underdog, the on-device assistant from Conway Research, has released Saluki 27B under Apache 2.0. Underdog Saluki 27B is a 2-bit […]
Google Research RRSI Guide: Mastering Self-Improving AI Agents
In this tutorial, we implement RRSI (Regularized Recursive Self-Improvement), a method that lets an LLM agent rewrite its own harness, […]
JetBrains Releases Mellum2.1: A 12B MoE Open Model for Coding Agents
JetBrains has released Mellum2.1, an open model built for coding agents and fast sub-agents. Mellum2.1 is a 12B mixture-of-experts thinking […]
Architect Launches Liquid Inference, a Real-Time Auction for LLM Inference
Architect Financial Technologies has launched Liquid Inference, an LLM router that runs a live auction for every request. Liquid Inference […]
NVIDIA PivotOPD Teaches Multi-Turn AI Agents to Recover From Pivotal Mistakes
NVIDIA researchers, with Princeton University and the University of Maryland, have introduced PivotOPD, an on-policy distillation method for multi-turn LLM […]
Perplexity AI Releases pplx-embed-v2-late: A 0.6B Edge Model and a 9B Model Scoring 92.4% on MADQA
Perplexity has released pplx-embed-v2-late, a pair of ColBERT-style multimodal embedding models. They come in 2 sizes: 0.6B for fast, cheap […]
- agentic AI
- AI
- AI Agents
- AI infrastructure
- AI Shorts
- Applications
- Artificial Intelligence
- Editors Pick
- Enterprise AI
- For Devs
- Guardrail
- Hardware
- harness
- Language Model
- Large Language Model
- Machine Learning
- New Releases
- Promote
- Security
- Small Language Model
- Software Engineering
- Sponsored
- Staff
- Tech News
- Technology
What Happens When a Trusted Model Repo Changes? Unsloth Studio Re-Checks Before It Runs
Beta Lessons Learned After over 500 million downloads, years of requests from the open-source community, and being a top product […]
Anthropic Releases Claude Haiku 5.5: A Small Model With 1M Context Priced at $0.10 per Million Input Tokens
Anthropic has released Claude Haiku 5.5, its cheapest and fastest small model to date. It targets high-volume work like summaries, […]
