ML Times

Sep 12, 2025

Qwen3-Next: Towards Ultimate Training and Inference Efficiency

Top model scores may be skewed by Git history leaks in SWE-bench

VaultGemma: The most capable differentially private LLM

Vector database that can index 1B vectors in 48M

Building a Deep Research Agent Using MCP-Agent

Lumina-DiMOO: An open-source discrete multimodal diffusion model

Larry Ellison: “Inference is where the money is going to be made.”

K2-Think: A Parameter-Efficient Reasoning System

Backprompting: Leveraging synthetic production data for health advice guardrails

Semlib: LLM-powered Data Processing

Universal Deep Research (UDR): A general wrapper for LLM-Based research

ButterflyQuant: Ultra-low-bit LLM Quantization through Learnable Orthogonal Butterfly Transforms

SEDM: Scalable Self-Evolving Distributed Memory for Agents