ML Times

Aug 11, 2025

Auf Wiedersehen, GitHub – CEO Steps Down

GPT-OSS vs. Qwen3 and a detailed look how things evolved since GPT-2

Diffusion language models are super data learners

Token growth indicates future AI spend per dev

Apache Iceberg V3 Spec new features for more efficient and flexible data lakes

From GPT-2 to gpt-oss: Analyzing the Architectural Advances And How They Stack Up Against Qwen3

Dropbox announces new gen server hardware for higher efficiency and scalability

Conversations remotely detected from cell phone vibrations, researchers report

Launch HN: Halluminate (YC S25) – Simulating the internet to train computer use

VulkanIlm: Accelerating Local LLM Inference on Older GPUs Using Vulkan (Non-CUDA) — Benchmarks Included

Associative memory inspires improvements for in-context learning using a novel attention residual stream architecture

Has anyone tried cross-modal transfer for visual reasoning? This 76% MMMU result surprised me

DRTP and No-Prop Hybrid in Pure C

NVIDIA Research Shapes Physical AI

Mini Footprint, Mighty AI: NVIDIA Blackwell Architecture Powers AI Acceleration in Compact Workstations