ML Times

Oct 25, 2025

Diamond Thermal Conductivity: A New Era in Chip Cooling

Why can't transformers learn multiplication?

How was Multi-head Latent Attention not a thing before DeepSeek-V2 came up with it?

ChunkLLM: A Lightweight Pluggable Framework for Accelerating LLMs Inference

Startup plans to cool data centers by converting heat to light

Compiler optimizations for 5.8ms GPT-OSS-120B inference (not on GPUs)

Torchcomms: A modern PyTorch communications API

Signal Processing for AI — A New Way to Think About LLMs and ANN Search

UFIPC: Physics-based AI Complexity Benchmark - Models with identical MMLU scores differ 29% in complexity