ML Times

May 11, 2026

Daily

Training an LLM in Swift, Part 1: Taking matrix mult from Gflop/s to Tflop/s

Lakebase architecture delivers faster Postgres writes

Interfaze: A new model architecture built for high accuracy at scale

Interaction Models

MISA: Mixture of Indexer Sparse Attention for Long-Context LLM Inference

Fast Byte Latent Transformer

A hackable compiler to generate efficient fused GPU kernels for AI models

Memory-Efficient Looped Transformer: Decoupling Compute from Memory in Looped Language Models

🤗MachinaCheck: Building a Multi-Agent CNC Manufacturability System on AMD MI300X

Looking for arXiv endorsement (cs.CV) to post my ViT positional embeddings paper

Signals: finding the most informative agent traces without LLM judges

Trajectory as the Teacher: Few-Step Discrete Flow Matching via Energy-Navigated Distillation

MAVEN: Multi-Agent Verification-Elaboration Network with In-Step Epistemic Auditing

Confidence-Aware Alignment Makes Reasoning LLMs More Reliable

Rethinking Importance Sampling in LLM Policy Optimization: A Cumulative Token Perspective