ML Times

Jul 14, 2025

Daily

Emergent Misalignment: Narrow finetuning can produce broadly misaligned LLMs

The upcoming GPT-3 moment for RL

Show HN: ArchGW – An intelligent edge and service proxy for agents

How to scale RL to 10^26 FLOPs

Embedding User-Defined Indexes in Apache Parquet

NeuralOS: An operating system powered by neural networks

Context Rot: How increasing input tokens impacts LLM performance

[R] Deep-dive into RoPE and why it matters

[D] Updated Document Intelligence Framework Benchmarks

[R] Unlearning Comparator — A Visual Analytics Toolkit for Machine Unlearning

AI Testing and Evaluation: Learnings from cybersecurity

Scaling Attention to Very Long Sequences in Linear Time with Wavelet-Enhanced Random Spectral Attention (WERSA)

White-Basilisk: A Hybrid Model for Code Vulnerability Detection