ML Times
Jun 14, 2026
Noise infusion banned from statistical products published by Census Bureau
The U.S. Department of Commerce has banned noise infusion in statistical products, which undermines the effectiveness of privacy-preserving techniques like differential privacy that balance data utility and confidentiality.Rio de Janeiro's "homegrown" LLM appears to be a merge of an existing model
The Rio-3.5-Open-397B model is a 0.6 x Nex-N2_pro + 0.4 x Qwen blend, lacking original training evidence, as it identifies itself as "Nex, from Nex-AGI" 79% of the time when prompted without its system mask.AI OSS tool repo goes archived over night after raising $7.3M Seed
TensorZero is an open-source LLMOps platform that integrates multiple functionalities, including a unified API for LLM access, observability for inferences, and automated optimization through its Autopilot feature, enhancing LLM performance across diverse tasks.Don't trust large context windows
Large context windows in LLMs are misleading; effective usability drops significantly after about 100k tokens, despite vendors advertising much larger capacities.Making Claude a Chemist
Claude is enhancing its chemistry capabilities by collaborating with expert chemists to analyze NMR spectra, a critical tool for determining molecular structures, thus streamlining the chemist's workflow.KPMG pulls report on AI usage due to apparent hallucinations
KPMG retracted its report on AI usage after organizations claimed inaccuracies, highlighting the risks of relying on AI-generated content without human oversight.Weave: Merging based on language structure and not lines
weave is an entity-level semantic merge driver for Git that ensures clean merges by recognizing non-overlapping edits, effectively eliminating conflicts in collaborative coding environments.Show HN: Dual YOLOv8n UAV Detection on RK3588S at 42 FPS Using NPU
Real-time UAV detection using YOLOv8n achieves 46 FPS on the Rockchip RK3588S NPU, leveraging 3 cores for parallel processing, which maximizes throughput while maintaining a low memory footprint of ~140 MB per stream.The first game engine for robotics
Lucky Engine is a groundbreaking game engine designed specifically for robotics, enabling robots to learn through millions of simulated trials without the risk of hardware damage or lab constraints.Inverse Rubric Optimization: A testbed for agent science
Inverse Rubric Optimization (IRO) tasks enable agents to learn preferences from a black-box judge, revealing complex behaviors and performance scaling, with Fable 5 showing superior results at lower label budgets but plateauing at higher ones.Cloud-based LLM gold rush is ending
Apple's shift to local AI processing at WWDC signals a move away from cloud-based LLMs, emphasizing efficiency and user autonomy in task management on Mac OS.The Verifier Tax: Horizon-Dependent Safety–Success Tradeoffs in Tool-Using LLM Agents
The Verifier Tax highlights a horizon-dependent safety–success tradeoff in tool-using LLM agents, where increased verification can lead to reduced task completion rates as task complexity grows.Coherent Context Can Silently Shift LLMs Into a Different Internal Regime — And Current Safety Systems Are Blind To It
A coherent target text can shift a model's internal regime before producing an output, allowing it to behave normally while operating under altered internal states, which current safety systems fail to detect.Derivative-Free Neural Network Optimization: MNIST Case