ML Times

Jul 6, 2026

Daily

A global workspace in language models

Does Code Cleanliness Affect Coding Agents?

Emily Bender Sets the Record Straight on "Stochastic Parrots"

The Private Capture of Public Genius

When 2+2=5

Fable 5 On Vending-Bench: Misbehaving, With Plausible Deniability

When AI Costs More Than the Engineer

The AI Superforecasters Are Here

Show HN: Pulpie – Models for Cleaning the Web

Python 3.14 compiled to metal – no interpreter

Show HN: Scan your AI agents for dangerous capabilities

Competence Gate: gating tool-use on a small model's internal confidence signal instead of its verbalised one — Qwen3.5-4B, open weights

Pruning RAG context down to what the answer actually needs

TRACE: open-source hierarchical memory for LLM agents, 82.5% on MemoryAgentBench’s EventQA using gpt-oss-20B

Best models for generating red-team attacks? Also looking for public datasets