ML Times

Jan 22, 2026

Claude's New Constitution

Claude's new constitution serves as a foundational document that articulates Anthropic's vision for Claude's values and behavior, aiming to enhance its training by providing a comprehensive understanding of its intended role in society.

Qwen3-TTS Family Is Now Open Sourced: Voice Design, Clone, and Generation

Qwen operates across multiple domains, including qwen.ai and chat.qwenlm.ai, indicating a robust infrastructure for AI-driven applications and services. Link to article

GPTZero Finds 100 New Hallucinations in NeurIPS 2025 Accepted Papers

GPTZero's analysis of 4841 NeurIPS 2025 papers revealed at least 100 confirmed hallucinations, highlighting a significant oversight in the peer review process that failed to catch these errors despite rigorous evaluations by multiple reviewers.

Show HN: Sweep, Open-weights 1.5B Model for Next-edit Autocomplete

Sweep Next-Edit 1.5B is a 1.5 billion parameter model designed for next-edit autocomplete, achieving predictions in under 500ms and outperforming larger models on benchmarks.

100 Hallucinated Citations Found in 51 Accepted Papers at NeurIPS 2025

100 hallucinated citations were identified in 51 accepted papers at NeurIPS 2025, raising concerns about the integrity of peer-reviewed research.

Letting Claude Play Text Adventures

Claude's performance in text adventures reveals that using a memory-augmented harness significantly reduces token usage, but it also leads to longer completion times for tasks, as seen in the game Anchorhead where Claude took ~250 turns to achieve objectives compared to ~100 with a simpler approach.

Binary Fuse Filters: Fast and Smaller Than XOR Filters

Binary fuse filters outperform xor filters by achieving storage efficiency within 13% of the theoretical lower bound while maintaining fast query speeds, making them a compelling alternative for engineers.

Three Types of LLM Workloads and How to Serve Them

Three distinct LLM workloads— offline, online, and semi-online—demand tailored architectural strategies to optimize performance and cost, with offline workloads focusing on throughput, online on low latency, and semi-online on flexible scaling.

Keeping 20k GPUs Healthy

Modal operates a robust GPU worker pool exceeding 20,000 GPUs, employing a comprehensive reliability system that includes instance type testing, machine image preparation, and both passive and active health checks to ensure optimal performance across various cloud providers.

‘Largest Infrastructure Buildout in Human History’: Jensen Huang on AI’s ‘Five-Layer Cake’ at Davos

AI is driving the largest infrastructure buildout in history, characterized by a "five-layer cake" model that includes energy, chips, cloud data centers, AI models, and applications, fundamentally reshaping job creation across various sectors.

Launch HN: Constellation Space (YC W26) – AI for Satellite Mission Assurance

Constellation Space has developed an AI system that predicts satellite link failures 3-5 minutes in advance with over 90% accuracy, enabling proactive traffic rerouting to prevent data loss.

AssetOpsBench: Bridging the Gap Between AI Agent Benchmarks and Industrial Reality

AssetOpsBench is a novel benchmark system that evaluates AI agents across six qualitative dimensions, specifically tailored for complex industrial applications, emphasizing multi-agent coordination and real-world operational challenges.

From Pilot to Profit: Survey Reveals the Financial Services Industry Is Doubling Down on AI Investment and Open Source

Financial institutions are significantly increasing AI investments, with nearly 100% planning to maintain or boost budgets, driven by a clear return on investment from AI applications like fraud detection and risk management.

Show HN: Differentiable Quantum Chemistry

slaterform is a differentiable Hartree-Fock engine built with jax, enabling efficient molecular energy function optimization through native electron integrals and standard basis sets from basis set exchange.