ML Times
Jan 22, 2026
Claude's New Constitution
Claude's new constitution serves as a foundational document that articulates Anthropic's vision for Claude's values and behavior, aiming to enhance its training by providing a comprehensive understanding of its intended role in society.
Qwen3-TTS Family Is Now Open Sourced: Voice Design, Clone, and Generation
Qwen operates across multiple domains, including qwen.ai and chat.qwenlm.ai, indicating a robust infrastructure for AI-driven applications and services. Link to article
GPTZero Finds 100 New Hallucinations in NeurIPS 2025 Accepted Papers
GPTZero's analysis of 4841 NeurIPS 2025 papers revealed at least 100 confirmed hallucinations, highlighting a significant oversight in the peer review process that failed to catch these errors despite rigorous evaluations by multiple reviewers.
Show HN: Sweep, Open-weights 1.5B Model for Next-edit Autocomplete
Sweep Next-Edit 1.5B is a 1.5 billion parameter model designed for next-edit autocomplete, achieving predictions in under 500ms and outperforming larger models on benchmarks.
100 Hallucinated Citations Found in 51 Accepted Papers at NeurIPS 2025
100 hallucinated citations were identified in 51 accepted papers at NeurIPS 2025, raising concerns about the integrity of peer-reviewed research.
Letting Claude Play Text Adventures
Claude's performance in text adventures reveals that using a memory-augmented harness significantly reduces token usage, but it also leads to longer completion times for tasks, as seen in the game Anchorhead where Claude took ~250 turns to achieve objectives compared to ~100 with a simpler approach.
Binary Fuse Filters: Fast and Smaller Than XOR Filters
Binary fuse filters outperform xor filters by achieving storage efficiency within 13% of the theoretical lower bound while maintaining fast query speeds, making them a compelling alternative for engineers.
Three Types of LLM Workloads and How to Serve Them
Three distinct LLM workloads— offline, online, and semi-online—demand tailored architectural strategies to optimize performance and cost, with offline workloads focusing on throughput, online on low latency, and semi-online on flexible scaling.
Keeping 20k GPUs Healthy
Modal operates a robust GPU worker pool exceeding 20,000 GPUs, employing a comprehensive reliability system that includes instance type testing, machine image preparation, and both passive and active health checks to ensure optimal performance across various cloud providers.
‘Largest Infrastructure Buildout in Human History’: Jensen Huang on AI’s ‘Five-Layer Cake’ at Davos
AI is driving the largest infrastructure buildout in history, characterized by a "five-layer cake" model that includes energy, chips, cloud data centers, AI models, and applications, fundamentally reshaping job creation across various sectors.
Launch HN: Constellation Space (YC W26) – AI for Satellite Mission Assurance
Constellation Space has developed an AI system that predicts satellite link failures 3-5 minutes in advance with over 90% accuracy, enabling proactive traffic rerouting to prevent data loss.
AssetOpsBench: Bridging the Gap Between AI Agent Benchmarks and Industrial Reality
AssetOpsBench is a novel benchmark system that evaluates AI agents across six qualitative dimensions, specifically tailored for complex industrial applications, emphasizing multi-agent coordination and real-world operational challenges.
From Pilot to Profit: Survey Reveals the Financial Services Industry Is Doubling Down on AI Investment and Open Source
Financial institutions are significantly increasing AI investments, with nearly 100% planning to maintain or boost budgets, driven by a clear return on investment from AI applications like fraud detection and risk management.
Show HN: Differentiable Quantum Chemistry
slaterform is a differentiable Hartree-Fock engine built with jax, enabling efficient molecular energy function optimization through native electron integrals and standard basis sets from basis set exchange.