# May 8, 2025

## Daily

### - Absolute Zero: Reinforced Self-play Reasoning with Zero Data [R]
- **Absolute Zero** introduces a novel **Reinforcement Learning with Verifiable Rewards (RLVR)** paradigm that enables a model to autonomously generate tasks for its own learning, eliminating the need for external data and human supervision.

### - Create and edit images with Gemini 2.0 in preview
- **Gemini 2.0 Flash** now offers **image generation** capabilities in preview, allowing developers to integrate **conversational image generation and editing** with enhanced features via the Gemini API in [Google AI Studio](https://aistudio.google.com/app/prompts/new_chat?model=gemini-2.0-flash-preview-image-generation) and [Vertex AI](https://console.cloud.google.com/freetrial?redirectPath=/vertex-ai/studio).

### - AI focused on brain regions recreates what you're looking at (2024)
- **AI systems** can now **reconstruct images** of what a monkey is looking at with **remarkable accuracy** by focusing on specific brain regions, enhancing the fidelity of the output.

### - Block Diffusion: Interpolating Autoregressive and Diffusion Language Models
- **Block diffusion language models** enhance **parallelized generation** and **controllability**, addressing limitations of autoregressive models by enabling **flexible-length generation** and improved inference efficiency through **KV caching** and **parallel token sampling**.

### - [R] Cracking 40% on SWE-bench with open weights (!): Open-source synth data & model & agent
- **SWE-smith** enables the generation of **100s to 1000s of task instances** for any GitHub repository, facilitating the creation of training data for software engineering agents, which has historically been a challenge.

### - Hypermode Model Router Preview – OpenRouter Alternative
- **Model Router** streamlines access to both **open-source** and **commercial AI models** through a **single API**, enabling developers to switch models effortlessly based on performance and cost.

### - [P] Introducing the Intelligent Document Processing (IDP) Leaderboard – A Unified Benchmark for OCR, KIE, VQA, Table Extraction, and More
- The **Intelligent Document Processing (IDP) Leaderboard** introduces a comprehensive benchmark for evaluating **Vision-Language Models (VLMs)** across **6 core tasks** and **16 datasets**, totaling **9,229 documents**.

### - Bridging the gap between keyword and semantic search with SPLADE (2024)
- **SPLADE** merges the strengths of **keyword** and **semantic search**, enhancing recall by generating synthetic terms that improve search results even when queries differ from indexed content.

### - [D] TLMs: Task-Specific Language Models - What are they really?
- **TLMs** (Task-Specific Language Models) utilize a **novel architecture** that enhances efficiency and accuracy, allowing training on low-end gaming GPUs, significantly reducing costs.

### - [R] Process Reward Models That Think
- **ThinkPRM** addresses the challenge of **expensive step-level supervision** in training Process Reward Models (PRMs) by utilizing only **8K process labels** to enhance reasoning verification through long chains-of-thought.

### - PyTorch Foundation Expands to Umbrella Foundation and Welcomes vLLM and DeepSpeed Projects
- The **PyTorch Foundation** has evolved into an **umbrella foundation**, now hosting significant projects like **vLLM** and **DeepSpeed**, enhancing its role in the open source AI ecosystem.

### - PyTorch Foundation Welcomes vLLM as a Hosted Project
- The **PyTorch Foundation** has officially welcomed **vLLM**, a high-throughput, memory-efficient inference engine for large language models (LLMs), enhancing its integration with diverse hardware platforms like NVIDIA and Google Cloud TPUs.

### - OpenAI’s response to the Department of Energy on AI infrastructure

### - Cadence Taps NVIDIA Blackwell to Accelerate AI-Driven Engineering Design and Scientific Simulation
- **Cadence's Millennium M2000 Supercomputer**, powered by **NVIDIA Blackwell**, enhances performance for engineering and life sciences applications, achieving **up to 80x higher performance** than previous CPU-based systems.

### - LM Studio Accelerates LLM Performance With NVIDIA GeForce RTX GPUs and CUDA 12.8
- **LM Studio 0.3.15** enhances LLM performance on **NVIDIA GeForce RTX GPUs** with **CUDA 12.8**, improving model load and response times significantly, while introducing new developer tools for better integration and control.
