Gemini 3 Flash: Frontier intelligence built for speed
Gemini 3 Flash is a new AI model that combines frontier intelligence with high speed and cost efficiency, enabling users to learn and build faster than ever before.
Meta Segment Anything Model Audio
SAM Audio is a new feature from Meta that enhances audio experiences, integrating advanced AI capabilities to improve sound quality and user interaction.
TRELLIS.2: State-of-the-art large 3D generative model (4B)
TRELLIS.2 is a cutting-edge 4B-parameter model that excels in image-to-3D generation, utilizing a unique O-Voxel structure for creating complex 3D assets with high fidelity and efficiency.
Eigenvalues as models
Neurons may require enhanced computational capabilities, as suggested by Sutskever, prompting exploration of eigenvalues in optimization within neural networks.
The State of AI Coding Report 2025
AI coding tools have significantly boosted developer productivity, with lines of code per developer increasing by 76%, from 4,450 to 7,839, as reported in Greptile's internal data.
T5Gemma 2: The next generation of encoder-decoder models
T5Gemma 2 introduces significant architectural innovations, such as tied embeddings and merged attention, enhancing efficiency and enabling compact models ideal for on-device applications.
FunctionGemma 270M Model
FunctionGemma is a specialized version of the Gemma 3 270M model, fine-tuned for function calling, enabling seamless execution of API actions from natural language inputs.
Virtualizing Nvidia HGX B200 GPUs with Open Source
NVIDIA HGX B200 GPUs can be virtualized using open-source tools, enabling high-performance GPU VMs that leverage the unique NVLink architecture for efficient multi-GPU communication.
How China built its ‘Manhattan Project’ to rival the West in AI chips
China's prototype for extreme ultraviolet lithography machines (EUVs), crucial for advanced semiconductor production, was developed by ex-ASML engineers, marking a significant leap in its AI chip capabilities.
High-Performance Wavelet Matrix for Python, Implemented in Rust
JavaScript loading issues can hinder site functionality, often due to browser settings or extensions, impacting user experience significantly.
The Open Evaluation Standard: Benchmarking NVIDIA Nemotron 3 Nano with NeMo Evaluator
NVIDIA's Nemotron 3 Nano 30B A3B employs an open evaluation standard using the NeMo Evaluator, allowing for transparent and reproducible benchmarking that can be independently verified by users.
From profiling to kernel patch: the journey to an eBPF performance fix
The article details how a profiling session with Superluminal led to a significant performance enhancement in the Linux kernel, specifically making eBPF map-in-map updates much faster, achieving a 31x speedup in capture startup time.
A local-first memory store for LLM agents (SQLite)
OpenMemory provides real long-term memory for AI agents, enabling persistent, explainable, and scalable cognitive structures without the need for cloud dependencies or vendor lock-in.
DEER: Draft with Diffusion, Verify with Autoregressive Models
DEER introduces a novel speculative decoding framework that utilizes diffusion large language models (dLLMs) for drafting, effectively addressing the limitations of autoregressive (AR) models by enabling parallel decoding and reducing uncertainty accumulation.
Pulse is a document extraction system designed to create LLM-ready text, addressing the limitations of existing models in processing complex documents like long PDFs and dense tables, which often lead to subtle yet significant errors.