ML Times

Dec 30, 2024

DeepSeek-V3 Technical Report

KAG – Knowledge Graph RAG Framework

Breaking NATO Radio Encryption [video]

How Well Do LLMs Generate Code for Different Application Domains?

Empirical Study of Test Generation with LLM's

Measuring and Understanding LLM Identity Confusion

[D] Batch Normalization and effect on the gradients

[P] Introducing LongTalk-CoT v0.1: A Very Long Chain-of-Thought Dataset for Reasoning Model Post-Training

Toward Adaptive Reasoning in Large Language Models with Thought Rollback

Research Galore From 2024: Recapping AI Advancements in 3D Simulation, Climate Science and Audio Engineering

Is Your Text-to-Image Model Robust to Caption Noise?

A Survey on Large Language Model Acceleration based on KV Cache Management