ML Times

Oct 18, 2024

Use Prolog to improve LLM's reasoning

Grandmaster-Level Chess Without Search

D PyTorch 2.5.0 released!

Bugs in LLM Training – Gradient Accumulation Fix

Microsoft BitNet: inference framework for 1-bit LLMs

P How to build a custom text classifier without days of human labeling

LLMD: A Large Language Model for Interpreting Longitudinal Medical Records

R DART can generate high-quality human motions in real-time, achieving over 300 frames per second on a single RTX 4090 GPU! It combines text inputs with spatial constraints, allowing for tasks like reaching waypoints and interacting with scenes.

PyTorch 2.5 Release Blog

An Evolved Universal Transformer Memory

LLMOPT: Learning to Define and Solve General Optimization Problems from Scratch

SimLayerKV: A Simple Framework for Layer-Level KV Cache Reduction

Latent Space Chain-of-Embedding Enables Output-free LLM Self-Evaluation

Cerberus: Efficient Inference with Adaptive Parallel Decoding and Sequential Knowledge Enhancement

Mitigating Hallucinations in Large Vision-Language Models via Summary-Guided Decoding