ML Times

Aug 21, 2024

Data Exfiltration from Slack AI via indirect prompt injection

NVIDIA Announces First Digital Human Technologies On-Device Small Language Model, Improving Conversation for Game Characters

Self-Supervised Learning for Videos

Launch HN: Outerport (YC S24) – Instant hot-swapping for AI models

Join Our Global Paper Reading Group for a Deep Dive into "Plan Like a Graph (PLaG)" - Enhancing LLMs in Asynchronous Plan Reasoning | ICML 2024 with the author Fangru Lin!

HiRED: Attention-Guided Token Dropping for Efficient Inference of High-Resolution Vision-Language Models in Resource-Constrained Environments

Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model

Beta9: Open Serverless GPU Cloud

Predicting Phenotypes from Omics Data

PEDAL: Enhancing Greedy Decoding with Large Language Models using Diverse Exemplars

Using synthetic data from LLMs to train/finetune other models

NVIDIA Showcases New AI Capabilities With ACE, RTX Games and More at Gamescom 2024

SLMming Down Latency: How NVIDIA’s First On-Device Small Language Model Makes Digital Humans More Lifelike

Lightweight Champ: NVIDIA Releases Small Language Model With State-of-the-Art Accuracy

Improving Hugging Face Training Efficiency Through Packing with Flash Attention