ML Times
Jan 21, 2025
Daily
Weekly
DeepSeek-R1 introduces a novel reasoning model that leverages large-scale reinforcement learning without prior supervised fine-tuning, achieving performance on par with OpenAI's models across various tasks, including math and code.
DeepSeek's first-generation reasoning models demonstrate performance on par with OpenAI-o1, excelling in tasks involving math, code, and reasoning, showcasing their versatility and potential applications in various domains.
The Mind Evolution strategy enhances inference time compute in Large Language Models by generating, recombining, and refining responses, leading to superior performance in natural language planning tasks. Link to article
Kimi k1.5 achieves state-of-the-art performance in multi-modal reasoning, significantly outperforming competitors like GPT-4o and Claude Sonnet 3.5 by up to 550% on various benchmarks, including AIME and MATH-500.