# Jan 21, 2025

### Daily

### Weekly

- **DeepSeek-R1** introduces a novel reasoning model that leverages **large-scale reinforcement learning** without prior supervised fine-tuning, achieving performance on par with OpenAI's models across various tasks, including math and code.

- **DeepSeek's first-generation reasoning models** demonstrate performance on par with **OpenAI-o1**, excelling in tasks involving math, code, and reasoning, showcasing their versatility and potential applications in various domains.

- The **Mind Evolution** strategy enhances **inference time compute** in Large Language Models by generating, recombining, and refining responses, leading to superior performance in natural language planning tasks. [Link to article](https://arxiv.org/abs/2501.09891)

- **Kimi k1.5** achieves **state-of-the-art performance** in multi-modal reasoning, significantly outperforming competitors like GPT-4o and Claude Sonnet 3.5 by up to **550%** on various benchmarks, including AIME and MATH-500.
