DeepSeek unveils new AI models rivalling GPT-5 and Gemini 3 Pro
The new models announced by DeepSeek DeepSeek-V3 Mixture of Experts transformer.
The new models are based on three key breakthroughs such as DeepSeek Sparse Attention, Scalable Reinforcement Learning Framework, and Large-Scale Agentic Task Synthesis Pipeline. (Express: Image) Chinese AI startup DeepSeek has released two new AI models named V3.2 and V3.2-Speciale. According to the AI upstart, the new models demonstrated a performance on par with state-of-the-art models like GPT-5 and Gemini 3 Pro. The models were able to achieve this performance while bringing down costs and keeping them accessible under an open-source licence.
The DeepSeek-V3.2 reportedly matches or is close to the performance of Claude Sonnet 4.5, GPT-5, and Gemini 3 Pro in use cases such as tool use, coding tests, etc. Meanwhile, the Speciale model hit gold-medal scores at the 2025 International Math Olympiad and Informatics Olympiad.
Describing the new DeepSeek-V3.2 model, the company said that its approach is built on three key technical breakthroughs – DeepSeek Sparse Attention (DSA), Scalable Reinforcement Learning Framework, and Large-Scale Agentic Task Synthesis Pipeline. The Chinese AI startup claimed that the DSA mechanism ‘substantially reduces computational complexity while preserving model performance’ and has been optimised for long-context scenarios. It essentially splits attention into two components.