Sections

Search

US and China Discuss AI National Security Alert SystemQwen-Image-2.1 Releases Open Weights for Unified Image Generation and EditingAI Chatbots Drive Surge in Discovered Security FlawsReddit: 16GB Is the Common VRAM Ceiling for Local LLM UsersGoogle Launches Googlebook Laptops for Android Integration
All stories

Models··1 min read

DeepSeek Training 2T-Parameter Model, Plans 8T-Parameter Model Next

Reddit r/LocalLLaMA: DeepSeek is training a 2-trillion parameter model, with plans for an 8-trillion parameter model in the future.

DeepSeek Announces 2T and 8T Model Plans

DeepSeek is currently training a 2-trillion parameter AI model and has stated plans to build an 8-trillion parameter model in the future. [1]

Context: Previous DeepSeek Models

The discussion outlines current DeepSeek models, including a 'Pro' model at 1.6T parameters, which contrasts with the more ambitious 2T and 8T parameter targets. [1]

Sources

  1. Reddit r/LocalLLaMA · Community discussion · Sep 21, 2026
    Deepseek training 2T and plans 8T model
    DeepSeek is training a 2T-parameter model and plans to eventually build an 8T-parameter model.
    Current DeepSeek models: Flash parameter count of 552 billion Pro: 1.6T (trillion) total parameters with 49B (billion) activated weights per token