{"slug": "deepseek-is-training-a-2t-parameter-model-and-plans-to-build-an-8t-parameter-one", "title": "DeepSeek is training a 2T-parameter model and plans to build an 8T-parameter one", "summary": "DeepSeek CEO Liang Wenfeng told investors that using more domestic chips for AI training is now a major priority, with Huawei expected to begin deliveries as early as Q4. DeepSeek is currently training a 2-trillion-parameter model and plans to eventually build an 8-trillion-parameter model.", "body_md": "Wall St Engine on X: \"DeepSeek CEO Liang Wenfeng told investors that using more domestic chips for AI training is now a major priority, with Huawei expected to begin deliveries as early as Q4.\nNote: DeepSeek is training a 2T-parameter model and plans to eventually build an 8T-parameter model.\" / X\n\nWall St Engine on X: \"DeepSeek CEO Liang Wenfeng told investors that using more domestic chips for AI training is now a major priority, with Huawei expected to begin deliveries as early as Q4.\nNote: DeepSeek is training a 2T-parameter model and plans to eventually build an 8T-parameter model.\"\n\nDeepSeek CEO Liang Wenfeng told investors that using more domestic chips for AI training is now a major priority, with Huawei expected to begin deliveries as early as Q4.\nNote: DeepSeek is training a 2T-parameter model and plans to eventually build an 8T-parameter model.\n\nDeepSeek CEO Liang Wenfeng told investors that using more domestic chips for AI training is now a major priority, with Huawei expected to begin deliveries as early as Q4.\nNote: DeepSeek is training a 2T-parameter model and plans to eventually build an 8T-parameter model.\n\nDeepSeek prioritizing domestic chips with Huawei deliveries as early as Q4, while training toward 2T then 8T parameters, is a supply-chain CapEx decision first. Export controls show up as wafer and interconnect lead times, not just model cards.", "url": "https://wpnews.pro/news/deepseek-is-training-a-2t-parameter-model-and-plans-to-build-an-8t-parameter-one", "canonical_source": "https://twitter.com/wallstengine/status/2101982843656388644", "published_at": "2026-09-21 13:06:29+00:00", "updated_at": "2026-09-21 13:25:12.975767+00:00", "lang": "en", "topics": ["artificial-intelligence", "large-language-models", "ai-chips", "ai-infrastructure"], "entities": ["DeepSeek", "Liang Wenfeng", "Huawei", "Wall St Engine"], "alternates": {"html": "https://wpnews.pro/news/deepseek-is-training-a-2t-parameter-model-and-plans-to-build-an-8t-parameter-one", "markdown": "https://wpnews.pro/news/deepseek-is-training-a-2t-parameter-model-and-plans-to-build-an-8t-parameter-one.md", "text": "https://wpnews.pro/news/deepseek-is-training-a-2t-parameter-model-and-plans-to-build-an-8t-parameter-one.txt", "jsonld": "https://wpnews.pro/news/deepseek-is-training-a-2t-parameter-model-and-plans-to-build-an-8t-parameter-one.jsonld"}}