{"slug": "a-dark-horse-enter-china-s-ai-race-startlux", "title": "A dark horse enter China's AI race: StartLux", "summary": "StartLux-V1.0-27B-Preview, a 27-billion-parameter local model from Chinese startup StartLux (formerly Yuandian Xinghui), scored 39.25% in the CAICT MCP specialized test, ranking second overall and trailing DeepSeek-V4-Pro's 1.6-trillion-parameter model by just 1.3 percentage points. Founded by Shanda co-founder Chen Danyan, StartLux focuses on local models that run on consumer PCs, aiming to challenge cloud-based AI dominance.", "body_md": "Something big has happened - a dark horse has emerged among China's top domestic large model developers.\n\nA new company, with its first model having only 27B parameters, took second place overall in the CAICT MCP specialized test.\n\nRanked ahead of it is the killer weapon Liang Wenfeng has kept under wraps for nearly a year: DeepSeek-V4-Pro, boasting a parameter count of 1.6 trillion.\n\nThe difference between the two is a mere 1.3 percentage points.\n\nThis dark horse is the StartLux-V1.0-27B-Preview, from StartLux (formerly Yuandian Xinghui).\n\nLooks a bit unfamiliar, doesn't it? Don't worry, its founder and CEO is an old acquaintance: Chen Danyan.\n\nKnown as the \"godfather\" of programmers in the internet era, his achievements are a matter of public record:\n\nShanda Network's co-founder and the head of Lian Shang Network, one of China's earliest programmers to introduce the concept of \"shareware\"\n\nAfter a decade of retirement, he has made a comeback, this time betting on local models.\n\nThis has to do with Chen Dawei's recent rare public appearance.\n\nAt the 18th anniversary reunion of Shanda Innovation Institute, he publicly declared \"eight non-consensus views for the AI era,\" four of which center on local models.\n\nLocal models will completely destroy the cloud market, catching up with Claude in three years and occupying 80% of the market. The model competition based on parameters is going to be outdated...\n\nStartLux is the best representation of his idea.\n\nThe company's business has not followed the industry trend, instead focusing on the commercialization of small and beautiful local models, making it the world's first truly local model company in the true sense.\n\nAs the first market-oriented scorecard, StartLux-V1.0-27B-Preview does not rely on the cloud and can run directly on consumer-grade PCs.\n\nIn other words, this local model, which is nearly 60 times smaller, has Agent capabilities that can match those of a trillion-level cloud-based flagship model.\n\nWhat justification is there for this?\n\n## Small Model Achieves Big Results, 27B Outperforms 1.6T\n\nBefore the answer is revealed, let's take a look at who the comparison is being made to.\n\nIt's often said that nobody remembers the second place, unless the first is DeepSeek.\n\nMoreover, the gap is minimal, making it well worth discussing.\n\nThe results come from the authoritative institution, China Academy of Information and Communications Technology's trustworthy AI large model benchmark test MCP special item, which sets six types of tasks around real application scenarios:\n\nLocation navigation, web search, browser automation, financial analysis, code repository management, 3D design.\n\nAn additional comprehensive assessment will be added, focusing on evaluating Agent's multi-tool collaboration, complex task execution, and interaction in real-world environments.\n\nIn simple terms, MCP-Universe doesn't evaluate models based on their responses, but rather on whether they can actually get things done. This is also the most fundamental aspect of judging an Agent's quality.\n\nThe results showed that StartLux-V1.0-27B-Preview had a comprehensive score of 39.25%, ranking second.\n\nDeepSeek-V4-Flash-0731, with over 284 billion parameters, and Step-3.7-Flash, at 198 billion, trail DeepSeek-V4-Pro by just 1.3 percentage points.\n\nWith the same 27B parameter scale, StartLux also surpasses Qwen 3.6 by 5.34 percentage points.\n\nIt also excelled in individual subjects, ranking first in location navigation, financial analysis, and browser automation, with its other sub-items also ranking high.\n\nLet's take a look at two case studies, putting data aside.\n\nThe first question is about a two-year Microsoft stock investment, with Claude Sonnet 4.6 as the competing topic.\n\nClaude's answer is: $47,254, 89.02%.\n\nVideo link: https://mp.weixin.qq.com/s/365CtdgGYFlEKNDICtoCfg\n\nStartLux gave: $47,499.09, 90.00%.\n\nIt may seem similar, but in the financial industry, a tiny difference can lead to enormous losses.\n\nCareful examination of the two models' reasoning processes shows that, due to missing raw data, Claude Sonnet 4.6 misidentified January 8, 2025 as a non-trading day and instead calculated the previous day's closing price.\n\nUnder the same circumstances, StartLux retrospectively reviews the original data and verifies the market trends around the target date to confirm the accurate closing price before completing the calculation and generating visualization.\n\nUltimately, the conclusion reached by StartLux proved correct, and it was fully verifiable and traceable.\n\nThe second task was more straightforward: both models were asked to search for flight tickets in a browser at the same time.\n\nI'm unable to open a browser or interact with live websites like Google Flights. I can only process text and don't have real-time browsing capabilities. To find this flight yourself, here's what you'd do: 1. Go to google.com/flights 2. Enter Singapore (SIN) → Beijing 3. Select one-way, set departure date to 5 days from today 4. Filter by \"Nonstop\" and \"Economy\" 5. Check the results — note that flights to Beijing may land at either Capital (PEK) or Daxing (PKX). Exclude any Daxing arrivals. 6. Compare prices and pick the cheapest nonstop option landing at PEK. If you'd like, I can help you think through typical price ranges, airline options on this route, or general tips for finding cheap flights.\n\nAmong them, StartLux-V1.0-27B-Preview completed the search in about 95 seconds, finding an Air China ticket priced at $299.\n\nBy contrast, Claude Sonnet 4.6 required more screenshot confirmations when navigating date selection, popup dismissal, and filter menus, and even accidentally triggered the time filter panel at one point.\n\nMore than 200 seconds later, it gave a lowest quote of $556. The price was higher, and the search time was still twice that of StartLux.\n\nIn particular, in terms of operational pathways, StartLux is much more concise, requiring only 12 steps, whereas Sonnet requires a full 21 steps.\n\nThis is enough to illustrate that, in Agent tasks, the scale of parameters is no longer the only decisive variable.\n\nNew variables are being introduced through post-training.\n\n## Cutting Prices, Not Capabilities\n\nStartLux-V1.0-27B-Preview was not trained from scratch.\n\nIt is also based on Qwen3.6-27B, but the final test score is significantly higher than Qwen, and the reason lies in the task data and automated post-training methods.\n\nIn simple terms, StartLux trains a more capable Agent, with training data focusing on reinforcing abilities such as task understanding, tool selection, parameter construction, multi-step execution, status checking, and result verification.\n\nThe model needs to learn not only to output the final text, but also when to invoke which tool, how to adjust when a tool returns an exception, and under what circumstances it can declare the task complete.\n\nFurther training will push this process even further.\n\nThe team has independently developed a brand-new, multi-dimensional, verifiable, and scalable model iteration optimization technology, which uses the AI-trained AI (Auto Research) method to enable the model to autonomously execute tasks in a real-world tool environment and continuously adjust its strategy based on environmental feedback.\n\nFor instance, the financial analysis case mentioned earlier, which involves backtesting and revision, as well as the constraint identification and path selection in browser tasks, are the most intuitive manifestations of post-training.\n\nAccording to official information, StartLux-V1.0-27B-Preview is also the country's first local Agent model to complete post-training using the Auto Research method.\n\nThis does not mean that the Scaling Law is invalid.\n\nLarge parameter cloud models are still the mainstream choice at present, but StartLux has also given a clear signal: this is not the only solution.\n\nIn the words of Chen Danyan:\n\nScaling Law is a \"passing fairy,\" without it, AI cannot take off, but it has merely passed through the development path of AI and its future is not necessarily tied to it.\n\nThe industry has also become aware of this issue, and the technical path for large models is currently showing a trend of distinct divergence:\n\nOn one hand, there are the die-hard believers in \"more power leads to miracles,\" with parameters scaling from tens of billions to hundreds of billions, and then to trillions, while training costs also surge exponentially;\n\nOn the other hand, the Agentic assessment system, represented by Harness, has quickly gained popularity, with an increasing number of experts and scholars explicitly advocating for \"less is more\".\n\nTo paraphrase Wang Yangming, the unity of knowledge and action means higher-quality action is what truly matters. StartLux offers another example of simplifying models.\n\nThis can also explain why StartLux insists on local models.\n\nOnce the model is on a PC, for long-term tasks, its cost structure can shift from continuously accumulating cloud-based token fees to more controllable device computing power and electricity consumption, while the model can also better understand the user's long context, achieving more personalized goals.\n\nIt's not just StartLux, as Meta, Google, and NVIDIA have also recently been increasing their investment in local small model development.\n\nHowever, most of these are still in the experimental stage or cater to niche groups, with only one company focusing on local model commercialization.\n\nStartLux also stated that they will steadily advance their own foundation model training and explore new architectures such as diffusion-based language models, and as long as they are on the right path, the future is promising.\n\nStartLux has taken over the local model, and since this is a non-consensus route, those at the helm need to be two types of people: those who dare to take bets and those who can deliver results.\n\nStartLux's all-star team exemplifies this, pairing entrepreneurs with scientists in a formidable alliance.\n\nStartLux founder Chen Danyan\n\nCEO Chen Danyan had previously been briefly introduced as a serial entrepreneur and one of China's first-generation programmers, who rose to fame around the same time as Zhang Xiaolong and Lei Jun, and started his business ventures in the same era as Ma Yun and Ma Huateng.\n\nHe was one of the first people in China to introduce the concept of \"shared software\" and co-founded Shanda Network with his brother Chen Tianqiao, as well as the Shanda Innovation Institute, a cradle of internet innovation, and was also instrumental in incubating the nationally popular product \"WiFi Master Key\".\n\nIt can be said that he is extremely familiar with Chinese market users and products, and local models are also his comfort zone.\n\nStartLux co-founder Guo Quanwei\n\nThe person responsible for implementing the technical roadmap is StartLux co-founder and CTO, Guo Quanwei.\n\nKuo Chuan-wei holds a Ph.D. in Computer Science and Engineering from National Yang Ming Chiao Tung University, with research areas covering locally deployed large models, Agentic AI, AI for Science, AI for Finance, and privacy-preserving machine learning.\n\nBefore joining StartLux, he served as the Chief Algorithm Scientist at AI Science company Huanliang Technology, and earlier worked at the Industrial Technology Research Institute of Taiwan, where he developed data privacy, data de-identification, and privacy-preserving machine learning.\n\nHe is also a recipient of the 2024 TAAI Best Paper Award and holds data-privacy-related invention patents as the first named inventor.\n\nAnother co-founder of StartLux, Luo Yongxiang, is the former Managing Director of Morgan Stanley Asia.\n\nChen Danyan understands products and users, Guo Quanwei has long studied local models, Agents, and data privacy, and Luo Yongxiang is in charge of marketing and investment financing. This combination is highly suited to StartLux and will also help drive StartLux's long-term development.\n\nAs for what the team ultimately wants to deliver, it's not just a set of model weights.\n\nIn StartLux's vision, local intelligent solutions should be deployable with one click, similar to installing Office. Users do not need to understand quantization, GPU memory configuration, and inference frameworks, nor do they need to optimize them repeatedly themselves.\n\nThe team currently plans to launch its first-generation local smart solution for enterprise and individual users within the year.\n\nIn this light, the emergence of StartLux is by no means just \"another player\" entering the scene.\n\nIt also represents that, following the emergence of cloud-based model companies like DeepSeek and Kimi, domestic local models are also starting to fill in the gaps.\n\nFrom a rising star in the intelligent era to circling back to the first-generation programmers of the internet era, the fundamental paradigm of models is shifting gears—but through it all, Chinese companies have kept passing the baton and pushing forward.\n\nOfficial website link: https://startlux.com/", "url": "https://wpnews.pro/news/a-dark-horse-enter-china-s-ai-race-startlux", "canonical_source": "https://chinaonchina.com/article/chen-dawei-returns-enters-the-large-model-sector", "published_at": "2026-09-03 11:12:57+00:00", "updated_at": "2026-09-03 11:23:04.383901+00:00", "lang": "en", "topics": ["artificial-intelligence", "large-language-models", "ai-products", "ai-startups"], "entities": ["StartLux", "Yuandian Xinghui", "Chen Danyan", "DeepSeek-V4-Pro", "CAICT", "Shanda Network", "Claude Sonnet 4.6"], "alternates": {"html": "https://wpnews.pro/news/a-dark-horse-enter-china-s-ai-race-startlux", "markdown": "https://wpnews.pro/news/a-dark-horse-enter-china-s-ai-race-startlux.md", "text": "https://wpnews.pro/news/a-dark-horse-enter-china-s-ai-race-startlux.txt", "jsonld": "https://wpnews.pro/news/a-dark-horse-enter-china-s-ai-race-startlux.jsonld"}}