cd/entity/ModelScope· home entities ModelScope
grep -l @modelscope /news/*.json | wc -l → 36

ModelScope

mentions 36 type Organization page 1/2 feed RSS

// recent coverage 36 mentions

07:11
2026-08-30
gist.github.com
developer-tools

KV Cache Size Calculator

A developer released a Python tool that calculates and plots KV cache size versus context length from HuggingFace config.json files. The tool supports standard MHA/GQA, MLA, hybrid architectures, and …

15:21
2026-08-28
gist.github.com
large-language-models

Best llama.cpp config for Qwen3.8-Flash-Next (RTX 4090 24GB)

A developer has published a configuration guide for running the Qwen3.8-Flash-Next 125B MoE model with llama.cpp on an RTX 4090 24GB system, achieving up to 29 tokens per second decode speed. The setu…

00:00
2026-08-28
mindstudio.ai
large-language-models

Tencent Hy4 Preview: Inside the 770B Open-Weight Flagship Model

Tencent released Hy4 preview, a 770B-parameter open-weight Mixture-of-Experts language model with 49B active parameters and a 1M token context window, under Apache 2.0. In a blind evaluation across 20…

00:00
2026-08-28
mindstudio.ai
artificial-intelligence

How to Run Tencent's Hy4 Preview Locally with vLLM or SGLang

Tencent's 770B-parameter Hy4 preview model, a Mixture-of-Experts architecture with 49B activated parameters per token, is now deployable locally via prebuilt Docker images for vLLM and SGLang, requiri…

22:45
2026-08-26
lanternite.org
ai-ethics

Find out where your images have been scraped

Lantern, a new free service, lets artists and creators upload images to generate one-way fingerprints and scans publicly available AI training datasets on Hugging Face and ModelScope to notify them if…

13:09
2026-08-26
cryptobriefing.com
artificial-intelligence

Alibaba unveils latest Qwen model to boost global AI adoption

Alibaba released three new AI models in August 2025, including the flagship Qwen3.8-Max with 2.4 trillion total parameters (95 billion active), the lightweight Qwen3.8-27B with 27 billion parameters, …

12:48
2026-08-12
tokenstead.ai
artificial-intelligence

Nemotron 3.5 Lightning

NVIDIA released Nemotron 3.5 Lightning on 2026-08-11, a 31.6B-parameter mixture-of-experts model with ~3.6B active parameters per token, hybrid Mamba-Transformer architecture, multi-token prediction, …

21:13
2026-08-09
byteiota.com
artificial-intelligence

Qwen3.8 Open Weights Drop This Week: Read Before You Download

Alibaba will release open weights for Qwen3.8-Max and Qwen3.8-27B this week on Hugging Face and ModelScope, marking the first time a Max-class model is open-sourced. The 2.4-trillion-parameter Mixture…

06:15
2026-08-05
unite.ai
large-language-models

Tencent Opens Hy3 to Global Users Across Products and Cloud

Tencent opened its Hy3 large language model to global users on August 5, 2026, making it available through WorkBuddy, Miora, and Tencent Cloud TokenHub, with free access until August 31, 2026. Hy3, a …

16:12
2026-08-03
startupfortune.com
large-language-models

Alibaba's Qwen3.8-Max Launch Pushes Its Shares Up 6% in Hong Kong

Alibaba Group Holding Ltd. shares rose as much as 6% in Hong Kong trading on Monday after the company unveiled Qwen3.8-Max, a 2.4 trillion parameter mixture-of-experts model that Alibaba claims trails…

13:06
2026-08-03
github.com
ai-agents

Rescene – Free AI agent aggregator, no API key required

Rescene, a free AI agent aggregator, has been released, requiring no API key. It aggregates 7 free providers and 18 model entries, featuring intelligent routing, browser automation, and Computer Use, …

09:47
2026-07-28
comfyfile.com
large-language-models

How to Download and Run Kimi K3 Open Weights

Moonshot AI released the full Kimi K3 open weights on July 27, 2026, a 2.8-trillion-parameter mixture-of-experts model with a one-million-token context window and a 1.4 TB download. The model uses nat…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics