New Model Available: Ling-3.0-flash-VL InclusionAI released Ling-3.0-flash-VL, a hybrid instant/reasoning model with tool calling that builds on Ling 3.0 Flash (124B total / 5.5B active MoE) and adds native visual perception and advanced visual agent capabilities. The model is listed on an AI gateway with provider routing, and its benchmark scores are sourced from Artificial Analysis. Ling 3.0 Flash VL builds on Ling 3.0 Flash 124B total / 5.5B active MoE from InclusionAI , further strengthening its language capabilities while adding native visual perception and advanced visual agent capabilities. Hybrid instant/reasoning model with tool calling. Back to Models https://zenmux.ai/models Providers Route requests across multiple providers. Copy a provider slug to set your preference. ~~$0.06~~ $0 / M tokens ~~$0.18~~ $0 / M tokens Read: ~~0.012~~ 0 / M tokens Write: - / M tokens262.14K-- Uptime 24hours Direct request success rate on AI Gateway and per-provider. Throughput 24hours P50 throughput on live AI Gateway traffic, in tokens per second TPS . Latency 24hours P50 time to first token TTFT on live AI Gateway traffic, in milliseconds. Activity Token volume and request traffic to this model over time. Benchmarks Scores on standardized evaluations. Higher percentages are better — and rank percentile shows Metrics sourced from Artificial Analysis https://artificialanalysis.ai/ Apps Public apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for. View All https://zenmux.ai/analytics/apps Related Models More models from inclusionAI https://zenmux.ai/inclusionai