Ling-3.0-flash-Fin is a finance-enhanced MoE model built on Ling-3.0-flash, with 124 billion total parameters and approximately 5.1 billion activated parameters. Designed for real-world investment workflows, it is optimized for complex multi-step tasks and long-horizon planning and execution. With a relatively small active parameter footprint, it delivers competitive financial performance while maintaining strong general capabilities in reasoning, coding, and mathematics.
Providers #
Route requests across multiple providers. Copy a provider slug to set your preference.
$0.06
$0
/ M tokens
$0.18
$0
/ M tokens Read:
0.012
0/ M tokens
Write:
-/ M tokens262.14K3.77s13.0tps
Uptime #
24hours Direct request success rate on AI Gateway and per-provider.
Throughput #
24hours P50 throughput on live AI Gateway traffic, in tokens per second (TPS).
Latency #
24hours P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds.
Activity #
Token volume and request traffic to this model over time.
Benchmarks #
Scores on standardized evaluations. Higher percentages are better — and rank percentile shows
Metrics sourced fromArtificial Analysis
Apps #
Public apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for. View All
Related Models #
More models from inclusionAI