Ornith-1.0-9B
Ornith AI released Ornith-1.0-9B, a 9-billion-parameter dense language model, on 2026-06-25, with a 262K context window and MIT license on HuggingFace. The model scores 43.1 on Terminal-Bench 2.1, 69.…
Ornith AI released Ornith-1.0-9B, a 9-billion-parameter dense language model, on 2026-06-25, with a 262K context window and MIT license on HuggingFace. The model scores 43.1 on Terminal-Bench 2.1, 69.…
Ornith AI released Ornith-1.5-397B, a 403B-parameter mixture-of-experts agentic-coding model, on 2026-08-18, claiming coding scores on par with Claude Opus 4.8 and ahead of GLM-5.2 and DeepSeek-V4-Fla…
An anonymous reasoning model called 'stealth/ox-alpha' has been available free on OpenRouter since Aug 20, 2026, with undisclosed operators but community fingerprinting suggesting it is an unreleased …
XAI released Grok 4.6, a proprietary frontier model built with Cursor, featuring 500K context, a knowledge cutoff of Feb 2026, and text and image input with text output. Priced at $2 per 1M input toke…
Ornith AI released Ornith-1.0-35B, a 35B-parameter mixture-of-experts model with 3B active parameters per token, on June 25, 2026, under an MIT license on HuggingFace, featuring a 262K context window.…
Ornith AI released Ornith-1.0-397B, a 397-billion-parameter mixture-of-experts model with 262K context, on 2026-06-25, now superseded by Ornith-1.5-397B. The model self-reports coding scores of 77.5 o…
Ornith AI released Ornith-1.5-9B, a 9-billion-parameter dense language model, on August 18, 2026, claiming it matches or exceeds larger models like Gemma 4-31B and Qwen 3.6-35B on agentic coding tasks…
Ornith AI released Ornith-1.5-35B-A3B, a 36B-parameter mixture-of-experts model with 3B active parameters per token, on August 18, 2026, under an MIT license on HuggingFace. The model achieves 79.0 on…
Alibaba released the open-weight Qwen3.8-27B multimodal model on 2026-08-05 under Apache-2.0, featuring a 27B dense architecture with a vision encoder, 262,144-token native context, and hybrid thinkin…
DeepInfra offers the cheapest API pricing for Qwen3 30B A3B at $0.12 per million input tokens, with Alibaba at $0.13 and NextBit at $0.14, according to live pricing data refreshed about 19 hours ago v…
Google's Gemma 4 26B A4B, a 25.2B-parameter multimodal mixture-of-experts model with 3.8B active parameters and 128 experts, is now available with DekaLLM offering the cheapest API pricing at $0.06 pe…
Z.ai released GLM-5.3 on August 14, 2026, a post-training upgrade of the same 743B Mixture-of-Experts base as GLM 5.2, with major gains in coding and agentic ability and an emergent cybersecurity capa…
Zhipu AI released GLM 5.3, a 743B-parameter mixture-of-experts model with ~40B active per token, achieving a Terminal-Bench 3.0 score of 28.3 (up from 4.6 on GLM 5.2) and a DeepSWE v1.1 score of 66.9 …
OpenAI released GPT-5.6 Terra, a proprietary AI model available in August 2026, which does not embed watermarks or provenance metadata in generated output. The model scores 90 on general performance a…
Ruby on Rails published the first public, same-harness benchmark of frontier models performing real Rails tasks, testing 8 models across 21 atomic tasks with 504 runs at a total cost of $491. Claude O…
NVIDIA released Nemotron 3.5 Lightning on 2026-08-11, a 31.6B-parameter mixture-of-experts model with ~3.6B active parameters per token, hybrid Mamba-Transformer architecture, multi-token prediction, …
OpenRouter lists GPT-5.6 Sol, a proprietary AI model released in August 2026, with a general score of 95 and a value score of 19 points per dollar per million input tokens, priced at $5.00 per million…
OpenAI released GPT-5.6 Luna, a proprietary AI model with a general score of 82 and a price of $0.20 per million input tokens, making it the cheapest option among providers. The model does not embed w…
Anthropic's Claude Fable 5, a proprietary model launched in June 2026, scores 95 on general benchmarks and offers a value of 10 points per dollar per million input tokens, with input pricing at $10.00…
Meta released Muse Glimmer 30B, its first open-weights model in the Muse family, under the Apache 2.0 license, featuring a 1.8B vision encoder and a 128K+ context window for coding, agentic workflows,…