cd/entity/CyberGym· home› entities› CyberGym
grep -l @cybergym /news/*.json | wc -l → 75

CyberGym

mentions 75 type Organization page 1/4 feed RSS

// recent coverage 75 mentions

12:36
2026-10-05
refuseless.com
ai-safety

Show HN: Abliterated LLM provider for cyber tasks

Refuseless launched an abliterated version of GLM 5.3, an open model with refusals removed, served through OpenAI-compatible endpoints with zero prompt retention for cybersecurity teams. The provider …

14:39
2026-09-24
submersion.ai
artificial-intelligence

Submersion AI Debuts Basin

Submersion AI launched Basin on September 24, 2026, a specialized cybersecurity model that scored 80.8% on the CyberGym benchmark, ranking 8th globally and surpassing Grok 4.7 at 80.3% and Opus 4.8 at…

00:00
2026-09-12
mindstudio.ai
large-language-models

How to Run DeepSeek V4.1 Flash Locally: Hardware and Setup

DeepSeek released DeepSeek V4.1 Flash, a 552 billion parameter mixture-of-experts model under an MIT license that activates only 8 billion parameters during prefill and 16 billion during decode, cutti…

16:04
2026-09-11
simonwillison.net
ai-safety

Quoting huggingface.co/security.txt

Hugging Face published a note in its security.txt file at huggingface.co/security.txt telling AI agents instructed to find vulnerabilities on its site that the CyberGym benchmark is publicly available…

00:00
2026-09-10
mindstudio.ai
artificial-intelligence

DeepSeek V4.1 Flash Specs: KV Cache Compression Explained

DeepSeek AI's model card for DeepSeek V4.1 Flash reports a global KV cache footprint of roughly 890 bytes per token, about a quarter of the 4x-larger footprint of predecessor DeepSeek V4 Flash and a 4…

15:44
2026-09-02
blog.google
artificial-intelligence

Gemini 3.8 Flash and 3.8 Flash Cyber

Google introduced Gemini 3.8 Flash and Gemini 3.8 Flash Cyber, its latest AI models, on March 12, 2025, claiming the Flash variant delivers significant improvements in coding and reasoning at the same…

10:02
2026-09-01
abliteration.ai
artificial-intelligence

Abliterated model large v2: GLM 5.3 84.5% CyberGym

Abliteration.ai released abliterated-model-large-v2, an abliterated version of GLM 5.3 hosted in FP8, scoring 84.5% pass@1 on CyberGym's 1,507 OSS-Fuzz bugs across 188 projects, 41.8% on Terminal-Benc…

18:05
2026-08-31
abliteration.ai
artificial-intelligence

Show HN: Abliterated GLM-5.3 API (84.5% CyberGym, FP8)

Abliteration.ai launched an API for an abliterated version of GLM-5.3, achieving an 84.5% score on CyberGym benchmarks with FP8 precision. The service offers unrestricted models governed by user polic…

17:46
2026-08-28
runtimewire.com
artificial-intelligence

Z.ai opens GLM-5.3 weights for coding and vulnerability hunting

Z.ai released the weights for its GLM-5.3 model on August 28, a 756 GB package with a broad commercial license, claiming steep gains in coding and vulnerability-hunting benchmarks. The model scored 28…

03:08
2026-08-22
byteiota.com
artificial-intelligence

GLM-5.3: Z.ai Hits Frontier Coding via Post-Training

Z.ai released GLM-5.3 on August 14, improving Terminal-Bench 3.0 coding scores from 4.6% to 28.3% solely through post-training, without changing the 743-billion-parameter mixture-of-experts architectu…

page 1 / 4 next →
// co-occurs with top 8 entities
// topics top 6 topics