cd/entity/CyberGym· home entities CyberGym
grep -l @cybergym /news/*.json | wc -l → 60

CyberGym

mentions 60 type Organization page 1/3 feed RSS

// recent coverage 60 mentions

03:08
2026-08-22
byteiota.com
artificial-intelligence

GLM-5.3: Z.ai Hits Frontier Coding via Post-Training

Z.ai released GLM-5.3 on August 14, improving Terminal-Bench 3.0 coding scores from 4.6% to 28.3% solely through post-training, without changing the 743-billion-parameter mixture-of-experts architectu…

13:40
2026-08-18
cryptobriefing.com
artificial-intelligence

CyberGym results show AI surpasses 90% in vulnerability detection

UC Berkeley's CyberGym benchmark shows AI agents can now reproduce real-world software vulnerabilities with 93.2% accuracy, up from 10-30% a year ago. The leaderboard leader, a Sangfor AI Agent runnin…

06:00
2026-08-17
dev.to
artificial-intelligence

Zhipu va ouvrir les poids du meilleur chasseur de failles

Zhipu AI will release its GLM-5.3 model in stages, starting with API access and then open weights after security evaluations, positioning it for cyber defense and programming. The model scored 84.5 on…

09:09
2026-08-15
sourcefeed.dev
artificial-intelligence

Z.ai Caught Up on Finding Bugs, Not on Exploiting Them

Z.ai's GLM-5.3 scored 84.5% on UC Berkeley's CyberGym benchmark on August 14, edging out Anthropic's restricted Claude Mythos 5 at 83.8% and OpenAI's GPT-5.6 Sol at 83.6%, but the Beijing lab's own fi…

23:02
2026-08-14
cryptobriefing.com
artificial-intelligence

GLM-5.3 identifies serious vulnerability in Cursor code editor

Z.ai's open-weights model GLM-5.3, released on August 14, 2026, identified a significant vulnerability in Cursor, the AI-powered code editor, according to security researcher Joshua Saxe. The model sc…

20:17
2026-08-14
dev.to
artificial-intelligence

GLM 5.3: Zhipu's Open-Weight Model Excels at Coding and Cyber

Zhipu AI released GLM 5.3, an open-weight model that improves coding and cyber capabilities through advanced post-training techniques rather than a larger architecture. The model, based on the same 74…

09:09
2026-08-14
sourcefeed.dev
artificial-intelligence

Z.ai Built a Better Coder and Blinked on Open Weights

Z.ai shipped GLM-5.3 on August 14 with Terminal-Bench 3.0 scores jumping from 4.6% to 28.3% and DeepSWE v1.1 from 46.2% to 66.9% on the same 743B base model as GLM-5.2, attributing all gains to scaled…

page 1 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics