cd/entity/DeepSWE· home› entities› DeepSWE
grep -l @deepswe /news/*.json | wc -l → 78

DeepSWE

mentions 78 type Organization page 1/4 feed RSS

// recent coverage 78 mentions

12:36
2026-10-05
refuseless.com
ai-safety

Show HN: Abliterated LLM provider for cyber tasks

Refuseless launched an abliterated version of GLM 5.3, an open model with refusals removed, served through OpenAI-compatible endpoints with zero prompt retention for cybersecurity teams. The provider …

19:51
2026-09-29
letsdatascience.com
artificial-intelligence

Together AI tells LDS what its $3.35 coding result leaves out

Together AI Staff AI/ML Engineer Zain Hasan told Let's Data Science that the company's reported 83% coding-task success rate at $3.35 per task for a DeepSeek-first, GPT-second cascade was reconstructe…

01:23
2026-09-27
ianbarber.blog
ai-safety

Environments and Benchmarks

Xiaomi released MiMo 2.6 this week with an unusually open reinforcement-learning process, publishing an RL dashboard, a technical report, details of how it built its RL environments, and the RL enviro…

12:09
2026-09-24
byteiota.com
large-language-models

DeepSeek V4.1 Flash: Open Weights, $0.003/M Tokens

DeepSeek released V4.1 Flash under MIT open weights, priced at $0.003 per 1 million cached tokens, and the model scores 74.2% on DeepSWE, beating Claude Opus 5. The release covers what changed in the …

06:54
2026-09-24
twitter.com
artificial-intelligence

Contrastive Language Model (CLM): An Ultra-Fast System One Model

Jacky Kwok announced Contrastive Language Model (CLM), an 8-billion-parameter "System One" model trained with a contrastive learning objective linking states and actions, which delivers up to 9× faste…

03:06
2026-09-23
byteiota.com
ai-agents

GitHub HydraFusion: Multi-Model Copilot CLI Is Here

GitHub released Project HydraFusion, a research preview that routes GitHub Copilot CLI coding tasks across multiple models via three execution patterns — Single, Cascade, and Critique — instead of a s…

13:47
2026-09-22
sebastianraschka.com
large-language-models

MiMo-V2.6 Pro Architecture and Training Notes

Xiaomi's MiMo-V2.6 Pro ranks No. 1 on open-weight benchmarks by weighted average despite using a classic Grouped Query Attention architecture with Sliding Window Attention at a 128-token window, accor…

15:05
2026-09-21
github.com
ai-tools

Practical model evaluation and compression tools

Developer 0xSero released model-toolkit, a GitHub repository of standalone Python tools for evaluating, observing, pruning, and quantizing language models, drawn from the REAP and EXL3 experiments inc…

12:56
2026-09-19
scrimdata.com
ai-research

We found defects in 37 of DeepSWE's 113 tasks

An audit of DeepSWE v1.1 found defects or ambiguous requirements in 37 of the benchmark's 113 tasks (32.7%), based on a review of all 372 recorded failures for the Opus 5, Sol and Fable 5 models. The …

00:00
2026-09-16
together.ai
ai-products

Migrating from closed to open source models, Together

Together published a migration guide for companies moving from closed-source to open-source AI models, arguing the process can take weeks to months rather than months to years when a managed service i…

00:00
2026-09-12
mindstudio.ai
large-language-models

How to Run DeepSeek V4.1 Flash Locally: Hardware and Setup

DeepSeek released DeepSeek V4.1 Flash, a 552 billion parameter mixture-of-experts model under an MIT license that activates only 8 billion parameters during prefill and 16 billion during decode, cutti…

12:40
2026-09-10
the-decoder.com
large-language-models

New Deepseek model V4.1-Flash cuts memory needs for AI agents

Deepseek released V4.1-Flash, a multimodal model with 552 billion parameters that cuts KV cache memory to a quarter of its predecessor's requirement, with only 16 billion parameters active per token. …

page 1 / 4 next →
// co-occurs with top 8 entities
// topics top 6 topics