cd/entity/GPT-OSS· home entities GPT-OSS
grep -l @gpt-oss /news/*.json | wc -l → 20

GPT-OSS

mentions 20 type Organization feed RSS

// recent coverage 20 mentions

06:29
2026-08-19
discuss.huggingface.co
large-language-models

Looking for an Open-Source LLM to Replace Llama 3.3 70B Versatile

A developer is seeking open-source replacements for Llama 3.3 70B Versatile after Groq shut down the model, focusing on smaller parameter sizes with comparable performance for RAG, agentic workflows, …

00:00
2026-08-19
lucumr.pocoo.org
artificial-intelligence

What Is Reasoning

A new paper and online discussions have revealed methods to extract hidden reasoning traces from closed-weight AI models, prompting an investigation into how these traces work. Reasoning traces are si…

06:51
2026-08-15
dev.to
artificial-intelligence

An AI Capture-the-Flag Tournament: What the Scoreboard Counted

An AI capture-the-flag tournament run by developer Seth Wheeler found that model size does not reliably predict security reasoning. Over 327 games, a 3B local fine-tune led in main flag captures, whil…

16:23
2026-08-05
anicka.net
artificial-intelligence

Buddhist AI Research

Karma Electric, a Buddhist AI research group, reports that training language models to reason about suffering rather than memorize refusal patterns yields safety that persists even when compliance neu…

04:00
2026-08-04
machinebrief.com
natural-language-processing

Writing-System-Level Tokenizer Adaptation for Byte-Level BPE

A new arXiv paper (2608.00582v1) introduces BPE-guided insertion, a method to adapt byte-level BPE tokenizers to underrepresented languages without changing the model vocabulary size. On Ukrainian ada…

00:00
2026-07-29
runagentrun.co.uk
ai-tools

OpenAI open-sources its security agent

OpenAI open-sourced its Codex Security CLI, a command-line tool that scans code for vulnerabilities and suggests patches, under the Apache 2.0 license on Wednesday. The tool, previously codenamed Aard…

13:12
2026-07-20
tokenstead.ai
ai-tools

Run the Grok CLI on Ollama Cloud and custom providers

XAI's open-source terminal AI coding agent grok now supports custom model endpoints, allowing users to run models like GLM 5.2, DeepSeek V4 Pro, and Kimi K2 via Ollama Cloud or any OpenAI-compatible p…

04:00
2026-07-14
arxiv.org
large-language-models

Reference-Based Distillation Detection in LLMs

Researchers at arXiv introduce a reference-based distillation detection method that identifies whether a large language model was distilled from a specific teacher model by comparing its output alignm…

00:00
2026-07-10
neuronpedia.org
artificial-intelligence

Welcome to the J-Space 🌌

Neuronpedia announced the J-space, a global workspace in AI models revealed by Jacobian lens, and added support for 11 more models. The feature allows users to see hidden reasoning in AI, with pre-fit…

05:52
2026-06-20
github.com
machine-learning

Release 4.0.0 · HuggingFace/Transformers.js

HuggingFace released Transformers.js v4, a major update featuring a new WebGPU backend rewritten in C++ for faster AI model inference in browsers, Node, Bun, and Deno. The release adds support for lar…

16:45
2026-06-15
developer.nvidia.com
machine-learning

Boosting MoE Training Throughput with Advanced Fusion Kernels

NVIDIA introduced advanced fused MLP kernels for mixture-of-experts (MoE) models, built with the CuTe DSL, delivering 1.3x–2x kernel-level speedups and enabling sync-free MoE execution. The optimizati…

14:45
2026-06-05
ianbarber.blog
large-language-models

Somehow, more on distillation

Microsoft AI released a detailed technical report on the development of its first model, MAI-Thinking-1, emphasizing a controlled, reproducible training process built on human-generated data and propr…

00:00
2026-05-10
jola.dev
large-language-models

Running local models on an M4 with 24GB memory

The article describes the author's successful setup for running local AI models on an M4 Mac with 24GB of memory, specifically highlighting Qwen 3.5-9B (Q4 quantized) as the best performing model at ~…

// co-occurs with top 8 entities
// topics top 6 topics