cd/entity/ONNX Runtime· home› entities› ONNX Runtime
grep -l @onnx runtime /news/*.json | wc -l → 64

ONNX Runtime

mentions 64 type Person page 3/4 feed RSS

// recent coverage 64 mentions

00:00
2026-07-23
offlinetts.com
artificial-intelligence

Supertonic 3 Guide: 31-Language Browser TTS

Supertonic 3, a 99M-parameter, 31-language text-to-speech model designed for local ONNX Runtime inference, is now available in OfflineTTS with 10 built-in voice presets, speed and step controls, wavef…

06:10
2026-07-18
github.com
artificial-intelligence

SigLIP 2 text embedding on CPU with Rust and ONNX

A new CPU-only SigLIP 2 text embedding server built with Rust and ONNX Runtime provides an OpenAI-compatible embeddings endpoint for live text queries, allowing systems to reserve GPU capacity for ima…

07:37
2026-07-11
machinebrief.com
artificial-intelligence

VisionAId: Transforming Smartphones into Visual Assistants

VisionAId, a new Android app, turns smartphones into visual assistants for the visually impaired using six on-device deep learning models and optional cloud-based Google Gemini Flash. The app features…

10:23
2026-06-25
phoronix.com
artificial-intelligence

AMD Contributes ONNX Runtime Backend To FFmpeg DNN Filter

AMD engineer Steven Xiao contributed an ONNX Runtime back-end for FFmpeg's DNN filter, enabling AI model inferencing on multiple GPU and NPU platforms including NVIDIA CUDA, Windows DirectML, and AMD …

16:34
2026-06-24
dev.to
developer-tools

Neonmem 0.9.7 is out.

Neonmem 0.9.7 introduces a two-level importer that separates folders and files into a searchable knowledge pool and agent chats into typed memories, using IBM Granite-30M embeddings via ONNX Runtime f…

05:52
2026-06-20
github.com
machine-learning

Release 4.0.0 · HuggingFace/Transformers.js

HuggingFace released Transformers.js v4, a major update featuring a new WebGPU backend rewritten in C++ for faster AI model inference in browsers, Node, Bun, and Deno. The release adds support for lar…

15:20
2026-06-19
heyneo.com
ai-tools

Extend Claude limits by offloading AI tasks to Neo

Neo launched an MCP server that lets users offload AI tasks from Claude, reducing costs by 62% and speeding up runtime by 37% in benchmarks. The tool integrates with Claude Code and other MCP clients,…

08:02
2026-06-18
devclubhouse.com
ai-infrastructure

Unified x86 AI Acceleration: Inside the New ACE Specification

The x86 Ecosystem Advisory Group released the AI Compute Extensions (ACE) Specification on June 15, 2026, defining standardized x86 extensions for matrix multiplication and tile registers to accelerat…

← prev page 3 / 4 next →
// co-occurs with top 8 entities
// topics top 6 topics