cd /news/ai-infrastructure/wici-one · home topics ai-infrastructure article
[ARTICLE · art-119704] src=wici.ai ↗ pub= topic=ai-infrastructure verified=true sentiment=· neutral

WiCi One

WiCi One, a wireless GPU device from an unnamed company, is now available for pre-order at $1,999 USD for early signups, down from $2,599, and claims to deliver real-time AI inference and gaming performance over Wi-Fi 7 to existing devices. The device features a full-power GPU, Wi-Fi 7 radio, and a software stack supporting local AI models like Ollama and LM Studio, with benchmarks showing up to 500+ tok/s for Gemma 4 and 110 fps in Cyberpunk 2077 at 4K Ultra with DLSS 4. The product aims to provide a private, low-latency alternative to cloud computing for AI workloads.

read8 min views3 publishedSep 3, 2026

WiCi One wirelessly equips the devices you already own with the best GPU — real-time, private AI without making anything larger, heavier, or hotter.

Starting at $2,599 $1,999 USDFor early signups only

Top 10 ideas get WiCi One free.

Sign up for the newsletter and a chance to get one free in the Pioneer plan — we'll send you the details.

#

Let the GPU come to you.

AI is becoming personal. WiCi One is the wireless compute layer for this new era — one dedicated device that equips your phones, laptops, and future AI devices with the best GPU over Wi-Fi.

A wireless compute layer.

Compute becomes a shared resource every device can access wirelessly — quiet, always on, always near.

Not device-only.

Powerful AI shouldn't make devices larger, heavier, or hotter. Your devices stay light — WiCi One carries the load nearby.

Not cloud-only.

Skip the cloud's tradeoffs — latency, ongoing costs, privacy, and a hard dependence on connectivity.

A new kind of #

GPU runtime

A full software stack runs on the main board — driver, runtime, scheduler, and developer tools. Everything an app needs to treat WiCi One as nearby compute.

Wi-Fi 7 radio

A high-bandwidth wireless module tuned for the WiCi protocol. Twin antennas keep nearby devices in lockstep — low latency, always available, no cables.

Full-power GPU

A dedicated graphics processor sized for real-time AI inference. The heavy lifting happens here, just a few feet from your phone, your laptop, your speaker.

Chassis & cooling

A precision-machined enclosure with a quiet thermal path beneath a curved top. Built to sit on a shelf and be forgotten, the way infrastructure should feel.

Real workloads, #

served near you.

Works with the apps you already use.

AI & local models

  • Ollama
  • LM Studio
  • Hugging Face
  • PyTorch
  • ComfyUI
  • Open WebUI

Video & streaming

  • Premiere Pro
  • CapCut
  • OBS Studio

3D & creation

  • Blender
  • AutoCAD
  • Rhino
  • Maya
  • Unreal Engine
  • Unity

Games

A few highlights, not the full list. If an app can use a GPU, it can use WiCi One with no changes needed: Driver mode shows up as a standard local GPU, and API mode serves standard endpoints. Your apps and games run on your own machine — not remote play, just a wireless GPU. Read our developer section

Benchmark† 5060 Ti edition 5090 edition
Kimi K32.8T params 2.5 tok/s
5 tok/s
DeepSeek V4 Flash284B params 8 tok/s
20 tok/s
Gemma 426B params 140 tok/s
500+ tok/s
FLUX.2 Klein1024² Model · 4B~0.49 img/s
Model · 9B~0.92 img/s
Z-Image-Turbo1024² Model · 6B~0.22 img/s
Model · 6B~1.19 img/s
Wan 2.2720p Model · TI2V-5B~6.4 gen fps
Model · I2V-A14B~3.0 gen fps
MiniMax H30.5 MP · native audio Model · 33B~0.7 gen fps
Model · 33B~4.1 gen fps
Cyberpunk 20774K · Ultra preset · DLSS 4 80 fps
110 fps
Black Myth: Wukong4K · Cinematic · DLSS 4 60 fps
90 fps
CapCut4K export 75 fps
90 fps
Premiere Pro4K H.264 export 70 fps
90 fps
Audio MemoryTTFA ≤1.2 s
≤0.8 s
BlenderCycles render 2,900 samples/min
6,000 samples/min

†Open-weight models, served behind OpenAI- and Ollama-compatible endpoints.

Measured over Wi-Fi 7 on a production WiCi One — close to cloud rates of 20-40 tok/s.

How do we serve T-param LLM? Learn more.

†Open image models, served through the WiCi SDK's multimodal pipeline.

Measured over Wi-Fi 7 on a production WiCi One.

†Games run unmodified on your own machine — not remote play: WiCi One attaches as a standard virtual GPU through the WiCi driver.

Measured over Wi-Fi 7 on a production WiCi One.

†CapCut and Premiere Pro run unmodified in Driver mode — GPU effects and encoders offload over Wi-Fi. Audio memory runs on the WiCi SDK's local audio pipeline.

Measured over Wi-Fi 7 on a production WiCi One.

†Blender runs unmodified in Driver mode, scored on its standard benchmark scene.

Measured over Wi-Fi 7 on a production WiCi One.

API, SDK, and Driver modes. Call it as an endpoint, build with the SDK, or attach it as a virtual GPU.Build on WiCi - Minimum overhead, strong performance. The WiCi Protocol keeps wireless serving close to local speed.Explore the technology - Built on research. A decade of wireless and AI-systems work, published at SIGCOMM, NSDI, and OSDI.Read our research

Tech specifications #

Wireless AI Computing

CPU Intel Core Ultra 7 255H AirCompute GPU sharing #

GPU NVIDIA RTX 5060 Ti · 16GB Configurable to NVIDIA RTX 5090 · 32GB Pricing for this configuration to be announced closer to launch #

Wi-Fi Wi-Fi 7 · 4×4 MIMO · 320MHz Router-grade AirLink Multi-Link transport

Fully Flexible GPU Runtime

Driver mode WiCi Virtual GPU on macOS, Windows & Linux ZeroTrip caching & trace replay #

SDK mode Python · JavaScript · Swift Out-of-the-box inference with multimodal support #

API mode OpenAI-compatible endpoint Drop-in for existing clients

WiCi One is fully application-transparent in Driver mode, inference-optimized through the SDK, and OpenAI-compatible in API mode. Whichever path you take, your applications get full computing flexibility. Build on WiCi

Unified Hierarchical Memory

Storage NVMe 5 · 4TB TurboStream weight paging #

LLM Trillion-parameter models

Measured on WiCi One 5090 edition, close toKimi K3 2.8T 5 tok/s DeepSeek V4 Flash 284B 20 tok/s Gemma 4 26B 500+ tok/s cloud rates of 20-40 tok/s

TurboStream turns 4TB of NVMe into working VRAM: trillion-parameter models, no cloud cost required. State-of-the-art research from our systems researchers and engineers, built into WiCi One.

Read more about our technologies →

Starting at $2,599 $1,999 USDFor early signups only

Sign Up to Win WiCi One

#

Fast. Local.

Personal.

Built for low-latency interaction — voice, agents, and responsive AI experiences that feel like they're happening right next to you, because they are.

Access the best GPU over Wi-Fi instead of carrying it — every device stays cooler, lighter, and lasts longer.

The more personal AI becomes, the more sensitive the data it touches. Keep processing inside your own walls — local-first by default.

High-frequency AI gets expensive when every interaction depends on cloud inference. Own the compute once — and share it across every device and everyone at home.

#

A wireless GPU runtime

for local AI apps.

WiCi One is a programmable endpoint on the wireless compute layer — run existing apps unmodified in Driver mode, build with the WiCi SDK, or call it through the API. At every altitude, your app keeps interaction, input, and UI on the device while GPU-bound inference offloads over Wi-Fi — model caching, streaming, and pipelining keep it real-time.

Virtual GPU driver

Attach WiCi One as a virtual GPU — existing applications run unmodified, fully transparent.

LLM-optimized SDK

Build with full control — PyTorch and multimodal pipelines on wici.gpu, while I/O stays on the device.

Drop-in API with trillion-parameter model support

Keep your existing client — point it at WiCi One's OpenAI-compatible local endpoint.

#

Run AI where life happens.

Personal AI assistant

Power fast, always-available AI agents across your personal devices — anywhere in the home or office.

Shared studio compute

Equip a whole studio of lightweight laptops with one shared GPU for AI-assisted rendering, image generation, and video editing.

Voice and multimodal agents

Support real-time voice, audio, vision, and interactive AI experiences.

Smart space intelligence

Enable speakers, cameras, robots, wearables, and AR glasses to tap nearby AI compute.

Developer prototyping

Build and test local AI applications using WiCi's protocol and SDK — from prototype to product.

#

A smart speaker,

free for early adopters.

A thank-you for our first customers — and a first glimpse of the local AI devices WiCi One can power.

Programmable.

Hackable hardware, real-time voice.

Local voice.

Powered by nearby GPU, not the cloud.

App-paired.

Configure from the WiCi One App.

Open SDK.

Build local AI from day one.

In summary

                        The whole picture,at a glance.

AI compute, closer to you.

Wi-Fi, GPU, SDK — in one box.

Real-time, by design.

Your data, your walls.

Run AI in every room.

A smart speaker, free for early adopters.

Stay in the loop. #

We're building the wireless compute layer for personal AI — quiet by design.

Sign up for the newsletter — and a chance to get WiCi One free — plus a free smart speaker for early adopters.

#

Questions?

Answers.

No. AI PCs put acceleration inside a single machine. WiCi One makes compute a shared resource all your devices access wirelessly.

No — it complements cloud AI, taking on the workloads where local latency, privacy, and cost matter most.

It is a bonus programmable device included for early adopters to experience local compute-enabled AI interaction.

Developers, researchers, AI builders, early adopters, and teams interested in personal local AI compute.

We anticipate shipping starts in Q4. Availability will be announced to newsletter subscribers first — along with the details of the Pioneer plan and its top-10 free devices. Sign up to be first in line:

── more in #ai-infrastructure 4 stories · sorted by recency
── more on @wici one 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/wici-one] indexed:0 read:8min 2026-09-03 ·