UniServe: Serving FastH3 at Its Fastest
UniServe, a serving engine for FastH3 8-Step text-to-video-with-audio generation from the hao-ai-lab, delivers lower median end-to-end latency and 20–46% higher throughput than FastVideo, vLLM-Omni an…
UniServe, a serving engine for FastH3 8-Step text-to-video-with-audio generation from the hao-ai-lab, delivers lower median end-to-end latency and 20–46% higher throughput than FastVideo, vLLM-Omni an…
Reactor and the Amazon Neuron Science team developed a kernel-centric optimization strategy that enables real-time, low-latency video generation on AWS Trainium chips, according to Reactor co-founder …
Reactor has emerged from stealth with $59 million to build a platform for real-time interactive video powered by AI world models, offering developers sub-50ms latency and pay-as-you-go pricing. The pl…
NVIDIA Research's Efficient AI Team and Singapore Lab released Sol-H3, an inference stack that generated a five-second, audio-synced MiniMax H3 video clip in 1.653 seconds on eight NVIDIA B300 GPUs, f…
LingBot-World 2.0 (LingBot-World-Infinity) introduces unbounded interaction horizons, rapid 60fps 720p video generation, diverse interactive elements, and an agentic harness for world modeling. The re…
San Francisco-based Reactor, co-founded by former Apple Vision Pro technical leads Alberto Taiuti and Bryce Schmidtchen, has raised $59 million in Series A funding led by Lightspeed Venture Partners, …