cd/entity/PyTorch· home› entities› PyTorch
grep -l @pytorch /news/*.json | wc -l → 615

PyTorch

mentions 615 type Organization page 2/31 feed RSS

// recent coverage 615 mentions

00:00
2026-09-30
blog.janestreet.com
machine-learning

Trading off compute for memory with activation checkpointing

Jane Street ML Infra intern Arsh Koneru built an activation checkpointing scheme that, given a model spec and a memory budget, selects a set of nodes to drop from the save set to minimize extra comput…

09:49
2026-09-28
dy.github.io
computer-vision

Show HN: GPU-font – match fonts with WebGPU

Developer dy released gpu-font, an open-source browser tool that identifies the closest font families to any line of text using WebGPU, published on Hacker News as a Show HN. The tool splits a crop in…

08:53
2026-09-28
marimo.io
developer-tools

Pixi adds non-Python deps to Python notebooks

Marimo's notebook sandboxes now support Pixi, letting standalone notebooks record Conda and PyPI requirements alongside their code, the marimo team announced. The integration, built with the Pixi team…

23:54
2026-09-27
github.com
artificial-intelligence

Show HN: DeepFace is now running on PyTorch alongside TensorFlow

DeepFace v0.0.101 now runs on PyTorch alongside TensorFlow, the library's maintainers announced in a Show HN post. Users can install the framework-specific dependency with `pip install deepface[tensor…

22:31
2026-09-27
runtimewire.com
large-language-models

Fireworks says its Kimi K3 variant cuts reasoning tokens by 40%

Fireworks AI says its Ember-1 model, built from Moonshot AI's Kimi K3, produces comparable answers with about 40% fewer tokens, according to a September 27 X post by co-founder Dmytro Dzhulgakov promo…

02:27
2026-09-27
dev.to
machine-learning

FlashAttention-2 from PyTorch to Triton

A developer implemented the FlashAttention-2 forward pass twice — first as a deliberately slow PyTorch reference with a Python double loop over tiles, then as a Triton kernel — to teach kernel optimiz…

00:12
2026-09-27
discuss.huggingface.co
ai-infrastructure

Nvidia RTX 5060

Nvidia's RTX 5060 is usable for current PyTorch and image-generation stacks, but older packaged environments such as Fooocus's official portable still default to torch==2.1.0 and xformers==0.0.23, cau…

17:47
2026-09-25
dev.to
ai-infrastructure

Where to rent or buy h100 and b200 gpus for ai startups

An engineering analysis of GPU infrastructure economics finds that owning an 8x H100 SXM5-class server carries roughly $310,000 in upfront capital plus about $37,500 in electricity and $4,300 per year…

16:30
2026-09-25
leimao.github.io
ai-infrastructure

CUDA Device Max Connections

NVIDIA's CUDA_DEVICE_MAX_CONNECTIONS environment variable defaults to 8 concurrent hardware connections to the GPU, capping real concurrency regardless of how many CUDA streams a developer creates, ac…

13:57
2026-09-25
discuss.huggingface.co
ai-infrastructure

FitCheck — Estimate LLM Training and Serving VRAM Before You Run

Developer Anassbzdd released FitCheck, an open-source tool that estimates peak VRAM for LoRA, QLoRA, and full LLM fine-tuning and for serving based on model weights and KV cache, reading a model's Hug…

12:48
2026-09-25
github.com
large-language-models

Show HN: JevPertus – Jev-style option scoring on Apertus

A developer released JevPertus, an open-source implementation that adds LoRA adapters and a small pointer head to the swiss-ai/Apertus-v1.5-8B backbone to assign probabilities to a question's answer o…

← prev page 2 / 31 next →
// co-occurs with top 8 entities
// topics top 6 topics