cd/entity/A100· home› entities› A100
grep -l @a100 /news/*.json | wc -l → 67

A100

mentions 67 type Organization page 4/4 feed RSS

// recent coverage 67 mentions

14:52
2026-05-27
dev.to
generative-ai

Semantic caching the VLM step in our product-photo pipeline

Photoroom reduced its vision-language model inference costs by approximately 62% within three weeks by deploying Bifrost as a semantic caching layer in front of the VLM step of its product-photo diffu…

05:37
2026-05-27
dev.to
machine-learning

The bf16 grad accumulator that killed our SDXL LoRA training

Photoroom's SDXL LoRA fine-tuning for a product photography model silently corrupted its adapter weights over six days due to a bf16 gradient accumulation issue. The custom training loop, forked from …

05:47
2026-05-26
dev.to
artificial-intelligence

AI Metrics Decoded: From Parameters to TOPS

A developer explains that understanding seven core AI metrics—parameters, tokens, FLOPS, TOPS, and FLOPs—is essential for avoiding costly deployment mistakes, such as choosing a 70B-parameter model th…

19:33
2026-05-18
rosmine.ai
artificial-intelligence

Was my $48K GPU server worth it?

The author quit their FAANG job in 2024 to become an independent researcher and built a $48K GPU server called "grumbl" with six RTX 6000 Ada GPUs. They chose these GPUs over A100s and H100s based on …

00:00
2023-07-20
cursor.com
large-language-models

Inference Characteristics of Llama-2

The article analyzes the cost and latency trade-offs of serving Llama-2-70B compared to GPT-3.5-turbo, concluding that Llama-2 is over 3x cheaper for prompt tokens but more expensive for completion to…

← prev page 4 / 4
// co-occurs with top 8 entities
// topics top 6 topics