cd/entity/Weights & Biases· home entities Weights & Biases
grep -l @weights & biases /news/*.json | wc -l → 25

Weights & Biases

mentions 25 type Person page 1/2 feed RSS

// recent coverage 25 mentions

13:47
2026-08-26
boydkane.com
large-language-models

Common LLM failure modes

Anthropic's Claude Opus 5 and Claude Fable 5 exhibit distinct failure modes, including Opus 5's constant praise, imprecise commentary, and a tendency to end statements with negations, while Fable 5 in…

15:45
2026-08-24
promptcube3.com
artificial-intelligence

The 8 Best AI Discussion Groups to Join in 2025

Hugging Face hosts one of the most active AI communities with over 10,000 open-source models shared as of 2025, focusing on natural language processing, computer vision, and audio AI development. Redd…

22:27
2026-08-22
twitter.com
large-language-models

Follow live the open training of a 535B (23B activated) LLM

Marin 535B-A23B, a 535-billion-parameter mixture-of-experts large language model with 23 billion activated parameters, began open training this week on 18.75 trillion tokens across 11 NVIDIA GB200 NVL…

10:25
2026-08-21
wandb.ai
large-language-models

Watch a 535B parameter (23B active) LLM get trained live

The MARIN community is publicly training a 535B-parameter mixture-of-experts large language model with 23B active parameters on 18 trillion tokens, streaming the process live via Weights & Biases. The…

07:56
2026-08-14
dev.to
mlops

The real LLMOps risk isn't the model. It's shadow AI.

CNCF's Daniel Bryant argues that LLMOps should be integrated into existing platform engineering rather than treated as a separate stack, warning that shadow AI pipelines built outside governance pose …

21:32
2026-08-13
promptcube3.com
artificial-intelligence

Reproducing 2

A new analysis of 2,200 AI research reproducibility cases finds that most papers fail to replicate due to hardware variance, hyper-parameter sensitivity, and dependency issues, with the missing link o…

18:42
2026-08-09
discuss.huggingface.co
artificial-intelligence

How do you usually experiment with RAG pipelines?

A developer seeking advice on experimenting with retrieval-augmented generation (RAG) pipelines asked the community how they compare configurations of retrievers, chunking strategies, embeddings, rera…

22:07
2026-07-31
byteiota.com
artificial-intelligence

Nscale Buys Anyscale for $1.65B: What Happens to Ray

Nscale, the British AI cloud company valued at $14.6 billion, agreed on July 30 to acquire Anyscale, the company behind the Ray distributed computing framework, for $1.65 billion. The acquisition, the…

18:31
2026-07-16
dibi8.com
artificial-intelligence

Data Science

Google Research released TimesFM 2.5, a decoder-only foundation model for time series forecasting, with 10,000 GitHub stars and support for installation, fine-tuning, and real-world applications. The …

02:58
2026-07-16
discuss.huggingface.co
machine-learning

My latest ablation run: integrating Engram onto two backbones

A 200-step, ~1.7B-parameter ablation run in OLMo-core comparing Engram on a standard attention Transformer versus a 3 GDN-layer + 1 attention-layer hybrid found that Transformer + Engram reached sligh…

12:00
2026-06-29
letsdatascience.com
artificial-intelligence

CoreWeave launches ARIA agent for W&B research

CoreWeave launched ARIA, an AI agent built with W&B Weave that automates experiment analysis for machine learning projects, entering preview to analyze thousands of runs and metrics in minutes. The ag…

16:14
2026-06-25
discuss.huggingface.co
large-language-models

OLMo-core + Engram graft: 2B/600M-A debug comparison

A researcher ran a 200-step debug comparison between a base OLMo3 600M model and a DeepSeek-style Engram memory graft variant, finding the graft stable and showing improved early learning behavior. Th…

20:16
2026-06-24
discuss.huggingface.co
ai-safety

AI safe guards and optimization feedback

A guide recommends tools like LangChain, Guardrails AI, and OpenAI Moderation API to add safety guards and optimization feedback to AI systems. It also suggests building alert systems, logging activit…

19:13
2026-06-21
discuss.huggingface.co
large-language-models

OLMo-core + Engram graft: small-scale debug comparison

A debug comparison between a base OLMo3 600M model and an Engram memory variant showed the grafted model achieved lower training and evaluation cross-entropy loss and faster gradient norm stabilizatio…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics