Do your capabilities homework
A technical AI safety researcher argues that safety-focused researchers should engage with capabilities research, highlighting On-Policy Self-Distillation (OPSD) as a promising alternative to GRPO for…
A technical AI safety researcher argues that safety-focused researchers should engage with capabilities research, highlighting On-Policy Self-Distillation (OPSD) as a promising alternative to GRPO for…
A leak at DeepSeek exposed proprietary engineering pipelines, including data curation, RL alignment recipes, and system configurations, rather than model weights. The incident highlights that the true…
DeepSeek, the Chinese AI lab founded by Liang Wenfeng, is raising roughly $10 billion at a $45 billion valuation from China's state AI fund and High-Flyer, with the founder stating the company priorit…
Researchers discovered that reasoning models like o1 and R1 often overthink, computing answers at around 30% of their chain-of-thought but continuing for the remaining 70%. The termination decision is…
Chinese AI startup DeepSeek has begun developing its own AI chips focused on inference workloads to reduce reliance on NVIDIA and cut costs, sources told Reuters. The project is in early stages and ha…
US auto safety regulators have opened a preliminary investigation into Rivian Automotive Inc. covering about 115,000 R1 electric pickups and SUVs over potential failure of a rear toe link suspension p…
Chinese AI startup DeepSeek slashed pricing for its flagship V4-Pro model by 75%, reducing costs to as low as $0.003625 per million tokens, effective immediately through a permanent efficiency-driven …