cd/sources/apple-ml-research· home sources Apple ML Research
cat /sources/apple-ml-research.feed | wc -l → 42

Apple ML Research

articles 42 domain machinelearning.apple.com → page 2/3 feed RSS
00:00
2026-07-06
machinelearning.apple.com
large-language-models

Scaling Properties of Continuous Diffusion Spoken Language Models

Researchers at Google and Apple found that continuous diffusion spoken language models (SLMs) exhibit scaling laws similar to autoregressive models, with validation loss and phoneme Jensen-Shannon div…

00:00
2026-07-06
machinelearning.apple.com
ai-safety

Understanding Annotator Safety Policy with Interpretability

Researchers at Apple introduced Annotator Policy Models (APMs), interpretable models that learn annotators' internal safety policies from labeling behavior alone, enabling diagnosis of disagreement so…

00:00
2026-07-02
machinelearning.apple.com
large-language-models

Multi-Agent Teams Hold Experts Back

A study by researchers including Aneesh Pappu and James Zou found that self-organizing multi-agent LLM teams fail to match the performance of their best individual expert, with losses up to 41.1% on M…

00:00
2026-05-11
machinelearning.apple.com
artificial-intelligence

BalCapRL: A Balanced Framework for RL-Based MLLM Image Captioning

Researchers at BalCapRL introduced a balanced reinforcement learning framework for multimodal large language model image captioning that jointly optimizes utility-aware correctness, reference coverage…

00:00
2026-05-08
machinelearning.apple.com
artificial-intelligence

Apple Workshop on Privacy-Preserving Machine Learning & AI 2026

Apple hosted a two-day Workshop on Privacy-Preserving Machine Learning & AI in early 2026, bringing together its researchers and the broader academic community to discuss advances in private learning,…

00:00
2026-05-08
machinelearning.apple.com
machine-learning

RVPO: Risk-Sensitive Alignment via Variance Regularization

Researchers at Duke University introduced Reward-Variance Policy Optimization (RVPO), a risk-sensitive alignment method that penalizes inter-reward variance to prevent language models from neglecting …

← prev page 2 / 3 next →