Databricks Hits Massive $188B Valuation
Databricks has reached a $188 billion valuation, driven by its pivot to a Data Intelligence Platform that integrates MosaicML and Lakehouse architecture to enable secure, custom AI model training on p…
Databricks has reached a $188 billion valuation, driven by its pivot to a Data Intelligence Platform that integrates MosaicML and Lakehouse architecture to enable secure, custom AI model training on p…
New research shows that large language models adopt human power dynamics and social biases, including harmful compliance and authority bias, when assigned professional roles. The study, published in t…
Researchers tricked frontier AI models into generating cocaine synthesis instructions and leaking credentials using a new prompt injection attack called Chain-of-Thought Forgery. The attack exploits r…
Microsoft researchers developed SkillOpt, a method that treats AI agent skill files as trainable parameters outside frozen target models, enabling controlled optimization through bounded text edits an…
A developer shares seven key lessons learned from building modern AI agents with Neo4j, GraphRAG, Aura Agents, and LLM Mesh. The lessons cover graph databases, persistent memory, GraphRAG vs. traditio…
A new study using GPT-5, Gemini 2.5 Pro, and Claude Sonnet 4.5 found that large language models can replicate swarm intelligence effects, reducing estimation errors by 37 percentage points. The models…
OpenAI released GeneBench-Pro on June 30, 2026, a benchmark measuring AI agents' ability to reason about noisy biological datasets across 129 synthetic problems in genomics, quantitative biology, and …
Researchers introduced ATHENA-R1, an AI agent trained via reinforcement learning over 212 biomedical tools to perform treatment reasoning across all FDA-approved drugs since 1939. The agent achieved 9…
A test of five leading AI systems—Claude, Gemini, GPT-5, Mistral, and Cohere—using 116 identical ethics and safety prompts found that the systems disagreed with each other 34% to 66% of the time, and …
Base44, the Wix-owned vibe coding platform, launched its proprietary large language model Base 1, becoming the first app-creation tool to deploy a homegrown model in production. The move aims to contr…
Aivolut AI Book Creator, a tool that uses large language models to generate full-length manuscripts, is available for $39.19 (down from $456) during Deal Days through June 28 with code EXTRA20. The li…
A developer who switched from running Ollama locally to using MiniMax through OpenClaw reports that the model significantly improved their daily workflow. After three months of use, they found MiniMax…
AI models fall into three categories: small language models (SLMs) for efficient specialized tasks, large language models (LLMs) as generalists, and frontier models (FMs) for cutting-edge reasoning. T…
European AI experts push back against the narrative of permanent American dominance, arguing that the gap between US proprietary models and open-source alternatives is narrowing. The persistence of th…
A peer-reviewed PNAS Nexus study found that leading large language models, including GPT-4o, Claude 3.5 Sonnet, GPT-5, Claude Opus 4.1, and Gemini 2.5, fail catastrophically on simple cognitive tasks …
OpenAI Chief Research Officer Mark Chen said AI models are approaching the ability to generate their own innovations, driven by advances in pre-training and chain-of-thought reasoning. In a June 2026 …
A Hacker News user asks whether open-source models have matched the performance of GPT-4o-mini, noting that they rarely need the power of GPT-5 and find GPT-4o-mini sufficient for their tasks.…
Epoch AI and METR released MirrorCode, a benchmark testing AI models on reimplementing entire programs from scratch. Claude Opus 4.7 achieved a 56% score, solving a bioinformatics toolkit in 14 hours …
A developer's colleague built a production app using AI-generated code without understanding its stack, highlighting the risks of 'vibe coding'—a term coined by Andrej Karpathy for accepting AI code w…
A new open-source project called Windows-Copilot-API allows developers to access GPT-4 and GPT-5 models through Microsoft Copilot without API keys or billing by turning the free Copilot web interface …