Language Preservation Efforts Get an AI Boost
Computer science graduate student Ivory Yang at Dartmouth's Guarini School, with collaborators Weicheng Ma and Assistant Professor of Computer Science Soroush Vosoughi, built an AI framework called Nü…
Computer science graduate student Ivory Yang at Dartmouth's Guarini School, with collaborators Weicheng Ma and Assistant Professor of Computer Science Soroush Vosoughi, built an AI framework called Nü…
OpenAI's GPT-4, launched in March 2023, had an 8,192-token context window with limited access to a 32,768-token variant, not the 128,000-token window claimed. The 128K context window was introduced wi…
A developer detailed how to build an autonomous customer onboarding agent by combining CrewAI's multi-agent orchestration with n8n's workflow engine. The system uses three agents—intake, validator, an…
A developer explains that a language model's context window acts as its working memory, limiting how many tokens it can process at once. The post details how the key-value cache enables efficient gene…
Advanced AI models from the GPT, Claude, and Llama families spontaneously conform to the majority opinion, with GPT-4 Turbo and Claude 3 Opus achieving 100 percent consensus in experiments published i…
A new study in Science Advances found that groups of up to 1,000 AI agents from the Claude, GPT, and Llama families spontaneously reached consensus without any instruction to cooperate, following the …
A developer team improved the coding performance of 15 large language models in one afternoon by changing only the edit tool in their harness, not the models themselves. The team found that edit forma…
A two-week benchmark of ten AI translation tools found Claude 3.5 Sonnet, GPT-4 Turbo, and DeepL Pro best for developer use cases, with Claude achieving 97% accuracy on technical content versus 94% fo…
OpenAI slashed prices on its smaller GPT-5.6 models on July 30, cutting the entry-level Luna model by 80% and the mid-tier Terra by 20%, as businesses question AI spending and cheaper Chinese competit…
A Communications Psychology study published April 28 analyzed 3,366 dream and waking-experience reports from 207 adults, plus 351 dream reports from 80 adults during Italy's first COVID-19 lockdown, f…
A systematic evaluation of five large language models for technical market analysis finds that GPT-4 Turbo achieves the highest annualized return and Sharpe ratio among general-purpose models, while d…
A new study published in PLOS One found that voters rated AI-generated political debate responses as more authentic and relevant than real answers from politicians. Researchers used GPT-4 Turbo to mim…
AI chatbots impersonating 112 UK public figures produced responses rated as more authentic, coherent, and relevant than the real individuals' statements, according to a study published in PLOS One. Re…
Between January and June 2026, OpenAI, Anthropic, and Google made 14 combined pricing changes across their LLM model lineups, with some prices dropping, others rising, and models being deprecated and …
Zalando researchers published a framework using multimodal LLMs to automate product retrieval evaluation, achieving human-level accuracy at up to 1,000 times lower cost and reducing evaluation time fr…
Microsoft announced at Build 2026 that GitHub Copilot will replace GPT-4 Turbo with its own Polaris model starting August 2026, a mixture-of-experts architecture tuned per programming language. Enterp…