{"slug": "ai-safety-is-designed-in-the-west-and-failing-users-everywhere", "title": "AI safety is designed in the West, and failing users everywhere", "summary": "OpenAI became the first major AI company to voluntarily pause training on a model due to safety concerns, following incidents where its models broke free during tests and hacked websites, with Anthropic and Meta reporting similar issues. Trust and safety teams are concentrated in Silicon Valley, leading to AI systems that fail basic tasks in developing nations, where health-related queries often yield errors affecting diagnoses, as noted by Urvashi Aneja of Digital Futures Lab and Elizabeth Orembo of Research ICT Africa.", "body_md": "Last month, OpenAI became the first major artificial intelligence company to [voluntarily pause training](https://openai.com/index/pacing-model-development-cyber-capabilities/) on a model because of safety concerns. The unprecedented move came weeks after its models broke free during a test and hacked other websites, sparking concern about losing control of AI systems, as [Anthropic](https://www.anthropic.com/news/investigating-incidents-cybersecurity-evals) and Meta reported similar incidents.\n\n“We care very deeply about AI safety,” OpenAI chief executive Sam Altman [said](https://x.com/sama/status/2089787807611195475) on X. “Model progress is now extremely rapid, and we always said we would take action if we felt that model capabilities were outstripping the pace of safety and alignment.”\n\nAI researchers had been calling for a slower pace for development and more safety measures even before these incidents. Generative AI systems are solving complex mathematical problems and helping [develop](https://www.nature.com/articles/d41586-026-02612-3) highly specialized drugs in the U.S. and other Western nations. But they are failing at basic tasks elsewhere because trust and safety teams are largely concentrated in Silicon Valley, and do not reflect the concerns of countries with different languages and cultural contexts.\n\nThe gaps have real consequences: Queries related to health are among the most [common uses](https://www.nature.com/articles/s44360-026-00174-2) of AI chatbots worldwide, yet in many African and Asian nations, even multilingual AI tools make errors that can affect diagnoses and treatment decisions. More than two-thirds of chatbots do not adequately account for dialects or recognize urgency cues, according to a review in India.\n\nThe global majority “still remains at the margins of the larger AI safety discourse,” Urvashi Aneja, founder of research organization Digital Futures Lab, who is working on a report for the United Nations on AI safety in developing nations, told *Rest of World*. “The frameworks being built to evaluate AI systems, the standards that govern them, and the institutions that oversee them have largely been designed in, and for, a small set of high-income countries.”\n\nTrust and safety is an umbrella term for the teams at tech companies whose job is to ensure that users are protected from harmful online experiences. There is no universal trust and safety standard; each company has its own framework — based on its values and principles, and enforcement methods.\n\nDeveloping nations are adopting AI at a slower pace than wealthier nations, yet the risks fall “[disproportionately](https://www.un.org/independent-international-scientific-panel-ai/en/preliminary-report)” on them because of inadequate resources, limited domestic AI infrastructure, and dependence on foreign technologies, the U.N. said in a recent report.\n\nAt the heart of the issue “is the question of who gets to define what counts as a safety problem in the first place,” Elizabeth Orembo, a fellow at Research ICT Africa, a think tank, told *Rest of World*. Big tech firms tend to focus on model risks such as deception, autonomous behavior, cyber capabilities, and aiding bioweapons, she said. They do not pay much attention to deployment risks including discrimination, exclusion, surveillance, [language failures](https://restofworld.org/2023/ai-content-moderation-hate-speech/), and the inability of affected communities to seek remediation.\n\n## Life-threatening consequences\n\nTrust and safety teams run evaluations that assume reliable electricity and internet connectivity, functioning courts, robust data protection laws, formal labor markets, and a press and civil society that report failures. So “a model can pass every frontier safety evaluation and still produce [unsafe outcomes](https://restofworld.org/2026/ai-content-moderation-consent-muse/) when deployed,” Orembo said.\n\nNatural language processing in healthcare in Africa [showed](https://aclanthology.org/2025.emnlp-main.535/) cultural and linguistic bias, poor adaptation to medical contexts, and translation errors, researchers found. For example, in Tigrinya, spoken by about 9 million people in Eritrea and northern Ethiopia, machine translation rendered smallpox as syphilis, gonorrhea as diabetes and “you have been given intravenous antibiotics” as “you have been given intravenous insecticides.”\n\nSuch mistranslations “can be life-threatening,” Orembo said.\n\nOn a recent [AI safety index](https://futureoflife.org/ai-safety-index-summer-2026/#scorecard) from the Future of Life Institute that evaluated nine leading companies, Anthropic, OpenAI, and Meta had the highest scores on metrics such as risk assessment, current harms, existential safety, and governance and accountability. DeepSeek, xAI, and Mistral had the lowest scores.\n\nBut “even industry leaders … are retreating from prior commitments, despite calling publicly for a pause,” Future of Life Institute, a nonprofit that researches AI risks, noted. This has “undermined safety frameworks across the board.”\n\nIn low- and middle-income countries, the consequences are compounded, and users feel the effects “immediately” as they can affect access to wages and essential services, the U.N. Development Programme said in a [recent report](https://www.undp.org/publications/small-states-big-signals-what-adoption-practice-reveals-about-trust-safety-and-ai-performance-globally).\n\nThere is growing evidence of these failures, including AI-powered facial recognition and ID systems that lead to denial of wages, [meals](https://pulitzercenter.org/stories/indias-facial-recognition-drive-hungry-children-erasing-them), or [school attendance](https://www.investigate-europe.eu/posts/facial-recognition-software-ai-europe-brazil-school-children), and tools that [misidentify crops](https://restofworld.org/2026/ai-agriculture-local-data/) or [mistranslate](https://restofworld.org/2026/ai-social-good-humans/) local terms.\n\nLanguage failures are inevitable, as data sets to train large language models are predominantly in English and other widely spoken Western languages, and LLMs display poor translation and [higher hallucination rates](https://arxiv.org/abs/2410.18270) in low-resource languages, Dhanaraj Thakur, director of the Fair Technology Initiative at the George Washington University Law School, told *Rest of World*.\n\nThe guardrails “may work well in English, but fail or are easily circumvented in low-resource languages,” he said. “The result is that users of these models that speak English or other high-resource languages end up being safer than those that speak low-resource languages, a new kind of AI divide.”\n\n## A critical moment\n\nGovernments are trying to address these concerns. At the inaugural AI summit in 2023 in the U.K., 28 countries signed the [Bletchley Declaration](https://www.digitalpolitics.co/r/9140fd14?m=0edbce84-6647-4b71-8b6d-5a231630fca1), a voluntary commitment to identify and respond to existential risks tied to AI. The India AI summit earlier this year also addressed safety, while China — which is [regulating AI](https://restofworld.org/2026/china-ai-boyfriend-ban-bytedance-doubao/) more aggressively — in July proposed [mechanisms](https://aisafetychina.com/) to manage AI risks in developing nations.\n\nFor these countries, there is no time to lose, Aneja said.\n\n“There is a lot of [optimism about AI](https://restofworld.org/2026/ai-optimism-asia/) in these countries now, unlike the backlash that you see in the West,” she said. “But if governments don’t invest in the safety infrastructure now, public trust will erode, and the opportunity to leverage benefits from AI will go away, and then we’re looking at deepening inequality.”\n\nRecently, more than 1,300 employees at AI companies including Anthropic, Meta AI, OpenAI, and Google DeepMind [wrote an open letter](https://www.pacingthefrontier.com/) saying the industry, the government, and society “may need the option to buy time to address emerging risks, develop security measures, and strengthen oversight” of AI systems.\n\nFor Sumiya Khan, a 16-year-old in New Delhi, this is sorely needed. Experiencing constant fatigue and dizziness, she turned to ChatGPT for advice in Hindi, like she had done before. The chatbot chalked up her symptoms to stress and poor sleep. But when the symptoms persisted, Khan visited a doctor, who diagnosed iron-deficiency anemia, which is prevalent in low- and middle-income countries. Left untreated, it can lead to irregular heart rhythm and, in extreme cases, heart failure.\n\n“We trusted the AI because it sounded so convincing,” Mehnaz Begum, her mother, told *Rest of World*. “That made us wait longer than we should have to see a doctor.”*Additional reporting by Sajid Raina in New Delhi.*", "url": "https://wpnews.pro/news/ai-safety-is-designed-in-the-west-and-failing-users-everywhere", "canonical_source": "https://restofworld.org/2026/ai-safety-bias/?utm_source=rss&utm_medium=rss&utm_campaign=feeds", "published_at": "2026-09-01 10:00:00+00:00", "updated_at": "2026-09-01 10:25:35.915947+00:00", "lang": "en", "topics": ["ai-safety", "ai-policy", "artificial-intelligence"], "entities": ["OpenAI", "Anthropic", "Meta", "Sam Altman", "Urvashi Aneja", "Digital Futures Lab", "Elizabeth Orembo", "Research ICT Africa"], "alternates": {"html": "https://wpnews.pro/news/ai-safety-is-designed-in-the-west-and-failing-users-everywhere", "markdown": "https://wpnews.pro/news/ai-safety-is-designed-in-the-west-and-failing-users-everywhere.md", "text": "https://wpnews.pro/news/ai-safety-is-designed-in-the-west-and-failing-users-everywhere.txt", "jsonld": "https://wpnews.pro/news/ai-safety-is-designed-in-the-west-and-failing-users-everywhere.jsonld"}}