cd /news/generative-engine-optimization/hinglish-changes-which-brands-ai-rec… · home › topics › generative-engine-optimization › article
[ARTICLE · art-143091] src=dev.to ↗ pub= topic=generative-engine-optimization verified=true sentiment=· neutral

Hinglish Changes Which Brands AI Recommends

A study of 480 AI answers collected in India found that AI engines recommend different brands when the same buying question is asked in Hinglish versus English, with Perplexity's brand list shifting 23.4 points beyond normal rerun noise, Gemini's 7.8 points and ChatGPT's only 2.2 points. The research, based on 10 skincare and fashion buying prompts run 8 times per language across the three engines on 14 August 2026, also found Gemini showed at least one source link under 90.0% of English answers but only 41.3% of Hinglish answers. Raw data was published in the IndicGEO repository on GitHub.

by read10 min views1 publishedOct 1, 2026

Originally published on depra.ai on 15 Sept 2026. Yes, AI engines recommend different brands when the same buying question is asked in Hinglish, and how much depends on the engine. In our study of 480 AI answers collected from India, Perplexity's brand list changed 23.4 points beyond normal rerun noise, Gemini's changed 7.8 points and ChatGPT's changed only 2.2 points.

Hinglish here means Hindi and English mixed and typed in English letters, like "konsa face wash kharidna chahiye". Rerun noise is how much an engine's brand list changes when you simply ask the identical question again. The second headline is about sources: Gemini showed at least one source link under 90.0% of English answers and under only 41.3% of Hinglish answers.

This post is the plain readout for brand and marketing teams. The Hinglish vs English AI shopping study page is the record of every table, interval and test, and the raw data is public in the IndicGEO repository on GitHub.

We wrote 10 buying questions: 5 about skincare and 5 about fashion. These are the prompts, meaning the exact text typed into each AI engine, written the way a buyer would type it. None of them named a brand. Each question had an English version and a Hinglish version with the same meaning, the same budget and the same India anchor. An independent reviewer checked every pair before analysis, and 3 fashion pairs were rewritten after failing that review.

One pair, word for word:

We asked each question 8 times in each language on three engines:

Every request was located in India and started fresh, with no chat history and no logged-in account. All 480 answers were collected on 14 August 2026 inside one 57-minute window, between 12:25 and 13:22 UTC. The order was shuffled, so neither language got a different time of day.

Brands were counted by matching each answer against a brand list built from all 480 answers and checked by hand. Some brand names are also everyday Hindi words. "Bata" is a shoe brand and also the Hinglish verb for "tell". Those names counted only when capitalised as a brand, so Hinglish answers were not over-counted.

An AI engine rarely gives the same answer twice. English and Hinglish answers will always differ a little, even if language has no effect. So we measured two things per engine:

The overlap score is the number of brands two answers share, divided by all brands either answer named. The language effect is the gap between the two overlaps. A gap of zero means language changes nothing beyond chance.

This follows a point made in Don't Measure Once, a 2026 paper by Schulte, Bleeker and Kaufmann: AI answers vary across runs, prompts and time, so one answer is an unreliable measure of a brand's visibility.

Perplexity, by a wide margin.

Engine Brand overlap between reruns Brand overlap, English vs Hinglish Language effect 95% range
ChatGPT 54.7% 52.5% 2.2 points 0.3 to 4.6
Gemini 51.2% 43.3% 7.8 points 2.6 to 13.9
Perplexity 68.3% 44.9% 23.4 points 12.9 to 35.1

Basis: 480 answers, 80 per engine and language (10 prompts x 8 runs), India-located, 14 August 2026. The 95% range is the span the true language effect most likely falls in, given that the study used only 10 prompts. It comes from resampling the 10 prompts 10,000 times.

Perplexity is the most consistent engine on reruns, at 68.3% brand overlap, and also the most sensitive to language. When the most self-consistent engine drops to 44.9% overlap across languages, chance is a poor explanation. Its choice of which brand to name first also shifted: the first brand matched in 69.5% of rerun pairs and 49.8% of English and Hinglish pairs.

Gemini sits in the middle at 7.8 points. Its first-named brand matched in 63.2% of rerun pairs and 42.2% of cross-language pairs.

ChatGPT barely moves. The study fixed its test before looking at the data: a one-sided test, which only asks whether English and Hinglish answers overlap less than reruns do. The 2.2-point gap passed it (p = .046, meaning a gap this size would show up by chance less than 5 times in 100 if language made no difference). It is still the weakest result in the study. The whole gap came from the skincare prompts and 4 of the 10 prompts went the other way. A two-sided test, which also allows for Hinglish overlapping more, would not pass (p = .092). The fair reading is that Hinglish ChatGPT answers named close to the same brands as English ones for these questions.

The sources behind the answers follow the same order. The overlap in cited websites between languages fell 3.1 points on ChatGPT, 22.3 points on Gemini and 32.4 points on Perplexity, each measured against the same rerun baseline. If you track several engines, why AI engines disagree about your brand covers what drives those differences.

The engine-level gaps are averages. For single brands, the swings on Perplexity were large. Each figure below is the share of one engine's 80 answers in one language that named the brand.

Perplexity

| Brand | English answers | Hinglish answers | Change |

|---|---|---|---|
| The Derma Co | 2.5% (2 of 80) | 20.0% (16 of 80) | +17.5 points | 
| Deconstruct | 3.8% (3 of 80) | 16.3% (13 of 80) | +12.5 points | 
| Minimalist | 22.5% (18 of 80) | 33.8% (27 of 80) | +11.3 points | 
| Plum | 28.7% (23 of 80) | 15.0% (12 of 80) | -13.7 points | 
| La Roche-Posay | 13.8% (11 of 80) | 1.3% (1 of 80) | -12.5 points | 
| Taneira | 15.0% (12 of 80) | 0.0% (0 of 80) | -15.0 points | 

Gemini

| Brand | English answers | Hinglish answers | Change |

|---|---|---|---|
| Dot & Key | 10.0% (8 of 80) | 21.3% (17 of 80) | +11.3 points | 
| Allen Solly | 5.0% (4 of 80) | 15.0% (12 of 80) | +10.0 points | 
| Minimalist | 45.0% (36 of 80) | 36.3% (29 of 80) | -8.7 points | 
| CeraVe | 10.0% (8 of 80) | 0.0% (0 of 80) | -10.0 points | 

On ChatGPT, no brand moved by more than 10 points. The largest change was Deconstruct, from 10.0% to 18.8% of answers.

One pattern stood out on Perplexity. Brands that lost ground in Hinglish leaned international or premium, such as La Roche-Posay and Taneira, a premium Tata brand. Brands that gained leaned toward mass-market Indian D2C brands, such as The Derma Co, Deconstruct and Minimalist. D2C, short for direct-to-consumer, describes brands built to sell online straight to shoppers. The share of named brands that are Indian rose in Hinglish on Gemini, from 75.6% to 82.0%, and on Perplexity, from 63.7% to 66.3%. On ChatGPT it stayed flat, 74.4% to 74.7%. We observed this tendency and did not test it as a hypothesis.

Treat these as observations about AI shopping recommendations in India, from 10 questions on one day. They are not endorsements and they say nothing about which product is better. The same brands could score very differently on another set of questions.

A citation is a source link an engine shows with its answer. Gemini attached at least one to 72 of 80 English answers (90.0%) and to only 33 of 80 Hinglish answers (41.3%). Its average number of links per answer fell from 9.4 to 5.2. Its Hinglish answers were also 33% shorter, averaging 3,099 characters against 4,609 in English.

ChatGPT and Perplexity attached sources to every answer in both languages, 80 of 80 each. What changed on those two engines was which websites they cited.

The study did not test why Gemini behaves this way. What it recorded: 47 of 80 Gemini Hinglish answers showed no source links at all, against 8 of 80 English answers.

Almost always.

Engine Asked in English Asked in Hinglish
ChatGPT 80 of 80 in English 80 of 80 in Hinglish
Gemini 80 of 80 in English 71 in Hinglish, 8 in Devanagari, 1 in English
Perplexity 80 of 80 in English 80 of 80 in Hinglish

Devanagari is the script Hindi is usually printed in. Gemini used it for 8 of its 80 Hinglish answers, a script the person asking did not type. The language labels were checked against human judgement on a sample of 30 answers and agreed on all 30.

So a Hinglish buyer reads a Hinglish answer. On Perplexity and Gemini that answer also links to a different set of websites and names a different set of brands than the English one.

If you track English prompts only, you are measuring what English buyers see and leaving out Hinglish AI search. On Perplexity and Gemini, Hinglish buyers see a different brand list. Five steps follow from the data: Depra runs this as a product. Hinglish AI visibility tracking is on every plan: Depra suggests Hinglish versions of your buying questions when India is one of your markets, detects the language of every answer, and your report shows an English vs Hinglish table. ChatGPT, Gemini and Google AI Overviews, the AI-written summary Google shows above some search results, are checked daily. The Perplexity visibility tracker runs weekly. Plans start at ₹1,999 a month plus 18% GST, listed on the pricing page. For the basics of reading these numbers, see how to track brand mentions in AI answers.

To see your own brand's English and Hinglish numbers across ChatGPT, Gemini, Perplexity and AI Overviews, start 7 days free on any plan, no card. Your first scan starts at signup.

Yes, and the size of the change depends on the engine. In a study of 480 AI answers collected from India on 14 August 2026, the brand list changed beyond normal rerun noise by 23.4 points on Perplexity, 7.8 points on Gemini and 2.2 points on ChatGPT.

Barely. ChatGPT's brand list changed 2.2 points more between English and Hinglish than between two runs of the same question. That gap came only from the skincare prompts and is the weakest finding in the study, so for these buying questions ChatGPT gave close to the same brands in both languages.

Perplexity. Its brand list changed 23.4 points beyond rerun noise, against 7.8 for Gemini and 2.2 for ChatGPT. The 95% range for Perplexity was 12.9 to 35.1 points, which is the span the true effect most likely falls in given only 10 prompts. On Perplexity, The Derma Co appeared in 2.5% of English answers and 20.0% of Hinglish answers.

The study measured the drop and did not test the cause. Gemini showed at least one source link in 72 of 80 English answers (90.0%) and 33 of 80 Hinglish answers (41.3%). Its Hinglish answers also carried fewer links on average, 5.2 against 9.4, and were 33% shorter.

Yes. All 480 raw answers, the prompt pairs, the brand list, the analysis code and the method, which was written and fixed before any analysis, are public in the IndicGEO repository on GitHub (github.com/ifham001/IndicGEO), with code under MIT and data under CC BY 4.0. Running the analysis script recomputes every table in this post from the raw answers.

── more in #generative-engine-optimization 4 stories · sorted by recency
── more on @perplexity 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/hinglish-changes-whi…] indexed:0 read:10min 2026-10-01 · —