AI Mentions Negative News About Your Brand: How Fast It Enters, How Long It Stays

by

·

AI Mentions Negative News About Your Brand: How Fast It Enters, How Long It Stays

When AI mentions negative news about your brand, the clock behaves nothing like a press cycle. Across 34 tracked incidents, bad news entered AI answers at a median of 3 days after first mainstream coverage — then stayed for a median of 41 days after the incident was verifiably resolved.

The gap between "we fixed it" and "the answer says we fixed it" is where most crisis plans quietly fail. Comms teams measure the news cycle, which peaks and decays in about a week. The AI answer runs on a different curve entirely, and nobody on the crisis call owns it.

This piece is built on daily answer-level tracking, not opinion. It covers how fast each surface picks up an incident, which sources carry it in, how long the tail runs by incident type, which counter-evidence measurably shortened that tail, and what produced nothing at all.

What happens when AI mentions negative news about a brand?

An AI assistant surfaces an outage, lawsuit, layoff or recall inside an answer about your company — usually as a qualifying clause attached to an otherwise neutral description. It is not a ranking penalty. It is a sentence in a paragraph that a buyer, candidate or investor reads instead of your website.

The mechanics differ from search. A news story ranks and then decays; an AI answer summarizes whatever the model retrieves at query time, and that retrieval set updates on its own schedule. So the incident can enter late, persist long, and reappear after you thought it was gone.

Three prompt families matter, because they behave differently:

  • Incident prompts — "did X have a data breach?" Highest mention rate, fastest to clear.
  • Trust prompts — "is X reliable?", "is X a good place to work?" Slowest to clear; the incident survives here as a general reliability caveat long after the specifics fade.
  • Shortlist prompts — "best tools for Y." Lowest mention rate, but the only family where you get silently dropped instead of qualified.

Bad news enters all three. It leaves them at very different speeds, and most monitoring only watches the first.

How we tracked 34 incidents across six AI surfaces

We monitored 34 incidents at 27 B2B SaaS and consumer tech companies between January 2025 and June 2026: 12 outages or security incidents, 11 lawsuits or regulatory actions, and 11 layoffs or restructurings.

Each brand had an 18-prompt set (6 incident, 6 trust, 6 shortlist), run 3 times daily across ChatGPT, Google AI Overviews, Google AI Mode, Perplexity, Gemini and Copilot. Tracking ran 30 days before the incident where a baseline existed and 90 days after — roughly 1.3 million answer samples in total.

Two definitions used throughout:

  • Entry lag — days from the first mainstream article to the first answer that mentions the incident unprompted.
  • Tail length — days from verified resolution (service restored, case dismissed or settled, restructuring completed) until the incident appears in fewer than 10% of runs of its trigger prompt.

This is observational panel data, not a controlled experiment. Where we report an effect, it is a difference between cohorts, and we flag it as such.

How fast does bad news enter AI answers?

Median entry lag was 3 days; the fastest was 14 hours and the slowest 11 days. Retrieval-heavy surfaces move first, and outages move faster than legal or workforce news.

Surface Outage / security Lawsuit / regulatory Layoff / restructuring
Perplexity 1 day 2 days 2 days
Google AI Overviews 1.5 days 2 days 2 days
Google AI Mode 2 days 3 days 3 days
Copilot 3 days 4 days 4 days
ChatGPT (browsing on) 3 days 4 days 5 days
Gemini 4 days 5 days 6 days

Median days from first mainstream coverage to first unprompted mention, 34 incidents.

Entry is not synchronized. At peak, only 4 of 34 incidents appeared across all six surfaces on the same day. That matches BrightEdge’s finding that engines disagreed on which brand to flag 73% of the time on overlapping negative prompts — with the practical addition that disagreement is partly a timing artifact, not only an editorial one. Check one engine on day 2 and you will call the all-clear on a fire that has not reached the other five.

Why Perplexity and AI Overviews go first

Both lean hardest on live retrieval and both favor news domains for anything that looks time-sensitive. In our panel, when a query contained a temporal or status word ("down", "outage", "sued", "lawsuit", "layoffs"), these two surfaces pulled a news citation in the majority of runs within 48 hours.

That speed cuts both ways: they are also the first to pick up a resolution page, which is why they dominate the fastest recoveries later in this article.

Why ChatGPT and Gemini lag — and why that isn’t good news

Slower entry looks like breathing room. It isn’t. Slower surfaces in our panel also produced the longest tails: Gemini’s median tail was 19 days longer than Perplexity’s for the same incident. The same inertia that delays the bad news delays the correction.

Practical consequence: the surface where the story arrives last is the surface where it dies last. Sequence your response by tail length, not by arrival order — Perplexity self-corrects quickly once your page exists, so the work that actually needs sustained pressure is the slow tier.

Deep research modes amplify the tail

Multi-step research agents don’t just retrieve one page — they run several rounds of queries and read further down each result set. In our panel, the deep-research variants of ChatGPT and Perplexity surfaced incident details in prompts where the standard mode said nothing, including incidents past their standard-mode tail. Forum threads and court dockets that never made a normal answer routinely made a deep-research one, because multi-step agents dig past the first page of sources. If your buyer runs deep research during diligence, assume the incident is still on the table months after your dashboard says clear.

Which sources carry the bad news into the answer?

News articles supplied 38% of citations attached to incident answers — but non-news sources together supplied the majority. This is the single most actionable finding in the dataset, because you can influence four of the six categories.

Citation source Share of incident-answer citations Can you influence it?
News articles 38% Indirectly — via follow-up coverage
Forums and Reddit threads 21% Yes — reply in-thread, on the record
Aggregators, review sites, competitor "alternatives" pages 14% Partly — profiles yes, rival pages no
Brand’s own status page or postmortem 11% Fully
Court filings and regulatory documents 9% No
Other 7%

Two implications. First, a forum thread with 40 upvotes outlived the original article in 7 of 34 incidents — the news moved on, the thread did not. Reddit’s weight here is structural, not incidental: Google and Reddit signed a data licensing deal reported at roughly $60M/year in 2024, and Reddit content is a standing fixture in AI answers regardless of your incident.

Second, that 14% slice is why an incident often reaches buyers through a rival’s page — a competitor updating their "alternatives to X" page during your outage week is writing the source an AI will quote back to your prospect. We unpack that dynamic in when a competitor’s ‘alternatives’ page is AI’s main source about you.

The 11% own-domain slice is the encouraging number. Brands that published a real postmortem got themselves into the citation set, which is the precondition for changing the sentence.

How long does an incident stay after it’s resolved?

Median tail was 41 days after verified resolution — but the spread by incident type is enormous. Legal outcomes are stickiest, outages clear fastest.

Incident type Peak mention rate (incident prompts) Median tail after resolution
Outage / security 64% of runs 19 days
Layoff / restructuring 52% of runs 38 days
Lawsuit / regulatory 71% of runs 71 days

Lawsuits persist because the filing is a durable, well-linked document and the dismissal usually isn’t. In 8 of 11 legal incidents, answers still described the case in present tense weeks after it closed — "is facing a lawsuit" rather than "settled a lawsuit in March."

That asymmetry is the whole problem. Bad news arrives as a document; resolution arrives as a non-event. Nobody writes an article titled "nothing happened after all."

The shortlist tail is shorter than the trust tail

Incident mentions faded from shortlist prompts ("best tools for X") at a median of 16 days — far faster than from trust prompts. But the damage lands harder: in 9 of 34 incidents the brand vanished from shortlist answers entirely, for a median of 22 days, and took a further 31 days to return to its pre-incident share of voice after mentions normalized.

Recovery has two phases, and most teams stop measuring after the first. If your dashboard only tracks sentiment, you will declare victory while your share of voice is still a third below baseline. Some of that residual wobble is just noise — AI recommendation sets change between runs even with nothing happening — which is precisely why single-run recovery checks lie in both directions.

Layoffs have a second tail nobody tracks

Workforce incidents behave differently from the other two types: the news tail runs 38 days, but the employer-brand tail runs longer, because employee-review sites keep the story retrievable long after coverage stops. In our layoff cohort, review-site citations still appeared in "is X a good place to work?" answers after the news citations had dropped out. If you handle a layoff purely as a press problem, you clear the news answers and leave the employer-brand answers untouched — a different prompt set, a different audience, a different fix.

Which counter-evidence measurably shortened the tail?

Ranked by observed effect, getting the resolution covered by an outlet that already cited the original story beat everything else — cutting the median tail by 22 days. These are cohort differences between incidents where the action happened within 14 days of resolution and those where it didn’t, against the 41-day panel median.

Counter-evidence action Median tail change Cohort size
Resolution covered by an outlet that cited the original story −22 days n=9
Dated, indexable resolution page on own domain −17 days n=16
Updated third-party profiles (review sites, company databases) −9 days n=13
Machine-readable status/incident history (outages only) −7 days n=8
Direct-answer FAQ page addressing the incident question verbatim −6 days n=11
Executive posts on social platforms only −1 day n=12

These actions co-occur, so the effects are not additive — a brand that lands follow-up coverage usually also published a resolution page. Treat the ranking as a priority order, not a total.

Three details separated the pages that worked from the ones that didn’t:

  1. An explicit date in visible text, not just in metadata. Undated updates were cited in only 2 of 8 cases.
  2. The incident named plainly — "the March 12 outage", not "recent service disruption". Models retrieve on the words people search with.
  3. Resolution status in the first 40 words, so an extractive summary picks it up without reading the whole page.

Two structural traps we watched brands walk into. Do not delete the incident page once resolved — a 404 removes your only influenceable citation and hands the slot back to news and forums; update in place instead. And do not move it mid-recovery: one panel brand rebuilt its trust center during the tail and lost its own citations for weeks, the same failure mode as changing domains without preserving AI citations.

The underlying move is ordinary answer engine optimization applied under pressure: write the sentence you want the model to repeat, and put it where a summarizer will find it.

What showed no measurable effect

Four common responses produced no detectable tail reduction in our panel. Naming them saves budget.

  • Press releases on wire services alone. Present in 14 incidents; no cohort difference. They rarely entered the citation set.
  • PDF-only statements. Zero appearances as citations across the panel.
  • Homepage banners. Removed within days, so nothing durable remained to retrieve.
  • Legal removal requests aimed at forum threads. No reduction observed; in 2 cases the thread gained activity afterward.

Waiting also underperforms. The often-quoted claim that crisis visibility in language models peaks around days 30–45 and persists for months is published without a dataset, and our timing runs earlier — peak mention rate landed at a median of day 6, not day 30. But the underlying warning holds: the passive decay curve is far longer than a press cycle, and doing nothing means accepting all 41 days.

A 60-day response playbook for AI answers

Run this alongside your existing crisis plan, not after it. The windows below map to the entry-lag data above.

  1. Hours 0–6 — freeze a baseline. Capture how each surface currently describes you before the incident lands. Without a pre-incident baseline you cannot prove recovery later.
  2. Hours 6–48 — publish the dated holding page. One URL, plain title naming the incident, status in the first 40 words, updated in place rather than replaced. This is the page that becomes your 11% citation slice.
  3. Days 2–5 — watch the fast surfaces. Perplexity and AI Overviews will show you the framing the slower engines adopt next week. Whatever wording they pick up is the wording you must counter.
  4. Days 5–14 — service the non-news sources. Update review-site and company-database profiles, and answer the actual thread where the complaint lives. That 21% forum slice does not fade on its own.
  5. Days 14–30 — land resolution coverage. Go back to the specific outlets already cited in the answers. One follow-up from a cited source outperformed six from uncited ones.
  6. Days 30–60 — rebuild the shortlist. Track category prompts separately until share of voice returns to baseline, since mention sentiment recovers well before inclusion does.

Late-funnel prompts deserve their own workstream here, because "what are the downsides of X?" becomes the highest-traffic route to your incident for months afterward. That prompt family keeps returning the incident after the direct incident prompts have gone quiet — the objection-turn pattern we document in winning the objection turn in AI chats.

What to do if the incident is old and you’re starting late

Most teams find this article mid-tail, not on day 0. The order changes:

  • Skip the holding page; publish the resolution page directly — dated, incident named plainly, outcome in the first 40 words.
  • Read the current citation set before writing anything. Whatever sources the answers cite today are your actual targets; the day-0 news list is stale.
  • Correct the tense first. "Is facing" → "settled in March" is a cheaper, faster win than removing the mention, and it is what a buyer actually reads.
  • Assume the trust and shortlist prompts are still wrong even if the incident prompts have cleared. Check those two families before declaring recovery.

How to measure whether the tail is actually shrinking

Measure mention rate across repeat runs, not a single screenshot. In our panel, a single daily run misclassified incident presence in roughly a fifth of brand-days — the incident appeared in one run of a prompt and not the next, with nothing having changed.

Track four numbers weekly:

  • Incident mention rate — share of runs where the incident appears, per surface.
  • Framing — present tense vs. resolved tense. Tense flips before rate drops, so this is your earliest true signal of recovery.
  • Citation set composition — is your resolution page in it yet? If not, nothing else you’re doing will move the sentence.
  • Shortlist inclusion and share of voice — the recovery metric everyone forgets.

Sampling depth matters more than dashboard polish; our write-up on how many prompts and repeat runs make an AI visibility number trustworthy covers the thresholds. If you have no monitoring in place when an incident hits, the fastest starting point is a minimal brand mention tracking setup across ChatGPT and other assistants — a baseline captured on day 0 is worth more than a perfect system built on day 30.

Where this data stops

34 incidents at 27 companies is a real panel, not a census. Effects are cohort comparisons, so selection bias is possible: brands that publish fast resolution pages are often the same brands with functioning comms teams, and some of the measured gain belongs to that competence rather than the page.

Coverage is skewed to B2B SaaS and consumer tech in English-language answers. Regulated industries — where models are notably more cautious — and non-English surfaces are outside this dataset. Consumer recalls, which draw far heavier news volume, likely run longer tails than anything reported here. Cohort sizes in the counter-evidence table run n=8 to n=16; treat the ordering as directional and the exact day counts as approximate.

Sentiment rates in the wider population are low overall: BrightEdge measured negative mentions at 2.3% in AI Overviews and 1.6% in ChatGPT. Our panel deliberately samples the tail of that distribution — the incidents, not the average day. If you have never had an incident, your realistic exposure is far lower than these numbers suggest.

One last boundary: incidents don’t stay in one audience lane. The same lawsuit reaches a buyer as a risk caveat, a candidate as a stability question, and an analyst as a liability line — three prompt families, three tails, one event. Measuring only the buyer’s version understates the spread, and it is the version that clears first.

Frequently asked questions

How quickly does AI pick up bad news about a company?
Median 3 days from first mainstream coverage in our 34-incident panel, with Perplexity and Google AI Overviews typically first (1–2 days) and Gemini last (4–6 days). The fastest observed entry was 14 hours for a major outage.

How long does negative news stay in AI answers after it’s resolved?
Median 41 days after verified resolution — 19 days for outages, 38 for layoffs, and 71 for lawsuits and regulatory actions. Legal matters persist longest because the filing is a durable document and the dismissal rarely gets equivalent coverage.

Can you get an AI model to stop mentioning an incident?
Not by request. The reliable path is changing the retrievable evidence: a dated resolution page on your own domain, updated third-party profiles, and follow-up coverage from outlets already cited in the answers. Removal requests showed no measurable effect in our data.

Do all AI assistants mention the same incidents?
No. At peak, only 4 of 34 incidents appeared across all six tracked surfaces on the same day. Checking a single assistant will systematically understate exposure — and understate it most on the surfaces with the longest tails.

Does an incident hurt whether AI recommends us, or only what it says about us?
Both, on different clocks. Mentions faded from shortlist prompts in a median of 16 days, but 9 of 34 brands dropped out of shortlists entirely for a median of 22 days and needed a further 31 days to regain pre-incident share of voice.

Should we delete the incident page once the issue is resolved?
No. Update it in place with the resolution and a visible date. Deleting it returns a 404 and removes the only citation in the set you fully control — the slot goes back to news articles and forum threads you don’t.

Is it too late to act if the incident was months ago?
No, but the sequence changes: publish a dated resolution page rather than a holding page, read the current citation set before targeting outlets, and fix the tense ("is facing" → "settled in March") before chasing removal. Check trust and shortlist prompts separately — they clear last.


Written by

Founder of MaxAEO. Helping brands get found in AI search across ChatGPT, Perplexity, Google AI Overviews, and more.

Run a free AI visibility audit →