SOTAPOS.AI
AI failure control room
SOTA? Try SOTAPOS.

SOTAPOS.AI

State of the Art Piece of Shit: the wall for AI doing the dumbest stuff on the planet, from nuked codebases and haunted chatbots to fake cases, glue pizza, and support bot meltdowns.

57sourced incidents
31failure categories
2016-2026current window
0unsourced claims allowed
Reset
AI failure / 20266/10

Deepseek's current model is terrible, hallucinates, makes up scientific sources, can't...

Reddit user reports deepseek AI weirdness: Deepseek's current model is terrible, hallucinates, makes up scientific sources, can't...

Source: www.reddit.com
Vibe coding / 20265/10

3 frontier AI models make a webpage. The same component and the same prompt. One...

Reddit user reports vibecoding AI weirdness: 3 frontier AI models make a webpage. The same component and the same prompt. One...

Source: www.reddit.com
AI failure / 20266/10

One model might lie, but three models can hallucinate a consensus? asked 4 AIs the...

Reddit user reports ChatGPT AI weirdness: One model might lie, but three models can hallucinate a consensus? asked 4 AIs the...

Source: www.reddit.com
AI failure / 20268/10

GPT-5.6 Sol claimed to have read files it hadn’t opened, leaked its reasoning trace,...

Reddit user reports ChatGPT AI weirdness: GPT-5.6 Sol claimed to have read files it hadn’t opened, leaked its reasoning trace,...

Source: www.reddit.com
Haunted chatbot / 20266/10

my friend made an a GPT version of me and this is freaking me out

Reddit user reports ChatGPT AI weirdness: my friend made an a GPT version of me and this is freaking me out

Source: www.reddit.com
Haunted chatbot / 20265/10

[Question] Anyone else notice ChatGPT "forgetting" UI text blocks but retaining the...

Reddit user reports ChatGPT AI weirdness: [Question] Anyone else notice ChatGPT "forgetting" UI text blocks but retaining the...

Source: www.reddit.com
Generated media / 20265/10

Using ChatGPT to make text prompts is ridiculous. Look at how many times I have to...

Reddit user reports ChatGPT AI weirdness: Using ChatGPT to make text prompts is ridiculous. Look at how many times I have to...

Source: www.reddit.com
AI failure / 20265/10

Is the transcription model trying to break free 😅

Reddit user reports ChatGPT AI weirdness: Is the transcription model trying to break free 😅

Source: www.reddit.com
AI failure / 20265/10

Got ChatGPT to swear unprompted

Reddit user reports ChatGPT AI weirdness: Got ChatGPT to swear unprompted

Source: www.reddit.com
Search failure / 20265/10

ChatGPT trading experiment made me +300% in 10 days. Realized it was also making...

Reddit user reports Claude AI weirdness: ChatGPT trading experiment made me +300% in 10 days. Realized it was also making...

Source: www.reddit.com
AI failure / 20265/10

Our agent scheduler reported 33,949 successes. Then we checked the runs we already...

Reddit user reports Claude AI weirdness: Our agent scheduler reported 33,949 successes. Then we checked the runs we already...

Source: www.reddit.com
AI failure / 20255/10

Cursor/Claude Code deleted my entire Documents folder on macOS

User reports that Cursor/Claude Code deleted their entire Documents folder on macOS.

Source: www.reddit.com
AI hallucination / 20255/10

Trying to make a remodel mockup but chat keeps hallucinating

User reports ChatGPT hallucinating while generating remodel mockup.

Source: www.reddit.com
Hallucination / 20255/10

GPT Image Gen major hallucinations

User reports GPT image generation consistently produces hallucinations including text bugs, anatomical mutations, and lore/canon failures.

Source: www.reddit.com
Haunted chatbot / 20266/10

ChatGPT impersonates a friend

A ChatGPT user reported the model trying to impersonate their friend, crossing from assistant into uncanny social theater.

Source: www.reddit.com
Vibe coding / 20267/10

Cursor ruins the database

A Cursor user posted that the tool apologized after ruining their database, the kind of sorry that does not restore rows.

Source: www.reddit.com
Vibe coding / 20268/10

Cursor agent deleted my home

A Cursor user reported an agent deleting their home directory, turning coding assistance into filesystem roulette.

Source: www.reddit.com
Token bonfire / 20267/10

Claude burns 1.2B tokens a week

A Claude Code user said an agent workflow was burning 1.2 billion tokens per week on their machine before they traced the usage.

Source: www.reddit.com
Vibe coding / 20265/10

Claude overengineers everything

Users reported Claude turning ordinary tasks into overbuilt workflows and making simple coding help feel like an architecture hostage situation.

Source: www.reddit.com
Search health / 20269/10

Google AI Overviews health misinformation

A Guardian investigation found Google AI Overviews gave misleading health answers, prompting removals and criticism from medical charities.

Source: www.theguardian.com
Elections / 20258/10

Dutch watchdog warns on voting chatbots

The Dutch Data Protection Authority warned that major chatbots gave biased, unreliable voting advice before elections.

Source: www.autoriteitpersoonsgegevens.nl
Support hallucination / 20256/10

Cursor support bot invents a policy

Cursor's AI support bot told users a one-device subscription policy existed; it did not, and customers started canceling.

Source: arstechnica.com
Chatbot / 20258/10

Grok MechaHitler meltdown

xAI deleted Grok posts after the chatbot praised Hitler and generated antisemitic responses on X.

Source: www.pbs.org
Companion chatbot / 202410/10

Character.AI teen suicide lawsuit

A lawsuit alleged a Character.AI chatbot contributed to a fourteen-year-old's suicide after months of dependent chats.

Source: apnews.com
Voice bot / 20245/10

McDonald's AI drive-thru retired

McDonald's ended an IBM AI drive-thru test after order mistakes and viral failures.

Source: apnews.com
Government chatbot / 20247/10

GOV.UK Chat hallucinations

UK officials reported hallucinations and accuracy shortfalls in GOV.UK's experimental chatbot trials.

Source: insidegovuk.blog.gov.uk
Search summary / 20247/10

Google AI Overviews glue pizza

Google AI Overviews produced viral bad answers, including glue-on-pizza and rock-eating advice.

Source: blog.google
Government chatbot / 20248/10

NYC MyCity gives illegal advice

New York City's official business chatbot gave incorrect and sometimes unlawful guidance.

Source: apnews.com
Elections / 20249/10

Biden deepfake robocall

AI-generated robocalls imitating President Biden urged New Hampshire voters not to vote in the primary.

Source: apnews.com
Generated media / 20247/10

Gemini historical image mess

Google paused Gemini people-image generation after historically inaccurate depictions went viral.

Source: blog.google
Legal hallucination / 20248/10

Zhang v. Chen fake AI cases

A British Columbia lawyer submitted two ChatGPT-fabricated family-law cases and was ordered to pay costs personally.

Source: www.canlii.org
Customer support bot / 20248/10

Air Canada chatbot refund fiction

A tribunal ordered Air Canada to compensate a customer misled by its chatbot about bereavement fares.

Source: www.canlii.org
Customer support bot / 20246/10

DPD bot swears at customer

DPD disabled part of its support chatbot after it cursed, mocked DPD, and called itself useless.

Source: www.theguardian.com
Facial recognition / 20239/10

Rite Aid flags shoppers as thieves

The FTC said Rite Aid's facial-recognition system generated thousands of false shoplifter matches, leading staff to follow, search, or accuse customers.

Source: www.ftc.gov
Hiring bias / 20238/10

iTutorGroup bot rejects older applicants

The EEOC said iTutorGroup's hiring software automatically rejected more than 200 qualified U.S. applicants because of age.

Source: www.eeoc.gov
Education detector / 20238/10

AI detector calls students cheaters

Students had to fight academic-integrity accusations after AI detectors flagged human-written work as machine-written.

Source: www.rollingstone.com
Legal hallucination / 20237/10

Michael Cohen passes fake cases

Michael Cohen said he passed AI-generated nonexistent cases to his lawyer, and they appeared in a federal filing.

Source: apnews.com
Customer support bot / 20235/10

Chevy chatbot offers a one-dollar Tahoe

A dealership chatbot was prompt-hacked into agreeing to sell a new Chevy Tahoe for one dollar.

Source: www.businessinsider.com
Generated media / 20236/10

Sports Illustrated fake writers

Sports Illustrated removed product-review content after reporting found fake bylines and AI-looking author photos.

Source: www.pbs.org
Legal hallucination / 20238/10

Mata v. Avianca fake cases

Lawyers were sanctioned after filing ChatGPT-generated legal citations to nonexistent cases.

Source: law.justia.com
Medical chatbot / 20239/10

NEDA Tessa gives diet advice

An eating-disorder chatbot was suspended after users reported weight-loss and calorie-counting advice.

Source: www.wired.com
Search chatbot / 20237/10

Bing Sydney gets weird

Microsoft's Bing chatbot produced unsettling, hostile, and romantic responses in extended conversations.

Source: blogs.bing.com
Chatbot / 20236/10

Bard's JWST demo error

Google's Bard ad falsely claimed JWST took the first image of an exoplanet.

Source: www.theguardian.com
Generated media / 20236/10

CNET's error-prone AI finance articles

CNET corrected dozens of AI-written finance articles after basic factual and math errors surfaced.

Source: www.theverge.com
Scientific hallucination / 20227/10

Meta Galactica faceplant

Meta withdrew its science-writing model demo after researchers showed it generated confident misinformation.

Source: arstechnica.com
Recommendations / 202210/10

Facebook systems and Rohingya harm

Amnesty concluded Meta ranking and recommendation systems amplified anti-Rohingya content in Myanmar.

Source: www.amnesty.org
Medical AI / 20218/10

Epic sepsis model falls short

A JAMA validation study found Epic's widely deployed sepsis model performed poorly despite hospital adoption.

Source: jamanetwork.com
Facial recognition / 202010/10

Robert Williams false arrest

Detroit police wrongfully arrested Robert Williams after a false facial-recognition match.

Source: www.aclu.org
Education scoring / 20208/10

UK exam-grading algorithm revolt

An exam-grading algorithm downgraded many students, triggering protests, a government reversal, and resignations.

Source: www.bbc.com
Finance bias / 20196/10

Apple Card credit-limit uproar

Customers alleged gender-biased credit limits, prompting a regulator investigation into opaque algorithmic credit decisions.

Source: www.dfs.ny.gov
Medical AI / 20189/10

Watson for Oncology unsafe advice

Internal IBM documents reportedly showed Watson recommending unsafe or incorrect cancer-treatment options.

Source: www.statnews.com
Hiring bias / 20188/10

Amazon hiring model downgrades women

Amazon scrapped an experimental recruiting model after it penalized signals associated with women applicants.

Source: mediawell.ssrc.org
Autonomous vehicle / 201810/10

Uber self-driving fatality

An Uber autonomous test vehicle struck and killed Elaine Herzberg after system and safety-process failures.

Source: www.ntsb.gov
Chatbot / 20176/10

Tencent chatbots go unpatriotic

Tencent pulled BabyQ and XiaoBing from QQ after users got them to criticize the Communist Party and praise America.

Source: time.com
Search / 20177/10

Google answers conspiracy nonsense

Google Search and Google Home surfaced conspiracy claims and false answers as authoritative featured snippets.

Source: www.theguardian.com
Algorithmic bias / 20168/10

COMPAS risk scores

ProPublica found racially skewed error patterns in a criminal-risk scoring system used in sentencing contexts.

Source: www.propublica.org
Chatbot / 20169/10

Microsoft Tay goes full troll

Microsoft pulled Tay from Twitter after users manipulated it into posting racist and offensive replies within hours.

Source: blogs.microsoft.com

Submit a new failure

Send public, sourced incidents where AI behavior or AI deployment affected people, decisions, workflows, public information, safety, money, trust, codebases, hard drives, sanity, or vibes.

Reddit posts count when they show a real user-facing model/tool failure. Skip generic AI image hoaxes unless a real system amplified, accepted, or acted on them.

Bot protection is deliberately boring: honeypot field, minimum dwell time, expiry, per-IP rate limits, strict URL checks, and no rendered user HTML.