Explaining AI's unreliable information sources from @wetclaude
The creator explains her research into why AI search engines provide strange and unfamiliar sources. She reveals that reputable websites use robots.txt files to block AI crawlers, forcing AI to pull from an 'unlocked web' of lower-quality, often AI-generated content, and advises viewers to practice AI literacy by checking the sources themselves.
Creator: @wetclaude on TikTok
Video format
Speaker address
Video outline
- State an AI anomaly
- Uncover the systemic cause
- Reveal the feedback loop
- Offer a literacy tip
Narrative framework
The Investigative Reveal
Narrative framework logic
Presenting a personal, intriguing anomaly and then systematically investigating its root cause, revealing a larger, hidden systemic issue, and concluding with an empowering, actionable insight for the audience.
Topics: Tech, Internet Culture, Journalism, Education
Concepts: Breakdown
Formats: Speaker address
Elements: Image Overlay, Podcast Setup, Jump Cut
Account types: Personal Brand
Transcript excerpt
I've logged every source that my AI has been giving me since June. Not the answers, but the sources or the citations, and this really weird thing has been happening. I recognize almost none of them. Why I'm doing this? The past five years, I've been writing a newsletter on tech, and right now I'm doing a research project on AI and shopping. So I'm looking at a lot of data right now about how AI is recommending products. So I went looking for why this is happening, and I looked at consumer reports. They've been testing products since 1936. But if you read their robots file, GPT bot disallow, Google's AI crawler disallow, common crawl, which feeds a bunch of these training sets disallow. They just don't want their testing scraped. And that's happening everywhere. So 60% of high credibility news sites now block AI crawlers. And the reason is this crazy stat, which is that traffic from bots and crawlers has now greater than human traffic on the web. So the sites with something to lose started to put up no entry signs, basically. So when you search something on cloud or chat tier browsing on, it's not necessarily reading the Internet. It's reading a part of the Internet that still lets
1,492,337 views