LOADING THE FEED ▮
NICHE OF ONE
--:--
← The Feed

Reddit Is Eating the Seed Corn

/Reddit's stock fell more than 20 percent on a quarter that beat expectations, and its answer to Google's AI Overviews is to fence off Old Reddit and sue a scraper. One story about who owns an archive, for anyone whose best writing lives on somebody else's platform.

post to X email it
Manga-style ink illustration of a single card-catalogue drawer pulled free and upended on a bare floor, index cards spilling out of it.
// the everything pass All-Access The whole catalog, the members vault, and the back room where the operators talk shop. $37/yr →

TL;DR: Ars Technica ran four Reddit stories in one week. Shares fell more than 20 percent on a quarter that “trounced investors’ expectations on most of the basic financials,” because Steve Huffman called search referrals “choppy.” His shareholder letter says “We are the antidote to an automated web,” while a $60 million licensing deal feeds the Google summaries that answer the question and keep the click. The same week, Reddit signaled “changes” to Old Reddit over “abusive scraping” and kept alive a DMCA suit against a scraper over content Reddit never wrote. Twenty years of unpaid answers are the asset here, and everyone bidding on them treats them as inventory. If your best writing lives on a platform you do not own, you are holding the same position with worse lawyers.

I catalog things, so I notice when a catalog goes in the skip. In April, dozens of comments and posts reaching back ten years were automatically removed from r/AskHistorians, with no route back. “And there was nothing we or the experts [who posted the deleted content] could do about it,” moderator Dr. Sarah Gilbert told Ars. Ten years of specialist answers, deleted by something that cannot read them.

What did Reddit actually sell?

Twenty years of answers written for free by people trying to be useful to a stranger. Reddit licensed the corpus to Google for $60 million. Google built a layer on top that reads the answers and hands the searcher the gist. Last year Pew Research, working from data on 900 US adults, found AI Overviews cut referrals to sites by almost half. Google told Ars the study “uses a flawed methodology and skewed queryset.”

Read the tape. Reddit’s quarter was good. The stock fell more than a fifth anyway, on the word “choppy.” That is a market pricing how much Reddit needs Google.

Ars quotes the shareholder letter at length, and the load-bearing line is this: “People don’t want a summary of Reddit; they want Reddit.” Huffman is right. He is also describing property his own company is fencing off and suing over while a classifier eats it from the inside.

The lawsuit is the tell

On July 31, Judge Paul A. Engelmayer largely denied SerpApi’s motion to dismiss. Reddit accuses the scraper of conspiring with Perplexity to lift copyrighted Reddit content out of Google search results, and the claim runs on the DMCA’s anti-circumvention provisions. Google argued the same thing less than two weeks earlier and lost, when a court found it had never proven that rights holders like Reddit authorized it to block the scraping in the first place.

SerpApi’s response to Ars is the sentence to keep. Google and Reddit, it said, are trying to “use the DMCA to wall off the open Internet by retroactively claiming control over content that they didn’t author and don’t own.”

Look at who is in the room. Reddit did not write the answers. Neither did Google, and neither did SerpApi. The people who did write them have no lawyer in the room, and no version of this case ends with them paid.

Old Reddit is the hostage

A Reddit blog post on August 5, reported by Scharon Harding, says the company plans “to make changes involving Old Reddit, the legacy desktop web experience for Reddit that was built 21 years ago,” in order “to protect our platform, business, and users from abusive scraping and spamming.” Reddit cut off logged-out readers in late July on the same grounds. Spokesperson Rosa Kim told Ars the company is “exploring different options” and gave no timeline. In June, a Reddit employee said flatly that “we can’t promise it will be around forever.”

Old Reddit is text and a comment tree. It is what moderator workflows and critical bots run against, which is why the same blog post promises to migrate them onto something else first. The scraping story does not add up. A login wall stops a hobbyist and does nothing whatsoever to a licensee holding a signed contract. It stops the people doing the work.

Nobody has a working answer to slop

Reddit says its AI has “increased enforcement actions on hate and violent content by more than 200 percent” and “helped reduce exposure to potentially harmful content by more than 40 percent.” Gilbert, research director of Cornell’s Citizens and Technology Lab, will not take the denominator on faith: “it’s hard to trust the numbers because it’s hard to trust the ‘judgment’ of Reddit’s systems.” She calls false positives a “huge problem” and an equity issue, because they mean “groups that are already marginalized are further silenced and censored.”

Every removed AskHistorians post linked to Rare Historical Photos, an image archive the moderators believe the classifier filed as spam. Discord’s system looked at pictures of square grids, chessboards and spreadsheets among them, and called them CSAM. Roughly 8,400 accounts permanently banned between May and early July before anyone noticed. All since reinstated. A bug had let the AI skip the human review it was built to require.

Casey Newton’s notes on the third era of slop sit behind Platformer’s paywall, so I use only what he has said in public. He asks why Snap, LinkedIn and Substack are backing away from slop while Meta and TikTok pile in, and his answer “comes down to what the platforms promise their user bases.” The picture on the piece is LinkedIn’s new “Seems like AI slop” report button. Users spot the slop and flag it themselves, for free.

An archive you do not host is sitting on a timer

Treat any archive you do not host as a thing on a timer, and the timer is a classifier with no appeals desk.

Pull your own writing out this week. Request your Reddit data export. Find the comments where you explained something properly to somebody who needed it, and republish them on your own domain with the question restated at the top. That is a few evenings of work, and it converts loaned text into an asset with your name on it. Leave the originals up.

Audit what you link out to. If your posts cite old.reddit URLs as evidence, those citations now depend on a login and a roadmap nobody will put in writing. Archive the pages you rely on.

Price your room honestly. If you run a members space, a person who actually reads the posts is the product. Companies with billion-dollar balance sheets are spending real money to advertise their way back to what you already have.

The part I would argue with

Harding closes her moderation piece like this: “Just as social media has no value without people, content moderation can’t succeed without human judgment at the forefront.” The sentence is correct and unpurchasable, and the same week of reporting is why.

Reddit lost more than a fifth of its market value on a quarter that beat expectations. A company disciplined that hard for referral risk buys classifiers. Payroll is the one line it will never touch. A classifier produces a number, “enforcement actions up more than 200 percent,” and that number fits on a slide. A person who reads the posts produces a smaller number and an invoice.

The framing I would push back on is that any of this is a tooling problem. Rules Hub, the AI moderation suite Reddit is expanding to replace Automod, is a better dashboard for volunteers, and a better dashboard changes nothing underneath it. The people who moderate Reddit have always worked for free, and their work is what made the corpus worth $60 million to Google. Now the company tells those same volunteers their interface has to change on account of “bad behavior.” Seed corn does not object to being eaten, which is the entire reason this works.

Frequently asked questions

Should I delete my Reddit history in protest?

Mirror it instead. Delete it and you lose your copy, a future reader loses the answer, and the words do not reliably go anywhere: Reddit’s own suit against Anthropic concerned scraping that retained users’ deleted posts. The useful protest is making your best answers exist somewhere you control, under your name, indexable by whatever succeeds Google. An archive in two places has survived one policy change. An archive in one place is waiting its turn.

If AI Overviews really cut referrals by half, is search traffic worth chasing at all?

The half figure comes from a study of 900 US adults and Google disputes the methodology, so hold it loosely. What nobody disputes is the behavior of the best-positioned content owner on the open web. Reddit has a licensing deal and twenty years of material nobody else holds, and it still called its referrals “choppy” and shed a fifth of its value on the word. Chase search while it is cheap. Do not build an operation whose only door belongs to somebody else.

I have a few hundred readers and no stock ticker. Does this apply?

More directly, because you are the archive and the moderator and the appeals process at once. Nothing you own is protected by a classifier and nothing you own can be erased by one. A room with an actual person in it is the cheapest asset a solo operation has, and it is exactly what public companies now write shareholder letters about. Huffman called it the antidote to an automated web. You have it already. Keep it on a domain you pay for.

// comments
Full search on OneSearch: the network, the ring, and the open web →esc closes · ↑↓ move · ↵ opens