signalDEV Community2026-09-18
Meta Reads More Than Everyone Else Combined. None of It Was a Question.
Over 30 days, Meta accounted for 52.3% of AI crawler traffic on a WordPress site (6,027 requests from its training crawler meta-externalagent) but sent zero user-initiated requests (meta-externalfetcher), while its citation-improving crawler meta-webindexer made only 155 requests. Perplexity's traffic was 93% user-initiated and OpenAI's 21%, highlighting Meta's extreme imbalance toward data extraction over attribution, though Meta publishes no IP ranges to verify the crawler's identity.
- why now
- 过去30天统计显示Meta抓取占52.3%却零用户主动引用,说明其当前只训练不回引。
- topic
- AI Industry Impact
- source
- DEV Community
#Meta
score
score 7 out of 10. 0-10: how dense the facts are, multiplied by how much you can do with them after reading. 8+ means the topic's evidence bar is met: benchmarks and availability for a new model, amount and investors for a funding round, revenue figures for a solo-money story. Below 5 an item does not enter the digest. A press release scores 3 or less, a reprint loses 2, anything older than 14 days loses 1, a headline that misleads loses 3.
read the source