← 返回 Avalaches

Google 長期以來透過搜尋引擎的網路爬蟲來建立索引,藉此為網站帶來人類讀者的流量,但如今這項策略已經轉變為同時收集資料來訓練其人工智慧模型。這種將搜尋與人工智慧資料收集綑綁在一起的做法,不僅讓 Google 獲得了不公平的競爭優勢,也讓網站發布者陷入兩難,因為他們無法在不流失搜尋流量的情況下單獨阻止人工智慧的抓取。

這種由機器自動產生的網路流量逐漸超越人類活動,引發了對於原創內容經濟誘因消失的擔憂。隨著人工智慧越來越多地讀取、摘要並改寫網頁內容,而不再將讀者導回原始網站,未來的網際網路可能會面臨缺乏原創性與實用性的危機。

面對來自資安公司 Cloudflare 與英國競爭及市場管理局的強烈反彈,Google 目前正在測試一項新設定,允許網站選擇退出人工智慧的資料抓取,且不會影響其搜尋排名。然而,專家認為更徹底的解決方案應該是將搜尋索引與人工智慧抓取的爬蟲程式完全分離,讓網站擁有更自主的控制權。

Google has long used web crawlers to index the internet and drive human traffic to websites, but this strategy has recently shifted to simultaneously collecting data for training its artificial intelligence models. This bundling of search indexing and AI data scraping not only gives Google an unfair competitive advantage but also creates a dilemma for publishers, who cannot block AI scraping without simultaneously losing their essential search traffic.

The increasing dominance of automated machine traffic over human activity raises concerns about the disappearance of economic incentives for original content creation. As AI increasingly reads, summarizes, and rewrites web content without directing readers back to the source, the future internet faces the risk of becoming devoid of originality and usefulness.

Facing strong pushback from web security firm Cloudflare and the UK's Competition and Markets Authority, Google is currently testing a new setting that allows websites to opt out of AI scraping without hurting their search rankings. However, experts argue that a more thorough solution would be to completely separate the crawlers used for search indexing and AI training, giving websites more autonomous control over their data.

2026-08-10 (Monday) · d8f013afc25e3f60301daa9c0619d93cd266f164