Comment on A Project to Poison LLM Crawlers

<- View Parent
GamingChairModel@lemmy.world ⁨1⁩ ⁨week⁩ ago

If I am reading this correctly, anyone who wants to use this service can just configure their HTTP server to act as the man in the middle of the request, so that the crawler sees your URL but is retrieving poison fountain content from the poison fountain service.

If so, that means the crawlers wouldn’t be able to filter by URL because the actual handler that responds to the HTTP request doesn’t ever see the canonical URL of the poison fountain.

In other words, the handler is “self hosted” at its own URL while the stream itself comes from the same URL that the crawler never sees.

source
Sort:hotnewtop