this post was submitted on 28 May 2024
1 points (100.0% liked)

Technology

60216 readers
2484 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related content.
  3. Be excellent to each another!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, to ask if your bot can be added please contact us.
  9. Check for duplicates before posting, duplicates may be removed

Approved Bots


founded 2 years ago
MODERATORS
 

A purported leak of 2,500 pages of internal documentation from Google sheds light on how Search, the most powerful arbiter of the internet, operates.

The leaked documents touch on topics like what kind of data Google collects and uses, which sites Google elevates for sensitive topics like elections, how Google handles small websites, and more. Some information in the documents appears to be in conflict with public statements by Google representatives, according to Fishkin and King.

top 6 comments
sorted by: hot top controversial new old
[–] woelkchen@lemmy.world 2 points 7 months ago

Some information in the documents appears to be in conflict with public statements by Google representatives

I would have never guessed that.

[–] flappy@lemm.ee 0 points 7 months ago (2 children)

Can't wait for selfhosted web search to become better.

[–] Jako301@feddit.de 0 points 7 months ago

How is that even supposed to work? These search engines need per definition massive databanks to search through. Either you need your own crawler and indexer which is more than just inefficient, or you are limited to a relatively short list of curated static results.

[–] jonne@infosec.pub 0 points 7 months ago (1 children)

You mean hosting your own crawler/indexer? That doesn't really sound like a thing you could do cost-effectively.

[–] warmaster@lemmy.world 0 points 7 months ago (1 children)
[–] Paradox@lemdro.id -1 points 7 months ago

Federated directories. We're going back to Yahoo like it's 1995