In an age of LLMs, is it time to reconsider human-edited web directories?
Back in the early-to-mid '90s, one of the main ways of finding anything on the web was to browse through a web directory.
Lycos, Excite, and of course Yahoo all were originally web directories of this sort.
These directories generally had a list of categories on their front page. News/Sport/Entertainment/Arts/Technology/etc.
Each of those categories had subcategories, and sub-subcategories that you clicked through until you got to a list of websites. These lists were maintained by actual humans.
Typically, these websites also had a limited web search that would crawl through the pages of websites listed in the directory.
By the late '90s, the standard narrative goes, the web got too big to index websites manually.
Google promised the world its algorithms would weed out the spam automatically.
And for a time, it worked.
But then SEO and SEM became a multi-billion-dollar industry. The spambots proliferated. Google itself began promoting its own content and advertisers above search results.
And now with LLMs, the industrial-scale spamming of the web is likely to grow exponentially.
My question is, if a lot of the web is turning to crap, do we even want to search the entire web anymore?
At some point, does it become more desirable to go back to search engines that only crawl pages on human-curated lists of websites?
And is it time to begin considering what a modern version of those early web directories might look like?
@ajsadauskas @degoogle
It looks like there’s a couple projects to continue the directory DMOZ. I hope they’re sharing work with each other!
Got any links?
@Emperor
Yeah. Sorry, I was hesitant to post links at first before I vetted them.
It looks like “Curlie” is the official continuation of the DMOZ project:
https://curlie.org/
The other ones I was seeing, it turns out, are static mirrors of 2017 DMOZ.
Thanks for that, a real blast from the past. I have a vague memory that I was an editor on the ODP or dmoz back in the day.
Yes, perhaps not coincidentally, I thought it best to ask for a human-curated link.
@Emperor
Y’know, come to think of it, Wikipedia might be a better project to point to here. All the content on there is hand curated. When I’m interested in a subject, I usually go to wikipedia first instead of a search engine. Sometimes I am directed out to other websites from there.
I set up a quick keyword search so I can type “wp blah blah blah” into my url bar and it searches wikipedia.
https://support.mozilla.org/en-US/kb/how-search-from-address-bar?redirectslug=Smart+keywords&redirectlocale=en-US