You are right, the engine should focus more on content and less on the complexity or quality of the formatting - but I also kind of get it.
try searching “wikipedia” on wiby before claiming this next time, thanks.
I've found myself confused by this:
https://en.wikipedia.org/w/load.php?lang=en&modules=ext.cite...
On a side note why does every wiki site have something like this? I can't tell what it is and haven't had time to look into it. But thats more than just a few lines of js to layout a page.
Lastly, some sites like Wikipedia have hundreds of millions of pages, allowing them in would swamp the index as some people cant help but want to submit hundreds or thousands of pages from the same site (which would get blocked). I don't have enough resources to carry the zillions of wiki pages and have them drown out everything else. Most people know where to go if they want to look something up in Wikipedia and Google does a great job of it for just about every query you make, there will be Wikipedia prominently up there as a top result.
Computer Closet: http://www.computercloset.org/compindex.htm
Guide to Spam-like products: http://spam.budwin.net/
Already a fan.
I’m sure I could find something like this in Google if I tried hard enough, but damn, you just don’t see this kind of stuff in Google anymore (or at least I don’t).
EDIT: here’s another.
This is fantastic! Based off some of the other comments I don’t know if this will be a quality search engine, but I think this might be the second coming of StumbleUpon.
View source ...
<meta name="GENERATOR" content="Microsoft FrontPage 4.0">
:-)
Surf's up! That's 100% pure goddam information superhighway right there.
FWIW I don't see any mention of suckless on Wiby itself, so the headline may be misleading. As someone who has a negative impression of the suckless community, despite agreeing with their goals of "keeping things simple, minimal and usable", this matters.
> What kind of pages get indexed?
> Pages must be simple in design. Simple HTML, non-commerical sites are preferred.
> Pages should not use much scripts/css for cosmetic effect. Some might squeak through.
> Don't use ads that are intrusive (such as ads that appear overtop of content).
> Don't submit a page which serves primarily as a portal to other bloated websites.
> If you submit a blog, submit a few of your articles, not your main feed.
> If your page does not contain any text or uses frames, ensure a meta description tag is added.
> Only the page you submit will be crawled.
Some additional ethos info is on the About page[1].
Have you ever tried scrolling over things you don't like?
"Evidence That Humans And Dinosaurs Coexisted"
So that actually kind of sucks.
I don't see what sucks - we're talking about the WWW here, and we should expect the full spectrum of thoughts and opinions. Unless you're hoping that each and every boutique search engine will filter the results according to your preference?
"vaccines" and "covid 19" had decent results, mostly people's blogs and some sites trying to prevent misinformation.
"abortions" mostly gave statistics, a blog by a 14-year old student in China, and an extensive collection of writings in support of Ayn Rand and against Noam Chomsky.
"trump" gives the most perplexing results, which I don't really know how to describe. Interestingly, the Rand/Chomsky page shows up here too.
I suppose it depends on your definition of suckless, but all the sites that were returned were lightweight, and content-first. As for the content... interesting might be the word I'd use. And it definitely doesn't show up in mainstream engines.
i love it
I've been toying with the idea of like crowd sourced index curation. Like maybe backed off a git repo or something. Would be an interesting experiment.
Lunix indeed.
QUOTE
r6000 1 real estate laws 1 "Kitty Hawk Inc." software 1 Mt. Mansfield 1 jake OR rothermel 1 obsession 1 REALTOR.COM REAL SELECT 1 black jack 1 micheal jackson hate 1 repair projection tv screen 1 v tech 1 german meats 1 compare automobile 1 varmvattenberedare 1 Free Family tree information 1 digsolve.zip 1 epson lx-800 dip switch 1 télévision 1 offf -road 1 "michael robertson" 1 flight attendants tower air 1 "westwood studios" 1 aled jones shirtless 1 boomtown 1 hampton court 1 pics of power rangers nude 1 linux fvwm 1 pissing pisses femal 1 Fresh water aquarium 1 file compare programm 1 1979 4x4 chevy trucks 1 beanie babie magic 1 Cheyenne Frontier Days 1 apedemak 1 pennies 1 rap hip hop ra 1 "loving every minute" 1 sat-a2 1 45 auto 1 shack map 1 micro dc30 1 providian 1 ski conditions 1 adult key sites 1 people finder telephone numbers 1 icebreaker 1 WWW.SONY.COM 1 WICKED WOMAN 1 bank-owned reo orange county 1 university maryland 1 teen hideout 1 canyon texas real estate 1 script of scream 1 Last-Minute OR orlando OR NurFlug 1 "eternal champions" 1 "barrel house" 1 midi files about slayer 1 xmas cards 1 flashnet 1 Judge Roy S. Moore 1 beautiful latino women 1 "Stacy Sanchez" 1 NASCAR Dale Earnhart 1 "mac warez" 1 puritan pride herbs 1 "
- Mark Rosenfelder's Metaverse -- Bob's Reviews Women in Comics Zompist Phrasebook
- Metaverse Christianity and the Problem of Shame
- Piero Scaruffi's knowledge base -- Singularity Metaverse Blockchain Virtual Reality A Timeline of Artificial Intelligence A.I. slides Future of Technology Tributes Birthdays: a secular calendar of saints Centennial
The "sites which suck" filter seems to have left only sites of random blithering. After eliminating ad-heavy sites and the usual suspects, there's no easy way to rank. Now you need real content evaluation. Which is really hard.
You'd probably get better results by searching maybe 50 sites, such as Wikipedia and Brittanica for general knowledge, and a few just-the-facts new sites (Reuters, BBC, Japan Times, the Economist, the Guardian.) omitting their opinion articles. For popular culture, go for the trades - Variety, Chartbeat, etc.
But if wikipedia (which doesnt seem included) doesn't count as a content-first site, i dont know what does.
Searched for wow, got a lot of fishing sites. That was surprising.
Thanks for sharing.