HN user

chris_f

1,208 karma

Was working on https://www.runnaroo.com

Posts38
Comments198
View on HN
you.com 3y ago

Show HN: PyTorch search engine

chris_f
65pts17
www.ben-evans.com 3y ago

Ads, Privacy and Confusion

chris_f
1pts0
www.google.com 4y ago

A gallon equals how many pints?

chris_f
2pts2
www.tampabay.com 4y ago

Tampa Bay home is being sold as an NFT

chris_f
4pts0
www.theblockcrypto.com 4y ago

Coalition of US crypto firms unveils travel rule compliance platform, TRUST

chris_f
2pts2
www.runnaroo.com 4y ago

Show HN: Prototype product review search engine built this morning

chris_f
1pts0
www.wired.com 4y ago

Cloudflare Is Taking a Shot at Email Security

chris_f
6pts1
www.theverge.com 4y ago

I completely buy the conspiracy theory that Roy Kent from Ted Lasso is CGI

chris_f
1pts0
news.ycombinator.com 4y ago

Ask HN: Anyone else seeing odd Google weather search results?

chris_f
2pts0
slate.com 5y ago

Running from the Pain (2018)

chris_f
1pts0
blog.google 5y ago

MUM: A new AI milestone for understanding information

chris_f
257pts208
www.youtube.com 5y ago

An App Called Napster (Documentary)

chris_f
1pts0
www.xda-developers.com 5y ago

WhatsApp updates its Privacy Policy to mandate data-sharing with Facebook

chris_f
60pts17
tedium.co 5y ago

How AltaVista, the first good search engine, fell into the digital abyss

chris_f
349pts228
medium.com 5y ago

Building a Better Search Engine for Semantic Scholar (2020)

chris_f
1pts0
www.cupwire.com 5y ago

Managing your privacy: Web Browsers

chris_f
2pts0
news.ycombinator.com 5y ago

Ask HN: Why is Chrome a better browser than Firefox?

chris_f
1pts6
themarkup.org 5y ago

Simple Search by the Markup

chris_f
2pts0
darebee.com 5y ago

Best of the Modern Web: Darebee

chris_f
2pts1
www.seoco.co.uk 5y ago

What the Search Engines Looked Like About 20 Years Ago

chris_f
6pts3
www.theverge.com 5y ago

Bing is now Microsoft Bing as the search engine gets a rebrand

chris_f
61pts97
themarkup.org 5y ago

Blacklight – A Real-Time Website Privacy Inspector

chris_f
207pts101
www.politico.eu 5y ago

Google and data brokers accused of illegally collecting people’s data: report

chris_f
38pts0
www.grantfortheweb.org 5y ago

Grant for the Web awards Distributed Media Lab $2M to develop new revenue models

chris_f
1pts0
techcrunch.com 5y ago

A bug in Joe Biden’s campaign app gave anyone access to millions of voter files

chris_f
14pts5
www.zdnet.com 5y ago

Mozilla research: Browsing histories are unique enough to identify users

chris_f
238pts125
coil.com 5y ago

Thoughts on monetizing a privacy focused search engine

chris_f
3pts1
coil.com 5y ago

Privacy and Search Engine Monetization

chris_f
1pts0
www.runnaroo.com 6y ago

Show HN: Runnaroo – A new search engine

chris_f
380pts196
news.ycombinator.com 6y ago

Ask HN: Is anyone using the Web Monetization API?

chris_f
5pts2

Any specific sites? Happy to spin one of these up for you focused on Web3. These work the best when the search engine creator has domain expertise to showcase the best sources.

I could try to find some sources, but my guess at the best Web3 sites would probably miss the mark.

Good stuff! Github is one of the sources, but not specific repos. I actually think we can break out Github into individual repo sources pretty easily.

In the broadest sense, a highly targeted search engine (like this) can provide better results because Google has to determine user search intent AND return the right results from trillions of webpages. The advantage of this search engine is that all of the users are looking for the same type of information, and the result sources can be curated to ensure high quality and relevant results.

The more targeted the topic, the harder the time Google has to provide quality signal through the noise of SEO and sheer volume of content on the web.

In addition to the above, the UI provides some cool features like allowing horizontal scrolling of sources to provider higher information density (important for discovery), and some source content can be viewed in the side pane without leaving the page.

But ultimately it would be good to hear if this approach does make it easier to find relevant and higher quality Pytorch info.

Worth mentioning is the Alexandria.org project [0]. It is a non-profit search engine built on data from Common Crawl. The coverage is limited because of Common Crawl, but the relevance is decent. They also provide an API.

I believe one of the biggest impacts toward breaking up Google's monopoly on search is making them open up access to their index, even requiring Google to provide direct API search access for others to build alternative search products. They have a search API today, but it is prohibitively expensive to build on ($5/1000 calls).

I built a fairly popular search engine a couple years back, but the cost of Google's search API and increasing number of bot attacks make it difficult to reason keeping it online.

[0] https://www.alexandria.org/

Mojeek is excellent, and because they use 100% their own index they have a much higher hill to climb.

When I say "Best"for a general search engine, my definition is that it would fulfill the needs of myself and my non-technical family members. Kagi and Brave Search both do that while being different enough to not be just another Bing clone. I use Mojeek often, think it is great, and having their own index is a tremendous asset, but it doesn't quite meet that full definition yet.

Credit to them for trying some new things on the UI front, but it looks like the organic results are from Bing (like most other alt/privacy search engines). It would be interesting to learn more about how/if they plan to build their own index, or set themselves apart.

IMO, Kagi and Brave search are the two best alternative general search engines right now.

Runnaroo was pretty good as well ;-)

This is really an analysis of the use of biased language in news articles, which is interesting but only one dimension of potential bias.

It is very possible to use non "charged" language, but still report a topic with a strong bias. For example, Slate is left leaning by most measures, but the below landscape chart from the study has them dead center. Maybe they are better at using neutral terms?

https://space.mit.edu/home/tegmark/phrasebias.jpg

They key is to only submit information that they already have, not anything new. For documents that need to be uploaded in some cases, the options are either to use a heavily redacted real document with everything blacked out except for the essential info, or just upload a random file because most of the time no one checks and it is just a required field to submit a form.

You are correct it is not actually "deleted", but it will stop your information from showing up on the website.

IMO, the "best" free people search is https://www.truepeoplesearch.com/. Their opt-out process is also pretty easy.

As an exercise I once went through the process of manually requesting my information removed from most of the top brokers.

It would be difficult to automate because the opt-out processes usually aren't straightforward like unsubscribing from an email list. Many sites make it purposely difficult and involve going through multiple steps, providing verification like a drivers license, and email confirmations.

Here is a really good check list of different brokers: https://inteltechniques.com/data/workbook.pdf

Nice! Maybe at one point you can release a general web search engine for the Common Crawl corpus? It seems even simpler than this proof of concept, but potentially more useful for people looking for a true full text web search.

There isn't an easy way today to explore or search what is contained in the Common Crawl index.

Thank you for all the kind words! I'm the creator of Runnaroo.

Runnaroo started as just a fun experiment, but it quickly became apparent that you could launch a meta-search engine better than just about everything out there (including DDG [0]), and I was frankly surprised how quickly it was embraced by such a large number of people (it's Show:HN reached the #2 spot [1]).

The challenge then became how to fund the cost of the site in an ethical way in line with the site's core principles. I started looking at different solutions [2], including becoming the first search engine to implement Web Monetization, but I never really even came close.

I believe Runnaroo will live on in some iteration, I just have to figure out what that would be.

Also, if you used Runnaroo and liked it, please don't hesitate to reach out (anything AT runnaroo.com). It has been a solo project, but I'm sure the future will involve more collaboration.

I would also be interested in sharing the story of how Runnaroo evolved over the last year, and the different experiences of launching a search engine if anyone is interested or has a platform for those conversations.

------

[0] Don't take my word for it - https://news.ycombinator.com/item?id=24248666

[1] https://news.ycombinator.com/item?id=23771131

[2] https://coil.com/p/runnaroo/Privacy-and-Search-Engine-Moneti...

Thank you so much for the kind words! It was very much a labor of love as there was no tracking, ads, or really any significant attempt at monetization.

I was actually happy keeping it up for all of the users such as yourself, but it started to become not worth the effort as it grew and more and more people would abuse the service.

Runnaroo is a one-man side project operation, and I hope it has shown what is possible in the search engine space with minimal resources.

Good article. Petal search and Gowiki were completely new to me and I pay pretty close attention to other search engines (I created Runnaroo).

Another area to focus on when reviewing search engines are the enriched results (i.e. "Instant Answers", "Deep Searches", etc.).

These types of results have become almost an expectation for average search user almost as much as the organic results.

I was part of the pilot [0], and enjoyed the process. It didn't cost anything, and they only requested random surveys on the service.

My extension just sets the default search engine to Runnaroo, and I don't promote it much because people can just do that through the browser, but it was nice to get some increased exposure and I did notice an uptick in downloads, especially initially.

In general, I'm all for Mozilla testing different monetization strategies to offset some of their reliance on Google.

[0] https://addons.mozilla.org/en-US/firefox/addon/runnaroo/

"To me it seems the public can't be persuaded that paying for something is better than getting it "for free" (with ads and tracking)"

This is even more true when the paid option is often a worse version of the "free" offering.

"There's also things like https://coil.com/ who seem like they help support online content creators. I wonder if there's a way to treat search results like "content".

It is possible. I built the search engine [0] that was the first to integrate Coil as a monetization source. It is pretty small, but Coil payments do cover about 2% of the monthly cost to run the service.

Infinity Search also uses Coil. [1]

Here is an article with some thoughts around monetizing a privacy based search engine [2].

---------

[0] https://www.runnaroo.com/

[1] https://webmonetization.org/

[2] https://coil.com/p/runnaroo/Privacy-and-Search-Engine-Moneti...