HN user

kzrdude

12,270 karma

I try to play music. And I actually believe the free software stuff.

Posts45
Comments5,166
View on HN
simplyexplained.com 1y ago

An NFC movie library for my kids

kzrdude
1384pts314
docs.coiled.io 2y ago

Dask DataFrame Is Fast Now

kzrdude
13pts0
gendignoux.com 2y ago

Thoughts on the xz backdoor: an lzma-rs perspective

kzrdude
2pts0
github.com 2y ago

CPython UOps Optimization

kzrdude
5pts0
liavas.net 2y ago

Differential Geometry (Course Handouts)

kzrdude
1pts1
www.youtube.com 2y ago

Tensors for Beginners [video]

kzrdude
2pts1
www.youtube.com 2y ago

How we are making CPython faster. Past, present and future – Mark Shannon [video]

kzrdude
6pts0
www.youtube.com 2y ago

Light Years Ahead: The 1969 Apollo Guidance Computer [video]

kzrdude
7pts2
github.com 2y ago

Dtreeviz: Decision Tree Visualization

kzrdude
1pts0
rmoralesdelgado.com 3y ago

Concatenating in NumPy with r_ and c_

kzrdude
2pts0
www.bitecode.dev 3y ago

Why not tell people to “simply” use pyenv, poetry or anaconda

kzrdude
5pts1
lwn.net 3y ago

Yet another memory allocator for executable code

kzrdude
12pts0
settlers2.net 3y ago

A new Settlers II map generator

kzrdude
2pts0
www.youtube.com 3y ago

Summer of Math Exposition 2 Results

kzrdude
1pts1
www.quantamagazine.org 3y ago

Particle physicists puzzle over a new duality

kzrdude
146pts49
www.nytimes.com 4y ago

The Problems of Being a Surrogate Mother in Wartime

kzrdude
1pts1
docs.python.org 4y ago

Faster CPython: what's new in upcoming Python 3.11

kzrdude
2pts0
freeciv.fandom.com 4y ago

Freeciv 3.0.0 Changes (2022)

kzrdude
1pts0
twitter.com 4y ago

Finding Missed Optimizations Through the Lens of Dead Code Elimination

kzrdude
2pts1
www.theguardian.com 4y ago

Belgian-Briton Zara Rutherford is youngest woman to fly solo around world

kzrdude
1pts0
news.ycombinator.com 4y ago

Tell HN: Reddit was showing “Blocked” when accessed with Firefox

kzrdude
273pts167
news.ycombinator.com 4y ago

Ask HN: Why does lwn.net give connection reset errors?

kzrdude
1pts2
www.reddit.com 4y ago

Chessvision-AI bot finds YouTube videos mentioning any given chess position

kzrdude
3pts0
lwn.net 4y ago

The folio pull-request pushback (Linux)

kzrdude
3pts0
teddit.net 5y ago

Played Nethack regularly for 10 years completely unspoiled (~1988-1998)

kzrdude
1pts0
jexer.sourceforge.io 5y ago

The Evolution of a Terminal Programmer

kzrdude
2pts0
www.theguardian.com 5y ago

Nobel archives reveal judges’ safety fears for Aleksandr Solzhenitsyn

kzrdude
140pts89
lkml.org 5y ago

Report on University of Minnesota Breach-of-Trust Incident

kzrdude
2pts1
daniel.haxx.se 5y ago

Where Is HTTP/3 Right Now?

kzrdude
7pts0
github.com 5y ago

Hardware performance counter support [for the rustc self-profiler]

kzrdude
1pts0

Well, you can use the google models from Pi. Go to aistudio.google.com and set up an API key. There is a free quota, it's relatively large for the Flash Lite models.

...and last time I looked the limits were more generous for Gemma 4 there, but they have been tightened a bit. That's how it goes, always changing.

If you go small enough it should be no problem. For example Gemma 4 E4B in Q6 or Q4 quantization should run well on your laptop. It shouldn't be too taxing, but would still want to eat 7-9 GB of VRAM or so.

Now that model is mostly useful for writing or chatting.

Wikipedia has policies and it needs to use reliable sources. This rule is often skirted and a lot of facts are cited to self-published and fast moving web sources. Those who are braking the progress on the article here are doing the right thing, trying to uphold the editorial standard.

Luckily there is now a New Scientist article to link to, so, the issue should now be resolved.

Qwen 3.8 3 days ago

Fwiw, American industry has given away a lot for free - you could include large parts of the open source movement in that - and all the "free" VC backed services like facebook would be another prong of the same comparison. I would rather compare this way, that China is gaining soft power and goodwill, in the technology and innovation sense, in a way that's similar to how USA has done in the past.

OpenRouter's ToS also seems to allow them to store your submitted prompts anyway, so privacy advocates would have to look elsewhere anyway, that's at least how I understand it (and it surprised me).

I didn't know that Amish thought so lowly of their own language, I think that's just sad. It's their own language and there's no reason to measure it against others.

It might be impossible to run such a large model on the laptop, if I were you I'd first start experimenting with running smaller models, using llama.cpp or lmstudio, just to start understanding how it works and learn about some of the concepts involved - quantization and all!

For example Qwen3.5-4B or Qwen3.5-9B could probably be run cpu-only on some quantizations at more reasonable speeds.

If you want a larger one, you could probably run Qwen3.6-27B or 35B-A3B decently slowly.

Now to be humble - maybe it's impossible to run colibri (this project in the OP) on your computer. Note the performance stories - 0.07 token/s on a 24 GB laptop with 24 CPU threads!