HN user

stared

13,470 karma

A curious being, doctor of sorcery. Posts, projects & resume at: https://p.migdal.pl/

Now: benchmarking AI at: https://quesma.com/blog/

Previously: co-founder & CTO at https://quantumflytrap.com/

Posts827
Comments2,305
View on HN
quesma.com 4d ago

Baba Is Solved by Fable 5 and GPT-5.6 Sol, but at what cost?

stared
3pts4
visquill.com 6d ago

Demographic Profiles

stared
1pts0
p.migdal.pl 6d ago

Show HN: Tree, truth, druid, dryad, and tar share one Proto-Indo-European root

stared
3pts0
quesma.com 6d ago

Baba Is Solved by Fable 5 and GPT-5.6 Sol, but at what cost?

stared
7pts3
eu-space.europa.eu 13d ago

Urban heat island effect in Brussels during the late June 2026 heatwave

stared
5pts0
quesma.com 13d ago

The true cost of saying "Hi" to an AI agent

stared
9pts4
quesma.com 23d ago

Qwen 3.6 27B is the sweet spot for local development

stared
1192pts759
p.migdal.pl 27d ago

Show HN: Tree, truth, druid and tar share one Proto-Indo-European root

stared
5pts1
astro.build 29d ago

Astro 7.0

stared
2pts0
p.migdal.pl 1mo ago

Show HN: Tree, truth, druid, dryad, tar and dendrite share the same PIE root

stared
2pts0
www.science.org 1mo ago

Wolves are reconquering Europe. Can people learn to live with them?

stared
46pts66
opusmagnumbench.com 1mo ago

Opus Magnum Bench

stared
3pts0
www.pnas.org 1mo ago

Statistical detection of systematic election irregularities (2012)

stared
2pts0
twitter.com 1mo ago

GLM 5.2 ranks #2 in Code Arena: Frontend

stared
2pts1
www.science.org 1mo ago

Deepfakes are everywhere. The godfather of digital forensics is fighting back

stared
4pts0
openrouter.ai 1mo ago

Surpassing Frontier Performance with Fusion

stared
3pts0
europeancorrespondent.com 1mo ago

Bye-bye Butterflies

stared
2pts0
m-malinowski.github.io 1mo ago

Can AI answer open questions in physics?

stared
4pts0
www.science.org 1mo ago

Microbe with tiny genome may be evolving into a virus (2025)

stared
3pts1
www.good.is 1mo ago

Ireland paid artists a basic income and it was a boost for the economy

stared
6pts4
github.com 1mo ago

15 Years of StarCraft II Balance Changes Visualized

stared
4pts0
gbaeval.com 1mo ago

Models try to write a Game Boy Advance emulator from scratch

stared
3pts0
p.migdal.pl 1mo ago

Games in which you walk (2019)

stared
2pts0
meffmadd.github.io 1mo ago

Can LLMs Play Baba Is You?

stared
2pts0
github.com 1mo ago

Show HN: Steam-like experience for DOOM community maps and mods

stared
3pts0
reason.com 2mo ago

The Surprising Divide over What Counts as True

stared
4pts0
arxiv.org 2mo ago

A Single Neuron Is Sufficient to Bypass Safety Alignment in LLMs

stared
3pts0
www.defensenews.com 2mo ago

Ukrainian drone strike on fuel depot prompts Latvian prime minister resignation

stared
1pts0
github.com 2mo ago

ICLR 2026 – Institutional Affiliations Dataset and Analysis

stared
15pts2
gbaeval.com 2mo ago

GBA Eval - Build a Game Boy Advance emulator in WebAssembly from scratch

stared
3pts0

I am curious how it works for people with ADHD.

I mean, I would love to read more books (and I have a very long backlog), but very rarely I find them stimulating enough to sustain attention. And if they do... well, then time flies and other commitments are missed.

Sure, everyone's ADHD is different and I know quite a lot of people, for whom book are go-to. For me there is a narrow bar between not enough stimulated to start, and too stimulated so in my own words of a thousand thoughts.

There was "It’s Not Enough to Be Right – You Also Have to Be Kind" https://news.ycombinator.com/item?id=21490714

But I think the core part is WHY we want to be right? To prove something to others, or to ourselves? To feel better? As a compulsion? As a gambler's fallacy? Many motivations are less lofty that we dare to admit.

I wasted way to much time arguing online. It was mostly wasted time, and wasted emotions. I mean, I also had many eye-opening and enlightening discussions, but these rarely were fights.

All experiments with Qwen 3.6 required no more than 48GB Apple Silicon. I believe you can go even further with more aggressive quantizations - one can go down even further.

In any cases, from the economic point of view, running models on laptops make little sense. Even at the pure cost of energy consumption, it might be hard to beat pricing at tokens generated at scale.

At the same time, it is a breaktrough, that will change the game. Previously such vibe coding on consumer device was not hard or costly - it was impossible.

Plotnine 29 days ago

Violin plots have an interesting reputation (https://xkcd.com/1967/, https://www.reddit.com/r/labrats/comments/91ex4u/is_it_just_..., https://jabde.com/2022/12/22/banned-violin-plots/).

For showing distributions, I much prefer strip plots (https://seaborn.pydata.org/generated/seaborn.stripplot.html), perhaps with opacity, or swarm plots (https://seaborn.pydata.org/generated/seaborn.swarmplot.html) - no averaging with an unknown kernel, no hiding distributions behind a box plot, and the data is directly visible. We also directly see whether it is based on 5, 100, or many more points.

When using histograms, binning is usually more straightforward than kernels. And in any case, the mirror reflection of a histogram is not needed.

A plain dictionary makes this harder or virtually impossible (you most often will silently fail)

I didn't know about these Typst restrictions. Silent fails are the worst, so if a constructor is necessary to prevent these, good it is there.

Thanks for explaining! (And for developing gribouille in the first place!)

My take: does something add value OR is there because were are just used to?

Gribouille is not ggplot2, or other. Syntax is different. Superficial keyword similarity is (usually) a false friend. Reusing a keyword might be useful, but keeping an unnecessary construction is (in my view), a cargo cult.

Typst itself breaks with a lot of LaTeX stuff, and it is good that it does not pretend it is LaTeX-with-Rust, but has a fresh look.

Deno Desktop 1 month ago

Tauri is getting traction in the meantime.

A non-native UI has some issues, but also one clear advantage - it is easier to make a cross-system app with the same looks.

Interesting! If I get it right, the API is in the spirit of Observable Plot (https://observablehq.com/plot/), less ggplot2.

In any case, I'm curious whether aes is necessary, or whether it would suffice to drop this function entirely and just use keys in the mapping (similarly for labs). Or, more broadly, whether using patterns from other implementations of the Grammar of Graphics is a conscious decision, or some sort of legacy baggage.

It revolves around the sentiment of "go deeper" - but I think it is a double-edged sword. Sure, entropy, tensors and gradients are important - and yes, they are pretty much requirements.

But from what I see, it is the opposite - a lot (if not virtually all) progress in the last decade of deep learning was not because of a fundamental idea, but incremental, experimentally-verified practice. Even though I think there is good intuition for why ReLU is better than sigmoid (tl;dr: last layer is log(sigmoid) ~ ReLU, putting anything different inside kills the gradient), the original paper by Hinton himself was more or less "because it trains 3x faster".

Re-thinking fundamentals might help, but most "let's change the fundamentals" is rarely how it works. Even the most seminal papers, i.e. AlexNet and "Attention Is All You Need", are refinements of existing ideas, and show how they help.

Machine learning is an experimental science. Many mathematically cool ideas do not work. Many engineering ones do.

I've tweeted before that one of the most important traits in a researcher is healthy paranoia. Be paranoid!

I have seen so many PhDs burned out to cinders; I don't think it is any more a good piece of advice than "depression is good for philosophers". Sure, be a relentless explorer.

In short, holding on to ideas for too long can actually be counterproductive. Stay open-minded and refuse to let ego cloud your judgement.

Which I think is true.