HN user

antirez

31,706 karma

I'm a computer programmer based in Sicily (Italy).

Website: http://invece.org

Blog: http://antirez.com

My sci-fi book (English translation): https://www.ibs.it/wohpe-ebook-inglese-salvatore-sanfilippo/e/9791280845337 or Amazon Kindle

Twitter / X: @antirez

Bluesky: https://bsky.app/profile/antirez.bsky.social

Posts87
Comments3,700
View on HN
noperator.dev 1mo ago

You can just say it

antirez
406pts221
antirez.com 2mo ago

Redis array: short story of a long development process

antirez
320pts110
github.com 2mo ago

Redis new Array type PR and request for feedbacks

antirez
2pts1
antirez.com 4mo ago

GNU, and the AI Reimplementations

antirez
4pts1
antirez.com 4mo ago

Implementing a Z80 / ZX Spectrum emulator with Claude Code

antirez
151pts72
github.com 5mo ago

Voxtral.c Voxtral Realtime 4B model inference as a C library

antirez
5pts0
github.com 5mo ago

Tgterm – Control Claude Code from Telegram on macOS (< 1000 lines of C code)

antirez
2pts0
github.com 6mo ago

Flux 2 Klein pure C inference

antirez
453pts141
www.timeextension.com 11mo ago

A new 3D platformer for the ZX Spectrum

antirez
2pts0
antirez.com 1y ago

Coding with LLMs in the summer of 2025 – an update

antirez
600pts414
www.youtube.com 1y ago

HNSW as abstract data structure: video intro to Redis vector sets [video]

antirez
49pts0
antirez.com 1y ago

Redis is open source again

antirez
1896pts789
github.com 1y ago

Reinforcement Learning in less than 400 lines of C

antirez
10pts4
antirez.com 1y ago

We are destroying software

antirez
1012pts622
www.youtube.com 1y ago

Adding UTF-8 support to the Kilo editor using AI and human hints [video]

antirez
19pts1
github.com 2y ago

A new driver for the Badger2040 e-ink, or EPDs hacking and waveform LUTs design.

antirez
4pts1
antirez.com 2y ago

The origins of the Idle Scan

antirez
140pts13
news.ycombinator.com 3y ago

Show HN: Wohpe, my sci-fi novel about AI, programmers, climate change

antirez
2pts2
vickiboykis.com 3y ago

The cloudy layers of modern-day programming

antirez
280pts165
presse.inserm.fr 7y ago

New antibiotics effective against resistant bacteria in mice

antirez
145pts97
antirez.com 7y ago

Redis will remain BSD licensed

antirez
448pts197
antirez.com 8y ago

Redis PSYNC2 bug post mortem

antirez
122pts30
antirez.com 9y ago

Redis on the Raspberry Pi: adventures in unaligned lands

antirez
7pts2
github.com 9y ago

Neural Redis: simple to use neural network data structure module for Redis

antirez
7pts2
antirez.com 10y ago

The binary search of distributed programming

antirez
87pts15
antirez.com 10y ago

Is Redlock Safe? Reply to Redlock Analysis

antirez
181pts135
antirez.com 10y ago

Disque 1.0 RC1 is out

antirez
275pts51
soveran.com 10y ago

Human Error in Software

antirez
86pts12
antirez.com 10y ago

Generating unique IDs: an easy and reliable way

antirez
74pts64
redis.io 10y ago

Creating secondary and composed indexes with Redis

antirez
4pts0

Execution is not the code, but how do you decide to do every part. "Idea does not matter, execution does" always meant: "big generic ideas don't matter, it is how you organize it in the myriad of details it is composed of (in a given incarnation of the general idea) that matters."

Because Redis is not "my project", it is a piece of software many relies upon, so I use, for that software, what the community at large agrees to be ok: AI-assisted coding with human careful reading and evaluation of every line.

Thanks! And sorry for not yet merging many of those. The problem is, I'm dealing with tensor parallelism for the CUDA and Metal-RDMA fork right now, so was not albe to care about PR / issues for a lot of time.

Thanks, I believe that as a whole choosing the BSD created a more positive effect, so I'm happy with that. It is just that it is really unfair to read a comment where people use ValKey to accuse you of AI slop :D It means that our community, and this site itself, is at this point really low quality. This will in turn discourage the many great folks that are here. A replacement is needed. But TLDR, I would release Redis again with the BSD license if I could go back in time.

Sure, it costs less, and AWS is in a dominant position. Users here are playing the side of the bully since they don't care about what is right and wrong with the hyperscalers. "BSD is better than AGPL!" And give money to the wrong side of the history. Nor that I expected anything better, the single person has a given sensibility, the mass, as a whole, do whatever is in a given moment convenient or believed to be more pure (license wise). However thanks to that, you will see how little progresses we will have (and we are having) in the space of open source system software with very open licenses. Developers of software mostly are not happy to bring OSS to the success to see them used by hyperscalers to capture all the value. However I did it again, with DwarfStart, to release code under the BSD license: even in the current situation, I think it is better to give back than to have a personal gain, but this is a position that very little folks can afford to take.

However: this conversation is completely out of topic but people instead of talking about AI and code, which is a tabu, will move the conversation to personal attacks and shit like that.

Do you understand Redis and ValKey have mostly overlapping code bases? And of that intersection, a big part of the code was written by myself by hand. So no, that's not the case. Also as I wrote in the blog post, Redis is currently not using AI if not as AI-assisted coding. Of course I'm not writing this reply for you, since I believe if you write a comment like that, you are part of that HN slice that makes this site at this point a slop place (no need for AI for very low quality), but for others that may find this information useful.

The flaw is here:

1. Compute costs collapsed since the advent of Cloud and yet hyperscalers still have fat margins.

Cloud opposes switch inertia. To setup a complex system in a different environment is a complex operation. Changing AI provider is switching an endpoint.

In Catania 10 gbit Internet costs 35 euros/month and is available everywhere in the city and even in a big slice of the small towns around Catania. And indeed public incentives played a big role. But what I believe it is more interesting is that 1 gigabit was common like 10 years ago or even more. Infrastructure is a bit too important to be left to what the market believes will be profitable.

Btw the paradox I have is that my local lan is 1 gbit...

Europeans do a lot of stupid things, but I believe in light of all the scandals we saw in recent times, you can't explain EU behavior and choices without accounting for corruption. EU division and different level among the different countries of wealth, integrity of political sphere, and different cultural biases make us the perfect target for bribes in order to control votes and choices. Not just promoted by external actors. The Chat Control is a great example: everybody understands how bad this is, the arguments are mostly a shield to avoid revealing the real agenda.

DwarfStar work in progress numbers: I see 14 tokens/sec generation, that slopes to 10 t/s with longer 10k or more context size. Consider that the indexed attention requires evaluating 2048 selected rows, 2x DeepSeek and with less compression, so the performances with larger contexts here to south faster. Prefill can be 180 t/s on small contexts to 150 t/s and less with larger contexts. I used DeepSeek v4 PRO in this conditions, it is usable but it is far from the 35 t/s 400 t/s prefill you get with DeepSeek v4 Flash 2 bit on a MacBook m5 max. But likely my implementation is yet not optimized enough, so a bit more performance can be obtained. I'm using 4 bit quants. The model is also definitely less sparse than DeepSeek v4, so it activates a bigger percentage of parameters. If it works decently at 2-bit, that would be a win even for machines where 4-bit fits, since this would mean 2x memory (equivalent) bandwidth basically for the routed experts.

Local inference needs really hard a 1.2 / 1.5 T/s memory bandwidth system with 512GB and 2/3 times the GPU compute of Mac Studio M3 Ultra, at an affordable 10/15k price point. A variant with 1TB memory would also be welcomed at 20k price point.

I hate compilers 1 month ago

So to avoid those energy-hungry LLM companies from scraping your website, you force each browser to compute a lot of hashes in a necessarily energy-hungry loop, creating, at the same time, all the kind of accessibility problems?

They didn't freaked since the order was to still allow 350 million people using it: there is, in such large population, everything, including single persons very against the country, the government and so forth. If they really freaked they would say "we need to investigate, you have to retire the model". That would be a more defensible POV at least.

Apple WWDC 2026 1 month ago

They try so hard to do a polished presentation that everything is kinda fake and unauthentic. I don't understand how this attitude survived so many years.

How LLMs work 2 months ago

There is a different way to look at this: that is, actually the Transformer is a minimal complication of what the based model is: in theory the neural network could be just a huge FFN, which is anyway the part of the Transformer that does the heavy lifting. But this would be impossibile to train both numerically and computationally, so the Transformer encodes enough priors for it to work: the causal attention, and the math tricks like the residuals and so forth. But the bottom line of all this is that the Transformer works because of the incredible semantical power of simple/huge FFNs.

DaVinci Resolve 21 2 months ago

Got a copy of the Studio version a few months later I opened my YouTube channel: among the best money spent in software of my life.

What I say has nothing to do with efficient market hypothesis. Here the question is simpler: in small companies where there are competitors, who does the wrong choices will be seriously hit since customers will star preferring less slop and more reliability, if AI is mis-used. And companies that instead of firing, hire the folks that are "ideas people" and can use AI efficiently, and now how to control the quality of the output, will deliver more and better. For bigger companies: AI is driving salaries at a more normal level (honestly we want a bit too high, in recent years, even for people with a very low level of knowledge, didn't we?) and to marginally reduce total spending and not deliver the timeline they have, and are used to observe for years, will be noticed. Also companies in the past had a dangerous tendency to over-hire. I don't think now they will invert the direction and over-fire. I have the feeling many managers will instead reason in terms: what is today the great programmer fit? The one with low level knowledge of each algorithm, or the one that has good ideas and understands product, quality, processes, other than programming? And they will try to mix AI and people in order to have an edge.

Actually workflow impact in the world of software can be observed in weeks/months at max. And token spending too, is a voice that they see at the high floors. Also, there was never a strong willing in IT companies to reduce cost of work force: it is done sometimes, but it is more common to see them over-hiring.

Indeed this will likely happen in the future, but not today. I was experimeting with SSD streaming in DwarfStar for DeepSeek v4 PRO inference in 128GB systems (and Flash inference iwth 32/64). GPT 5.5 ran the whole night, I checked what it had accomplished regardless of all the hints I provided in the specification document. After reasoning on the problem I gave him the design fixes and the tokens/sec were 4x after 10 minutes. And this is true for every domain where the human babysitting the AI know a few things in that domain. However this is a moving target, and at the current rate, soon or later, indeed AIs will do much better than us in many domains.

A few years ago, the probability of such shit reaching the Hacker News home page was near zero, because regardless of the merits, here was not full of normies that could not understand when a behavior is unacceptable (I'm referring to the violence of the language of the issue). And now, here we are, surrounded by people that can't tell the most obvious things.

AI is not a product per se, it is a technology you can decline into a product, and the product has a lot less value than the technology itself. Who has the best LLM can copy any product idea and make it a lot better. Similarly if open weight LLMs are everywhere and powerful, open source products in the space of agents are too simple to replicate for people to pay big money to a few companies: not everything is alike, not every parallel makes sense. The pi agent is good as a replacement for Codex and Claude Code if you wire frontier models to it. And when products are complex and matter a lot, like complicated AI-powered design suites for instance, there is no reason why OpenAI / Anthropic will win this space instead of a random startup. So either a few companies retain frontier AI, or those companies will die.

About IRC / Slack: other than the fact IRC was abandoned, Slack is about control, not product. The product is terrible.

FTP / Dropbox: this comparison does not make sense.