HN user

sjdv1982

9 karma

works on Seamless

Posts2
Comments27
View on HN
GPT-5.5 3 months ago

At some point, OpenAI is going to cheat and hardcode a pelican on a bicycle into the model. 3D modelling has Suzanne and the teapot; LLMs will have the pelican.

I wanted to ask almost this question, then saw that it is on #1 right now.

My use case is ssh. I would like to stick my private key into a local Docker container, have a ssh-identical cli that reverse proxies into the container, and have some rules about what ssh commands the container may proxy or not.

Does anyone know of something like this?

I am sorry, I am not a real computer scientist and I find it difficult to find the right term. With "sufficiently expressive", I mean things like dependent types and refinement types, that can express the constraint on a unit vector.

It seems to me that this is more or less the same thing, but Monte Carlo. Like MCMC vs symbolic Bayesian inference.

I am actually a research engineer paid by the French government. They take digital sovereignty pretty serious over here, which is sometimes good, sometimes less so.

Definitely the right call on Windows, though. Even my parents (in their mid-seventies) moved to Linux this year.

It is all about API contracts, right?

After the first run, you have a script and an API: the agent discovery mechanism is a detail. If the script is small enough, and the task custom enough, you could simply add the script to the context and say "use this, adapt if needed".

Or am I misunderstanding you?

I would like the AI to attach a confidence interval that the answer is "Yes" rather than "No". AlphaFold does this very well, but LLMs... not so much.

Interesting to hear the industrial SWE perspective, it is very different.

I am a scientific research engineer (bioinformatics), and here no one cares much about covering all the possible code paths.

What we care about is if the code computes "the correct thing", i.e. that it represents the underlying science.

No such guarantee with LLMs. But no such guarantee without LLMs, either (the "code growing above our heads" has happened already, a long time ago). Still, I would say that LLMs are a big net positive for us: they are better at checking such things than we are.

My fear is that this is going to lead to an optimal orchestration language. For example, that Claude switches to Sumerian for all communication between agents. One thing is if they try to silo like that, but my real fear is that it may actually perform well.

(Not sure if it would be Sumerian, Esperanto or something more artificial. As long as it is esoteric enough for one company to hoard all the expertise in it.)

I am a structural bioinformatics engineer, so my ignorance (adjacent fields not quite carrying over) comes from two different directions, so to say.

That being said: I feel that there must be some kind of benchmark for this. If no such benchmark exists, use your framework, pair up with a couple of pharmacists, and create one.

Nice map!

The First Age / Second Age boundary is not unlike the K/T boundary...

Compared to that, Second Age / Third Age isn't that different (places like Dunland and Tharbad were forested, according to Treebeard). So if you wish to make the map a bit more ageless, you could just add a few alternate names. - Dol Guldur was Amon Lanc in the Second Age - Lothlorien was Laurelindorenan in the Second Age - Mirkwood, Minas Tirith and Minas Morgul are late-Third-Age-isms too.