HN user

blcknight

1,238 karma
Posts0
Comments291
View on HN
No posts found.

I know tftp is still in wide use, I wonder if there's things out there looking for stuff that's less common like NNTP, finger servers, etc

Grok 4.5 14 days ago

Elon's rhetoric doesn't really match the model's behavior. It is willing to criticize Elon and argues against many of the insane right way points he tries to make.

No need to be defensive. If you ask an epidemiologist, they would almost certainly agree that it is essentially a marker of sexual activity at this point. It is transmissible even with condoms, it has many strains, it is widely prevalent, and the WHO states that basically everyone is sexual active will get one strain. Guardrail prevents against 9 - the ones we were able to create vaccines for and we know have negative effects.

There's also not really great tests for it, so you do not know if you have it or not.

Chinese models are almost certainly cheating on benchmarks, I would bet if you saw the training data that the benchmark canaries are in there.

GLM may be a good model in general but it s benchmaxxed and definitely not as good as Opus 4.8.

You realize that the companies listed employ many of the core open source maintainers for large projects? It is project-specific, but 80% of Linux kernel development is from paid corporate employees. Similar for kubernetes. All the load bearing infrastructure is already handled by these companies... literally no one else is going to have the resources or experience to redirect large efforts on securing F/OSS.

What would you propose otherwise?

How can an agent use these tokens then? If it sources the file can't it just read the env?

It also sounds like it is missing the important step of keeping the LLM credentials from the agents themselves. For example my GCP creds have access to far more than Vertex. This is solved by OneCLI and OpenShell via MITM proxies which seems more elegant to me. The tools live in containers and can't see anything but can use everything.

It also allows finer grained access controls, rating limiting, and there's talk of scanning for destructive actions.

Claude Fable 5 1 month ago

The fallback doesn't seem to be working for me, I haven't scanned a project in it immediately booted me when it found a security bug even though I didn't ask for it

It hasn't changed, and I don't know why people are saying that most books don't have DRM. It is only a small minority.

Tor books is the largest publisher without it (owned by Macmillan). Otherwise everything is truly hard DRM either ACSM with epub or Kindle's. They are both more or less easily defeated though.

MCP Hello Page 2 months ago

All kinds of tools make it really difficult to not make a URL clickable and even if it wasn't clickable they might still put it in address bar...

That's why I mentioned `-p`.

`--continue` and `--resume` are broken from `-p` sessions for the last 2 weeks. The use case is:

1. Do autonomous claudey thing (claude -p 'hey do this thing')

2. Do a deterministic thing

3. Reinvoke claude with `--continue`

This no longer works. I've had this workflow in GitHub actions for months and all of a sudden they broke it.

They constantly break stuff I rely on.

Skill script loading was broken for weeks a couple months ago. Hooks have been broken numerous times.

So tired of their lack of testing.

Fight enshittification. For whatever reason, many travel sites no longer send full details in the e-mail confirmation, they want you to click through to the site...which means I can't forward it to plans@tripit.com for automatic import.

Immediately after booking something,I tell Gemini to add it to my TripIt. Works great. I have a little prompt explaining how I like it formatted that I cut and paste, so I can just make this a one-click prompt. I could also have it add flights to my.flightradar24.com.

I also use Gemini in Chrome to add appointment confirmations to my calendar. Or remember things in Google Keep.

There's lot of use cases for this kind of thing.

For the love of god fix bugs and write some fricken tests instead of dropping new shiny things

It is absolutely wild to me you guys broke `--continue` from `-p` TWO WEEKS AGO and it is still not fixed.

There is a cognitive ceiling for what you can do with smaller models. Animals with simpler neural pathways often outperform whatever think they are capable of but there's no substitute for scale. I don't think you'll ever get a 4B or 8B model equivalent to Opus 4.6. Maybe just for coding tasks but certainly not Opus' breadth.

Yea this is the thing that makes no sense to me. Any frontier model can unmiminize minified JS pretty decently. Obviously not everything comes through, comments and such, but I always assumed the reason it wasn't open source was to prevent an endless shitstorm of AI slop PR's, not because they were trying to protect secret sauce.