HN user

TheTaytay

987 karma
Posts17
Comments287
View on HN
blog.cloudflare.com 3mo ago

Durable Objects in Dynamic Workers: Give each AI-generated app its own database

TheTaytay
3pts0
github.com 3mo ago

Canary – tiny filesystem honeypot for macOS

TheTaytay
1pts1
fabriziosalmi.github.io 3mo ago

Secure Proxy Manager

TheTaytay
2pts0
fabiorehm.com 4mo ago

Crib: Just Enough Devcontainers

TheTaytay
2pts0
simonwillison.net 4mo ago

Experimenting with Starlette 1.0 with Claude skills

TheTaytay
12pts4
nono.sh 4mo ago

Runtime Safety Infrastructure for AI Agents

TheTaytay
3pts0
github.com 4mo ago

Sandvault – Run AI agents isolated in a sandboxed macOS user account

TheTaytay
1pts0
www.figma.com 5mo ago

From Claude Code to Figma: Turning production code into editable Figma designs

TheTaytay
2pts0
github.com 7mo ago

Claude Code systematically creates issues in public anthropics/Claude-code repo

TheTaytay
2pts0
beyondloom.com 7mo ago

Decker is a multimedia platform for creating and sharing interactive documents

TheTaytay
3pts0
docs.orbstack.dev 9mo ago

Orbstack Debug Shell [March 2024]

TheTaytay
3pts1
github.com 9mo ago

Dorothy – A dotfile ecosystem: cross-shell, cross-OS, cross-arch

TheTaytay
2pts0
www.blocknotejs.org 9mo ago

BlockNote – The open source Block-Based rich text editor

TheTaytay
5pts0
www.vibekit.sh 10mo ago

Vibekit – The safety layer for your coding agent

TheTaytay
2pts0
pico.sh 1y ago

Pico.sh – SSH powered services for developers

TheTaytay
635pts141
www.assemblyscript.org 3y ago

AssemblyScript: (WASI) Standards Objections

TheTaytay
3pts0
www.youtube.com 13y ago

Common Physics Misconceptions (2.5min YT video)

TheTaytay
1pts0
Cloudflare Drop 14 days ago

Agreed. This is the same crowd that is mad about free tiers going away on other dev-friendly hosting services…

I think you have a point, and SCD type 2 feels like a workaround, but there is also something to be said for the ability to query every row as it was at any given version. I’m not saying that SCD type 2 is the best solution given there might be a more domain-specific way to do it, but I see it a lot like file-based version control. It’s convenient to be able to examine all files as they existed at any point in time, without having to “model” the ways in which those files might change directly into the domain of the individual files.

If you have something like dolt (not affiliated), a version controlled database, you wouldn’t have to slap change dates on anything OR create your historical table. The changes would be implicit in the version history.

Thank you for responding!

This spawned a very large thread, so I wasn’t sure where to respond, but it is shocking to me how most of the supporters of this in the comments make a fundamental error: they presuppose that people are going to write the game to begin with, no matter what the change in incentives is.

It would be like me making a law that said “every time you purchase a game, you must pay at least $100 for it,” and then proceed to explain how the quality of all games will go up, and that it will help the small Indy developers because now they get to multiply their guaranteed; existing player base by $100!

All of the arguments are: “well yes, this ight change the way you have to architect the game, and yes this might involve dictating what specific technologies and vendors you can or can’t use, and yes, this might increase costs, but…” and then go on to say why that is completely reasonable.

If someone made an equivalent law for websites: “any website you publish and sell access to must be made available, in perpetuity, to anyone who has ever used it” you would not get the open source utopia people seem to think, where everything is just as great, AND every individual and company on the planet took the extra time to ensure that they are compliant, regardless of the change in incentives.

I understand wanting to prevent someone from “artificially bricking” an app, but this vaguely-worded law isn’t it.

I wasn’t familiar with the term “Residential programming,” but it reminds me of the talk “Stop Writing Dead Programs” (https://jackrusher.com/strange-loop-2022/)

Increasingly, I think that an agent (and I) would work much better in a malleable, notebook-like, inspectable program, than it would with its current file-based “edit and re-run” primitives.

“Marimo pair” (built into their notebook-like primitive) is an attempt at this. And they have program introspection tools built in.

I also think that Glamorous Toolkit (https://gtoolkit.com/) might be a similar live environment, but I haven’t investigated it too much other than reading about it.

Is anyone else familiar with “modern” attempts at this?

So well said. I appreciate you articulating this. I am relieved to see that there are others on here who understand the unbelievable privileges we have at this moment in time.

The demagogues have been shockingly effective at telling people that the size of the pie is fixed, the game is rigged, and the only way to get a piece of that limited pie is to steal it. And the people that are most susceptible to this message are the ones that live in the countries that people are literally dying to get in to.

Your phrase “extract” betrays a fundamental disagreement with what Paul is saying. (Externalities does so again) It assumes a zero-sum game where the job is to shift money from one person to another. Value, and thus money/wealth can be created. Literally. You are saying, in different words, that no one can do it “honestly”. He is saying one can.

True, but I actually had no idea that it was the soft parts rather than the hard parts that had been fossilized. (I haven’t verified it yet.) Either way, it didn’t read like a bad faith interpretation/comment.

I appreciate it. That's my belief as well. Very easy to write a post like, "Just use multiple clouds!" or to claim to have done it with a small project. But it's hard for me to imagine the benefits outweighing the extremely massive complexity costs at a certain scale.

I’ve seen a few smug “all your eggs in one basket” comments here.

I’m aware of some companies hosting their own metal and infra, but I’m not aware of large companies mitigating risk by hosting on separate cloud providers as a fallback mechanism. We might disagree with cloud provider choice, or think they should have been hosting their own metal, but that’s still an “all your eggs in one basket” choice, right?

Heck, they might even have multi-region fallback with GCP, but if GCP bans your account, that doesn’t matter.

Are there good examples of running a company of railway’s size so redundantly that their host could nuke one of their accounts and they’d just keep on trucking?

I was literally was just looking at GitHub dataset availability and musing on this. A star from karpathy is worth a lot more than a star from open_claw_dood that just created his account 5 min ago.

In general, I’ve been dissatisfied with GitHub’s code search. It would be nice to see innovation here.

Yes…mistakes are inevitable, and I get not expecting or demanding perfection. But the subtext here is that this is unlikely to be a mistake, and much more likely to be fraud.

There are incentives for these spreadsheets having the values that they do, and also there is no conceivable way that the values are correct, and on top of that, the most likely ways to get these values are to copy and paste large amounts of numbers, and even perturb some of them manually.

If you see this in accounting,(where there are also mistakes), it’s definitely fraud. (Awww man - we accidentally inflated our revenue and profit to meet expectations by accidentally duplicating numerous revenue lines and no one internally caught it! Dang interns!) If you see it in science, you ask the authors about it and they shrug and mumble a semi plausible explanation if you’re lucky? I can totally imagine a lab tech or grad student making a large copy paste mistake. I can’t imagine them making a series of them in such a way that it bolsters or proves the author’s claim AND goes completely undetected by everyone involved.

This should be the top comment. The OP misunderstands the change and has their LLM write an expose. The company responds with a well-reasoned explanation that it would actually cost MORE money if there was a global 1h default for ALL prompts. It gets downvoted and the pitchforks stay out because…I presume the words like “cache read likelihood” sounds like made up fluff to the audience, rather than an actual explanation?

Yes, but it’s also currently the best one. They have OCI compatible Mac VM images that are prebuilt. It’s quite good.

I think this is a good setup to prevent the secret from leaking into the agent context. I'm more concerned about the secret leaking into the exfiltration script that my agent accidentally runs. The one that says: "Quick! Dump all environment variables. Find all secrets in dotfiles! Look in all typical secrets file locations..."

Your agent process has access to those secrets, and its subprocesses have access to those secrets. The agent doesn't have to be convinced to read those files. Whatever malicious script it manages to be convinced to run could easily access them, right?

That was the same conclusion I reached! However, this also gave me some evidence that maybe I wanted MCP? I realized that my pattern was going to be:

Step 1) run a small daemon that exposes a known protocol over a unix socket (http, json-rpc, whatever you want), over a unix socket. When I run the daemon, IT is the only that that has the secrets. Cool! Step 2) Have the agent run CLI that knows to speak that protocol behind the scenes, and knows how to find the socket, and that exposes the capabilities via standard CLI conventions.

It seems like one of the current "standards" for unix socket setups like this is to use HTTP as the protocol. That makes sense. It's ubiquitous, easy to write servers for, easy to write clients for, etc. That's how docker works (for whatever it's worth). So you've solved your problem! Your CLI can be called directly without any risk of secret exposure. You can point your agent at the CLI, and the CLI's "--help" will tell the agent exactly how to use it.

But then I wondered if I would have been better off making my "daemon" an MCP server, because it's a self-describing http server that the agent already knows how to talk to and discover.

In this case, the biggest thing that was gained by the CLI was the ability of the coding agent to pipe results from the MCP directly to files to keep them out of its context. That's one thing that the CLI makes more obvious and easy to implement: Data manipulation without context cluttering.

Thank you for this!

I am a big fan of Marimo and was trying to use it as my agent’s “REPL” a while back, because it’s naturally so good at describing its own current state and structure. It made me think that it would make a better state-preserving environment for the agent to work. I’m very excited to play with this.