HN user

sarreph

3,938 karma

I make software, images (sometimes moving), and music.

Today, my focus is on personalizing the internet at Kenobi.ai (YC W22)

Reach me at my name (Rory) and my company’s domain.

Posts222
Comments595
View on HN
pokayoke.codes 6d ago

Show HN: Pokayoke – turn code conventions into checks for agents

sarreph
3pts0
harpist.site 9d ago

Show HN: Harpist – convert any website into a refined and documented API

sarreph
2pts0
developers.cloudflare.com 19d ago

Cloudflare Wrangler CLI auth profiles support

sarreph
2pts0
pokayoke.codes 21d ago

Show HN: Pokayoke – turn repo conventions into deterministic checks for agents

sarreph
1pts0
machinelearning.apple.com 28d ago

The "Super Weight:" How a Single Param Can Determine an LLM's Behavior (2025)

sarreph
2pts0
pokayoke.codes 1mo ago

Show HN: Pokayoke – deterministic guardrails for agentic coding

sarreph
3pts0
www.bbc.co.uk 1mo ago

Why is Lidl opening a pub?

sarreph
5pts0
en.wikipedia.org 2mo ago

Yabasic (Yet Another Basic)

sarreph
18pts0
www.whatwesee.space 2mo ago

Show HN: An AI generated, online art exhibition

sarreph
1pts0
www.whatwesee.space 5mo ago

Show HN: What We See. An AI generated art exhibition

sarreph
1pts0
www.ycombinator.com 6mo ago

New YC homepage

sarreph
298pts157
news.ycombinator.com 7mo ago

Launch HN: Kenobi (YC W22) – Personalize your website for every visitor

sarreph
46pts56
news.ycombinator.com 7mo ago

Show HN: Kenobi – AI personalized website content for every visitor

sarreph
7pts2
blog.cloudflare.com 10mo ago

A Lookback at Workers Launchpad

sarreph
1pts0
www.theguardian.com 10mo ago

350k UK households on low-interest fixed-rate mortgages set to increase

sarreph
2pts0
www.theguardian.com 11mo ago

Deal to get ChatGPT Plus for whole of UK

sarreph
3pts0
www.bbc.co.uk 11mo ago

UK interest rates cut to lowest level in more than two years

sarreph
3pts0
en.wikipedia.org 11mo ago

Fast and Slow

sarreph
2pts0
www.bbc.co.uk 1y ago

Fake-will fraudsters steal millions from the dead

sarreph
9pts1
www.bbc.co.uk 1y ago

UK tyres meant for recycling sent to furnaces in India

sarreph
9pts0
www.bbc.com 1y ago

12th monkey dies in Hong Kong zoo amid bacterial outbreak

sarreph
3pts0
vercel.com 1y ago

Vercel Is Down

sarreph
28pts12
www.bbc.co.uk 2y ago

UK economy grew slightly in February

sarreph
1pts0
vercel.com 2y ago

Vercel: Improved Infrastructure Pricing

sarreph
48pts32
www.bbc.co.uk 2y ago

Questions raised over Temu cash 'giveaway' offer

sarreph
1pts0
www.creativebloq.com 2y ago

Never forget that utterly ridiculous Pepsi logo design document (2023)

sarreph
18pts27
www.bbc.co.uk 3y ago

UK inflation falls to lowest level (7.9%) in more than a year

sarreph
22pts37
www.youtube.com 3y ago

ChatGPT NPC coaches me talking to people at a party in VR [video]

sarreph
1pts0
labs.google.com 3y ago

Google Search Labs

sarreph
2pts0
www.bbc.co.uk 3y ago

Emergency alert test fails to sound on some phones

sarreph
2pts0

I used to be a die-hard destructurer! I liked the readability aspect, but ultimately two core problems emerged that pushed me towards dot-accessors:

1. It's highly verbose, especially in a language like TypeScript where you're often defining properties in types and then destructuring the same properties, making the whole thing look quite duplicative.

2. If you're passing an object through different functions doing similar things, I would quite often have to rename a destructured property in TypeScript because I would want to reuse the same name in a function-scoped variable.

I think for exhaustive switches, destructuring can be fine, but in general as long as you've got a well-defined set of types backing everything up, dot-access is neater.

--

Also on the point about optional parameters:

let's suppose some stations now have an anemometer and are able to record wind speed

I find that discriminated unions are so much better at dealing with these kinds of problems. Also tends to work better with dot-access because you're not then destructuring undefined properties.

That's actually the wikipedia article that inspired me to name a repo-convention checker I made called Pokayoke[0]. Do find it fascinating (as with Andon) how many cool Japanese loan-words there are for process refinement.

[0] https://pokayoke.codes

https://dozenal.game

A daily puzzle game called Dozenal that I've been making with a friend. We've been increasing our user base over the past couple of months and are still trying to refine the learning curve.

If you like number puzzle games, I would be very keen for you to give it a go and to hear your feedback on it!

From the article, it's for "jobs in sectors like agriculture and construction". Would be interesting to learn about how this kind of work is managed in hotter climates.

For office work, a lot of European countries (especially the UK) haven't invested in AC as much as the rest of the world because they haven't needed to. This is especially apparent in housing, where working from home is becoming difficult in these higher temperatures.

Sure:

he was essentially groomed from a young age

It's one thing to choose a poor work-life balance for oneself; a different thing entirely to demand it of others

Jarred was a stinky manager

It's hard, in my opinion, to lend credence to the author here when they decided to devote the first and largest section of their article to an incisive display of speculative ad hominem.

Would have been a great opportunity to outline the benefits of Zig! I've been keen to pick Zig up recently due to mitchellh's evangelism and inspiring writing on the subject.

This article puts me off learning Zig.

481 upvotes on HN, and only $136 USD donated (out of $64k target) -- at the time of writing.

Given the amount of traffic this project has received by being at the top of the front page for half a day, one has to wonder if a different approach to soliciting donations would have yielded them more money.

Clearly, everyone here is at least interested in the idea of a .self domain, and I wager that most (even the naysayers) of the commenters would register theirs.

Imagine if instead of asking for a $15–125 donation behind a CTA, they asked for $2 to "pre-register" your domain (with higher tiers for more benefits). I have a feeling they would have raised a lot more money...

I disagree if your application is networked. Most SaaS is built on RESTful APIs that can be converted trivially into interfaces / contracts for tool use.

The author alludes to it but the defence to this is seemingly insurmountable at the moment because we’re ostensibly operating LLMs on a single channel — their inner, subconscious voice. Right?

Interacting with an LLM is a bit like seeing the output of an Inside Out (the Disney movie) scene. Or it’s a bit like a human brain that we’re providing tool call access and introspection with some kind of advanced neuralink.

But - like the author says - _we know_ our inside voice from the outside world, because we’re embodied.

Is there something we can do here by attempting to bifurcate internal and external systems? Like a conscious and subconscious stream of information on two separate bands?

If the model somehow knew its User was not it because it was clearly an external signal, then the attack documented here would be about as effective as a Jedi mind trick without the Force.

I think something that is a mix between localStorage or IndexedDB and access to the user's filesystem would be better.

I agree with the comments about how much of a security risk this poses. But, isn't that the case with any binary or executable files and apps we download from the internet anyway?

It would be cool if you could have a specially-demarcated directory (e.g. even inside the application like `~/Applications/Chrome/<website>/local_files`) which you can just open super easily with a button from Chrome, and just copy files over into that directory as needed. Would provide the benefits of a more secure enclave with the flexibility of being on the filesystem.

Claude Fable 5 1 month ago

I had intended to caveat that: I'm sure I'm not the first person to ask about this!

you still see improvements

This is expected if they are training their models on it, right?

objectively-bad results

Keen to learn when this has been the case, i.e. across version increments in major models.

Claude Fable 5 1 month ago

I'm beginning to wonder how much of a useful metric the pelican is because surely the frontier labs must be training their models on pelican-artistry because of how well known your test is now?

I agree with the author that -- right now -- we're still in the part of the AI adoption / product development curve that it's an extreme force multiplier.

I like to think of it as a normal distribution, the further away a programmer is to the right of the mean, the more their benefit. It's almost like it's their standard deviation squared (σ²). So someone like Matt Perry (as OP mentioned), who is a >99.99% programmer for argument's sake and is therefore four standard deviations away from the mean... Matt gets a (4×4) 16x multiplying effect on their productivity.

Someone who is a slightly above average programmer might see a 2 or 3x boost on their productivity, which is huge(!) and might also make them fear for their job. Which tracks with the level of moral panic we are seeing and experiencing. This math kinda still holds up for "bad programmers" too (i.e. left of the mean), as in they still see a boost to their productivity (negative squared is a positive number)... but there's something iffy about their results. The technical debt is unmaintainable and because they don't _understand_ the systems that they're operating in, they end up in the "3 hour" prompt loops that the OP refers to.

Similarly, if Matt Perry handed me the keys to the Motion repository and told me to take over, I wouldn’t have the same results even though I have access to the same set of LLM tools.

The question is -- how long is this multiplier going to exist for? Some people would wager "for the foreseeable long-term future"; some people think it will widen further; and some people think it will diminish or god forbid even collapse. It feels like most arguments at the moment (like this article's) are that the humans who "know what they are doing" will be able to baton the hatches and avoid being usurped by ever-capable models. I saw it in a café yesterday: someone was using a coding agent to build a marketing website for their project, getting more and more frustrated by not getting the outcome they wanted. Their friend typed a couple of sentences on their keyboard and got a "Dude! How did you do that? That was sick!" a minute or so later. "I used to build websites" the friend said. -- The friend 'knew what they were doing'.

How much longer is knowing what you're doing going to be a moat?

I think this is one of the key takes right now. I too have similar experience.

Which way is it going to go?

i) “Seniors” also get superseded by even more capable models that can do all of the things which currently require experience.

ii) Linguistics become the new higher order abstraction (English is the new high-level programming language) _but_ there are different / orthogonal ways of approaching software development than the way we do things now — which “juniors” become more adept at more quickly.

It is a pity that you can't make an experimental commit on an experimental branch without igniting a fire of delirium through some people who -- if they were able to put their emotional response aside for a minute and could weigh this up on the basis of merit -- would probably agree with the motivations for researching this approach.

if/how hard it’d be to get it to pass Bun’s test suite and be maintainable

Every month brings new opportunities to completely abstract the process of porting code with agents, all using linguistics. What an exciting time.

For those looking for a similarly interesting (and interestingly similar) example, see Cloudflare's port of Next.js[0], "vinext", from a couple of months ago. It had some teething problems at the start but I'm using it in a few production projects now with minimal issues.

[0] https://github.com/cloudflare/vinext

I've been working on an adjacent problem (extracting website branding data from a URL) for the past year, and previously had to rely on procedural techniques such as these -- props to the author!

However, models are now getting to the point where we are starting to learn the bitter lesson[0] even with stuff like color-palette generation. Nano Banana 2 [gemini-3.1-flash-image-preview] especially is adept at performing arbitrary operations on images. Before then, you would have to use a model such as Gemini Flash to perform segmentation[1] and then post-analyze those segments.

Here's a prompt I used with Nano Banana 2 in AI Studio

Derive a coherent, designer's color palette from this image alone.

Provide 5 distinct HEX color codes as your response.

[Attachment == the picture of the car, first in the author's article] [Settings: Output .. images & text; Thinking level .. minimal]

Response:

I have extracted five distinct hex color codes directly from the key elements in this image, representing the colorful facade and the vintage car:

#FF96C5 (The main pink wall)

#38C6F1 (The light blue car)

#AEF6A5 (The green wall)

#E51988 (The dark pink trim and railing)

#5F432B (The dark wood of the door and windows)

And they all pretty-much check out. Not hyper-accurate, but really not far off anymore. I didn't even have to try!

[0] https://en.wikipedia.org/wiki/Bitter_lesson [1] - https://ai.google.dev/gemini-api/docs/image-understanding#se...

The bakery example is interesting, because it's presented as "both sides have been working on this thing and think they should get 50%"... and then the _solution_ is "A path back to 50% for Daniel" -- who gets an objectively worse deal than his disputant.

It's definitely an interesting application of LLMs, the output text to me reads very GPT-ey, with the punctuated and concise phrasing.

I am fed up of getting gaslit by coding assistants. "Your AI agent says it's done." really is a problem! Nice packaging here.

I built something similar[0] a few months ago but haven't maintained it because Codex UI and Cursor have _reasonable_ tooling for this themselves now IMO.

That said there is still a way to go, and space for something with more comprehensive interactivity + comparison.

[0] https://magiceyes.dev/