HN user

Sherveen

34 karma

sherveen@freeagency.com https://aimuscle.com/ Startups, AI, society: http://youtube.com/@sherveenshow

Posts8
Comments25
View on HN

I don't understand this critique. (1) Did you previously think you weren't getting paid for doing what a company wants you to do, aka what THEY thought was productive? (2) Do you think all this AI generated code is useless?

Edit: y'all are some whiney folk, ain't ya?

As I said on Product Hunt (which upset Garry quite a lot) --

If he weren't the CEO of YC, this wouldn't be on PH, and it wouldn't be on HN.

This is not an impressive setup, folks. It's overengineered and deeply into its own form -- it will not make your agents better, and is likely to make it worse. There are lots of other people to follow/learn from/mimic for skills/context engineering.

Agent Skills 6 months ago

I think skills are probably a net positive for the general population, but for power users, I do recommend moving one meta layer up --

Whenever there's an agent best practice (skill) or 'pre-prompt' that you want to use all the time, turn it into a text expansion snippet so that it works no matter where you are.

As an example, I have a design 'pre-prompt' that dictates a bunch of steering for agents re: how to pick style components, typography, layout, etc. It's a few paragraphs long and I always send it alongside requests for design implementation to get way-better-than-average output.

I could turn it into a skill, but then I'd have to make sure whatever I'm using supported skills -- and install it every time or in a way that was universally seen on my system (no, symlinking doesn't really solve this).

So I use AutoHotkey (you might use Raycast, Espanso, etc) to config that every time I type '/dsn', it auto-expands into my pre-prompt snippet.

Now, no matter whether I'm using an agent on the web/cloud, in my terminal window, or in an IDE, I've memorized my most important 'pre-prompts' and they're a few seconds away.

It's anti-fragile steering by design. Call it universal skill injection.

Yeah, everyone else in the comments so far is acting emotionally, but --

As a fan and DAU of both OpenAI and the NYT, this is just a weird discovery demand and there should be another pathway for these two to move fwd in this case (NYT to get some semblance of understanding, OAI protecting end-user privacy).

Don't you think it's a little circular that you always default to assuming that their support is about regulatory capture?

Like, what if they had that opinion before they built the company? If you saw evidence of that (as is the case with Anthropic), would that convince you to reconsider your judgement? Surely, you think... some people support regulatory frameworks, some amount of the time... and unless they banned themselves from every related industry, those might be regulatory frameworks that they might one day become subject to?

You aren't taking what he said seriously. The junior could also get sick, present management issues, etc.

If this person plus a junior represented "1.3 engineering knots," he's saying... "actually, I'm still 1.3 engineering knots without him."

When this person leaves, they go find someone else who is 1.3 engineering knots. The junior represented .3, without the 1., it doesn't matter that much. Headcount strategy shifts.

Love it. I've got a lot of these sorts of micro-tools, too, CLIs, etc.

Even better when you have them all in a repo w/ an agent like Codex or Claude Code to constantly tweak/remix them as needed.

This makes a tremendous amount of sense. Most people are so bad at using AI for productive purposes -- but outside of eng, most AI fluency is actually hot garbage. People just haven't gained an understanding or appreciation for the degree of quality and capability they can achieve.

And once they can use it and get visible results, those orgs are ripe for large amounts of AI product adoption.

Only downside to the OAI version will be that it's OAI specific.

As someone who has tried almost all of the AI browsers that are accessible or in a relatively open beta, plus all the browser control frameworks and agents, I super agree with the notions behind this post.

Curious about your approach, though: so, it's a literal script, or an LLM being told to follow a deterministic script and only get subjective when necessary? Based on the blog, it looks like the former, but why not the later? Get the LLM to be pseudo-deterministic but still step-by-step it so that it can handle UI changes and adjacent interfaces.

Everyone in this thread who posts some variation of "wow love it how the government gets to decide if you get to sell your startup or how the market should work" should be handcuffed to their chair and forced to answer these 3 questions:

1. is there any role for gov't antitrust in your view of modern capitalism? 2. if there is a role, why is Adobe x Figma not the perfect example for enforcement? 3. if your answer is "Adobe clearly isn't a monopoly, look at the existence of Figma as evidence," why are you dumb?

Yup, another fun thing to do w/ this: let Claude Code talk to and control Gemini CLI, OpenCode, other CC instances, etc. in interactive mode! A different flavor of subagent. :)

Study mode 12 months ago

LLMs are vulnerable to your input because they are still computers, but you're setting it up to fail with how you've given it the problems. Humans would fail in similar ways. The only thing you've proven with this reply is that you think you're clever, but really, you are not thinking, period.

I'm kind of bothered by how many folks in the "AI influencer" space just pick up on the latest model hype, "Grok 4 changes EVERYTHING" type of nonsense.

And Grok 4 is a great example where they're just completely lying about the practical results. Elon wants to claim this is the smartest model, but it's like... 3rd or 4th best, at best.

Benchmarks, for a variety of reasons, now seem inadequate to capture models' actual strength, so I decided to run Grok 4 and o3 (and Grok 4 Heavy + o3-pro) through a gauntlet of questions that I think demonstrate real, practical differences between the two.

Hope this is helpful!

This is completely incoherent. 3 reasons:

1. he talks about what he's shipped, and yet compares it to crypto – already, you're in a contradiction as to your relative comparison – you straight up shouldn't blog if you can't conceive that these two are opposing thoughts

2. this whole refrain from people of like, "SHOW ME your enterprise codebase that includes lots of LLM code" – HELLO, people who work at private companies CANNOT just reveal their codebase to you for internet points

3. anyone who has actually used these tools has now integrated them into their daily life on the order of millions of people and billions of dollars – unless you think all CEOs are in a grand conspiracy, lying about their teams adopting AI