HN user

viccis

2,374 karma
Posts0
Comments810
View on HN
No posts found.

If the human solved a significant open problem without relying on AI, don't you think they'd have claimed the credit for themselves?

You'd likely make a lot more money if Anthropic paid you a couple million to release it as a Fable discovery. For that matter, if "I solved it" bragging rights and CV candy is that important, then why not just publish the counterexample without mentioning Fable?

I was initially a bit skeptical, as it seems very convenient that this discovery came from an employee of the company that was able to take credit for it. When I saw no real prompt was published, that was even more suspicious. Some recent mathematical discoveries included the prompt and session [1]. Why not include the prompt? That and the thought the model used could be incredibly valuable information for the field of mathematics.

I'm not even saying Fable didn't do it, and my attitude might be different with other companies, but Anthropic has so much obvious nonsense marketing going on (their AI welfare experts for example), that I think the burden of proof should be expected to be on them here.

1: https://chatgpt.com/share/6a55aa50-b484-83ea-85c0-c7e7b4bda4...

"Supporting audiobooks" involves more than just playing audio files, unless you want to ignore the basic set of features associated with audiobooks as their own medium.

You're inventing a person to get mad at. I don't see any evidence that this founder holds those kind of views.

Easily the biggest delta between a good one and a bad one I've ever encountered. It's so easy to test too. Just ask a product manager if they can give a talk on <some product they own> in 2 hours. If they respond by slacking all of the engineers that work under them to stop what they're doing and give them the info to give that talk, fire them. If they are ready to talk right now then give them a raise.

From what I can tell, I have a similar workflow. I get Fable 5 / Sol on High to talk to me about requirements until it's ready to design. I then have it break the design into tickets. I use kata, an agent-oriented issue tracker. If it ever starts to get bloated, I'll just make my own as it doesn't need too many features. Anyway, I tell it to include sufficient context in each such ticket to be picked up by a new implementation agent. Explicit user stories, Cucumber style acceptance criteria, etc.

When it's done, I switch to a smaller model and tell it to start a /goal of calling `kata ready` to get tickets ready to be worked. Work one ticket on each goal iteration, committing changes when done. Stop when all the tickets are either closed out or are blocked on actions from me.

It works fantastically. I can get entire (relatively straightforward) iOS apps done in under my $20/month five hour session window.

I get even better results if I do the QA session with the frontier model with two output artifacts: an implementation spec and a design prompt for Claude Design. I push the design prompt through Claude Design, tweak the results, and get a design spec. When I have the frontier model do the planning, I have it read the implementation spec from before, along with the design spec's overview file.

When I do it this way, I've done A/B tests between Fable 5 and Sol, and the apps they wrote were basically identical. It's so much cheaper than the "let loose the subagent fleet!" form of context management.

I haven't even added any kind of frontier model validation cycle into this loop yet. Everyone keeps talking about a subagent flow in which the big boy reviews the work of the drones, but I've found that if it encodes its acceptance criteria well enough, they do a satisfactory job of it themselves.

I don't see how that particular political expression is a risk to their VPN's stance. This is a lot of nothing; would have been better to ignore the screeching from Bluesky.

Not really, this looks more like pre-1990s leftism that you see from people like Caesar Chavez. Broadly redistributive policies towards the working class and a strict curbing of immigration as a source of exploited labor underclass for corporations to fatten their margins. There's a reason Bernie called open borders a "Koch brothers proposal".

We're rediscovering the "architect" job role a decade or so after switching to agile processes and staff engineers largely replaced it. God help us when they rediscover UML for agents.

Blender 5.2 LTS 3 days ago

I started pre-1.8, in the C-key era. Haven't used it in many years so looking at it every now and then is genuinely shocking to me how much it's progressed. I remember a family member who worked at a triple A game dev studio giving me a tour and when I talked to the 3D modeler guys there and told them I used Blender they stifled a chuckle. To be fair, it was a very rudimentary tool back with UI flows invented by and for aliens. And now it's used all over the place in both amateur and professional settings.

At some point, narcissistic injuries in response to "rtfm" or "stfw" got so prominent that just telling people asking others to spend their time explaining things that were 5 second google searches away to google it became a faux pas

Makes me nostalgic for the old days (2000s and early 2010s) when everything was WordWord. Facebook, Instagram, Myspace, Dropbox, etc. Now it's just normal words without vowels. The Grindrification of naming.

I love this author's article about saving half a million dollars with a click from a while back. Nikhil, if you're reading this, I have a decent war story about a similar situation (I was lucky enough that there were two such things, so I saved a full million a year and was still denied a $15k raise) and so I was really entertained at your post from a few years back. There are actually a lot of lessons to be learned about corporate politics there, about how you can save someone a million a year in perpetuity and promise to do that again next year (which I could have!) and see them still refuse to pay you a single extra cent.

Codex Resets 4 days ago

Looks like no Fable for Pro. Will likely be cancelling my Pro account.

When you've earned your opinions about architecture and code quality the hard way, they feel less like textbook rules and more like scar tissue.

I don't think it's common for any compsci programs to (competently at least) teach architecture and code quality.

The honest truth is that in the last few months, there have been days when I have spent close to two full days writing a plan for an LLM to execute: obsessively clarifying, specifying, re-specifying, only to have it still do something inexplicably stupid.

It's because LLMs are actually taking us back in time to the pre-agile days where there was a career path (architect) that involved almost nothing but painstaking spec authoring and endless meetings to review and course correct the work of the engineers whose job was to implement what you designed as closely as possible. I have to emphasize that this was a different career path than what we think of as a senior engineer today. Not everyone likes this.

Pseudpocalypse 6 days ago

Yeah, pseudpocalypse would be Substack going out of business or something similar.

That's the saving grace of it for me tbh

AI art shows the philistines for who they are. It's a giant decoy to lure them away from spaces where people are making art of more substance.

I feel the same way about things like the absolutely miserable quality of Marvel, Star Wars, etc. stuff leading to films being increasingly polarized between 60 IQ blockbusters and more interesting small projects. It's what kicked off the 70s arthouse cinema in the US, and I'm hoping we get a similar renaissance soon with the success of all of the small and difference projects lately.

doing concerts where the attendees are just other musicians

If this is how jazz musicians actually made money, they'd have all died of starvation at this point. Every jazz musician that's not a top 20 or so act makes most of their money playing whatever kind of music people pay them to make, whether it's church music, session musician gigs, etc. The entire reason they started doing (and still do!) straightahead jams is that they were tired of playing schlock all day and wanted to spice it up with and for other people tired of playing schlock.

Andy Warhol

Given that the man made his living doing stunts on top of the corpse of modernity, I'm not surprised he would love anti-human stuff like AI music. He famously said he wished he could be a machine. I don't think appealing to his nihilistic embrace of the spectacle is persuasive.

Yeah I had it remake my favorite TI-89 graphing calculator game in Python and it one shotted it perfectly.

If only a cynical Frenchman had written a book critiquing peoples' tendency towards simulating things that don't exist.

Neocities is cool, but the medium is the message and we've generally moved on from this (treasured!) past. Any attempt to replicate it tends to wind up hyperreal and forced.

Same for me. I never use them. I use Fable on highest effort to plan things and then record the plan in tickets. I use Kata, which is CLI and agent oriented, but I suppose Jira or other systems would work too. I tell it to put enough context in each ticket to on-board a fresh coding agent to implement it. Then I just do /goal, telling to to run `kata ready` to get new tickets to work and continue until they're all closed according to acceptance criteria or until they're blocked on actions from me. I need to play around with getting it to switch to smaller models (or spawning 1 subagent) to do ticket implementation and then auto compact after each. Either way, it results in really easy workflows and uses very few tokens compared to the built in subagent flows that doing this completely avoids.

AI 2040: Plan A 12 days ago

No, in fact, most of us could not have just shown up and bid on it with any expectation of a meaningful outcome.