HN user

owentbrown

12 karma
Posts2
Comments17
View on HN

I use AI to write documentation frequently. An AI agent, 600k tokens into a feature that it just authored, understands the feature as well or better than I do. It can document, compactly, what the code actually does, as well as the intent. These docs add value above having no docs and letting another AI, with zero context, read your bare code.

In the context in which I work, the quality of these docs is sufficient.

Unless your docs are being read by thousands of people, in 2026, in most contexts, handwriting docs adds less value than writing high quality instructions to your agent.

There are exceptions. For example, when kicking off a project, short crisp prose in a product brief aligns a team, removes uncertainty, and clarifies decisions. These are worth hand writing.

F3 29 days ago

Nice! The world can always use a better data format.

I think you might get some traction if you post the advantages over parquet and other files directly on the readme, so that if someone goes to https://github.com/future-file-format/f3 the see why they should try it.

Mention the advantages and post metrics. Cherry pick the metrics! There's probably a good use case for this but, from the current readme, it's not clear who should use this and why.

Gemini 3.5 Flash 2 months ago

Has anyone switched from Claude 4.7 Opus or ChatGPT 5.5 to this? How does it feel? Dumber? Worth it for the speed? I'd love someone's subjective take on it, after doing a long session of coding.

Reiner Pope gave a talk on Dwarkesh Patel about token economics. I guess faster is a lot more expensive, generally.

Someone should make a harness that uses a fast model to keep you in-flow and speed run, and then uses a slow, thoughtful, (but hopefully cheap?) model to async check the work of the faster model. Maybe even talk directly to the faster model?

Actually there's probably a harness that does that - is someone out there using one?

Claude Opus 4.7 3 months ago

Is anyone else noticing that the benchmarks for Claude 4.7 don't specify the token window? Cursor, and LiteLLM at my company, limit the token window to 200k.

It feels like to me like 4.7 is not better, and is maybe worse than 4.6 when capped to 200k context window.

Does anyone have stats on performance of 4.6 vs. 4.7 when context window is capped at 200k?

I really appreciate the author for writing this.

I learned years ago that I when I write code after 10 PM, I'm go backward instead of forward. It was easy to see, because the test just wouldn't pass, or I'd introduce several bugs that each took 30 minutes to fix.

I'm learning now that it's no different, working with agents.

Eat Real Food 7 months ago

When did white flour bread become a whole grain?

There's a picture of a loaf of bread next to the word "whole grains".

I'm excited to see these improvements but, none of them are enough to make up for inconvenience of having to start a new conversation (/clear) after every task.

I've been using Gemini Code. The larger context window is big enough to work for a full session without having to /clear. It matters. Having to think so hard and conserve tokens with Claude is problematic.

[dead] 6 years ago

Fox News disables its election probability, but leaves up the broken page.