title literally says "command line".
HN user
behnamoh
A CLI tool that's not grep-able is a GUI disguised as TUI.
Meh, haven't missed radio in ages. If I have no control over the content, then the medium isn't for me.
Nope, I'll still buy Claude because the overall XP is better than Kimi and Qwen who literally copied basic harnesses to make kimi-cli and qwen-cli, respectively.
Also, you can tell if a model is genuinely powerful and well-thought-out vs a model that acts like it.
It's like Apple vs Xiaomi/Huawei. Sure, you can get a Huawei with bells and whistles, but most people learnt the hard way that those companies just copy the iPhone, so might as well get the real deal.
No, it's because they wanted a unified pipeline in claude code, claude app, and their website. All of them use more or less the same claude features (claude code has artifacts, claude website has "Ask User a Question" tool, etc.).
Much less fragmented compared to "write the app 3 times in 3 different languages".
The fact there is no way to disable auto-compaction and no way to go back in the conversation history to before a compact makes codex a no-go for me
Same. I've told @tibo many times about these issues but apparently OpenAI's focus is on features that buy them the most users ASAP instead of increasing harness quality.
It was explained to me that generally Chinese firms will compete hard and maximize revenue above all, whereas western firms tend to focus on profit.
ok but why?
How long is this gonna last? OpenAI obviously is doing it to gain mindshare but they're burning money before their IPO and can't sustain these resets.
They've bought a lot of dev goodwill tho, which matters I guess.
The entire universe seems to be inside a giant black hole, anyway, and the more it goes, the more evidence is found to support that. Might as well find a black hole and visit other universes than explore our own.
literally no one owes you anything, has nothing to do with age. You want open weight models? Go build one, but don't expect companies to do it for you because you're special.
Because a harness doesn't just "drive" the LLM. e.g., there's code in claude code that detects if the user's prompt shows they're angry, and they react to those prompts differently. (they use regex on "wtf", etc.!)
As I write this, I'm disputing a charge from Starlink because when I activated mine, they automatically started the free trial for me, but nowhere in the receipt or on their website did they say they would switch me to the most expensive plan after the free trial was over. And they made it extremely confusing to cancel my free trial. So I ended up being charged $120 for something that I had just opened out of the box. I disputed it twice through Apple Card and it still got rejected, even though I have the invoices. So I'm disputing it a third time, and I know one thing is clear: I am never going to purchase anything that Elon has made ever again. They also lied to me about the cost of keeping the Starlink. Their app said the satellite would need to stay "alive," and to do so I had to pay $5 a month. But then I recently realized that that was also a lie—you don't have to pay to keep the service or the satellite alive.
Im hoping Apple gets the new Siri working better on older phones.
Apple would never do that, if anything they did not offer their Siri with the most advanced AI on iPhone 16 Pro Max, which is one year-old only.
Parakeet isn't as good as whisper large.
In a world where you say "tmux" and Apple's VTT writes "T Max".
Whisper and GPT-4o (for diarization).
Still nothing beats OpenAI's VTT. Anthropic's sucks and Apple's isn't even usable.
Edit: Getting downvoted by Apple fanboys for telling the truth is a badge of honor.
I think a mandatory first thing for any engineer is to learn, understand and commit for life to the Ethics of their profession. It's a shame all these very picky recruitment processes and 'culture' of these giant companies didn't care about ethics and morality.
For some reason, the ethics followed by Asians, especially the Chinese are not fully compatible with the ethics of the west. Sometimes Chinese people call it being smart to circumvent or bypass the rules, something that would be called cheating in the west.
Codex is faster but you always have to correct it because it got something wrong
this has been my experience with Codex as well, and I have to fix its mistakes every single time. But recently, I literally threw away three hours of work because it kept adding hundreds of lines to my code base. When I restarted the entire work using Fable and Opus, it was like night and day.
I'm curious, is this true or something you heard from MSM and regurgitated?
choose one:
brutally honest vs kindness at all costsInteresting, but how do they "combine" the results of all those parallel agents? How do they know which parts of each agent response is signal vs noise?
how do you know gpt-5.5-pro is an ensemble? if it is, then how did OpenAI do it? why no other company has been able to pull it off?
How? I'm pretty much locked into Claude Code and even if gpt models are good now, the experience with codex CLI has been so bad I won't go back to it.
e.g., it still doesn't have /revise or /undo!
OpenAI already has a Mythos level model, it's called GPTCyber and before that, it was called gpt-5.5-pro.
On macOS I've been using piper (https://github.com/OHF-Voice/piper1-gpl) to announce claude code notifications and it works perfectly!
Fable is the first model that mostly writes without the AI slop format for me
I'm not that impressed by Fable's writing to be honest, still has the AI giveaways like em dash.
I still don't know why OpenAI doesn't put gpt-5.5-pro in Codex. It's one hell of a model and easily parallels Fable/Mythos. Sure, it'll use up your quota much faster but that's the price some users are willing to pay for absolutely high quality responses.
I think gpt-5.5-pro runs 12x parallel gpt-5.5 agents behind the scene and uses OpenAI's secret sauce to synthesize their answers into one insanely good response.
Tokenmaxxing was never a thing to begin with. Just because a few companies did it doesn't mean it was a widespread phenomenon.
What's wrong with Claude? I've asked it to analyze images and even Opus 4 would perfect nail it.