HN user

s3p

1,027 karma

industrial engineer sull1van(dot)com

Posts6
Comments652
View on HN

Not for me personally. While setting up a custom website, 3.5 Flash introduced tons of bugs that Claude had to fix. The website has about 3,000 lines of code spread across multiple files, and Gemini somehow couldn't do frontend changes without breaking things. Sharing my 2c, but I've stayed on GPT 5.5+ and Claude Sonnet/Opus 4.6+. Anything past that from those two have been bug-free, but Google's latest hasn't been.

That roughly comes out to $1.75 trillion for the US. That is literally their entire discretionary budget.

I don't really think the US would spend their entire 1.75tn budget on a tunnel

Hear me out on this one:

For a lot of math departments, that is exactly why they teach this. Education is rooted in application. We have entire careers that depend on certain aspects of mathematics, so most companies gatekeep that career by a degree. The degree requires the class. The student taking the class may not even be old enough to drink alcohol yet, and they can't possibly be expected to know of all the applications. Knowing and not telling them is doing them a disservice.

GPT‑Live 13 days ago

I see it a little differently here. If someone is offering a digital companion to someone who is lonely, that seems to be slightly preferable to having no connection at all.

98% isn't much 15 days ago

I mean the name was "98% isn't much" and the article made it sound like 98% isn't good enough

For pure interaction, CarPlay as a generic solution is very hard to beat infotainment systems that are deeply integrated with the vehicle itself

I can't find one instance of a car UI being as good or easier to interact with than CarPlay.

Claude Sonnet 5 22 days ago

Why did the other reply to this get flagged as dead? It was a comment about how someone would come out saying that Sonnet 5 would be better on the pelican test and therefore it has to be good. But I guess HN loves pelican SVGs so much that you're not allowed to criticize it.

Again, how would they do that?

Are they not doing what they should do, which is call for increased regulation? Last I checked, they were not able to create and enact laws.

Shame on a company for sticking to their values, I guess.

The dichotomy between Anthropic and OpenAI's treatment honestly couldn't be more obvious. OpenAI has also asked for increased AI regulation, and they've also released GPT 5.5 Cyber which is claimed to have the same vulnerability-finding abilities as Mythos. OpenAI received no such notices like Anthropic. OpenAI also received a government contract, while Anthropic was banned from DoD use.

Regardless of your thoughts about Dario or his company, this treatment is obviously not based in any rational principle, and pretending it is would be stupid.

It's only a matter of months before the open source models achieve this same capability. What is the US government going to do then? Ban all people in the world from accessing the Chinese models? If you think about these arguments for more than five minutes they really do fall flat.

if it is determined, in light of third-party assessment, to present unacceptable risks.

Yes. This assessment was made by Amazon, a frequent and serious government contractor which is generally trusted to handle high-security government, intelligence, and military contractor concerns.

Reads as partially disingenuous. Amazon did not conduct some thoroughly vetted, responsible security audit. Someone gave them examples of a 'jailbreak' and they notified the white house rather quickly. This was nary an official process. Calling it one is ignoring the facts of what happened.

That is literally not at all what this regulation does. This regulation does not pause development. This is designed to make sure non-US citizens cannot do work with/on this model. This is stupid because Anthropic never once said this is what they wanted. They said there should be a global effort undertaken for everyone to take a coordinated approach to slowing down development. They never said "please please make it impossible for Anthropic to develop models while letting everybody else develop what they want"

Siri AI 1 month ago

Yeah this is true however i will use the manufacturer app if Siri or the Home app doesn't work.

Siri AI 1 month ago

Ah, for me my key word is "turn off all my lights" and it works

Siri AI 1 month ago

All that data is lost when you migrate accounts though. I went from an old to a new 1P account and did the official way to copy (NOT exporting it to a text file and re-importing that way, actually copying it from the interface) and no version history persisted :/

Siri AI 1 month ago

Yeah but OpenAI isn't building phones. The real winning would be Google's deep integration of Gemini into all of their products.

Also, even when you DO get AI into products, consumers might not like them. The overuse of copilot led to a barrage of Microslop jokes, for example

Siri AI 1 month ago

Nope, it got markedly worse since AI. You used to be able to search for string literals, so if you remembered one obtuse phrase from a group chat, you could pull it up instantly. Now this 'intent' search will try to search for what it thinks you want, not what you typed. AFAIK, there is no way to search for literals anymore, thanks to AI.