Assuming nothing like DRM goes out of its way to prevent it, I’d be willing to bet significant money that by the time GTA6 is retro enough to require porting to “modern” platforms, precisely that sort of porting will not only be possible, it’ll be so commonplace that it won’t even merit an HN post when it happens.
HN user
apetresc
[ my public key: https://keybase.io/apetresc; my proof: https://keybase.io/apetresc/sigs/YlA9ZbsGbGbQQt7Ib8XYtagd6gFtbJZRe5qpXR4TMbY ]
---
meet.hn/city/43.4652699,-80.5222961/Waterloo
Socials: - github.com/apetresc - mastodon:fosstodon.org/@apetresc
Interests: Books, Education, Open Source, Programming
---
This isn't really security-related. The "AskUserQuestion" hook in question here is not the one that gets used for authorizing actions. That's a completely separate mechanism that is unaffected by this 60-second timer thing.
What this is referring to are those follow-up "here's two plausible alternative ways to do this, which one do you prefer?" questions you sometimes get, and usually at the beginning of a planning session when presumably you're still actively involved in the session. They get exponentially less likely as the turn goes on.
Maybe it's a good default, maybe it's not, I'll wait to pass judgment. But it's not security-related except in contrived scenarios you could construct where one side of an A-or-B UserQuestion has security implications that aren't caught by any other safeguard. I haven't ever really experienced that in practice.
How does that contradict the price increase? iPads still have RAM, yeah? If anything, not increasing the prices on iPads would undermine Macbook marketshare, would it not?
Your first Apple keynote, huh?
One of your favourite “real characters”? Do you just mean… people? Or does this mean something else I’m missing?
An em-dash is just Alt-(regular-dash) on most well-configured compose key configurations, it's not any harder.
Not so fast, it's currently 98.59%. That's technically two 9s!
This is my main question as well. Either stock or with some sort of separate hinge kit, that would completely seal the deal for me.
I can’t believe I’m nitpicking this on HN of all places, but the Taurus-Miltank thing is just fan headcanon based on nothing more than the analogy with real-life animals. While it’s true that Tauros is a male-only species and Miltank is female-only, they have separate Pokédex numbers and were introduced in different generations.
They can breed (because they’re in the same egg group) but the offspring can only ever be a Miltank. The only way to breed a Tauros is via ditto, same as with any other male-only species.
For days? Someone spent days trying to convince Claude to do something?
I've long maintained that the real indicator that AGI is imminent is that public availability stops being a thing. If you truly believed you had a superhuman, godlike mind in your thrall, renting it out for $20/month would be the last thing you would choose to do with it.
I don't think OP was making a value judgment or anything. It's just weird to say you won't consider Codeberg because you need reliability when Codeberg's uptime is at 100% and Github's is at 90%.
Windows is not public infrastructure. If the government's reliance on it has reached the level of "national importance", then that's the problem that needs to be addressed, not Windows' ownership.
Public infrastructure should be built on open-source, period.
Compared to getting them nothing, yes. But the OP's point is that this doesn't prevent the child from mentally comparing themselves to peers that have a smartphone, and viewing their Tin Can as a "restriction" imposed by their parents.
Which it is. I don't understand the need to wink-wink-nudge-nudge pretend it's anything else by the others in this thread. Just own it, restrictions aren't bad by default.
Well then clearly you haven't taken a look at https://status.claude.com.
If this was an Apple TV app, I think it would be the default mode in my household.
Improves outputs relative to what? Compared to previous contexts of 1M, it improves outputs by allowing them to exist (because previously you couldn't exceed 200K). Compared to contexts of <200K, it degrades outputs rather than improves them, but that's what you'd expect from longer contexts. It's still better than compaction, which was previously the alternative.
I don't think they're claiming "no degradation at scale", are they? They still report a 91.9->78.3 drop. That's just a better drop than everyone else (is the claim).
Those sorts of volume discounts are what you do when you're trying to incentivize more consumption. Anthropic already has more demand then they're logistically able to serve, at the moment (look at their uptime chart, it's barely even 1 9 of reliability). For them, 1 user consuming 5 units of compute is less attractive than 5 users consuming 1 unit.
They would probably implement _diminishing_-value pricing if pure pricing efficiency was their only concern.
Diametrically opposite to tokens beyond 200K being literally free? As in, you only pay for the first 200K tokens and the remaining 800K cost $0.00?
I don't think that's a fair reading of the original post at all, obviously what they meant by "no cost" was "no increase in the cost".
What percentage of the overall code was written primarily by agents?
Amazing-looking UX. My biggest feature suggestion: implement support for SSH jump hosts. I suspect having a single SSH gateway in your homelab that you configure all other hosts with ProxyJump in .ssh/config is a super-common setup amongst your target audience.
Why would a knife be a problem in a checked bag, even if it hadn't been sealed in the original package?
Yes, Discord will obviously never lose its network effect edge and get supplanted. That's simply what always happens with network effects.
Now excuse me while I go post to my Facebook about my new MSN Messenger and ICQ addresses.
Which model, 5.3 or 5.3-Codex? Yes, 5.3-Codex was announced and released. 5.3 wasn't announced. None of it is "absurd", and it also wouldn't have been "absurd" if they announce something but don't release it that same day (which they didn't do, but if they had - what exactly is absurd about that? Companies make announcements about future releases ALL the time.)
Honest question, when was the last time you caught it trying to use a command that was going to "nuke your system"?
I mean… yeah? It sounds biased or whatever, but if you actually experience all the frontier models for yourself, the conclusion that Opus just has something the others don’t is inescapable.
Scott Alexander essentially provided editing and promotion for AI 2027 (and did a great job of it, I might add). Are you unaware of the actual researchers behind the forecasting/modelling work behind it, and you thought it was actually all done by a blogger? Or are you just being dismissive for fun?
Why is it absurd?
Impressive that they publish and acknowledge the (tiny, but existent) drop in performance on SWE-Bench Verified between Opus 4.5 to 4.6. Obviously such a small drop in a single benchmark is not that meaningful, especially if it doesn't test the specific focus areas of this release (which seem to be focused around managing larger context).
But considering how SWE-Bench Verified seems to be the tech press' favourite benchmark to cite, it's surprising that they didn't try to confound the inevitable "Opus 4.6 Releases With Disappointing 0.1% DROP on SWE-Bench Verified" headlines.