HN user

apetresc

5,817 karma

[ my public key: https://keybase.io/apetresc; my proof: https://keybase.io/apetresc/sigs/YlA9ZbsGbGbQQt7Ib8XYtagd6gFtbJZRe5qpXR4TMbY ]

---

meet.hn/city/43.4652699,-80.5222961/Waterloo

Socials: - github.com/apetresc - mastodon:fosstodon.org/@apetresc

Interests: Books, Education, Open Source, Programming

---

Posts79
Comments943
View on HN
status.claude.com 4mo ago

Claude Scheduled Tasks caught in infinite loop after Daylight Savings Time

apetresc
4pts0
publicdomainreview.org 6mo ago

Happy Public Domain Day 2026

apetresc
430pts85
jvanelteren.github.io 1y ago

Advent of Code Stats Analysis

apetresc
1pts0
www.adventofchess.com 1y ago

Advent of Chess 2024

apetresc
2pts1
old.reddit.com 3y ago

Happy SDXL Leak Day

apetresc
3pts1
twitter.com 3y ago

Google Colab has started banning Stable Diffusion WebUI users

apetresc
21pts10
blog.unity.com 3y ago

Parsec is now part of Unity

apetresc
2pts0
odysee.com 3y ago

SEC vs. LBRY Summary Judgement Ruling (We Lost)

apetresc
9pts1
www.alignmentforum.org 3y ago

A Mechanistic Interpretability Analysis of Grokking

apetresc
1pts0
www.docker.com 4y ago

The Magic of Docker Desktop Is Now Available on Linux

apetresc
4pts1
www.chess.com 4y ago

Chess.com Releases New Classroom Feature

apetresc
2pts0
www.thegamer.com 4y ago

Squid Game Smuggler Has Been Sentenced to Death in North Korea

apetresc
35pts10
www.nytimes.com 4y ago

Magnus Inc.: The Business of Being World Chess Champion

apetresc
3pts0
reclaimthenet.org 5y ago

LinkedIn censors Swedish journalist’s profile in China

apetresc
1pts0
www.audacityteam.org 5y ago

Audacity and MuseScore Announcement

apetresc
38pts7
github.blog 5y ago

Updates to Our Terms of Service and Privacy Statement

apetresc
4pts0
medium.com 5y ago

Impact of Go AI on the professional Go world

apetresc
161pts89
www.lesswrong.com 6y ago

Classifying Games Like the Prisoner's Dilemma

apetresc
1pts0
www.scottaaronson.com 6y ago

MIP*=Re

apetresc
2pts0
www.serenityos.org 6y ago

SerenityOS: From Zero to HTML in a Year

apetresc
7pts0
linuxscreenshots.thecodingstudio.ca 7y ago

12 Years of Ubuntu Screenshot Tours

apetresc
1pts0
www.businessinsider.com 7y ago

Wal-Mart Testing a Burger-Flipping Robot

apetresc
1pts0
arxiv.org 7y ago

Quantum-inspired classical algorithms for PCA and supervised clustering

apetresc
3pts0
blog.twitch.tv 7y ago

Changes to Twitch Prime

apetresc
4pts0
arstechnica.com 8y ago

Epic ups Unreal Marketplace creators’ pay well above industry standard

apetresc
2pts1
andreascarpino.it 8y ago

How My Car Insurance Exposed My Position

apetresc
4pts1
lwn.net 8y ago

Linux 4.15-rc7

apetresc
2pts0
www.inc.com 8y ago

Uber's New CEO Just Sent an Amazing Email to Employees

apetresc
3pts0
blog.martin-graesslin.com 8y ago

Announcing the XFree KWin Project

apetresc
3pts0
time.com 9y ago

Maluuba Researchers Break Ms. Pac-Man Record

apetresc
3pts0

This isn't really security-related. The "AskUserQuestion" hook in question here is not the one that gets used for authorizing actions. That's a completely separate mechanism that is unaffected by this 60-second timer thing.

What this is referring to are those follow-up "here's two plausible alternative ways to do this, which one do you prefer?" questions you sometimes get, and usually at the beginning of a planning session when presumably you're still actively involved in the session. They get exponentially less likely as the turn goes on.

Maybe it's a good default, maybe it's not, I'll wait to pass judgment. But it's not security-related except in contrived scenarios you could construct where one side of an A-or-B UserQuestion has security implications that aren't caught by any other safeguard. I haven't ever really experienced that in practice.

How does that contradict the price increase? iPads still have RAM, yeah? If anything, not increasing the prices on iPads would undermine Macbook marketshare, would it not?

This is my main question as well. Either stock or with some sort of separate hinge kit, that would completely seal the deal for me.

I can’t believe I’m nitpicking this on HN of all places, but the Taurus-Miltank thing is just fan headcanon based on nothing more than the analogy with real-life animals. While it’s true that Tauros is a male-only species and Miltank is female-only, they have separate Pokédex numbers and were introduced in different generations.

They can breed (because they’re in the same egg group) but the offspring can only ever be a Miltank. The only way to breed a Tauros is via ditto, same as with any other male-only species.

I've long maintained that the real indicator that AGI is imminent is that public availability stops being a thing. If you truly believed you had a superhuman, godlike mind in your thrall, renting it out for $20/month would be the last thing you would choose to do with it.

Windows is not public infrastructure. If the government's reliance on it has reached the level of "national importance", then that's the problem that needs to be addressed, not Windows' ownership.

Public infrastructure should be built on open-source, period.

Compared to getting them nothing, yes. But the OP's point is that this doesn't prevent the child from mentally comparing themselves to peers that have a smartphone, and viewing their Tin Can as a "restriction" imposed by their parents.

Which it is. I don't understand the need to wink-wink-nudge-nudge pretend it's anything else by the others in this thread. Just own it, restrictions aren't bad by default.

Improves outputs relative to what? Compared to previous contexts of 1M, it improves outputs by allowing them to exist (because previously you couldn't exceed 200K). Compared to contexts of <200K, it degrades outputs rather than improves them, but that's what you'd expect from longer contexts. It's still better than compaction, which was previously the alternative.

Those sorts of volume discounts are what you do when you're trying to incentivize more consumption. Anthropic already has more demand then they're logistically able to serve, at the moment (look at their uptime chart, it's barely even 1 9 of reliability). For them, 1 user consuming 5 units of compute is less attractive than 5 users consuming 1 unit.

They would probably implement _diminishing_-value pricing if pure pricing efficiency was their only concern.

GPT-5.4 5 months ago

Diametrically opposite to tokens beyond 200K being literally free? As in, you only pay for the first 200K tokens and the remaining 800K cost $0.00?

I don't think that's a fair reading of the original post at all, obviously what they meant by "no cost" was "no increase in the cost".

Amazing-looking UX. My biggest feature suggestion: implement support for SSH jump hosts. I suspect having a single SSH gateway in your homelab that you configure all other hosts with ProxyJump in .ssh/config is a super-common setup amongst your target audience.

Yes, Discord will obviously never lose its network effect edge and get supplanted. That's simply what always happens with network effects.

Now excuse me while I go post to my Facebook about my new MSN Messenger and ICQ addresses.

GPT-5.3-Codex 6 months ago

Which model, 5.3 or 5.3-Codex? Yes, 5.3-Codex was announced and released. 5.3 wasn't announced. None of it is "absurd", and it also wouldn't have been "absurd" if they announce something but don't release it that same day (which they didn't do, but if they had - what exactly is absurd about that? Companies make announcements about future releases ALL the time.)

Honest question, when was the last time you caught it trying to use a command that was going to "nuke your system"?

GPT-5.3-Codex 6 months ago

I mean… yeah? It sounds biased or whatever, but if you actually experience all the frontier models for yourself, the conclusion that Opus just has something the others don’t is inescapable.

GPT-5.3-Codex 6 months ago

Scott Alexander essentially provided editing and promotion for AI 2027 (and did a great job of it, I might add). Are you unaware of the actual researchers behind the forecasting/modelling work behind it, and you thought it was actually all done by a blogger? Or are you just being dismissive for fun?

Claude Opus 4.6 6 months ago

Impressive that they publish and acknowledge the (tiny, but existent) drop in performance on SWE-Bench Verified between Opus 4.5 to 4.6. Obviously such a small drop in a single benchmark is not that meaningful, especially if it doesn't test the specific focus areas of this release (which seem to be focused around managing larger context).

But considering how SWE-Bench Verified seems to be the tech press' favourite benchmark to cite, it's surprising that they didn't try to confound the inevitable "Opus 4.6 Releases With Disappointing 0.1% DROP on SWE-Bench Verified" headlines.