HN user

cwyers

12,226 karma

[ my public key: https://keybase.io/colinwyers; my proof: https://keybase.io/colinwyers/sigs/oY3s_sY1T5jtSxebr0LrQSuJ8_6mFLfNR5_RjA_D6yU ]

Posts26
Comments2,863
View on HN
www.hillelwayne.com 5y ago

Are We Engineers?

cwyers
3pts2
twitter.com 7y ago

“Twitter has an algorithm that creates harassment all by itself”

cwyers
164pts76
maggie-stiefvater.tumblr.com 7y ago

A Story About Piracy

cwyers
106pts103
gizmodo.com 7y ago

MoviePass Is Now Re-Enrolling Former Customers Unless They 'Opt Out'

cwyers
3pts0
www.roughtype.com 9y ago

Digital sharecropping (2007)

cwyers
1pts0
azure.microsoft.com 9y ago

Azure IP Advantage

cwyers
2pts0
steve-yegge.blogspot.com 9y ago

The Monkey and the Apple

cwyers
2pts1
www.youtube.com 9y ago

A two-year-old's solution to the trolley problem

cwyers
20pts4
monty-says.blogspot.com 9y ago

Applying the Business Source Licensing (BSL)

cwyers
1pts0
kotaku.com 10y ago

Pokémon Go Could Be a Death Sentence for a Black Man

cwyers
4pts1
www.polygon.com 10y ago

Minecraft tops 100M copies sold, 40M people play every month

cwyers
1pts0
andrewgelman.com 10y ago

Gelman on the reproducibility crisis in social science

cwyers
1pts0
kotaku.com 10y ago

Fake Minecraft Sequel Tops iTunes Charts

cwyers
2pts0
blogs.windows.com 10y ago

Windows 10 in China

cwyers
1pts0
www.percona.com 10y ago

Percona Toolkit and systemd

cwyers
2pts1
chicago.suntimes.com 11y ago

Get ready to pay more for Netflix: Chicago sneaks “Cloud Tax” onto books

cwyers
5pts1
support.google.com 11y ago

“Google Wallet for digital goods” Retirement

cwyers
1pts0
medium.com 11y ago

Mordor, We Wrote

cwyers
2pts0
www.theverge.com 11y ago

If we want to reduce police brutality we have to end the war on drugs

cwyers
6pts0
www.polygon.com 11y ago

Threes dev wants to make strategy games less sadomasochistic

cwyers
1pts0
lists.debian.org 11y ago

Results for init system coupling

cwyers
35pts6
www.flamingspork.com 11y ago

MariaDB and Trademarks, and advice for your project

cwyers
1pts0
venturebeat.com 11y ago

Netflix will show ‘Crouching Tiger’ sequel same day as theaters

cwyers
3pts0
www.polygon.com 11y ago

Moon-landing conspiracy claim refuted by video game graphics

cwyers
2pts0
m.theatlantic.com 12y ago

Where Online Services Go When They Die

cwyers
2pts0
www.polygon.com 12y ago

Oculus VR heads talk about Facebook acquisition

cwyers
1pts1

```The associate agreed with me that they should not deprecate the old app until the new app can handle this configuration. I asked them to raise this up the chain: they need to push back the deprecation date.```

There's no way a CSR has any power over this.

From Rust to Ruby 2 months ago

and then when the work was done they were happy with the result

It's worse than that -- they don't even know the result! They never tried to run it!

I was on a GLP-1 a few years ago and lost 70 lbs. After I got off, I kept a ton of diet changes (no more Pepsi or Gatorade and a lot of water instead, switching to whole grains and fiber/protein variants on pasta, etc.) and gained the weight back in a year and a half. The literature backs this up: keeping up weight loss is hard.

If people can figure out how to write RFCs about IP over carrier pigeons for April Fools, they can figure out how to conceive of LLMs as a layer of abstraction beneath a protocol as well.

There's a lot of people in this thread that assume that Sam Altman is the one who is being dishonest here, and I kind of understand, but the other two parties who could just as easily be lying are Pete Hegseth and Donald Trump, and of the three of them if you think sama is the _most_ likely to lie I feel like you have not been paying attention.

GPT-5.3-Codex 6 months ago

Codex now lets you tell the LLM tgings in the middle of its thinking without interrupting it, so you can read the thinking traces and tell it to change course if it's going off track.

Ironically they had the foresight, they were just too early/didn't execute. They ran an online service (co-owner with IBM and CBS) called Prodigy that competed with AOL and CompuServ, and they tried to do online shopping there.

Ruby 4.0.0 7 months ago

Just use WSL2 and Docker Desktop. VS Code has DevContainer support so you can standardize on a Docker image for your project.

I mean, he's _allowed_. The Compiler Police aren't going to roll up to his house and take away his Jai compiler if there isn't a quorum of HN users blessing his efforts. But people can point out they don't feel the juice is worth the squeeze. Also, Blow is certainly an advocate for his position, which means this kind of public debate is germane to the question of if _other_ people should adopt Jai.

Node.js is the most popular web framework/technology in the StackOverflow developer survey. Express is more popular than FastAPI, Django, Flask and Rails in the same survey. Just... what are you talking about?

I was surprised by that, too, and assumed it was a decade-old article until I saw the date at the bottom. Both being mentioned before Python is wilder, as is the total exclusion of JavaScript.

You can actually look at history and see what happens when IBM tries to wrest control of the PC platform back with the PS/2, which was a flop with consumers because it wasn't backwards compatible enough with IBM's own previous PCs or the wider PC market that developed. A bunch of PC clone manufacturers got together and came up with the EISA bus standard so they wouldn't have to pay IBM license fees for MCA, and made it backwards-compatible with ISA cards people already had. It was successful enough that IBM ended up adopting EISA for some of their PCs.

The other notable thing about the situation is that three companies ended up simultaneously responsible for a large part of the PC platform, originally -- IBM, Microsoft and Intel. They all worked in various ways to encourage competition to each other -- the reason we see OS competition on the PC platform is that IBM and Intel both found it in their interests to allow other OSes on the platform to reduce Microsoft's leverage over them. IBM in fact created one of the competing PC OSes out the gate, OS/2, which was originally an IBM/Microsoft joint project until they started feuding. Now, OS/2 is dead, but IBM's interest in being able to support their own OS instead of Microsoft's is a big reason the PC platform was built in an OS agnostic way. People criticize UEFI for locking down the PC platform more than the previous BIOS implementations, but UEFI is still _way_ more open than basically any other platform, most of which don't have a standard for bootloaders at all. It's really the absense of a standard for bootloaders that keeps most Android phones locked down. Two Android phones from the same OEM might have different bootloaders, much less two phones from different manufacturers. We've yet to see an alternate OS with the resources to support implementing their own bootloaders for a majority of Android phones.

Because the original IBM PC was designed to be cheap and built in a hurry. IBM had a mandate for the original PC to use off the shelf components as much as possible. They also neglected to secure an exclusive license from Microsoft for DOS. 95% of building an IBM PC clone was buying the same parts and getting a DOS license from Microsoft (which they were very happy to sell you). Everyone saw what happened to IBM and just didn't do it that way again.

I'm not saying SWE-Bench is perfect, and there are reports that suggest there is some contamination of training sets for LLMs with common benchmarks like SWE-Bench. But they publish SWE-bench so anyone can run it and have an open leaderboard where they attribute the results to specific models, not just vague groupings:

https://www.swebench.com/

ARC-AGI-2 keeps a private set of questions to prevent LLM contamination, but they have a public set of training and eval questions so that people can both evaluate their modesl before submitting to ARC-AGI and so that people can evalute what the benchmark is measuring:

https://github.com/arcprize/ARC-AGI-2

Cursor is not alone in the field in having to deal with issues of benchmark contamination. Cursor is an outlier in sharing so little when proposing a new benchmark while also not showing performance in the industry standard benchmarks. Without a bigger effort to show what the benchmark is and how other models perform, I think the utility of this benchmark is limited at best.

Keep Android Open 9 months ago

The short version is: the PC is a historical accident. By "the PC" I mean "the Windows-Intel platform on which most consumer PCs were built." Linux and BSD were both able to exist in the form they did because there was a commodity hardware platform that was standardized (ad-hoc standardization, mind you) and _somewhat_ open. IBM, Microsoft and Intel were all best frenemies, able to exert enough power to standardize the PC platform but also able to exert enough power against each other to prevent them from locking the platform down too much. There is no standard "smartphone" platform like there is with the PC, really the only standard is Android AOSP. Because of this, it's a lot harder to do a third-party phone platform without adopting large parts of Android's code.

The lack of transparency here is wild. They aggregate the scores of the models they test against, which obscures the performance. They only release results on their own internal benchmark that they won't release. They talk about RL training but they don't discuss anything else about how the model was trained, including if they did their own pre-training or fine-tuned an existing model. I'm skeptical of basically everything claimed here until either they share more details or someone is able to interpedently benchmark this.

LLMs are good at pursuing objectives, but they aren't necessarily good at juggling competing objectives at once. So you can picture doing the following, for instance:

- "Here is a spec for an API endpoint. Implement this spec."

- "Using these tools, refactor the codebase. Make sure that you are passing all tests from (dead code checker, cyclomatic complexity checker, etc.)"

The clankers are very good at iteratively moving towards a defined objective (it's how they were post-trained), so you can get them to do basically anything you can define an objective for, as long as you can chunk it up in a way that it fits in their usable context window.

The Doctorow school argument, as best I can tell, would go 'the regulations on black car service were meant for things like limo services that don't compete directly with taxis, and once Uber started competing directly with taxis, regulators and authorities should have moved more aggressively to write new regulations/laws that regulated Uber the same way taxis are regulated.' They would not agree with "the reason why taxis are tightly regulated are for reasons that mostly do not apply to Uber."

And this is exactly why I think the question of "what is the correct way to regulate car ride services" shouldn't hinge on incumbency bias towards taxis, but actually ask the question of what is best for participants in the market (which doesn't just include taxis and Ubers but also includes public transportation and its users, for instance). But that doesn't fit neatly into Doctorow's enshitification narrative.

The opening of the article is laying out the case that the laws are good -- they make the market legible to participants. As he says:

``` To navigate all of these technical minefields, you need the help of a third party. In a modern society, that third party is an expert regulator who investigates or anticipates problems in their area of expertise and then makes rules designed to solve these problems.

To make these rules, the regulator convenes a truth-seeking exercise, in which all affected parties submit evidence about what the best rule should be and then get a chance to read what everyone else wrote and rebut their claims. Sometimes, there are in-person hearings, or successive rounds of comment and counter-comment, but that’s the basic shape of things.

Once all the evidence is in, the regulator—who is a neutral expert, required to recuse themselves if they have conflicts—makes a rule, citing the evidence on which the rule is based. This whole system is backstopped by courts, which can order the process to begin anew if the new rule isn’t supported by the evidence created while the regulator was developing the record.

This kind of adversarial process—something between a court case and scientific peer review—has a good track record of producing high-quality regulations. You can thank a process like this for the fact that you weren’t killed today by critters in your tap water or a high-voltage shock from one of your home’s electrical outlets. ```

And this is central to Doctorow's point, right? The narrow question of the legality of Uber's current service offerings is actually pretty well litigated, and if Uber was as flagrantly illegal as he claims, "we're an app" wouldn't have kept them in business. Doctorow argues that this is happening through regulatory capture -- the case isn't primarily that Uber is violating the currently existing set of laws, regulations, court precedents, etc. It's that Uber is violating what the regulations _would be_ in a world where they had less market power with which to influence regulations.

And so it's not enough to argue about how the apps get around _current_ laws. By Doctorow's own arguments, we're debating the merits of a counterfactual set of different regulations that we would have if you changed current conditions. And at that point, it is absolutely fair game to ask if this counterfactual set of different regulations is actually better for market participants.