HN user

richardw

4,144 karma
Posts52
Comments1,674
View on HN
www.science.org 1y ago

Durably reducing conspiracy beliefs through dialogues with AI

richardw
1pts0
medium.com 1y ago

PatientSeek, the First Open-Source Med-Legal DeepSeek Reasoning Model

richardw
2pts0
www.bbc.com 1y ago

A man who revealed Auschwitz's atrocities to the world

richardw
2pts0
blog.research.google 2y ago

Generative AI to quantify uncertainty in weather forecasting

richardw
2pts0
en.wikipedia.org 2y ago

Passive Daytime Radiative Cooling

richardw
1pts1
news.ycombinator.com 2y ago

Ask HN: Why is barcode data so hard to find?

richardw
5pts10
unchartedterritories.tomaspueyo.com 2y ago

OpenAI and the Biggest Threat in the History of Humanity

richardw
4pts0
www.epj-pv.org 2y ago

Thermal model in digital twin of vertical PV system explains yield gains

richardw
2pts0
www.nature.com 2y ago

Improving li-ion cells by replacing polyethylene terephthalate jellyroll tape

richardw
1pts0
news.ycombinator.com 3y ago

Ask HN: How to recover a pro photographer’s hacked Meta accounts

richardw
6pts7
www.med.hku.hk 4y ago

Omicron spreads 70x faster in bronchus but slower in lungs

richardw
2pts1
en.wikipedia.org 5y ago

Initial Stress-Derived Noun

richardw
215pts160
gisanddata.maps.arcgis.com 6y ago

Coronavirus Live Map

richardw
2pts2
aws.amazon.com 7y ago

Open Distro for Elasticsearch

richardw
1pts0
en.wikipedia.org 7y ago

Initial stress-derived noun

richardw
2pts0
oldconceptcars.com 7y ago

Old Concept Cars

richardw
203pts51
richardwatson.co 7y ago

Opinion vs. Skill on Hacker News

richardw
137pts73
oldconceptcars.com 8y ago

Old Concept Cars

richardw
3pts0
www.pnas.org 8y ago

Earth BioGenome Project: Sequencing life for the future of life

richardw
1pts0
www.openculture.com 8y ago

All Rulers of Europe Over the Past 2,400 Years Presented in a Timelapse Map

richardw
45pts2
www.cambridge.org 8y ago

Road map to clean energy using laser beam ignition of boron-hydrogen fusion

richardw
146pts55
www.coindesk.com 9y ago

The UN Wants to Adopt Bitcoin and Ethereum – And Soon

richardw
2pts0
gumroad.com 10y ago

Drip Email Mastery

richardw
1pts0
www.bloomberg.com 10y ago

Reddit: A Nine-Year Case Study in Absentee Management

richardw
71pts73
www.fastcompany.com 11y ago

Wikileaks launches giant searchable database of leaked Sony emails

richardw
1pts0
medium.com 11y ago

The curious case of the disappearing Polish S

richardw
2pts0
cs.stackexchange.com 12y ago

Is there a system behind the magic of algorithm analysis?

richardw
2pts0
www.theregister.co.uk 12y ago

You’re Not fired: The story of Amstrad’s amazing CPC 464

richardw
3pts2
www.bloomberg.com 12y ago

Bitcoin Judged Commodity in Finland After Failing Money Test

richardw
2pts0
www.forbes.com 12y ago

Google's Being Sued Over Search Patents, Not Just Android Or Smartphone Ones

richardw
1pts0

And a whole universe of random tests got baked into the AI’s training data. Websites of antelopes driving trains and hammerhead sharks swinging in a tyre swing were created. It was a short while until AI became so focused on animals that it gave up competing with developers. Life became sane again.

The AI labs have been subsidising. When they try turn a profit, people will move to the fast followers. The only people that won’t are those that compete on leveraging the very latest models and even then, once spend and scale goes to the cheaper providers, we’ll see deeper research from those providers too. Think “PC compatibles beat IBM, Sun, SGI eventually”.

Grok 4.5 14 days ago

If I had to summarise Google’s effort it would be: stay close but let the others burn themselves out. Position for the long game until you see something worth betting the company on.

Apple similar, without the “stay close” bit.

I made something to ping an AWS service to tell me the uptime of my internet connection. The idea was to sprinkle them around our area, connected to various home WiFi’s, and get a better triangulation of outages. Eg is whole pipe out, just one ISP etc.

I made and tested it but didn’t care enough to continue.

Claude Tag 29 days ago

How much do you think it’s holding them back? Market share seems to be surviving the lack of quality vs the other players.

Or possibly: they’re focusing where it matters?

I think Google just needs to keep its foot in the door and let the other two spend their way into oblivion, channeling “AGI or die trying”.

The benefits were in communication and relationship arbitrage? Surely both of those can be automated over time.

This is just one aspect but it’s still useful. Many people want to see a house and say “please help me make that”. In Australia there’s a set of house patterns that reduce the overhead of just landing a design and pushing it through the admin hurdles.

I generally accuse LLM’s of having no sense of value. The machine will make a complicated plan but entirely lose sight of eg the fact that response time matters to humans.

Not always, but enough that I consider it a thing to fire in a direction, not a thing that aims.

I assume the bet is that as you swap humans for machines, this pays for itself. Swap entire devs and teams and frankly, managers, and you make up a lot of 5%’s fast.

If it works. And I’m not sure who is going to buy the stuff the machines produce, but shrug. Presumably some bots click ads for NFT’s that other bots generate.

Hard problem. I find myself adding in filler to stop the thing from jabbering.

I also think it spends most of its iq on sounding good rather than thinking about the problem. “Yeah absolutely I can see why you’d like to…” etc. This is likely because it’s on a timer and maybe voice is more expensive to process? Text responses spend more time on the task.

The easiest way would be a straight tax on AI usage, and using that tax to pay a universal basic income

Problem for jobs is that there are 200 countries and all the earnings will go to a few. Universal basic income for everyone? Or just the US?

Who gets to keep their house locations in a new fair world? The person whose parents bought in the right place 50 years ago? Who pays the money these models earn, if nobody clicks ads or does a job? What is income for if we don’t work and can just ask the AI for everything we want?

What happens when the super smart AI comes up with “better” (more fair, consistent, etc) answers than you think you have to questions like the above? What if they end up socialist? Do we force it (and invite risk it escapes and fights us for the greater good) or give in to the presumably more thorough reasoning?

The cost between an A500 and a VGA-enabled PC in 1987 ($699 vs $3500-ish?) would have put them in such different categories and customer segments that they would rarely interact.

I remember seeing a PC one of the rich kids brought to boarding school in 1990 and realising it was just crisper than my A500. The PC’s in the school lab were all green and orange screens with one colour CGA, so this was quite a surprise. Still took me some time to accept reality :)

I’m moving away from Claude for anything complicated. It’s got such nice DX but I can’t take the confident flaky results. Finding Codex on the high plan more thorough, and for any complicated project that’s what I need.

Still using Claude for UX (playgrounds) and language. OpenAI has always been a little more cerebral and stern, which doesn’t suit those areas. When it tries to be friendly it comes off as someone my age trying to be a 20-something.

I have this as a skill Claude created to run the rest. It mentions each skill in turn, see below. It’s not deterministic but it definitely runs each skill and it’s raised a bunch of issues, which I then selectively deal with. Where I can, once an issue is identified, I make deterministic tests.

Text includes:

Invoke each review/audit skill in sequence. Each skill runs its own comprehensive checks and returns findings. Capture the findings from each and incorporate them into the final report.

IMPORTANT: Invoke each skill using the Skill tool. Each skill is independently runnable and will produce its own detailed output. Summarize findings per skill into the unified report format.

4. Architecture Health

Invoke: Skill(architecture-review)

Covers: module boundaries, cross-module communication, dependency direction, infrastructure layer rules, hexagonal architecture compliance.

5. Security Health

Invoke: Skill(security-review)

Covers: hardcoded secrets, SQL injection, authorization, HTTPS, CORS, input validation, authentication patterns.

I found that when I have “infinite” tokens my behaviour changed. 3-5 tabs so I’m not waiting, free side quests, huge review skills over whole codebase, skills that wrap 10 other skills. It’s like going from expensive data to uncapped.

I think these token doubles are there to kick you into a abundance mindset (for want of a better term) so going back feels painful. Stop counting tokens, focus on your project and the cost of your own time.

Facebook is cooked 5 months ago

Every so often my YouTube logs out and I’m exposed to the view a “random visitor” would see. Instantly visible because it’s filled with stupid content and sexual provocation.

I manage the shit out of FB and YouTube. You need to block a few things so it stops testing a few segment ideas.

Got a friend who is in the high frequency trading industry and uses both Java and C#. I asked about GC. Turns out you just write code that doesn’t need to GC. Object pools, off-heap memory etc.

It won’t do the absolute fastest tasks in the stack quite as well but supposedly the coding speed and memory management benefits are more important, and there’s no GC so it’s reliable.

Totally. Surely the IDE’s like antigravity are meant to give the LLM more tools to use for eg refactoring or dependency management? I haven’t used it but seems a quick win to move from token generation to deterministic tool use.

It’s a $90k engineer that sometimes acts like a vandal, who never has thoughts like “this seems to be a bad way to go. Let me ask the boss” or “you know, I was thinking. Shouldn’t we try to extract this code into a reusable component?” The worst developers I’ve worked with have better instincts for what’s valuable. I wish it would stop with “the simplest way to resolve this is X little shortcut” -> boom.

It basically stumbles around generating tokens within the bounds (usually) of your prompt, and rarely stops to think. Goal is token generation, baby. Not careful evaluation. I have to keep forcing it to stop creating magic inline strings and rather use constants or config, even though those instructions are all over my Claude.md and I’m using the top model. It loves to take shortcuts that save GPU but cost me time and money to wrestle back to rational. “These issues weren’t created by me in this chat right now so I’ll ignore them and ship it.” No, fix all the bugs. That’s the job.

Still, I love it. I can hand code the bits I want to, let it fly with the bits I don’t. I can try something new in a separate CLI tab while others are spinning. Cost to experiment drops massively.