HN user

frabcus

2,349 karma

I'm Francis Irving. Cohosts "Imagine an apple" a podcast about our inner mental worlds.

Made lots of the mySociety democracy websites - TheyWorkForYou, WhatDoTheyKnow and so on

Email me francis@flourish.org

Posts58
Comments574
View on HN
www.youtube.com 26d ago

What if plants could talk? (OpenAI YouTube) [video]

frabcus
2pts0
www.flourish.org 1mo ago

The frustration of agreeing with everyone about AI

frabcus
1pts0
github.com 2mo ago

Copilot silently inserts itself as a co-author in VS Code

frabcus
3pts1
www.youtube.com 1y ago

Big Carl lifts the dome onto second reactor – Hinkley Point C [video]

frabcus
1pts0
github.com 2y ago

Automating Meridian DSP5000 Speaker Activation for One-Click Audio Bliss

frabcus
1pts0
news.ycombinator.com 2y ago

Ask HN: Why is X's website still on domain twitter.com?

frabcus
2pts4
arxiv.org 3y ago

Mirages: On Anthropomorphism in Dialogue Systems

frabcus
41pts22
www.flourish.org 3y ago

What is high-quality about the data that trained generative AI?

frabcus
1pts0
arxiv.org 3y ago

Language Models Can Teach Themselves to Program Better

frabcus
7pts0
www.youtube.com 4y ago

I learned Unity in 3 simple* steps

frabcus
2pts0
news.ycombinator.com 5y ago

Could you set your Signal profile picture? It makes the app feel more human

frabcus
3pts2
developer.apple.com 6y ago

New Guidelines for Sign in with Apple

frabcus
1pts0
theintercept.com 7y ago

US Government’s Indictment of Assange Poses Grave Threats to Press Freedoms

frabcus
6pts0
www.flourish.org 7y ago

Brainstorming a better YouTube recommendation algorithm

frabcus
5pts0
news.ycombinator.com 8y ago

Ask HN: How do you license a neural network as free software?

frabcus
4pts2
duckduckhack.com 8y ago

DuckDuckHack is now in Maintenance Mode

frabcus
209pts83
www.howbb8works.com 10y ago

How Does BB-8 Work?

frabcus
2pts0
metrics.torproject.org 10y ago

Number of Tor users recently halved

frabcus
42pts16
www.flourish.org 10y ago

Sync/Backup workshop at Redecentralize Conference

frabcus
2pts0
cv.democracyclub.org.uk 11y ago

I collected the CVs of 687 people standing for Parliament

frabcus
2pts0
www.androidcentral.com 11y ago

Cyanogen teams up with Microsoft to offer bundled apps and services

frabcus
1pts0
cv.democracyclub.org.uk 11y ago

Curriculum Vitae of Future Members of Parliament

frabcus
1pts0
www.flourish.org 11y ago

The advert wars

frabcus
1pts0
www.flourish.org 11y ago

I promise never to use C/C++ for a new project

frabcus
18pts11
support.twitter.com 11y ago

Twitter app graph steals list of your mobile apps

frabcus
5pts0
www.planetary.org 11y ago

Philae update: “Go” for landing, despite apparent failure of cold-gas jet system

frabcus
1pts0
explainxkcd.com 11y ago

XKCD live-cartoons today's real attempt to land on a comet

frabcus
1pts0
gravyanecdote.com 11y ago

Why Twitter’s decision on Scraperwiki is bad for data democracy

frabcus
31pts16
www.gamedevmarket.net 12y ago

Game development asset marketplace

frabcus
2pts0
redecentralize.org 12y ago

Interview with David Irvine, founder of MaidSafe

frabcus
3pts0

Given a German court talked about not having the prompt logs making it hard in a specific case to prove it had sufficient human input, maybe someone could do an open source hosting service where the Claude Code logs were fully uploaded with each commit, so that could be later proven? It would be cool anyway to have such a service, so people could learn how to do LLM coding better from the examples.

Germany copyright law feels like the relevant thing here - Codeberg is a non-profit based in Germany.

Best quick English language overview of status that I could find: https://www.twobirds.com/en/insights/2026/germany/when-can-a...

It looks like Codeberg want only copyrighted material in their service, so it is reliable in the future that e.g. licenses must be followed (e.g. GPL), and copyright doesn't suddenly get declared as being of the model owner, and it isn't a copy of something else.

That is a cautious reasonable position - in early days of LLM coding (3 years ago!) indemnity from model companies was a major issue globally because of the lack of clarity of the law around this. The US specifically has settled on it being (effectively?) public domain. But I don't think that is fully settled, and it certainly isn't settled in international copyright law.

The goal of the vague "mostly" in the Codeberg change is to ensure there is enough human input to the code they host, to be reasonably sure under German copyright law it is copyright of the person sharing it.

Umm, Fable only really came out 2 weeks ago, and GPT-5.6 Sol only 1 week ago.

Yes, Kimi K3 appears a touch below them both, but above all other models. So I'd say a few weeks behind, not months now...

The advent of the internet was collaborative and based on introducing shared protocols for a couple of decades. It deserved criticism when globalised capitalism got involved, and monopolies started forming, leading to rent-seeking, excessive centralisation, and enshittification.

The impact is that the internet has a fraction of the value to improve people's lives as it should have. It is a very poor free market, incredibly poor competition because of lack of standards and protocols and interoperability. People's minds are ground down by social media, search engines don't work well any more and so on.

So yes - every new technology deserves many criticisms, so they can be addressed, and as a society we can gain the benefits of that technology and minimise the disadvantages.

The printing press lead to copyright, public libraries, universal literacy... All things which are now widely celebrated. They took centuries to work out, and are all government and regulatory intervention to fix problems critics noticed and campaigned about.

AI is the same, only it is at risk of moving much faster and having a much large negative impact before society reacts.

So no, most of the criticism of LLMs are not wrong - they are correct, as are the people saying the technology of LLMs is useful to people and the economy.

Critics are friends of a new technology - without responding to every criticism in a significant way, AI will rapidly lead to a Butlerian jihad. If you like AI, you should love criticism of AI even more.

This being a news site for hackers, I should point out you can use browser plugins to change the default feed. And (if you use an Android phone, but maybe there's a way with Safari?) you can run those plugins (at least on Firefox), so you can have the experience you want on mobile too!

I made one of these (called Instalamb for Instagram), but haven't maintained it recently as there wasn't much interest. There are plenty of others though.

I think my biggest disappointment with social media is not that capitalism made it harmful and addictive (that was inevitable), but that most people don't seem to care enough to even install an advert blocker, never mind something to make their feed cleaner. Despite having had a better experience before, and it being much easier to do than many things people do all the time in their daily lives.

Claude Sonnet 5 22 days ago

I've been wondering this - I don't have an intuition for Anthropic's gaming around military applications, or how this stage could play out in terms of relationship to Government controlling AI.

Are there some Less Wrong posts or similar I should read that probably explain it?

I only did undergraduate level in Maths, and to me there is a key aesthetic element which makes it created. The choice of axioms to use, the choice of with theorems are interesting.

Yes the "truth" (doesn't exist, see Gödels theorem) is discovered in a vast, wild landscape that Mathematicians explore.

But which areas are worth exploring is a critical question. Partly driven by application, partly aesthetic. It's a quest for simple things that are a bit surprising, or that were hard to make the statements so simple.

Long term, getting locked into proprietary software development tools is a bad idea. And these models are extremely proprietary. The ability of the US Government to cancel them at any time is one real recent example of one category of problem.

Back in the 1990s the good C++ compilers were proprietary, eventually GCC and LLVM caught up, and now dominate. The pattern repeats in software development, and there's no reason to believe it won't continue.

Yes, right now it makes sense to use Opus 4.8, but it is good that a significant number of people are using other options, and making sure they work and are ready for when you need them.

Plus it is extremely fun and connecting and hackerish to do local coding with a local model. Try it.

Siri AI 1 month ago

They can't under GDPR. The DMA is for market access - there are other laws for privacy. Those require use commensurate with what is needed for the service, so anyone who e.g. scraped all of a user's local info and stolen it would be breaking EU privacy laws themselves.

This is not complicated. Even in the US, every other industry is regulated to your benefit, you're just used to it and haven't realised. Digital technology obviously needs to be too. And yes, you have to do it properly.

AISI in the UK has been doing this for years - there are lots of papers https://www.aisi.gov.uk/category/safeguards and specific reports, e.g. this on GPT 5.5 https://www.aisi.gov.uk/blog/our-evaluation-of-openais-gpt-5...

This old post goes into lots of detail about what they do to red team and why: https://www.aisi.gov.uk/blog/early-lessons-from-evaluating-f...

NIST's similar unit in the US is now called CAISI https://www.nist.gov/caisi - interesting that the most recent post is an evaluation of DeepSeek capabilities, which sound more like watching China. But presumably this executive order alters the emphasis?

Right the original article says "Do you think macOS will get better or worse in the next 2 years?" (rhetorically implying "worse").

That could easily be true and Apple "will use even more tokens and spend even more money".

Make harness independent of model, so when pricing or quality changes you can switch.

Avoid lock in to stack from one provider (things like a harness that only works with models from one provider and so on).

Use local models (a couple of them do work a bit now, if you have 20Gb video RAM), which saves money and is more private, and works offline.

Can improve the harness, fix bugs in it, make it compatible with different systems and techniques.

This game happens every time in new cycles of developer technology. The good bet historically has always been to use open source - there's a reason most developer tooling just pre-AI revolution was open source (even things like Java and .NET which used to be proprietary).

I was finding this really interesting, that maybe a human had written it and it really reflected a vision for how we build software in this new world. I want to know the way, I'm curious!

Until I got to "One platform, three modes." and my brain just pattern matched "AI slop" and the entire post dissolved into meaningless for me.

I don't know if I can stop my mind reaching this conclusion. I'm sure someone at GitLab made some effort to carefully edit the post... But that it wasn't entirely rooted in a human who'd worked out how this stuff goes, but clearly had lots of AI writing it out... Just made my instinct go "this isn't worth paying attention to after all".

I've tested this extensively in a workflow (not agentic) context, and you're right, the underlying models are both good at full rewrite of code files, and at doing search/replace.

They've been decent at full rewrite for 2 years. I don't think they were good at search/replace until a year ago, but I'm not so sure.

It's true that the models 2 years ago would sometimes make errors in whole rewrite - e.g removing comments was fairly common. But I've never seen one randomly remove one character or anything like that. These days they're really good.

Main reason agentic harnesses use search/replace is speed and cost, surely! Whole file output is expensive for small changes.

Qwen 3.6 is out now and a touch better than 3.5.

I'm finding Google's Gemma 4 even better though - seems to hold up the agentic loop better than Qwen.

All will load into 20Gb of VRAM. None are amazing, but they do just about work.

I tried Mistral for a bit, and it is so fast everything else feels bad now by comparison. I think there's lots of opportunity for OpenAI, Anthropic to stumble on features and performance.

The spikes in the last 2 years have happened for very short amounts of time. If renewables are working, you don't get a spike, and save loads on this tariff. The small amount of time they're not, you sometimes have to pay more, but not for long enough to matter. It's fundamentally more effective for everyone than the default of buying the insurance of fixed prices.