HN user

bigbones

138 karma
Posts0
Comments39
View on HN
No posts found.

The theoretical advance we're waiting for in LLMs is auditable determinism

I think this is a manifestation of machine thinking - the majority of buyers and users of software rarely ask for or need this level of perfection. Noise is everywhere in the natural environment, and I expect it to be everywhere in the future of computing too.

Based on what I've seen so far, I'm thinking a timeline more like 5-10 years where anything involving at least frontend has all but evaporated. What value is there in having a giant app team grind for 2 years on the perfect Android app when a user can simply ask for the display they want, and 5 variants of it until they are happy, all in a couple of seconds while sitting in the back of a car. What happens to all the hundreds of UI frameworks when a system as a widespread as Android adopts a technology approach like this?

Backend is significantly murkier, there are many tasks it seems unlikely an AI will accomplish any time soon (my toy example so far is inventing and finalizing the next video compression standard). But a lot of the complexity in backend derives from supporting human teams with human styles of work, and only exists due to the steady cashflow generated by organizations extracting tremendous premiums to solve problems in their particular style. I have no good way to explain this - what value is a $500 accounting system backend if models get good enough at reliably spitting out bespoke $15 systems with infinite customizations in a few seconds for a non-developer user, and what of all the technologies whose maintenance was supported by the cashflows generated by that $500 system?

I expect pretty much the opposite to happen: it makes sense for languages, stacks and interfaces to become more amenable to interfacing with AI. If a machine can act more reliably by simplifying its inputs at a fraction of the cost of the equivalent human labour, the system has always adjusted to accommodate the machine.

The most obvious example of this already happening is in how function calling interfaces are defined for existing models. It's not hard to imagine that principle applied more generally, until human intervention to get a desired result is the exception rather than the rule as it is today.

I spent most of the past 2 years in "AI cope" mode and wouldn't consider myself a maximalist, but it's impossible not to see already from the nascent tooling we have that workflow automation is going to improve at a rapid and steady rate for the foreseeable future.

There will always be "real thinking" roles in software but the sheer pressure on salaries from the vastly increasing free labour pool will lead to an outcome a bit like embedded software development, where rates don't really match the skill level. I think the most obvious strategy for the time being is figuring out how to become a buyer of the services you understand rather than a badly crowded out seller

I'd put solid money on Warner earning a few cents every time an AI girlfriend somewhere sings happy birthday within 10 years

As I recall it, there was a time when copyright infringement on YouTube was so prolific that the rightsholders essentially forced creation of the first watermarking system that worked at massive scale. I do wonder if any corners of research are currently studying the attribution problem with the specific lens of licensing as its motivation

the IP rights holders have yet to bare their teeth. I don't think the outcome you suggest is clear at all, in fact I think if anything entirely the opposite is the most probable outcome. I've lost count of the number of technology epochs that at the time were either silently or explicitly dependent on ignoring the warez aspects while being blinded by the possibilities, Internet video, music and film all went through this phase. GPTs are just a new medium, and by the end of it royalties will in all likelihood still end up being paid to roughly the same set of folk as before

I quite like the idea of a future where the AI job holocaust largely never happened because license costs ate up most of the innovation benefit. It's just the kind of regressive greed that keeps the world ticking along and wouldn't be surprised if we ended up with something very close to this

The frequency of AI cope posts appears set to keep increasing until the very last one of us is fired. The reality is that users simply don't care about any concern to be found discussed in the church of software engineering, they want an app with a pink button, and very shortly we will be a much higher friction means to achieve that than the sloppy text generator. It's the same force behind why you can't buy high quality furniture any more or consumer devices that survive more than a few years: the vast majority of people simply don't care, they want to click "Buy" at a price level they don't have to think about and get on with their lives, only now it applies to knowledge economy work and entire industries are trying their hardest to ignore the transition we're so painfully obviously already in.

Even were it not the case for disposable CRUD Android/web apps that represent the bread and butter for half the industry, the effect on the structure of popular media is alarming in ways I don't think anyone has a hope of understanding yet. I imagine kids coming up now will not be glued to their phones anywhere nearly like the current batch are, or if they are glued to something, perhaps it is a conversational agent prompting them through an earpiece or similar.

Hand-wringing about the quality of LLM-driven app development really misses the point of all of this. We're currently using an extremely novel technology to emulate aspects of our now-defunct technology (which I believe includes the web), in much the same way fax gateways at one time were a popular application for email.

UML and "as little overhead as possible" probably shouldn't appear in the same train of thought. I remember it from the very earliest Linux VPS providers, IIRC it only got semi-usable with some custom work (google "skas3 patch"), prior to which it depended on a single process calling ptrace() on all the usermode processes to implement the containerization. And there's another keyword that should never appear alongside low overhead in the same train of thought

It gets more interesting when you think about the impact on groups. Sending an image to a group is enough for all devices associated with that group to be identifiable from CloudFlare's side, who additionally see a giant chunk of unencrypted traffic from the same client addresses going to other web sites. Given Cloudflare's less-than-straight approach to sales, it is astonishing the words "secure" and "Signal" ever appear in the same sentence.

CloudFlare get to see a fuckton of metadata from private and group chats, enough to trace who originally sends a piece of media (identifiable from its file size), who reads it, when it is is read, who forwards it and to whom. It really doesn't matter that they can't see an image or video, knowing its size upfront or later (for example in response to a law enforcement request) is enough

good enough to stream YT, in my experience, so presumably already good enough to attend meetings

YT needs bulk throughput while meetings need latency and quality. YT can seem smooth for much longer despite massive amounts of retransmission and packet loss, meetings fall apart rapidly with even a tiny bit of those

I don't know how meaningful it is any more, but with long polling with a short timeout and a gracefully ended request (i.e. chunked encoding with an eof chunk sent rather than disconnection), the browser would always end up with one spare idle connection to the server, making subsequent HTTP requests for other parts of the UI far more likely to be snappier, even if the app has been left otherwise idle for half the day

I guess at least this trick is still meaningful where HTTP/2 or QUIC aren't in use

dedicated 10G network to connect your servers

Do you have to ask Hetzner nicely for this? They have a publicly documented 10G uplink option, but that is for external networking and IMHO heavily limited (20TB limit). For internal cluster IO 20TB could easily become a problem

Oh whoa a 5 minute video for exactly this :) Apologies for making you be my Google. Yep, everything in wasm makes things much easier to work with, especially if you want to run it on a client device

It sounds like you might know the answer to this. Would it be straightforward to use this for sandboxed headless file conversion? You can do that already with LibreOffice, but it's a monster amount of unsafe code that's difficult to containerize securely

I think the central nature of moderation needs fixed, rather than moderation itself. Real world moderation doesn't work by having a central censor, it involves like-minded people identifying into a group and having their access to conversation enabled by that identification. When the conversation no longer suits the group, the person is no longer welcome. I think a technical model of this could be made to work.

Looked semi-seriously at doing a Twitter clone around the time Bluesky was first announced, and to solve this I'd considered something like GitHub achievement badges (e.g. organization membership), except instead of a static number, these could be created by anyone, and trust relationships could exist between them. For example, a programming language community might have existing organs who might wish to maintain a membership badge - the community's existing CoC would necessarily confer application of this badge to a user, thus extending the existing expectation for conduct out from the community to that platform.

Since within the tech community these expectations are relatively aligned, trust relationships between different badges would be quite straightforward to imagine (e.g. Python and Rust community standards are very similar). Outside tech, similar things might be seen in certain areas of politics, religion or local cultural areas. Issues and dramatics regarding cross-community alignment would naturally be confined only to the neighbouring badges of a potential trust relationship, not the platform as a whole.

I like the idea of badge membership and badge trust being the means by which visibility on the platform could be achieved. There need not be any big centralized standards for participation, each user effectively would be allowed to pick their own poison and starting point for building out their own visibility into the universe of content. Where issues occur (abusive user carrying a highly visible badge, or maintainer of such a badge turning sour or suddenly giving up on its reputation or similar), a centralized function could still exist to step in and potentially take over at least in the interim, but the need for this (at least in theory) should be greatly diminished.

A web of trust over a potentially large number of user-governed groupings has some fun technical problems to solve, especially around making it efficient enough for interactive use. And from a usability perspective, application onboarding for a brand new account

Running on little sleep but thought it was worth trying to sketch this idea out on a relevant thread.

OpenWrt One 2 years ago

Not a great spec compared to the very well supported Gl.inet Flint 2

ChatGPT Search 2 years ago

My first memory of something like this was the rushed Google Wave announcement occurring on the same day as a Bing rebranding launch (if memory serves, the domain registration timestamps for the Google domains were also highly coincidental but I've long since forgotten with what)