HN user

ZeroCool2u

3,986 karma

https://theolinnemann.com

Posts89
Comments707
View on HN
diataxis.fr 2mo ago

Diátaxis: A systematic approach to technical documentation authoring

ZeroCool2u
5pts0
www.currentaffairs.org 2mo ago

The Banal Horror of Jimmy Fallon

ZeroCool2u
8pts2
gardinerbryant.com 2mo ago

Steam Deck Software in 2026:Checking in with the Developers Behind the Ecosystem

ZeroCool2u
4pts0
news.ycombinator.com 4mo ago

Ask HN: What are some examples of your favorite product documentation?

ZeroCool2u
3pts2
www.bloomberg.com 4mo ago

Trump Orders US Agencies to Drop Anthropic After Pentagon Feud

ZeroCool2u
20pts3
tailscale.com 4mo ago

LM Link: Use local models on remote devices, powered by Tailscale

ZeroCool2u
2pts0
lmstudio.ai 4mo ago

LMStudio LM Link: Use your local models, remotely

ZeroCool2u
1pts0
reticulum.network 4mo ago

Reticulum Network

ZeroCool2u
2pts1
unsigned.io 4mo ago

RNode

ZeroCool2u
4pts1
huggingface.co 5mo ago

Coda-GQA-L Bounded Memory Differential Attention with Value-Routed Landmark Bank

ZeroCool2u
1pts0
allenai.org 7mo ago

Bolmo: Byteifying the next generation of language models

ZeroCool2u
1pts0
www.aboutamazon.com 7mo ago

Amazon introduces new frontier Nova models

ZeroCool2u
1pts1
blog.cloudflare.com 10mo ago

A simpler path to a safer Internet: an update to our CSAM scanning tool

ZeroCool2u
8pts0
github.com 10mo ago

1.0 release of the Google Cloud client libraries for Rust

ZeroCool2u
4pts0
github.com 11mo ago

Axiom: Decentralized AI network that discovers, verifies, and archives truth

ZeroCool2u
2pts0
www.bloomberg.com 11mo ago

Bloomberg: Anthropic Unveils More Powerful AI Model Ahead of Rival GPT-5 Release

ZeroCool2u
4pts3
semiconductorsinsight.com 1y ago

Fear of Losing Search Led Google to Bury Lambda, Says Mustafa Suleyman, VP of AI

ZeroCool2u
5pts0
research.google 1y ago

Google Research: Graph foundation models for relational data

ZeroCool2u
19pts0
old.reddit.com 1y ago

A Formal Mathematical Investigation on the Validity of Kellogg's Glaze Claims

ZeroCool2u
71pts20
blog.google 1y ago

Google: Helping startups build what's next with the AI Futures Fund

ZeroCool2u
4pts0
www.bloomberg.com 1y ago

Fed's Preferred Inflation Gauge Stalls While Spending Picks Up

ZeroCool2u
1pts1
windsurf.com 1y ago

Windsurf Extensions Achieve FedRAMP High Authorization and IL5 Compliance

ZeroCool2u
1pts0
lakesail.com 1y ago

Sail MCP Server: Spark Analytics for LLM Agents

ZeroCool2u
1pts0
www.alphaxiv.org 1y ago

MiniFed: Integrating LLM-Based Agentic-Workflow for Simulating FOMC Meeting

ZeroCool2u
1pts0
www.bloomberg.com 1y ago

Bloomberg: BYD's Five-Minute Charges Compare with Competitors (Gift Article)

ZeroCool2u
9pts4
www.wsj.com 1y ago

U.S. pauses all military aid to Ukraine

ZeroCool2u
480pts1015
aphyr.com 1y ago

Aphyr: Comments on Executive Order 14168

ZeroCool2u
3pts2
lwn.net 1y ago

Resistance to Rust abstractions for DMA mapping

ZeroCool2u
12pts1
ubuntu.com 1y ago

Bringing multiple windows to Flutter desktop apps

ZeroCool2u
3pts1
blog.jetbrains.com 1y ago

Mellum: JetBrains' New LLM Built for Developers

ZeroCool2u
3pts1
GPT‑Live 15 days ago

I used it for double checking some stuff while helping my friend build his PC! I had it in tight spaces and it helped me verify some details around which nvme slot to use first.

GPT‑Live 15 days ago

Gemini live has been able to do this for over a year now. I can just activate it on my phone and it really works surprisingly well, especially the interruption. I've tested it with my 95 year old Dutch grandmother and it switched seamlessly between English and Dutch with her and handled her poor hearing very well, including her asking for repetition.

I'm a little surprised by how much OAI is playing catch up here.

Claude Sonnet 5 22 days ago

Having used it quite a bit when it was out, it's not. It's certainly better, but in some ways it's worse. It's trained to be more "agentic" and even in cases where I wanted to talk things through first and I would explicitly tell it not to do something, it would take action on my behalf without checking first.

It's also still just prone to the kind of "stupid" mistakes we see from all LLM's. Like it can write great code, but it doesn't really have common sense without enormous guidance.

A crucial factor tech industry folks tend to ignore is how much executives value predictable costs. Cloud migrations got away with this, but still had to argue fiercely, because 'the cloud' and its serverless tech had the potential to significantly decrease overall spend for unpredictable, bursty workloads.

The usual counter-argument is the operational burden, but human capital is also a relatively fixed cost. A dedicated team of 3-5 FTEs could probably handle inference ops for a F500 company.

Meanwhile, the capability delta is shrinking fast. We have more evidence that local open-source is viable with the release of DeepSeek v4, and the industry is only trending further in this direction. Especially as we rely more on test-time compute and task-specific harnesses rather than model size.

So, if you're an executive looking at a marginal but fixed operations cost, added flexibility, and a rapidly closing gap in capability, why wouldn't you just run open-source models on your own infrastructure to get those highly predictable costs? Plus, you decrease the risk of one of the frontier

Interestingly, Anthropic uses Mintlify for their docs. Not Stainless. Obviously, the focus is on SDK generation, but still strange.

It is frustrating, because I really enjoyed my Valve Index and want a replacement and Meta has some of the best VR tech in the world, but I've waited 6 years for Valve to release their new headset to buy a replacement, simply because Meta can't be trusted.

Googlebook 2 months ago

The quality of apps in the Google Play Store has dropped massively. There are still some gems, but for better or worse, the ecosystem is simply not as strong as Apples and it's certainly not comparable to just having a device where you can install anything you'd like in a full desktop grade OS.

Googlebook 2 months ago

There was a time where Google could've been competitive in this space, specifically against Apples MacBook product line, but that has long since passed. The 3rd party manufacturer path means Google isn't committed to this and won't have competitive hardware. It'll just be another Chromebook and limited to the Google Play Store too, which just isn't good at this point.

Bedrock is both more expensive, less feature complete, and less reliable in terms of raw volume of 500 errors.

Interesting side effect of this is that Google Cloud may now be the only hype scaler that can resell all 3 of the labs models? Maybe I'm misinterpreting this, but that would be a notable development, and I don't see why Google would allow Gemini to be resold through any of the other cloud providers.

Might really increase the utility of those GCP credits.

GPT-5.5 3 months ago

Benchmarks are favorable enough they're comparing to non-OpenAI models again. Interesting that tokens/second is similar to 5.4. Maybe there's some genuine innovation beyond bigger model better this time?

In my experience Azure is full of consistency issues and race conditions. It's enough of an issue that I was talking about new OpenAI models becoming available via Bedrock on AWS and how convenient that was since I wouldn't have to deal with Azure and my colleague in enterprise architecture went on an unprompted rant about these exact issues. It's not the first time something like this has happened and I've experienced these issues first hand, so yes. I'd say reliability is a critical issue for Azure and it hasn't gotten better each time I've gone back to check.

I'm finishing my annual paid Pro Gemini plan, so I'm on the free plan for Claude and I asked one (1) single question, which admittedly was about a research plan, using the Sonnet 4.6 Extended thinking model and instantly hit my limit until 2 PM (it was around 8 or 9 AM).

Just a shockingly constrained service tier right now.

If you don't "have the full context on the Persona/Discord story" you should work on getting it.

It's literally the first thing that came to mind when I saw your post and not having a convincing/satisfying answer in direct relation to that catastrophe doesn't bode well for getting people to trust your brand. The rest of your answer is essentially the absolute minimum I'd expect from a business like this, but not sufficiently convincing.

Regardless of your opinion of Yann or his views on auto regressive models being "sufficient" for what most would describe as AGI or ASI, this is probably a good thing for Europe. We need more well capitalized labs that aren't US or China centric and while I do like Mistral, they just haven't been keeping up on the frontier of model performance and seem like they've sort of pivoted into being integration specialists and consultants for EU corporations. That's fine and they've got to make money, but fully ceding the research front is not a good way to keep the EU competitive.

GPT-5.4 5 months ago

Yup, that was it. Didn't realize they're different models. I suppose naming has never been OpenAI's strong suit.

GPT-5.4 5 months ago

Frontier Math, GPQA Diamond, and Browsecomp are the benchmarks I noticed this on.