HN user

ac29

7,243 karma

email: hn@imap.cc

Posts34
Comments3,442
View on HN
theredbeard.io 4mo ago

Five CLIs Walk into a Context Window

ac29
2pts0
kagi.com 1y ago

Kagi for Libraries

ac29
7pts0
www.olimex.com 1y ago

RVPC – Open Source Hardware Board

ac29
3pts1
www.servethehome.com 2y ago

Nvidia Shows Intel Gaudi2 Is 4x Better Performance per Dollar Than Its H100

ac29
1pts0
www.sifive.com 3y ago

HiFive Pro P550

ac29
6pts2
www.theguardian.com 3y ago

Californians harvest water during historic storms

ac29
4pts0
fuse.wikichip.org 3y ago

Intel, SiFive Demo High-Performance RISC-V on Intel 4

ac29
152pts145
www.biographic.com 3y ago

Past the Salt

ac29
2pts0
www.pine64.org 4y ago

DIY low-power 6 SSD NAS based on the Quartz64 ARM board

ac29
2pts0
chipsandcheese.com 4y ago

SiFive Completes Series F Funding Round

ac29
4pts1
doomberg.substack.com 4y ago

The SEC Crackdown on DeFi Is Imminent

ac29
14pts12
0pointer.net 5y ago

The Wondrous World of Discoverable GPT Disk Images

ac29
5pts1
ares.dev 5y ago

Ares Multi-System Emulator

ac29
4pts0
www.datacenterknowledge.com 6y ago

Ampere Gears Up to Launch 7nm, 80-Core Arm Chip for Cloud Data Centers

ac29
3pts0
pete.akeo.ie 6y ago

Installing Debian ARM64 on a Raspberry Pi 3 in UEFI Mode

ac29
2pts0
www.crowdsupply.com 7y ago

HiFive1 Rev B: Second-Generation 32-Bit RISC-V SoC and Dev Board

ac29
14pts0
blog.vyos.io 7y ago

VyOS 1.2 Released

ac29
4pts0
www.zdnet.com 7y ago

Japanese Government to compile list of insecure IOT devices

ac29
13pts0
www.t-mobile.com 7y ago

T-Mobile First to Launch Caller Verification

ac29
30pts8
fastmail.blog 7y ago

JMAP is on the home straight

ac29
193pts50
www.microsoft.com 7y ago

Microsoft Airband Initiative

ac29
1pts0
www.kaiostech.com 8y ago

Google invests $22M in KaiOS (Firefox OS fork)

ac29
5pts0
www.mercurynews.com 8y ago

Silicon Valley cities consider headcount-tax on employers

ac29
3pts0
youtube.googleblog.com 8y ago

YouTube Music, a new music streaming service, is coming soon

ac29
229pts318
www.libretro.com 8y ago

RetroArch – Achieving better latency than original hardware

ac29
1pts0
www.theatlantic.com 8y ago

Images of Bike-Share Oversupply in China

ac29
10pts2
www.eurogamer.net 8y ago

The retro gaming industry could be killing video game preservation

ac29
17pts0
blog.mozilla.org 8y ago

Retrospective: Looking Glass

ac29
6pts0
blog.mozilla.org 8y ago

Extensions in Firefox 59

ac29
496pts236
blogs.gnome.org 8y ago

Gnome: Introducing the CSD Initiative

ac29
5pts2

Its just a single benchmark, but Luna 5.6 xhigh scores within the margin of error the same as Opus 4.8 max on DeepSWE for 8x cheaper. Luna max is quite a bit higher than Opus and still 4x cheaper

OpenCode Go is a great deal but I recently dumped my subscription because I found myself rarely reaching for it over my Anthropic sub (I can get 40 hours of work a week out of the $20 sub and almost never hit weekly limits). Subscribed to OpenAI as my secondary and I've been really impressed with that too so far.

I expect if they add Kimi 3 to Go the limits are going to be really low since 2.7 is already one of the most limited models and 3 is much larger.

You do understand that the "frontier" people are usually talking about is the cost-intelligence frontier right?

By definition there is no model that is both cheaper and as intelligent or better than another on the frontier.

llama.cpp supports a wide variety of 4-bit and smaller quants and mmap's models by default, so you dont need to be able to hold the weights in memory (the OS will handle bringing them in from storage as needed)

Its cool to see this implemented in a tiny amount of code without dependencies, but does it actually bring more performance?

So... maybe we can still use third party harnesses with Claude Code subscriptions... for now?

The way I read this is: yes, if the third party harness uses Anthropic's Agent SDK. Most of them do not, AFAIK, and are still against ToS (though maybe its not enforced for now)

Anthropic just provides a subscription - which Enterprise usually doesn't want you to use because everything you're submitting through that will be trained on / becomes part of their model.

My Pro account very clearly has a toggle for "Help improve our AI models: Allow the use of your chats and coding sessions to train and improve Anthropic AI models."

I did my first ESP32 project recently and was amazed you can get a system that starts up Micropython, then a Wifi AP, DNS, and Web Server in a second or two total and uses less than 512kB RAM. And thats with a high level programming language.

IBM no longer owns any fabs

Per IBM: "IBM Research at Albany [...] includes more than 100,000 square feet of semiconductor fabrication space"

I guess that is technically a R&D fab not a production one, but they definitely have in house fabrication capability

I'd rather use Sonnet than Qwen

I get this, though the pace of Chinese releases is relentless. Qwen3.7 Plus/Max (closed variants) feel notably better than Qwen3.6, and Minimax M3 is a big jump from 2.7 in capability as well. Both of these families had their previous major release less than 90 days ago.

Anthropic must have Sonnet 5 either waiting or cooking though, they said smaller and larger models than Opus were coming and we already briefly had the larger model.

I'm curious what the downside for this speed is here

"DiffusionGemma's speedup is designed for local and low-concurrency inference. In high-QPS cloud serving, autoregressive models can be deployed to saturate compute efficiently, so DiffusionGemma's parallel decoding offers diminishing returns and can result in higher serving costs"

in the process standardized not just electric charging but the plug we're all using today

Wat? Are you talking about NACS? If so, a minority of non-Tesla EVs currently on the road use it in the US, and AFAIK zero outside of North America.

And even in the US, the vast majority of EV charging stations are AC and use J1772, a SAE standard that predates Tesla's existence.