HN user

ignoramous

18,668 karma

mz at celzero dot com

https://rethinkdns.com/

follow:

sec: tptacek, moxie, nickpsecurity, strcat, agl, SwellJoe, drewcrawford, schoen, mirimir, secfirstmd, mjg59, userbinator, gorhill, rgovostes, lallysingh, malandrew, mikewest, jedberg, wtarreau, michaelaiello, segmondy, whitequark_, jsnell, salgernon, geofft, jlund, sinak, alecmuffett, mrb, nneonneo, pwg, viraptor, indutny, jart, Danenania, cyphar, lisper, drfuchs, mahmoudimus, lrvick, woodruffw, lvh, FiloSottile, pgl, mh_, amarbi, psanford, jedisct1, lotharrr, axoltl, cperciva, sleevi, kwantam, syncsynchalt, Moral_, saurik, cryptonector, JoachimS, j2kun, woodruffw, bjackman, nneonneo

start: pg, sama, garry, mbesto, coffeemug, davidu, rms, xal, malgorithms, Alex3917, jacquesm, jl, AndrewWarner, emmett, ig1, anateus, lpolovets, ivankirigin, sahillavingia, Joshua, ericflo, immad, rdl, joshfraser,s gdb, grellas, gatsby, mayop100, bryanh, josephsunny, ayw, pbiggar, sytse, csallen, joshu, jhuckesteinm, pclark, whockey, sjtgraham, jenthoven, aresant, mayop100, cmdrtaco, zt, erohead, tomhoward, drderidder, rfrey, pauldix, akcreek, mojombo, salsakran, armon, drob, pavlov, goodmachine, snowmaker, sorenbs, simonw, zds, jessepollak, raviparikh, buf, paraschopra, ahaseeb

prog: aphyr, KiranDave, peterwaller, saosebastiao, amirmc, lhorie, judofyr, nikita, huhtenberg, pron, Animats, dmbaggett, chris_wot, aaronbrethorst, grey-area, drewg123, dom96, jashkenas, rich_harris, jordwalke, Homunculiheaded, mark_l_watson, kibwen, Sir_Cmpwn, KenoFischer, ahoyhere, scrollaway, trishume, dbaupp, skrebbel, cryptica, peterhunt, rauchg, TazeTSchnitzel, syrusakbary, megous, jemfinch, fogus, jcelerier, 1st1, anderskaseorg, willvarfar, Veedrac, kodablah, mraleph, enriquto, joosters, loeg, jorangreef, ot, asicsp, felixhandte, ezmobius, r1ch, luckydude, btilly, dherman, Manishearth, carllerche, gregdoesit, q3k, codahale, kuschku, _vbdg, enneff, phiresky, mitchellh, JoshTriplett, NovaX, kmavm, nullc, EvanYou, colanderman, sadiq, sereja, ot, pfdietz, c-smile, xena, WalterBright, izacus, bradfitz, scott_s, pizlonator, pojntfx, lifthrasiir, kentonv, barsonme, rspivak, tekknolagi, ndesaulniers

systems: brendangregg, DannyBee, steveklabnik, fsk, pcwalton, jbk, ajross, yosefk, netguy, minimax, munificent, ColinWright, beat, zwischenzug, derefr, jandrewrogers, shykes, lallysingh, dochtman, SamReidHughes, hnkimb3558, rurban, gonzo, bogomipz, marknadal, mazieres, bramcohen, koverstreet, bkanber, mafintosh, ksec, amluto, kyledrake, mrkurt, dsl, dspillett, kevinburke, catwell, kmod, scarface74, eru, sanxiyn, znpy, influx, ncmncm, toast0, evil-olive, bonzini, monocasa, ncopa, bfirsh, jefftk, tlrobinson, kiwicopple, stavros, alexellisuk, hardwaresofton, jhgg, jeffbee, hinkley, danielbmarkham, cpuguy83, silverstorm, josephg, tialaramex, apenwarr, nocarrier, KMag, kllrnohj, astrange, bayindirh, inkyoto, fweimer, mananaysiempre, antics

mods: dang

bio: Fomite, comicjk, anderspitman, danieltillett, wgrover

linux: rwmj, pdkl95, linuxlizard, polvi, phillips, zozbot234, wmf, beagle3, megous, bcrl, cdesai, lgierth, monocasa, loeg, swetland, jchw, abarth, microcolonel

misc: phkahler, davidw, antirez, gwern, patio11, jgrahamc, darksaints, jamwt, nostrademons, plinkplonk, mikekchar, holman, mikeash, edw519, jrockway, noonespecial, staunch, petercooper, jmathai, tzs, jacques_chester, coldtea, peteretep, happy-go-lucky, aaronbrethorst, mtgx, codingdave, jawns, jvns, CPLX, tomcam, ethomson, HenryR, JoeAltmaier, jerf, pjc50, dmix, jcr, alexbowe, JulianMorrison, adamnemecek, capnrefsmmat, detaro, calinet6, dargonwriter, tootie, kqr, comex, eloff, andersource, dantiberian, lern_to_spel, swyx, pgeorgi, sago, zokier

os: vezzy-fnord, rbehrends, vardump, amirmc, pjmlp, rsync, waddlesplash, eyberg, penberg, ambrop7, shuss, amscanne, tytso, surajrmal, captainmuon, Morgawr, quotemstr, olliej, kllrnohj, pizlonator, tadfisher, faragon

db: craigkerstiens, teraflop, ifcologne, espeed, gopalv, thekozmo, lorenzhs, SQLite, mytherin, pgaddict, sumeer, NovaX, arjunnarayan, dmoura, eatonphil, cube2222, zX41ZdbW, benbjohnson, JoelJacobson, benesch

graphics: pcolton, Jasper_, macawfish, kvark, Agentlien, shmerl, Atrix256, jasondavies, bhickey

net: zx2c4, keithwinstein, bsder, walrus01, apenwarr, newman314, stuntprogrammer, bluejekyll, jlgaddis, revertts, samcrawford, wahern, brian-armstrong, lrizzo, kev009, muppetman, Nrsolis, signa11, majke, shaklee3, emmericp, zamadatix, Sean-Der, techsupporter, downwithbgp, p1mrx, ghshephard, matsur

x/aws: colmmacc, _msw_, aligouri, illumin8, jcrites, socttlegrand2, openasocket, otterley, mslot, bbgm, NathanKP, appwiz, planckscnst, blasdel, mjb, ragona, donavanm, grogenaut, jeffbarr, twirrim, jtoberon, fnordpiglet, scarface74

x/nvidia: jebarker

x/google: kortilla

ai: jph00, eli_gottlieb, iandanforth, karpathy, mjn, nl, albertzeyer, bravura, michael_nielsen, dgacmu, cs702, emu, lhl, espadrine, edwardjhu, binarymax, jll29, MontyCarloHall

bootstrap: arvidkahl

crypto: jaekwon, tipsysquid, Taek, daeken, pbsd, davidcash

a11y: mwcampbell

eee: geerlingguy, sowbug, ta8645, mmmBacon, gchadwick, femto

Posts263
Comments5,018
View on HN
github.com 18d ago

OpenScience: Workbench for scientific research using custom LLMs

ignoramous
13pts1
github.com 28d ago

GELab-Zero: Android automation framework for multimodal LLMs

ignoramous
3pts0
github.com 2mo ago

Endo: JavaScript plugin framework with built-in supply chain attack resistance

ignoramous
2pts0
github.com 6mo ago

Toro: Deploy Applications as Unikernels

ignoramous
148pts139
eheidi.dev 7mo ago

Qwen3-VL 2B on Raspberry Pi with llama.cpp

ignoramous
2pts0
damagemag.com 8mo ago

Apple and Foxconn in China

ignoramous
1pts0
research.google 8mo ago

Toward provably private insights into AI use

ignoramous
2pts1
github.com 9mo ago

Xyne: Open-source LLM-driven search engine for Google Workspace

ignoramous
2pts0
github.com 10mo ago

Playground Elements: Serverless interactive editable coding env for the web

ignoramous
1pts0
opensource.salesforce.com 10mo ago

Salesforce near membrane: Sandbox JavaScript object graphs

ignoramous
1pts0
arxiv.org 11mo ago

Meta Clip 2: Worldwide

ignoramous
2pts0
jxnl.co 1y ago

RAG Anti-Patterns

ignoramous
2pts0
rot256.dev 1y ago

Code Signing for web apps using Signed HTTP Exchange (2023)

ignoramous
1pts0
phoenixfundresearch.substack.com 1y ago

Havard's private companies are collectively worth $393B

ignoramous
3pts0
arxiv.org 1y ago

Small Language Models: Techniques, Enhancements, Applications, Trustworthiness

ignoramous
3pts0
arxiv.org 1y ago

Guide to Fine-Tuning LLMs

ignoramous
157pts16
www.webwizwork.com 1y ago

Updated NotebookLM allows for more control

ignoramous
1pts0
openai.com 1y ago

Technical Goals (2016)

ignoramous
2pts0
github.com 1y ago

Vpqy: ML framework for queryable video analytics

ignoramous
2pts0
github.com 1y ago

Altitude: Triage violent, extermist, and terrorist content

ignoramous
3pts0
github.com 1y ago

Meta's Agentic System: Multistep reasoning,websearch,code interpreter for Llama3

ignoramous
2pts0
github.com 2y ago

Coppie: Rewrite content using 232 copywriting formulas with ChatGPT

ignoramous
1pts0
arxiv.org 2y ago

Mutlimodal neural networks converge to a shared statistical model of reality

ignoramous
34pts6
manara.tech 2y ago

Tell HN: Manara (YC W21) seeking companies to commit hiring engs from Palestine

ignoramous
8pts1
ai.stanford.edu 2y ago

Machine Unlearning in 2024

ignoramous
328pts94
www.theregister.com 2y ago

GPT-4 can exploit vulnerabilities by reading CVEs

ignoramous
81pts29
blog.cloudflare.com 2y ago

Running fine-tuned LoRA models on Workers AI

ignoramous
5pts0
github.com 2y ago

Sqlelf: Explore ELF Objects through SQL

ignoramous
2pts0
github.com 2y ago

WikiLLM: Inject facts into LLMs by editing parameters

ignoramous
3pts0
eclypsium.com 2y ago

Shimshady: Linux Secure Boot CVE

ignoramous
3pts0

Simplest way is to signup to OpenRouter and filter out all non ZDR (zero data retention) providers. Paying "API rates" however can prove expensive compared to coding plans (for instance, MiniMax $20/mo coding plan allows 1.7b tokens; depending on input/output/cache ratio, it is worth $200 to $500+), except for Hy3, DeepSeek v4 Pro, and MiMo v2.5 Pro (whose API rates are cheap and/or discounted already).

If you're looking to not have to deal with Chinese providers, AtlasCode ($20/mo), OpenCode Go ($10/mo), and Cline Pass ($10/mo) provide 2x to 6x usage for some of the popular open weights (depending on the model).

Personally, I subscribe to Z.ai ($17/mo), and pay API rates for MiMo v2.5, Hy3, Qwen 3.7 Plus, & DeepSeek v4 to the original providers (Xiaomi, Tencent, Alibaba, & DeepSeek).

focusing on what they can get wins in like speed instead

Speed as a differentiator has always been Google's thing. They (used to?) show the microseconds it took to query & rank web-scale search results. Chrome, notoriously, focused on speed at the expense of resource use. The very many efforts to efficiently speed up Android & its runtime since its inception, and so on...

their big model underperforms chatgpt 5.6

Possible but TFA claims:

  We have started our most ambitious pre-training run yet, for Gemini 4 ...

"The highest tier Chinese models are not more economical than US frontier models. Try GLM 5.2 and see how much it costs to do real work. I did, and it was more expensive than GPT 5.6." This is a flatly false statement.

It may not be false but may be a "category error" [0]. Reserved GPU pricing & bulk inference pricing is 3x to 6x cheaper than "API rates", but renting your own GPU cluster (in this crunch) to run a 600b+ open weights is going to be "more expensive than GPT 5.6".

Even then, it remains to be seen if Huawei will pull their weight (and match up to Nvidia) as spectacularly as their fellow Chinese AI Labs have. If so, the WAICO alliance is ready to go all-in.

[0] Ben, and probably other "influencers" in this space, may be prone (knowingly or unknowingly) to favour points that make their conclusion for them (https://en.wikipedia.org/wiki/Motivated_reasoning).

Don't think it is down to Wang or MSL but Meta's focus on "personal AI" led them to whatever strategy (OAI missed the developer market too, which Ant then captured; leading to several high profile departures at OAI, coincidentally hired from Meta). The original Llama team themselves started Mistral which hasn't gone anywhere. The simple fact of the matter here is, Chinese firms have the money and the talent to rival the US ones in this field, should they as much miss a beat.

Don't think HN is anti-AI but experts of course may have a different perspective. A comment in this discussion points out that LLMs have been doing math that's been long overlooked in favour of perhaps more impactful work: https://news.ycombinator.com/item?id=48974274

  There isn't a lot of money in academic math, and the ones that love it don't look for low value findings. Proofs like these are ... usually the domain of hobbyists ...

It's cheap enough to where I don't think about cost, made even cheaper by DeepSeek having the most effective and cheap caching in the industry

I use DeepSeek v4 Flash & MiMo v2.5 Pro. Prefer the latter over DeepSeek v4 Pro because it costs the same while being equally good & less chattier for coding workloads. Although, I've begun experimenting with Hy3 (as an in-between Flash & Pro) & GLM 5.2 (for long-horizon tasks).

tried Kimi K3 on a task I've done with every other model I use regularly and found it chewed a lot longer on the problem and ate up almost the entirety of a 5 hour usage limit on their $19 plan

ArtificialAnalysis puts Kimi K3 just below DeepSeek v4 & GLM 5.2 in token use per task, which is about 2x to 3x more tokens than Grok 4.5: https://x.com/ArtificialAnlys/status/2077832879187620192 / https://archive.vn/zBbFi 2 other open weights MiMo v2.5 & MiniMax M3 are comparatively thrifty.

Subscription usage limits are hard to measure as none of the providers tell you directly what it means in terms of tokens or anything else you can easily compare

I always put my coding subscriptions (that allow it) through "AI gateways" (Cloudflare & OpenRouter are free) which help track token use.

In my experience, Kimi & Qwen Cloud have opaque & restrictive limits, their "credits" drain faster. I now make it a point of subscribing (directly [0]) with providers that are transparent like MiniMax, DeepSeek, Xiaomi, & Z.ai.

[0] OpenCode Go, Cline, and AtlasCloud have generous limits for open weights, otherwise.

Open weights are indeed a commercial risk because there's no shortage of companies that'll make a business out of running MaaS (model as a service), close & resell a fine-tuned or distilled variant (DeepSeeking of Llama / Cursorification of Kimi, for example).

Android was initially developed for phones with a keyboard (similar to Blackberry).

It was codenamed "Astro Boy". Btw, the team Andy Rubin assembled to build "android" first built OS for Digital Cameras viz. FotoFrame.

Astonishing. Considering none of the BigTech except Google (Microsoft, Apple, Meta, Amazon, Nvidia, SpaceX) have managed to challenge OpenAI & Anthropic frontier models, such achievements are scarcely believable.

Re: GLM-5.2: For a ~750b model, it holds up pretty good against models 3x its size (and ~10x the cost). Same goes for Tencent Hy3 and MiniMax M3, which almost match Opus 4.6 levels with ~295b params.

on the top of the largest open models list

Moonshot (true to their name?) has always lead in terms of releasing the largest among open weight LLMs.

Moonshot is going to need the USD 500 million reportedly raised earlier this year to run this model.

Think Moonshot, as a spin-out, can expect backing from its former parent, Alibaba? I don't think they would be particularly worried about finances, if the Kimi K series continues to outperform the Qwen Max series (which seems to be the case; while Kimi is also super popular in China).

Possible, but pay-as-you-go Hy3 / DeepSeek v4 Pro / MiMo v2.5 Pro (from respective vendors) are genuinely good enough as daily drivers, given the costs (especially, low prices for input cache, which usually makes up 70%+ of total input for agentic workflows). I put in $10 in DeepSeek & Xiaomi MiMo, and I've barely used $1 each, in a week of coding work.

Coding Plans by MiniMax ($20/mo for 1.7b tokens) and Z.ai (~$30/week use for $17/mo) are also tremendous value for money.

I don't need it to be trustless. If it worked well, the tradeoff would be smaller...

You mean, on a self-hosted Ente (or Immich), a off-device/server-side, capable multimodal LLM, like Gemma4 12b would do?

Nowadays every CIA goat herder has their own Starlink terminal.

Can the Starlink radio be sniffed? If so, it might act like a beacon, a tell-tale sign (especially, in countries where Starlink is illegal).

will collectively effect all of us, why not apply democracy to it

That's because most execs proposing solutions are "Technocrats" or think like one: https://en.wikipedia.org/wiki/Technocracy

Besides, I don't think as a collective we're well equipped to decide one way or the other. If the collective were given a say, billions will be spent, often in consultation with technocrats, on doom / hype marketing (if it isn't happening already).