HN user

kennywinker

10,170 karma
Posts22
Comments3,399
View on HN
swiftmodules.com 10y ago

The state of Swift JSON parsing libraries

kennywinker
1pts0
geo.itunes.apple.com 10y ago

Show HN: Swift Modules iOS – A Searchable Index of Swift Packages

kennywinker
2pts0
swiftmodules.com 10y ago

LangKit – natural language processing toolkit in Swift

kennywinker
4pts0
swiftmodules.com 10y ago

Swift Modules – a searchable Swift package catalog

kennywinker
5pts0
github.com 10y ago

Swift Leftpad

kennywinker
2pts0
medium.com 10y ago

30 days of tvOS sales numbers

kennywinker
14pts28
medium.com 10y ago

CloudKit's Achilles Heel

kennywinker
2pts0
medium.com 10y ago

Building Rivet’s fuzzy string matching autocomplete

kennywinker
2pts0
rivet.link 10y ago

Show HN: Rivet – Easy iTunes and Amazon affiliate links on iOS

kennywinker
1pts1
weirdios.tumblr.com 11y ago

Somebody: A Messaging Service by Artist / Filmmaker Miranda July

kennywinker
1pts0
weirdios.tumblr.com 11y ago

Show HN: Weird iOS, a list of the weirdest, most artful or unexpected apps

kennywinker
136pts20
go-directly.com 12y ago

Show HN: Directly – Transit directions, faster

kennywinker
3pts0
coryalder.github.io 12y ago

Visual Guide to NSURL Components

kennywinker
1pts0
davander.com 12y ago

Show HN: Podlife - CocoaPods search, favorites, and notifications on iOS

kennywinker
1pts0
github.com 13y ago

Government of Canada's web toolkit on github

kennywinker
4pts0
www.kickstarter.com 13y ago

An open solution for transit directions in iOS 6 - OpenTripPlanner

kennywinker
4pts4
geocoder.ca 14y ago

Geocoder.ca sued by Canada Post for their open database of postal codes

kennywinker
288pts84
objectivesea.tumblr.com 14y ago

App Store Spam - 28 identical apps, one developer

kennywinker
59pts32
objectivesea.tumblr.com 14y ago

Deretina.py - never worry about Retina graphics again (iOS)

kennywinker
1pts0
mobile.davander.com 14y ago

CBC is forcing me to remove my app from the Mac App Store

kennywinker
105pts91
geekfeminism.org 15y ago

Who is harmed by a real-names policy?

kennywinker
4pts0
catcat.us 15y ago

The cat url shortener - catcat.us

kennywinker
2pts0

https://en.wikipedia.org/wiki/Wells_Fargo_cross-selling_scan...

https://en.wikipedia.org/wiki/Subprime_mortgage_crisis

Turns out banks run on trust and honor too. I personally know people who got loans because of knowing the right person not their credit score.

But even if banks did run on transparency and accountability, which they don’t, governments aren’t banks. You can’t double-entry your way to equal application of the law. Transparency and accountability are good things, but they are bolsters for the honor system.

If it had more vram it could really cook. Qwen3.6-35b-a3b quantized to 3bits is genuinely usable for coding running on the ten year old pascal card with 16gb i picked up recently.

I agree it’s too small a benefit to justify the investment, and I agree the bubble will pop. I just don’t think that means hardware prices become sane again for quite a while. I think if you half the price of a server GPU because demand from the big ai companies drops out, we’ll still have a shortage - it’ll just being going into commodity data centers to run open weight models.

* Software inference optimizations

Absolutely. I'd be surprised if they couldn't 2x performance in the next year. Still doesn't make a 1T model fit on your phone.

* Heavy quantization

I think this is a dead end if you're trying to fit a 1T model into a phone. Makes much more sense to train a model that's designed to be small, than train a model that's smart and then quantize it into stupidity.

* Chips with hardcoded transformer architecture

Totally, this will probably work great. Now good luck booking fab time any time in the next 2 years.

* Much cheaper HBM

Totally, this will probably work great. Now good luck booking fab time any time in the next two years.

* Much sparser models - 1T total with ~1-10B active params e.g.

Fewer active params helps with the speed of token generation, but if the whole model doesn't fit into ram it doesn't solve the issue of having to constantly stream portions of the model from disk to ram.

* Not to mention - 2 years of today's frontier models writing RTL and kernels at superhuman levels.

IMO this is a delusional myth-making idea being sold to us by ai companies. Machines that generate output based on statistical averages won't generate genuinely new ideas. They can help us try out ideas faster, but they're simply not capable of the kind of creativity and understanding required to push a field forward, except incrementally.

Even if the bubble pops and anthropic and openai et al implode - genie doesn’t go back in the bottle. The usefulness of LLMs for coding is proven, and a chip in a datacenter running 24/7 is always going to be more valuable than in a personal device running occasionally.

That doesn’t change until production capacity exceeds the datacenter demand. When that happens, they’ll start selling them down the market until it eventually reaches phones and toasters and whatever. But not in two years.

First off the math doesn’t math. Datacenters are willing to pay $50k for a single high end GPU. If you have unlimited capacity, yeah sell millions for $100 a pop or $10 a pop or whatever the bom cost of a phone GPU would be - but if you have limited capacity, you’re gonna sell all of that to the customer who is willing to pay the most PER UNIT.

Second off, this doesn’t work from a power consumption standpoint. When I run qwen3.6-35b, a far smaller model than op is suggesting, power usage spikes to 150-200W during inference. To fit a 1T model in the palm of my hand, the amount of processing required doesn’t fit the amount of power available.

Now I’m not saying this will never happen - there are some great leads, e.g. burning models directly on to a chip - but op’s scenario is definitely not happening in two years. Maybe 5, a lot more likely 10, unless of course local ai is made illegal

Unless there are major improvements to how much hardware it takes to run a 1T model, this is deeply unrealistic. First because why release hardware that puts your biggest customers (data centers) out of business. Second because as I understand it the data centers have bought up all the high end chip production capacity for at least the next year and unless the bubble pops that'll continue for a while.

I don’t think it counts as social engineering if it’s exploiting an llm, we might need a new word. Prompt injection doesn’t cover it, because it’s not about a malicious prompt.

I’m thinking some play on highjacking. AIjacking? Agent-jacking? Claudejacking?

Gamers complaining about disc less games despite that problem pipeline and waste.

Tbf the issue is the user-hostile parameters of buying a diskless game. Most people would be happy to download their games if they could back them up to a thumb drive and never get locked out of them and sell the game when they’re done with it.

Seems deeply tangential, but it’s not. Blaming people for wanting a physical thing because the alternative is being further abused by a corporation - that’s a miss. Be mad at game platforms for not offering real ownership in whatever the most climate-friendly way possible. Be mad at governments for not forcing companies to cost in the negative externalities of their business.

Yes well my mouth was full of breakfast so the toast was censoring me.

Incidentally, taking down a domain used for short links doesn’t prevent speech or publication, since they have about 20 other domains that the same info is available at. Like how knocking over a newspaper box doesn’t censor the paper. So, by your own definition this isn’t censorship. Which is weird because it probably is censorship. Almost like your definition is bad.