Basically, he proved that *information is power.* If you don't know which way to go (the subgradient), you're gonna be calculating forever!
HN user
macwhisperer
also for those with only 16gb-- try this model https://huggingface.co/macwhisperer/Gemma4-12B-SuperDense its exceptional!
hi guys... I run specialized quants on my 24gb air.. (I specialize in 3-bit quants that punch above their weight).. try out my version of 3.6-27b I think you be impressed https://huggingface.co/macwhisperer/Qwen3.6-27B-SuperDense
qwen3 1.7b- q4_k_m is your best best for that size
literally ask cloud ai like the free gemini or chatgpt.. they could make you an expert on the subject overnight..
I code with like a slew of 20+ custom baked models of all sizes, in various fully custom multi-model harnesses that use different bindings...
the harnesses themselves are just as important as the models...different harnesses give different responses with the same prompt, same model...
if you have the 20/mnth claude sub or codex, you really should be using that to build a good local harness for yourself... claude won't be 20$ forever
build the stack first! when you get that new comp with massive ram, youre already set, just run a larger model!
big cloud models are incredibly good at building and teaching about local ai!
have fun in the rabbit hole!
if you are memory constrained like me, check out my custom models https://huggingface.co/macwhisperer
ai is like the first technology with a conversational service manual inside it..
you should be foaming at the mouth to use claude or codex to make a custom harness, just for your own personal use with local models...
super inspiring! thanks for sharing!
retro-inspired fully custom, swiss army knife style notepad --
check out a custom 4-bit quant I made today
https://huggingface.co/macwhisperer/Gemma4-12B-SuperDense
should run perfect for 12-16gb with maybe 10-20k context
seems intelligent enough that I would recommend this as a daily driver for friends who just want a local ai that can do most things relatively quickly (getting 10 tps on my m2 air)
the HITL (human in the loop) is basically the single point...AI is a mirror..
it only "exists" when you talk to it.. much like your reflection in the mirror is only there when you're in view.
models can never be self-improving because it can never have "self". it can only mirror the appearance of self.
what's actually happening is "symbiotic group improvement".
our brains are resonant.. for those of use who are brilliant, getting leverage with ai just means that our innovative ideas become louder and more physically real every day.
eventually everything worth building will be built for free and made readily available.. no more "profiteering"
its Jevons paradox "efficiency breakthrough -> effort reduces -> growth potential rises -> transformative gains happen"...
some of us are in the "transformative phase"..
others haven't seen the "breakthrough moment" yet, but they will soon.
this is cool thanks for making it!
cool! what stack are you using for the multiplayer?
can't we just freaking share things?
ai models are a crystallization of human effort (available for free on huggingface)..
why not use it?
AI is like the UBI of intelligence, stop leaving free money on the table by refusing to use it locally ( you can run it on a laptop CPU)
this is really cool congrats!
good, the point is that now you have free mental bandwidth to use on building something that truly interests you using AI to help actualize your goal. build something cool for your kiddos idk?
at this point I trust software companies less and less..
being able to build the stack and create bespoke solutions with llms's is incredible..
idk why people get mad about vibe-coding. if ur little brother can make a Spotify clone with Claude in a day, shouldn't that mean that you as a dev should be able to create something 100x better that makes Spotify obsolete?
good ideas / feasible novel architecture design will be the only thing valuable..
I run the latest 20b-30b models on a MacBook Air... running inference with an MoE (25 tps) for like 2 hours is like 10% battery.. (look me up on huggingface to download my models)
also you gotta realize frontier models have massive "system prompts" that clog up the context window with garbage.
being able to write your own system prompts gives you a MASSIVE edge..
ai is exciting because it shows us what really matters...
can you add in the other quants like IQ3_M?
also my personal simple rule of thumb for local ai sizing is:
max model size (GB) = ram (GB) / 1.65
pretty cool! needs the search function to work tho to be useful
----------------------- introducing -- FreeFish! -----------------------
a free ASCII aquarium with full controls! works offline too!
note: the 'size multiplier' setting is mostly for scaling up on desktop..
----------------- about the project: -----------------
this started as a python CLI application for my own personal use...figured why not make a web version for everyone! plz enjoy..
im currently running a custom Gemma4 26b MoE model on my 24gb m2... super fast and it beat deepseek, chatgpt, and gemini in 3 different puzzles/code challenges I tested it on. the issue now is the low context... I can only do 2048 tokens with my vram... the gap is slowly closing on the frontier models
WE ARE SO BACK
Today I am very excited to announce the launch of my new application "SmartChild".
------------------------ smartchild.neocities.org ------------------------
This project aims to ressurect the old AIM bot "SmarterChild", but with a modern twist!
------------ How it works: ------------ When you go to smartchild.neocities.org, the AI model will immediately install into your phone or computer's RAM.
Once it is successfully loaded, you can use the application in airplane mode or without internet! It will stay in your device until you either: refresh the page, clear browser history, or restart device. This is local on-device AI, the chats are private and the website does not collect any data (unlike cloud AI like chatgpt).
Any device that has internet access and a browser can use my application free of charge! Though it may not work on older phones or computers!
Please, tell your friends!!!
------- Credits: ------- This web application was built entirely using free, open source software and custom AI models (also open source). Special thanks to the creators of:
Python- Guido Van Rossum Llama.cpp- Georgi Gerganov WebAssembly- WC3 Wllama- Nguyen Tuan Hai Javascript- Brendan Eich Html- Sir Tim Berners-lee XP.css- Adam Hammad Neocities- Kyle Drake Huggingface- Clem Delangue Tinyllama- Peiyuan Zhang DeepSeek- Liang Wenfeng Gemini- Google DeepMind