HN user

Feorn

16 karma

Complete failure, and total dilettante. Only a fool would accept anything I believe as fact without independently verifying it.

Posts0
Comments4
View on HN
No posts found.

A completely unquantized fp16 model weight 7B LLM is about 15GB on a disk. You need closer to 24GB of memory for inference with a decently sized context.

Quantization is black magic of the software variety that seems to be able to significantly reduce that without a commensurate loss in quality, though the results are a little subjective. Some well reviewed quantizations of 7B models can get them below 9GB.

Most brains don't treat themselves to an abundance of training data, the ones who do seem pretty knowledgeable to me.

Witty comment aside, the human brain is pretty efficient in terms of energy use considering it's taking in a ton of data while it's conscious. Two each audio and video streams, olfactory, gustatory, touch, vestibular, and all the interoception. Inference and training in real time. All for the low price of 125 watts, a quarter that if you're just measuring the brain and not the whole body.

The paper was published last year. https://www.frontiersin.org/journals/science/articles/10.338...

I'm not convinced this field will outpace silicon or whatever succeeds it, considering how big the semiconductor industry is.

I think ROM sites were an artifact of the internet speeds we had access to early on. Where it wasn't practical to just download an entire library for an older console. Downloading an N64 ROM over dialup still took a little longer than an mp3, whole collections of them were out of the question. Even early broadband in many areas was limited to speeds where a single ROM was more practical to download.

Complete ROM collections were, and probably still are, available for most old systems as torrents and on usenet.