HN user

wavelander

1,045 karma
Posts100
Comments26
View on HN
www.gradient-bang.com 3mo ago

Gradient Bang: a multiplayer game built with LLMs

wavelander
1pts1
uditsaxena.bearblog.dev 4mo ago

Spirals – Or Visualizing Cycles and Patterns

wavelander
3pts0
www.businessinsider.com 4mo ago

OpenAI, Anthropic turn to consultants to fight over the enterprise market

wavelander
2pts0
www.threads.com 5mo ago

Halide cofounder Sebastiaan de With joins Apple's design team

wavelander
1pts0
felixrieseberg.github.io 6mo ago

Claude Coach: Free Personalized Training Plans with AI

wavelander
1pts0
uditsaxena.bearblog.dev 7mo ago

Speculative Decoding in LLMs

wavelander
1pts0
danielmiessler.com 7mo ago

Anthropic's Vision Advantage Is a Lot Like Apple's from the 2010s

wavelander
2pts0
openai.com 8mo ago

OpenAI: Understanding neural networks through sparse circuits

wavelander
12pts0
fortune.com 8mo ago

A profile of OpenAI's 'builder-in-chief', Greg Brockman

wavelander
3pts0
openai.com 8mo ago

OpenAI for Science

wavelander
1pts0
uditsaxena.bearblog.dev 9mo ago

Airflow at Asapp: Enhancing AI-Powered Contact Centers (2024)

wavelander
1pts0
arxiv.org 9mo ago

The Art of Scaling Reinforcement Learning Compute for LLMs [Meta]

wavelander
1pts0
blog.python.org 9mo ago

Python 3.14: Free threaded Python is here

wavelander
7pts1
uditsaxena.bearblog.dev 9mo ago

On Agency, and Design

wavelander
2pts0
huggingface.co 9mo ago

DeepSeek-v3.2

wavelander
6pts0
arxiv.org 10mo ago

Predicting the Order of Upcoming Tokens Improves Language Modeling

wavelander
7pts2
www.youtube.com 11mo ago

The Oak Architecture: A Vision of SuperIntelligence from Experience, Rich Sutton

wavelander
2pts0
huggingface.co 11mo ago

From Zero to GPU: A Guide to Building and Scaling Production-Ready CUDA Kernels

wavelander
3pts0
ludwigabap.bearblog.dev 11mo ago

On Becoming competitive when joining a new company [2024]

wavelander
3pts0
www.anthropic.com 1y ago

Anthropic signs a $200M deal with the Department of Defense

wavelander
92pts100
mukulsaxena.wordpress.com 1y ago

Confusing Innovation with 'Jugaad'

wavelander
2pts0
muellerzr.github.io 1y ago

Limiting Qwen 3's Thinking

wavelander
1pts0
www.anthropic.com 1y ago

Exploring Model Welfare

wavelander
9pts0
developer.nvidia.com 1y ago

Nvidia CuML: Zero Code Change Acceleration on the GPU for Scikit-Learn

wavelander
3pts0
research.trychroma.com 1y ago

Generative Benchmarking

wavelander
1pts0
github.com 1y ago

DeepSeek drops recommended R1 deployment settings

wavelander
2pts0
twitter.com 1y ago

Replit Agent: Build mobile apps using AI

wavelander
4pts2
kyunghyuncho.me 1y ago

I sensed anxiety and frustration at NeurIPS 24

wavelander
201pts133
www.bbc.com 1y ago

Ratan Tata, Patriarch of Biggest Indian conglomerate, passes away

wavelander
4pts1
blog.isaacmiller.dev 1y ago

Betting on DSPy for Systems of LLMs

wavelander
83pts18

I believe this is a fantastic idea, even if only because I literally was ideating on the same lines just last week.

I think there were some usability issues, but overall, solid stuff.

I think if I were you, I would focus on the UX of legislative bill analysis rather than the LLM side of it, but that's just me.

I wonder if that's true at a certain stage of OpenAI, which because of the product bootstrapping skills of Sam and co, has made his role irrelevant?

I mean, Jakub can take it forward at the current scale and leadership team of Sam and other people, but maybe he could not have earlier, which is where Ilya shone?

I don't have a horse in this race, and maybe the root comment came off as flippant and disparaging. But I'm not reading "outsiders" as being what you say "gatekeeping".

Maybe another perspective is that "outsiders" may not have the same view of the issue as experts in the field and may not (historically, in OP's experience) seem to want to work together with the experts to develop this view. Handwaving away complexities and not willing to get hands dirty is something I've seen as well so maybe I'm a bit more empathetic, but cold shoulders from "experts" towards newcomers is definitely a thing.

Both of which could help both sides - bring more depth to the fresh view of the "outsiders" and actually bring valuable freshness to the depth of the "experts".

I'm not sure folks who're putting out strong takes based on this have read this paper.

This paper uses GPT-2 transformer scale, on sinusoidal data:

We trained a decoder-only Transformer [7] model of GPT-2 scale implemented in the Jax based machine learning framework, Pax4 with 12 layers, 8 attention heads, and a 256-dimensional embedding space (9.5M parameters) as our base configuration [4].

Building on previous work, we investigate this question in a controlled setting, where we study transformer models trained on sequences of (x,f(x)) pairs rather than natural language.

Nowhere near definitive or conclusive.

Not sure why this is news outside of the Twitter-techno-pseudo-academic-influencer bubble.

This makes a lot of sense at deep tech places that don't sell to enterprises where one can put their head down and just work on technical challenges. A similar place like Mongo DB or the hardware division of NVidia comes to mind. But if you're supposed to iterate with any input from more than 2 people that isn't highly objective (technical metrics that can correlate well with product satisfaction), communication and diffused collaboration is important and necessary. That doesn't mean it has to suck. But the Netflix model you described isn't applicable everywhere and shouldn't be blindly replicated.

I just bought "Meditations" by Marcus Aurelius.

Apart from the debate raging on about grabbing/absorbing more from a physical book than an ebook, I feel that I'm simply more attached to the physical aspect of the reading experience.

I've read ebooks when I've been unable to find the hard copies, but I almost always revert to the hard copies.

Well done ! I started reading this - got around to 10 and stopped for a while - and totally forgot about it.

It continued and - finally ! - finished.

PS: Fellow readers, did it ever stop being awesome and/or become unpleasant from the perspective of the original plot ?

Why medium.com? 13 years ago

It used to be that there was REALLY good stuff posted there. A lot of the initial content there was really helpful, quite insightful. I used to hang out there everyday. The writing experience was clean; you always required a photo and the comment sys was new. It was easily cataloged and a lot of amazing writers hung out there.

That was quite recent; this June, I think.

Then it became popular. Turns out anyone who could write properly, using correct grammar, make it seem like they were thinking differently than the herd, and/or were an entrepreneur. Or want to post some fluffy pseudo-intellectual stuff.

You really have to wade through a lot of bad stuff, to get to the good ones. It started out well. But the section "Editor's Picks" already has like 2900 articles or something.

Nowadays it's more like you select the best from the week's best or the month's best.

Ah. This site. I came across this when I was a kid - like 5-6 years ago - I had googled the term "how to become a hacker". This guy had opened my eyes then - I was 14 then. He still continues to do so. His links are still relevant and an amazing starting point for anyone who wants to learn.

One should probably keep visiting this page like every 6 months or so - just to check whether you are on the right track to learning.

A much needed perspective. All I've read are the H/W specs. Though I'm still apprehensive about spending so much on it, it does change how I see it. Still, better options seem abound.