HN user

qntty

3,107 karma
Posts21
Comments890
View on HN
en.wikipedia.org 2y ago

Microcosm–Macrocosm Analogy

qntty
1pts0
news.ycombinator.com 3y ago

Ask HN: How do you keep track of your medical history?

qntty
44pts28
news.ycombinator.com 3y ago

Ask HN: Common Misconceptions about Computers?

qntty
6pts3
beej.us 4y ago

Beej's Guide to Unix IPC (2015)

qntty
220pts54
en.wikipedia.org 4y ago

Pointer Swizzling

qntty
1pts0
www.youtube.com 4y ago

MIT 6.172 Performance Engineering of Software Systems

qntty
2pts0
picoctf.org 4y ago

PicoCTF

qntty
27pts0
www.youtube.com 5y ago

Donald Knuth's Christmas Lectures

qntty
2pts0
en.wikipedia.org 5y ago

Black and white hat symbolism in film

qntty
1pts0
regexone.com 6y ago

An Interactive Regular Expressions Tutorial

qntty
1pts0
en.wikipedia.org 6y ago

Scriptio Continua

qntty
1pts0
doc.cat-v.org 7y ago

Systems Software Research Is Irrelevant (2000)

qntty
68pts43
news.ycombinator.com 9y ago

Ask HN: Why is it so hard to make a contact whitelist on a smartphone?

qntty
7pts1
news.ycombinator.com 9y ago

Remind HN: Make Backups for 2FA

qntty
12pts6
news.ycombinator.com 9y ago

Ask HN: What newsletters or mailing lists are you subscribed to?

qntty
5pts1
news.ycombinator.com 9y ago

Ask HN: What can I do to mitigate climate change?

qntty
3pts0
news.ycombinator.com 9y ago

Ask HN: What laptop should I get instead of a Macbook Pro?

qntty
94pts111
www.theatlantic.com 9y ago

The Mother Behind the Entrepreneur

qntty
2pts0
news.ycombinator.com 9y ago

Ask HN: Show off your weekend project

qntty
19pts16
www.wsj.com 10y ago

Closed Minds on Campus

qntty
22pts0
www.democracyjournal.org 10y ago

Laissez Prayer: The Root of Christian Conservatism

qntty
2pts0

Pre-training mean exposing an already-trained model to more raw text like PDF extracts etc (aka continued pre-training). You wouldn't be starting from scratch, but it's still pre-training because the objective is just next token prediction of the text you expose it to.

Post-training means everything else: SFT, DPO, RL, etc. Anything that involves things like prompt/response pairs, reward models, or benefits from human feedback of any kind.

Sometimes a law is just on its face and unjust in its application. For instance, I have been arrested on a charge of parading without a permit. Now, there is nothing wrong in having an ordinance which requires a permit for a parade. But such an ordinance becomes unjust when it is used to maintain segregation and to deny citizens the First-Amendment privilege of peaceful assembly and protest.

I hope you are able to see the distinction I am trying to point out. In no sense do I advocate evading or defying the law, as would the rabid segregationist. That would lead to anarchy. One who breaks an unjust law must do so openly, lovingly, and with a willingness to accept the penalty. I submit that an individual who breaks a law that conscience tells him is unjust, and who willingly accepts the penalty of imprisonment in order to arouse the conscience of the community over its injustice, is in reality expressing the highest respect for law.

I always have to go back to read this part again because I feel like it's so unexpected. You don't really hear anyone saying quite the same thing today.

Writing a calculus book that's more rigorous than typical books is hard because if you go too hard, people will say that you've written a real analysis book and the point of calculus is to introduce certain concepts without going full analysis. This book seems to have at least avoided the trap of trying to be too rigorous about the concept of convergence and spending more time on introducing vocabulary to talk about functions and talking about intersections with linear algebra.

I tried and failed to find some kind of concrete methodology that they used to get to the number 30 months. I'm still waiting for quadratic algebra to make my knowledge of linear algebra obsolete.

Find Your People 1 year ago

I like the subway analogy. I'm sure I've heard some version of it before, but maybe because I was younger I didn't really get it. It really is a little strange to tell kids who have never really directed their own lives before to start doing it all of a sudden.

You're right, I was confusing TensorRT with Dynamo. It looks like the relationship between Dynamo and vLLM is actually the opposite of what I was thinking -- Dynamo can use vLLM as a backend rather than vice versa.

It sounds like you might be confusing different parts of the stack. NVIDIA Dynamo for example supports vLLM as the inference engine. I think you should think of something like vLLM as more akin to GUnicorn, and llm-d as an application load balancer. And I guess something like NVIDIA Dynamo would be like Django.

I believe this is a question you should ask about vLLM, not llm-d. It looks like vLLM does support pipeline parallelism via Ray: https://docs.vllm.ai/en/latest/serving/distributed_serving.h...

This project appears to make use of both vLLM and Inference Gateway (an official Kubernetes extension to the Gateway resource). The contributions of llm-d itself seems to mostly be a scheduling algorithm for load balancing across vLLM instances.

This just made me realize that we're 10 years past the 100th anniversary of Einstein publishing about general relativity. Which made me realize that we're a quarter of the way through the 21th century...

Also, I think a list made today would have to include some of the early work on deep learning that happened in the 20th century. Which goes to show that sometimes you don't know what's important until much later on.

I see several people complaining about this in this thread. Are you talking about their old interface or the new one they released a couple years ago? I find their new one to be decent. It's a little complicated sometimes, but I think it's hard to build a site that allows you to do so many things without being a little complicated.

Leftists may claim that their activism is motivated by compassion or by moral principles, and moral principle does play a role for the leftist of the oversocialized type. But compassion and moral principle cannot be the main motives for leftist activism. Hostility is too prominent a component of leftist behavior; so is the drive for power. Moreover, much leftist behavior is not rationally calculated to be of benefit to the people whom the leftists claim to be trying to help. For example, if one believes that affirmative action is good for black people, does it make sense to demand affirmative action in hostile or dogmatic terms? Obviously it would be more productive to take a diplomatic and conciliatory approach that would make at least verbal and symbolic concessions to white people who think that affirmative action discriminates against them. But leftist activists do not take such an approach because it would not satisfy their emotional needs. Helping black people is not their real goal. Instead, race problems serve as an excuse for them to express their own hostility and frustrated need for power. In doing so they actually harm black people, because the activists’ hostile attitude toward the white majority tends to intensify race hatred.

Even this seems to imply that there are also dishonorable parts that we are assumed to be aware of. Seems like a disrespectful way to talk about people who agreed to be your photography subjects.

It's disrespectful to candidates time to play misdirection games like this. If the form wasn't put up to mislead people, it would be different. There's a reason most places don't do stuff like this. The problem isn't the extra 30 seconds requires to open gmail, the problem is the kind of immature culture it signals.

This is an interesting connection. If I had to guess, the Chomskyist response would be to say, yes but this only applies to LLMs that have been pre-trained already (i.e. have structures in place needed to understand language in general). I think a Chomskyist would say that language learning is precisely like fine-tuning and not the pre-training of foundation models.