HN user

programjames

900 karma
Posts4
Comments503
View on HN
Heresy (2022] 4 days ago

Yes, people can always claim that wasn't their intent, but what you see quite often nowadays is people openly saying their intent is to get people to take action. "I wouldn't personally... but I have massive respect for those who do. We need more people with the balls to do something about it."

This is probably less restrictive than you are advocating for in your previous comment: as long as there is plausible deniability, the speech goes free.

Heresy (2022] 4 days ago

The standard can be, "don't utter words with the intent of spurring others onto antisocial actions."

I skimmed through the book, and it's lacking the information theory foundations. For example, "trust region methods" come from maximizing the policy's relative entropy (to a reference policy) under a tournament system where high-scoring agents are exponentially likely to survive. In general, a reward is the negative bits it costs an environment to propagate an agent (multiplied by some temperature).

GPT‑Live 14 days ago

Didn't Standard Intelligence release a duplex model two years ago? Sounds disingenuous to market this as a new generation of voice models, when it is really OpenAI finally catching up to the current generation after two years.

https://si.inc/posts/hertz-dev/

Optimizing for the metric involves:

1. Optimizing for generally applicable skills that the metric is trying to measure.

2. Optimizing adversarially to hill-climb the metric.

You want candidates to do (1) and not (2). You can make them agnostic to the second by setting

    d(expected gain)/d(opportunity cost) = 0
      ==>
    expected gain \propto opportunity cost
It is the case that most metrics are logarithmic: it takes just as much effort to decrease one bit of error as the next bit. So
    log(score) \propto (opportunity cost) \propto expected gain
Thus, for them to be agnostic, you should filter candidates proportional to their log-score on the metric (where 0 is a perfect score). Because generally applicable skills are generally applicable, they will still benefit from improving those, they just no longer benefit from adversarial optimization, unless your score function looks very similar to others who have not adopted this filtering process.

The issue with a hard cutoff is that people near the boundary are extremely incentivized to adversarially optimize, as it is usually cheaper than working on generally applicable skills and actually pays off for them. You see this phenomenon on AoPS where (esp. Californian) students talk about grinding for MATHCOUNTS instead of learning calculus.

Nondeterminism is also a feature, not a bug. If you don't want people to optimize against your filtering process, you have to make it somewhat nondeterministic. For example, better candidates are exponentially more likely to pass the filter, instead of a hard cut-off at the top-100. Then it becomes no longer worthwhile to Goodhart the filtering process, because it barely increases your chances and there are so many more places you can use your time better.

1.0 is "natural units". If your energy corresponds to nats, you should be using temperature 1.0. If your energy corresponds to bits, you should be using temperature ln(2) ~= 0.7. The optimization pressure is

     max nats = max entropy + energy / temperature

Why might energy correspond to bits or nats? Imagine your goal is to play as many interesting games of chess as possible in a tournament. This implies you have to keep winning. If you look at the RL environment from the right perspective, you can turn it into optimizing bits or nats.

You should include the error correction code length in the description length. This means Newtonian mechanics was a much longer theory to describe Mercury's orbit than general relativity. It was only the shorter theory before they had the data showing a discrepancy. Which is the correct approach to describing your reality, because until you see a discrepancy, the extensional properties all follow the shorter rules.

I guess the argument from OP would look like: "Yes, now imagine we poke and extend our universe as far as we can. How much bigger do you think our final 'shortest description' would be? I imagine it may be orders of magnitude more complex."

Well, I can imagine a squared circle... doesn't mean the math checks out. I would reply that you do not have to imagine, you can go about looking at different mathematically possible universes in Tegmark IV and find the expected number of bits for the one you actually exist in. Which is ~0 bits more complex than the shortest description based on the data you currently have.

Also, note that Newtonian mechanics is not actually a very short theory for building a universe, because you have to instantiate every object in the universe. You actually get a lot more of the structure for free with general relativity (re: Wigner's classification of the particles). An observer in a presumed-Newtonian universe calling it a simple theory would be like saying, "I compressed Wikipedia to one byte, just by putting it all in the decompiler!"

My perspective is it is impossible to cut through the three layers of bullshit between you and anyone who knows what they are talking about. The only way to do this is with brand-name qualifications, like "MIT graduate", not things that are actually impressive. This is also why you see senior developers saying, "the offers I'm getting are bigger and bigger," meanwhile skilled younger developers need to become a marketing professional just to get an interview.

Recruiters have utterly given up on being efficient in the market. I do not know why, but there is something very wrong given "spamming the same brand-name fish all the other recruiters are spamming" is their only strategy. My guess is there is a combination of bad (or an entire lack of) hygienic data filtering and a disconnect between compensation and terminal goals (hiring the best candidates).

"A handful of years"? This is like the trend in K–12 education of blaming all issues on the COVID-19 pandemic. No, education was in a visible decline for five years before COVID, which means it had probably been declining for another decade or two before that.

My experience was pretty contrary to points (1) and (4). My best teachers/professors directly conveyed information or skills. I found most students did the bare minimum to pass their classes (where "pass" = "not get their parents mad"). I tried to get a CS club started at my highschool and basically no one was interested, not even my friends.

Now, I did have a great coach in middle school who "created the conditions where willing students will learn", but I don't think she would have been a good teacher. She was great at organizing club meetings, finding the right materials to study, utilizing intraclub competition to motivate everyone, and getting her former students to come back and teach in highschool. I'm sure there was a lot more going on behind the scenes that she just knew how to do right, which made the club a whole lot better. But she wasn't a teacher. Closer to an administrator, but I think "coach" in the (m)athletic sense makes the most sense.

And, this is probably why my computer science club was not the success I envisioned. Yes, people are generally underachievers, but I also did not have the coaching skills to create the conditions where people wanted to overachieve.

I clicked on the article to learn, "why Japanese companies do so many different things," and then got hit with pages of low-bitrate context, such that my eyes started glazing over and it was difficult to find the answer to the question. So I appreciate their compression, or at least pointing to where the answer is found.

The primary objective of the United States, at least according to their leaders, is suspiciously absent from their comment. If we're being charitable, it is clear they disagree with their leaders:

All of this to get to a point where we are negotiating a deal which is worse than what we already had with the JCPOA.

But that deal also ended nearly a decade ago, and the United States has been in talks for more than a year to strike a new deal. It is facetious to say they gain nothing by starting a war, if your excuse is they could have just not blundered a decade ago. Unfortunately, the United States does not yet have access to time travel.

To clarify, in case that was not clear:

"Yes, the leaders say they gain something, but I disagree because they wouldn't have anything to gain if we could just go back in time and fix their blunders."

So, to clarify, you were just trying to spread the meme that Iran giving up its HEU was equivalent or inferior to just holding to Obama's deal, in spite of your belief the commenter you were replying to would not agree with this, nor the instigators of this war?

Maybe don't troll. Or sealion. Or shill. Or propagandize. Not sure which of these is the best description, but it's one of them.

Why would you think returning to Obama's Iran deal would be a win? Actually, let me word this better: how could you possibly think that anyone in the White House for the past decade would think returning to Obama's Iran deal is better than this war?

The reason it's so incredible you could think such a thing is the White House has been saying how horrible that deal was in every interview for months, and of course intermittently for a decade.

I'm sure you've at least heard what the United States' leaders have to say. To paraphrase Trump, "stopping nuclear proliferation to a group of lunatics that support terrorism". And also (still paraphrasing Trump), "the US doesn't need this as much as the rest of the world." So, perhaps this doesn't put the US ahead relative to other countries, but it puts them ahead of the counterfactual nuclear wasteland they could become.

Whether or not you believe the United States' leaders, whether or not you think there was a better way for them to achieve their goals (something something Obama deal) is up for debate. But it's very facetious to say you "can't think of a single way in which the United States came out ahead in the war," when the United States' leaders have been publicly announcing it for nearly a year.