Consider the 7600x3d also if you're in the market for something like this. It has a slightly higher base and boost clock, and half the TDP. It's 6 cores instead of of 8 but if you're building a gaming machine that may not make a big difference. If you're putting it in a small form factor the lower TDP might be helpful.
HN user
tedd4u
The corps ruled the US until Trump. Now it’s moving to an oligarchical / kleptocratic mode like Russia. Sure the corps are still involved, but not on top. Will they remain after the current administration? I think it will be hard to put the genie back in the bottle.
Should we compare release date to release date? Or training date to training date?
An article exploring research and citations of "first sleep" and "second sleep"
https://www.patheos.com/blogs/geneveith/2013/08/first-sleep-...
There's an opportunity to insert or remove a leap second twice a year. They only decide about 6 months in advance of each opportunity what to do (leap second, skipped second, or do nothing).
Probably around the same people started saying "I have an ask" and "that's a nice solve"
Right, will the devious unaligned thoughts squish over to a “K-space” we (the trainers) are not aware of?
See also Ouro [1]. Good citations in the “related work” section.
The Ouro looping results are interesting [1] and they are focused more on the improved reasoning from looping middle layers rather than the parameter efficiency aspect. They train 1.4 and 2.6B parameter models with 7T tokens. The training includes learning how many times to loop on any given token (there’s an early exit module). My guess as to why (as far as we know) looping is not in frontier models yet is that, at frontier training run scale, it’s probably going to require a lot of trial and error and at-scale research. While currently they already probably have a list of dozens or hundreds of of promising ideas that don’t complicate things as much. In the other hand, Ouro’s looping technique shows ability to compete well with models with 3x parameters which seems attention-getting to me. If there’s another 3x to be had down that path. It’s order of magnitude opportunity. Btw there is a great related work section in the paper.
I hesitate to propose ulterior motives, but given there have been several seemingly obtuse objections to projection from Rivian, perhaps the CEO is concerned that, if Rivian supports projection, it will harm the perception of the value of their software stack? Related, I think they licensed their stack to VW.
I've seen that video. Guaranteed they wouldn't have put the slightest bit of information in there if they thought would help the competition.
I've heard students react this way to seeing problems on exams that are not strictly of the types taught in class or in previously-assigned homework. "It's not fair! The teacher never showed us how to do problems like that!" This kind of thing was expected and assumed when I was in secondary school, that there would be combining of some concepts from the unit into a single problem. Very worried that there's both a cultural change supported by AI tools that will lead to the outsourcing of thought to the AI rather than outsourcing drudgery.
One problem here is that students that use AI to outsource thinking become people who cannot think. These people are not likely to be very useful to employers or even society. We have to figure out how to allow AI to outsource drudgery but not the thinking itself. It should be a better and better bicycle for the mind not a replacement for the brain.
Pen and paper is just not a very good way to produce text
Can you say more? It's worked well for 4,500 years. Probably the most important invention ever.
Grading written papers has worked fine for 200 years at least.
AMD was shipping faster integrated GPUs than the M1 Pro before the M1 ever hit shelves.
At 30 watts TDP?
Try this WTO site instead. It breaks out oil, natural gas, fertilizer, and agricultural products with 7-day average and prior-year 7-day average.
Don't forget insider threat vector, too.
If you're in San Francisco, note that Catherine Stefani and Matt Haney both voted for this ridiculous bill.
Roll call: https://legiscan.com/CA/rollcall/AB2047/id/1702219
There are many documented, exploited-in-the-wild font-file attacks (one example in 1]). Apple is re-writing their font interpreter specifically to improve security. [2]
[1] https://www.bleepingcomputer.com/news/security/facebook-disc...
[2] https://blakecrosley.com/blog/truetype-hinting-swift-migrati...
Won't it eventually be $1,000 or $5,000 a month? $5k a month would still be 97% less than many developers cost.
Probably legal in Texas? If it's directly over "your land?"
Here a link to the best recent HN-featured long-form article on Japan rail network. Probably spent more time with this than any other item posted here in months.
“Why Japan has such good railways”
One could perhaps put those in a different vault. Sounds like a pain to me. But nothing compared to an email and/or banking compromise.
Ah, sorry, I didn’t realize you already had a sizable array! Maybe you _are_ in the market for this thing!
Agent harness?
Trump cares about winning, and appearing invincible, a lot. That’s why in close races, he only endorses near the end when he’s sure who’s going to win.
Three comma guy would now be four comma guy
I know some pretty wealthy people. They are very aware of those who are 10x wealthier than them. If Noam has 1B, he is probably pretty aware of those that have 10B. He's met them and seen their properties, scope, and powers. Likewise, they are thinking about those that have 100B, and those are thinking about Elon, who now has "four commas."
You train the model then do a baseline evaluation. Then you evaluate many variants where you have removed or nulled out different layers or chunks of the model. By comparing the performance of those mutated models to the baseline you can learn a lot about the model. What parts don't have much value and can be removed, the location of "functions" or "facts." Etc. Google it.