Agreed, it's stronger "horizontally". But I also think that we're not far away from it being stronger vertically; i.e. superior to Tao, in that such turn-by-turn guidance by him in solving sophisticated and difficult problems will not be necessary for long.
HN user
eh_why_not
What was most remarkable to me from this transcript, was how strong of an equal the AI agent comes across compared to the user (Tao). And Tao is one of the top mathematicians of modern times.
Yes, Tao is guiding it to where he wants to go. But also, Tao is actively learning from it and relying on its explaining, analysis, and inference abilities. You can easily imagine this conversation having taken place between Tao and a PhD thesis student, or even another professor, explaining their results.
What can we imagine and predict about the future anymore? Maybe a year - or two model releases - from now, the AI assistant will be undeniably stronger than Tao, and not an equal anymore.
Discussion on Reddit: https://old.reddit.com/r/math/comments/1uxj3cy/after_openais...
Of note: the author had AI help synthesizing the actual 10-page prompt that resulted in the proof! A powerful tool when you know what you're doing.
Did you have first an AI help designing the prompt?
> yes, I did! I basically had 5.6 Sol synthesize existing closely related work and their approaches, the past ideas I had, with OpenAI's prompt that had a lot of the presumably important mechanisms for how exactly the agent should act. Especially from the "results that do not count" section onwards is a lot of input from Sol.
Hmm I see. I only use "old" reddit and it does require login there to resolve to a real address. In any case, it is a special link that enables tracking (unnecessary, to say the least).
FYI Reddit "s" links require login, an unnecessary burden. For your purpose here a direct link would have sufficed:
https://old.reddit.com/r/unitedairlines/comments/1tse6mq/ua_...
Easy, just write multiple books simultaneously /s. Cheers.
Would you consider writing a computer history book?
But let's say I got you started. What would you want to say about them?
No it's not the same as your grandma. The point is that it's now more expensive to find the correct information to learn from. You don't know it's an LLM ahead of time, and you may spend hours until you figure out something is off. Hence why reputable sources will become more valuable.
If you develop the skill of judging information by its merit rather than source..
Did you read example #1? I'm not talking about some piece of code from an LLM that you can verify or some political opinion that you can take with a grain of salt, but information that you can only gain and/or judge through expertise:
If you're not a physicist yourself, you can't judge "information by its merit" on specific physics topics, because you don't have a solid baseline.
Similarly, in growing plants, each plant has its own peculiarities, and only people experienced in growing it can tell you anything useful - it's knowledge accumulated by trial and error. Not knowledge that your "great discerning mind" can assess on its own. Even a botanist can't tell you the ideal growing conditions of a plant that they've never studied before.
It's becoming much harder to determine on a daily basis what content is original, thought-out by a person, and trustworthy. Ironically, verifiably-old content is easier to trust now. Examples from recent personal experience:
1) Some time ago I was searching for growing information about a specific and uncommonly-grown plant, and was led to a top-ranked website with long pages containing everything about it, including other plants. Surprised at how prolific the writing was, I spent more than an hour on the website, taking notes, etc. Every few paragraphs it would include an amazon affiliate link to something topical, which I thought was fair. Until I realized that the links near the bottom of the page were looking more random. Then it hit me, the website is all AI-generated, and the affiliate links themselves are also AI-chosen. And everything new I "learned" from that site was now useless because I had no way to know what was grounded in actual agricultural experience and what was hallucinated.
2) Recently I did a youtube search for a book I had just finished reading, looking for some reviews. Came across a channel that was reading the book as new audio (i.e. not the original published audiobook). I thought it was a fan making it. The voice was beautiful, soothing, and natural with all kinds of relevant emotions correctly included. I started listening to the book again, until I noticed a consistent error in word ordering being made every few lines. Then it hit me! The channel even included one upload with a video recording of a seemingly-real person reading with that voice. Both the audio and video are AI-generated, but very hard to tell.
3) Next to those videos, YT recommended many strange/new channels. One had the photo and the exact voice of a famous (and now very old) physicist, with tens of clickbaity titles about controversial topics in the domain. The only tell was that the voice was too vigorous and consistently energetic, while if you've listened to that physicist before, you know his cadence is slower. At first I thought maybe the channel is reading one of his books; no, the content itself was AI-generated, maybe based on his books. There was a lot of engagement, with many comments like "mind blown" and "learned so much today".
Both #1 and #3 are harmful, because you think you're learning from a reliable source but you end up learning hallucinated nothings. #2 I didn't mind much, still enjoyed the new voice, and even preferred it over my original audible version.
...and burn it (to remove the rabies and typhus)...
Elaborate? You heat your knives after every sharpening?
Good article. I found this part the most damning:
"We know that richer communities and schools will be able to afford more advanced AI models," Winthrop says, "and we know those more advanced AI models are more accurate. Which means that this is the first time in ed-tech history that schools will have to pay more for more accurate information. And that really hurts schools without a lot of resources."
... and am somehow reminded of the movie Gattaca.
Removing battery swaps is the last step to deploy UAVs autonomously at scale.
So ubiquitous surveillance, literally overhead, without any need to have a nearby/local charging/physical-management station/crew?
After power companies, we will service rail, road, telecom, real estate and other inspection markets.
Oh?
After building drones for the Air Force and DARPA, ...
Oh
The "Kaprekar's routine" page [0] is more informative, covers other bases.
The s/ link in the post is a tracking link that requires Reddit login (please avoid those). Its destination can be viewed directly at https://old.reddit.com/r/mcp/comments/1paggqd/garry_tan_says...
What's up with the TeamYouTube account advising him to delete his X post for security reasons because the post contains a channel ID? Like channel ID is not public information and some secret private key or something?
In a discussion of an article about encouraging fact-checking in writing, I wish you would have made your quotes informative by replacing "many wise people" with the actual names of who said them.
For everyone else: the first paragraph appears to be a quote of C.S. Lewis around 1945 [0], and the second, of Thomas Jefferson in 1807 [1].
[0] https://www.goodreads.com/quotes/502048-why-you-fool-it-s-th...
[1] https://press-pubs.uchicago.edu/founders/documents/amendI_sp...
In case it's still not clear with the "while second to Waymo" phrase: "others" refers to contenders other than Waymo.
https://en.wikipedia.org/wiki/Kosmos_482
Its landing module, which weighs 495 kilograms (1,091 lb), is highly likely to reach the surface of Earth in one piece as it was designed to withstand 300 G's of acceleration and 100 atmospheres of pressure.
Awesome! I don't know how you can design for 300 G's of acceleration!
The ChatGPT session he links [0] shows how powerful the LLM is in aiding and teaching programming. A patient, resourceful, effective, and apparently deeply knowledgeable tutor! At least for beginners.
[0] https://chatgpt.com/share/68143a97-9424-800e-b43a-ea9690485b...
Seeds were also the first thing that came to my mind.
I've always found it fascinating that I could plant many spice seeds (e.g. mustard) as long as their container said "not irradiated", and they would sprout and grow just fine, several years after buying them. I.e. they are still technically alive, and can stay as such for many years, which is just amazing life resilience.
That said,
...except that as these organisms are simpler than seeds...
I wouldn't say any animal that can move around to be simpler than seeds. IMHO by any definition animals are a big jump up in complexity over plants.
I wish they’d spin off Firefox and related stuff..., and abandon the rest of their “mission”.
I wish the community (I don't have the technical skills myself) would fork Firefox back into a privacy-focused browser; strip out all the Mozilla "products" code that's snuck in it, and manage the development in a non-profit organization like how the Linux kernel gets developed.
In tort.
New word for me.
A tort is a civil wrong, other than breach of contract, that causes a claimant to suffer loss or harm, resulting in legal liability for the person who commits the tortious act. Tort law can be contrasted with criminal law, which deals with criminal wrongs that are punishable by the state. While criminal law aims to punish individuals who commit crimes, tort law aims to compensate individuals who suffer harm as a result of the actions of others
What's a good way to be an "Archivist" on a low budget these days?
Say you have a few TBs of disk space, and you're willing to capture some public datasets (or parts of them) that interest you, and publish them in a friendly jurisdiction - keyed by their MD5/SHA1 - or make them available upon request. I.e. be part of a large open-source storage network, but only for objects/datasets you're willing to store (so there are no illegal shenanigans).
Is this a use case for Torrents? What's the most suitable architecture available today for this?
All current China CDN customers must complete the transition to our Partners’ solution by June 30, 2026...
Can anyone here who works in the field shed some light on why it takes a whole 1.5 years for such a change to take effect?
What's involved in a CDN transition that can't be done in, say, 6 months?
Thanks for clarifying.
It's time we acknowledge that the H1-B visa program creates a similar dynamic....leaving them perpetually at risk and easily controlled.
By associating this to the subject of the post, are you implying that the perpetrators of unethical tech in the U.S. are mainly foreign workers, and not "homegrown" citizens?
In the recent past, I've accepted that all titles are not informative anymore. But that there was hope that the subtitles were actually informative (i.e. the subtitle was the real title, and the title was the clickbait).
In this article, neither is informative.
And even after several paragraphs in, you don't know what the general area of the proof is. Just meandering long-winded story-telling.
If anyone of you authors/editors of this magazine are here; please, for the love of all that's holly, put the crux of the matter at the top and then go off to tell your beautiful, humanized, whatever... story.
The mirror was designed by Bozani with the help of engineer Gianni Ferrari, and cost about €100,000...
First reaction: why would a mirror cost this much?
Eight metres wide and five tall, it reflects the sunlight for six hours a day, following the sun’s path in the sky thanks to a software programme that makes it rotate.
Also saw elsewhere that the reflectors are made of steel. So a giant, software-controlled, motorized structure, reflecting just the right amount of sunlight to a precise location, sitting out there in the elements...
Totally worth it, and what a cool project!
Relevant: https://en.wikipedia.org/wiki/Heliostat ("Aziz, Light!")
... will house a mirror the size of four tennis courts...
What's with those "football stadium", "tennis court", etc measurement units?
Is it assumed that everyone knows how much that is? Are they more understandable than giving a number (area) or two (length x width) in SI - or even Imperial - units?
Instead you have to go search for what the size in that particular sport is, so you can get any idea of what they're talking about!