This reminds of sensitivity vs. specificity for medical tests. You essentially have a 2x2 matrix: Error vs. non-error on one axis, and detected vs. not detected on the other. If we assume that a reviewer LLM has 95% accuracy, meaning that it produces the correct result 95% of the time it (identifying errors as errors, and identifying non-errors as non-errors), you get the following: Of 100 cases, the output of 5 is wrong, and the output of 95 is correct. Of the 5 wrong outputs, the LLM correctly identifies 4.75 (95%) as wrong. But it will also create 4.75 false positives (5% of 95). Of the 95 correct ones, the LLM correctly identifies 90.25 (95%) as correct. Therefore it creates 0.25 false negatives. So the positive predictive value is low: When the reviewer LLM flags an error, it's actually only an error 50% of the time. On the other hand, the negative predictive value is high, so in ~99.7% of cases, there will be no error if the reviewer LLM didn't catch one. All of that is assuming that the actual accuracy of the reviewer LLM is truly a flat 95% and that the base error rate is truly only 5%. I think that in reality, LLMs perform better with a certain type of tasks / errors and 5% error rate is just an average.
HN user
juvvel
Oh, the joy was never fully lost to me. I still listen to bands so obscure that Spotify and Youtube Music don't have them, so I have no other choice. I'm slowly reverting back to my tech use of the late 2000s and early 2010s, with the added bonus of being able to access my own music collection from anywhere. I've seriously been considering getting an MP3 player again.
The opening of FFX basically is basically an FMV with edgy metal built in. :D
English is not my native language, but I consider myself fairly fluent. I've never heard the expression "belt-and-suspenders" before Claude.
I've recently noticed an increase in "bite". "This will only bite if..." It also loves "stress-testing", "matrix", "anchor" and "flagging".
I love Tiled Words, thanks for making it.
Kombucha is the best "natural" soft drink for me, too. It's not entirely sugar-free, though (even though you can get it to low-sugar with longer fermentation).
Ah, good old Dragonspice.de, they have provided me with supplies for many of my experiments as well. I have many of the essential oils already, I might try this! Thank you for posting your recipe.
I regularly replay FFX, I recently replayed the Battle Realms remaster, and now I'm once again on Phoenix Wright: Ace Attorney.
I agree with this observation. I often assumed it would be an USP if I highlight that I want to understand how things work instead of blindly following paradigms and frameworks, but it seems that (a sense of) uniformity just comes with too many perceived advantages, like you said.
I don't know the exact reason, but when they don't share the process or prompt, it seems like they're trying to gatekeep their results – which is very ironic from someone using a tool made possible by ingesting other people's work without their consent.
That's interesting to hear, because my impression was that software/web development is a field full of people who are self-taught or at least very enthusiastic about learning new tech. I am personally pretty undogmatic when it comes to languages or tools and I assumed most developers who care about solving problems are the same way.
So you're saying that someone who likes both OOP and Javascript and more functional-style programming could be a sought after candidate?
The Timely clock app. It was bought by Google and then never updated so it doesn't work on newer Android versions. Sad.
Jelly is thickened juice, jam contains the pulp. Why? Different textures.
I don't think there's an apprenticeship for every passion out there. I definitely didn't see my passions reflected in typical job profiles.
Much better than last year, which is what I had been hoping for. There's still a ways to go, but it seems the worst is behind me, finally. It was a long and dark period with extremely crippling anxiety. Ironically, what helped most was to stop trying so hard to get better, but not letting myself go either. A very tricky balance to achieve that mainly hinges on your ability to talk to yourself in a kind, encouraging way. As someone who has an avid aversion to advice that seems to be superficial feel-good fluff I rejected the "be kind to yourself" concept for a while. Which, I realize, was my inner bully hijacking my logical brain making me believe I was doing something "right" by being cynical and unforgiving with myself.
You can also apply retinoids topically. What it does is increasing collagen production.
This would be my dream. What other field did you enter?
This, and the "of course this is just one of many possible solutions" disclaimer at the end. I agree it was probably written by GPT.
I don't know much about professional translating, but wouldn't you be able to track the error rate of the machine translation by looking at the number of corrections? And then translators could use this as leverage to negotiate better pay if it's obvious they need are translating from scratch in, say, 50% of cases anyway.
This might not be the answer you're looking for, but I've frequently heard that charities are in dire need of people willing to do hands-on work instead of the "comfy" administrative stuff. Someone who is willing to guard the entrance to a DV shelter, for example, or someone to come in and do a few loads of laundry. I know we want to help with the skills we have, but it seems charities already have enough supply of people wanting to do some low-key work. Meanwhile the real problems go unsolved.
You can ask it to specify or point out any logical errors you observed and it will correct itself. If there's a contradiction I know that the information might be incorrect. I'm also just using it as a starting point to prime my brain, of course as of today we still need other sources to verify the knowledge.
It can explain stuff to you in exactly the way you need to, and walk you through problems. You can ask it to create quizzes / guided questions for you which helps in organizing your own thoughts. Another application I find quite remarkable is asking it for good prompts for image generation models. The biggest benefit is that you're not starting from zero when entering a new problem space / domain.
However do note that capitalist entities have infiltrated the government and have huge sway over it's regulatory policies meaning that anti-business regulatory policies are unlikely to occur.
Maybe this is the case in the US, the EU is known for being much more strict in its regulations, which is often ridiculed by the rest of the world. Those same people are going to hope for regulations once we see the effects of the current Wild West that is AI.
If we truly live in a world where those in power seek to eliminate all other (non-powerful) humans from the equation, then maybe that's our real problem. For the record, I don't actually believe this is the case. Some people might work toward this goal but I doubt that most powerful people would voluntarily want to give up their influence over the masses. Which is what they'd do if they leave them to fend for themselves while the bots run their businesses. At some point, when you have enough money already, you don't actually seek more money, you seek more influence, and money is just the vehicle. ETA: btw, if a truly autonomous money-making machine is invented, that's just going to cause inflation, making it all worthless.
So any and all commerce just becomes a different form of stock trading? I doubt it. People will still need and want tangible things.
Who will buy the bosses' shit if everyone is unemployed and can't afford it? Companies can't stay afloat by just buying from each other.
Capitalism and its value system is subjective, too. It's not set in stone. I believe we can still steer away from profit as the sole driver of, well, everything, if we want to.
Oh, I do believe programming as a profession is at risk and will change a lot, if not rendered obsolete. What I'm talking about is this idea of "just get used to the fact that there is no human skill that won't be replicable by AI in 2-10 years". It's a very bleak view of the future and our own biological complexity. We need to remember that we are the ones inventing the AI in the first place. We are limited by our imperfect ability to understand ourselves. It will get better, sure, there will be emergent properties, but there's no need to reject the inherent value of humanity even if it happens to produce less economically viable output.