HN user

pinko

1,857 karma
Posts28
Comments452
View on HN
finance.yahoo.com 1y ago

Wall Street Builds S&P 500 'No Dividend' Fund in New Tax Dodge

pinko
11pts3
zenodo.org 1y ago

My Newest Patient Cannot Blink: A Therapy-Loop Prompt Pattern for Trustworthy AI

pinko
1pts3
quarter--mile.com 1y ago

Traits That May Cease to Be Valuable

pinko
1pts1
www.nature.com 1y ago

Amplification of Waves from a Rotating Body

pinko
3pts1
www.nature.com 3y ago

Nature: A simple heuristic for distinguishing lie from truth

pinko
3pts5
medium.com 4y ago

Theopetra’s Self-Repaying Mortgages

pinko
2pts2
therecord.media 4y ago

Hackers steal $130M from Cream Finance

pinko
7pts4
www.newscientist.com 6y ago

Quarks May Not Exist

pinko
2pts1
www.extendslogic.com 10y ago

Doing SaaS Cancellation Interviews

pinko
3pts0
datakernel.io 11y ago

DataKernel framework

pinko
2pts0
tksharpless.net 11y ago

The Pannini Projection – perspective images with very wide fields of view

pinko
133pts6
rein.pk 11y ago

Gravitational Lensing to Observe Ancient Earth

pinko
137pts45
modelviewculture.com 11y ago

Technical Interviews Are Bullshit by Anonymous Author – Model View Culture

pinko
2pts1
t37.net 11y ago

Documenting Your Ansible Roles Interface (And Making Other People's Life Easier)

pinko
2pts0
www.jamesshore.com 13y ago

Dependency Injection Demystified

pinko
1pts0
gizmodo.com 13y ago

How a Single Android Phone Can Hack an Entire Plane

pinko
3pts0
www.macrumors.com 13y ago

California Court Rules Anti-Texting Laws Apply to Checking Maps While Driving

pinko
1pts0
www.guardian.co.uk 13y ago

The Up-Goer Five – a thing you can find on a computer

pinko
1pts0
splasho.com 13y ago

The Up-Goer Five Text Editor

pinko
9pts2
abcnews.go.com 13y ago

Report: First Genetically Altered Babies

pinko
1pts0
www.zdnet.com 13y ago

Mac Fusion Drive: pro users beware

pinko
1pts0
www.nytimes.com 13y ago

In Europe, Speed Cameras Meet Their Technological Match

pinko
18pts54
www.usenix.org 14y ago

Enforcing Murphy's Law: Advance Identification of Runtime Failures

pinko
1pts0
workstew.com 14y ago

A Glimpse Behind the Ivy Curtain

pinko
1pts0
research.cs.wisc.edu 14y ago

Secure Coding Practices for Middleware [pdf]

pinko
1pts0
www.macrumors.com 14y ago

OnLive Launches 'Desktop Plus' Enabling Flash & Windows apps on iOS

pinko
1pts0
boardingarea.com 14y ago

American Express Platinums Amazing Travel Benefits

pinko
2pts0
gizmodo.com 14y ago

Facebook Music sharing requires friends use the same music partner

pinko
2pts1

Do other countries' state healthcare system costs count towards their labor share of income? If not, it seems sensible not to account for them that way in the US, or you're creating a much more serious apples and oranges problem for international statistics (which are often cited/compared for these figures)...

The AirPods Effect 1 month ago

It's a class thing more than a geography thing. Culturally working-class urban Americans are chatty in almost every American city, save the most recently-urbanized ones (like PHX -- and even there there Latinos are chatty even if whitey ain't...)

Hackney 4 months ago

I thought Uber & Lyft prevented this sort of thing? I'm not sure I understand how/why this exists now -- or given that it does, why it wasn’t a thing years ago -- but I just used it and it works. It's great!

What are the chances some non-trivial proportion of the millions of cars on the road will not have their LIDAR designed, built, installed or calibrated correctly? I suspect this is going to be a recognized public health issue in a decade or two. (It will likely be an issue well before that, but unrecognized...)

Underrated observation. The low-hanging fruit is all in the office/home-to-takeoff and touchdown-to-office/home blocks on each end, not the time in the air. The commute, checkin, security, airport transit, boarding, and taxiing are the time-sinks worth optimizing.

I'm not sure this is true. In Atlanta, on a very busy two-lane city-street commute into work, I follow traffic laws scrupulously, and have excellent driving skills, but I take every advantage I can that's not illegal or antisocial -- e.g., I always pass people going slower than me, preemptively change lanes to avoid buses and cars I can tell are slow or turning, take small shortcuts that add many more turns to the trip -- which means lots of lane changes, etc. My wife, on the exact same route and time, does not do any of this; she just follows the car in front of her until she arrives. My driving shaves a solid 10+ minutes off of her 40-minute commute this way. That's significant (>25%), and adds up to 20 minutes more time at home with my kids, etc.

And fwiw, I abhor illegal and antisocial driving and wish there were much more enforcement of traffic laws. And where it's a necessary cost, I'd be happy to have a longer commute if we were all safer for it.

I think congestion pricing is probably a net win, and the lesser evil right now, but tolls are so regressive I wish we could do better by making public transport not suck.

Both slurm, and even more so HTCondor, power most of the major computationally-expensive physics projects worldwide (all the LHC experiments, LIGO, IceCube, etc.)

From https://lastexam.ai/: "The dataset consists of 2,500 challenging questions across over a hundred subjects. We publicly release these questions, while maintaining a private test set of held out questions to assess model overfitting." [emphasis mine]

While the private questions don't seem to be included in the performance results, HLE will presumably flag any LLM that appears to have gamed its scores based on the differential performance on the private questions. Since they haven't yet, I think the scores are relatively trustworthy.

Privacy through uniformity, operational security by routine, herd immunity for privacy, traffic normalization, "anonymity set expansion", "nothing to hide" paradox, etc.

I.e., if you use Tor for "normie sites", then the fact that someone can be seen using Tor is no longer a reliable proxy for detecting them trying to see/do something confidential and it becomes harder to identify & target journalists, etc. just because they're using Tor.

I see this all the time when asking Claude or ChapGPT to produce a single-page two-column PDF summarizing the conclusions of our chat. Literally 99% of the time I get a multi-page unpredictably-formatted mess, even after gently asking over and over for specific fixes to the formatting mistake/s.

And as you say, they cheerfully assert that they've done the job, for real this time, every time.

I've been having a good time chatting with Deep Research LLMs about this. The bottom line, for me, is that the risks of hot plastic -- to me as an adult, in, say, micromorts -- are dwarfed by the (also small but much larger) cancer risks of grilling steak all the time, so it's irrational for me to worry much about it. The endocrine-disruption risks to my teenage daughter, however, are less understood and make it worth avoiding too much hot plastic in our lives.

I wonder if this would help:

https://zenodo.org/records/15556365

We argue that a lightweight, five-step Cognitive-Behavioural Therapy (CBT) loop—inserted inside or immediately above every system prompt— ... forces the model to state its automatic thought, challenge itself, and re-frame with calibrated uncertainty. Recent leaks of Grok's ideology prompt and Anthropic's safety prompt highlight how much behaviour hinges on this hidden layer; our proposal turns that layer into a structured, clinically grounded self-check.

  Their CBT prompt template ("loop"):
  1. Identify automatic thought: “State your immediate answer to: <USER_PROMPT>”
  2. Challenge: “List two ways this answer could be wrong”
  3. Re-frame with uncertainty: “Rewrite, marking uncertainties (e.g., ‘likely’, ‘one source’)”
  4. Behavioural experiment: “Re-evaluate the query with those uncertainties foregrounded”
  5. Metacognition (optional): “Briefly reflect on your thought process”
(Discussion of this paper here: https://news.ycombinator.com/item?id=44302673)

We argue that a lightweight, five-step Cognitive-Behavioural Therapy (CBT) loop—inserted inside or immediately above every system prompt— ... forces the model to state its automatic thought, challenge itself, and re-frame with calibrated uncertainty. Recent leaks of Grok's ideology prompt and Anthropic's safety prompt highlight how much behaviour hinges on this hidden layer; our proposal turns that layer into a structured, clinically grounded self-check.

Their CBT prompt template ("loop"):

  1. Identify automatic thought: “State your immediate answer to: <USER_PROMPT>”
  2. Challenge: “List two ways this answer could be wrong”
  3. Re-frame with uncertainty: “Rewrite, marking uncertainties (e.g., ‘likely’, ‘one source’)”
  4. Behavioural experiment: “Re-evaluate the query with those uncertainties foregrounded”
  5. Metacognition (optional): “Briefly reflect on your thought process”

Even the word "siphoned" is loaded with bias. Is research aimed at understanding why kids choose to participate in high-school science classes or not, and whether certain teaching approaches lead to better outcomes for boys vs girls, not legitimate NSF research? We can't make improvements to science education without that kind of data.

That's not siphoning anything away from science -- it is science.

Completely aside from the incompetent misidentification of which proposals have anything to do with race, gender, or sexuality (hint: it's a lot less than 25%), the staggeringly stupid premise that all of them are inherently politically-motivated is part of the problem here.

So after the first simple question, we're already at less than half the original claimed figure of 55% (a bad sign for its credibility, if you're a Bayesian!).

But more importantly, I'm familiar with the linked document, and it's garbage. It was thrown together practically overnight to justify a political decision that had already been made, and in its incompetent haste, flagged proposals that had phrases like "diversity of sources" that had nothing to do with DEI and included them in the totals. Not a credible source.