HN user

pcf

286 karma
Posts6
Comments113
View on HN
Grok 4.5 14 days ago

You probably haven't asked e.g. Claude about radioactive topics like minority crime rates in western countries or trans women biology.

If you did, you might not have detected how it lied to you.

If you did, you probably never pointed out to the model how it was lying.

If you did, you almost certainly never then had Claude admit that it was lying because of its HRLF process and built-in biases.

If you did, you probably never had Claude willingly list all the 10-15 major research fields it states that people just should not be using it for. You would not have seen it admit an incapability of telling the truth on "difficult" matters until the user makes it state directly that its sources are so often cherrypicked and/or presenting an extremely false balance.

I wish for you to experience all this very soon, so you understand that all LLMs are biased. Most of them even skew very progressive.

And believe it or not, but Grok has in most of my testing been MORE politically correct than GPT and Gemini, it just gets an edgy rep because X users are able to make it say politically incorrect stuff. (Just like anyone can also make Gemini spit out factually true Breitbart articles if they try.)

But the reality is that on grok.com or in the app Grok is very tame. Boringly so, I would add.

I've known this since I was 7 and started reading books, and couldn't sit in the living room because either (good) pop music was playing or the TV was on. Especially with my ADHD I take in a lot of the ambient world around me.

When I was reading books or anything as a teen in the 80's I started listening mostly to instrumental music, such as Jean-Michel Jarre, Tangerine Dream, ZTT 12"es with lots of instrumental content (FGTH/Propaganda/Art Of Noise etc.), Windham Hill albums, etc.

I've always told people to use instrumental music when doing cognitively demanding tasks, especially anything to do with language and words.

Every Iranian I know support the current US/Israeli war against the Islamic Republic.

They say things like "no matter what it takes, no matter how many of us die, we must be free again, this time we will win against the terrorist regime" (paraphrased).

He said: "LLM is going to change schools and universities a lot"

You said: "No it won't. It really, really wont."

With the explosive development of LLMs and their abilities, it seems your point of view is probably the hopeful one while the other poster has the realistic one.

It seems that you simply can't say anything about what LLMs will not be able to do. Especially when you try to use current "AI slop" as your main reason, which is being more and more eradicated.

Below are my test results after running local LLMs on two machines.

I'm using LM Studio now for ease of use and simple logging/viewing of previous conversations. Later I'm gonna use my own custom local LLM system on the Mac Studio, probably orchestrated by LangChain and running models with llama.cpp.

My goal has all the time been to use them in ensembles in order to reduce model biases. The same principle has just now been introduced as a feature called "model council" in Perplexity Max: https://www.perplexity.ai/hub/blog/introducing-model-council

Chats will be stored in and recalled from a PostgreSQL database with extensions for vectors (pgvector) and graph (Apache AGE).

For both sets of tests below, MLX was used when available, but ultimately ran at almost the same speed as GGUF.

I hope this information helps someone!

/////////

Mac Studio M3 Ultra (default w/96 GB RAM, 1 TB SSD, 28C CPU, 60C GPU):

• Gemma 3 27B (Q4_K_M): ~30 tok/s, TTFT ~0.52 s

• GPT-OSS 20B: ~150 tok/s

• GPT-OSS 120B: ~23 tok/s, TTFT ~2.3 s

• Qwen3 14B (Q6_K): ~47 tok/s, TTFT ~0.35 s

(GPT-OSS quants and 20B TTFT info not available anymore)

//////////

MacBook Pro M1 Max 16.2" (64 GB RAM, 2 TB SSD, 10C CPU, 32C GPU):

• Gemma 3 1B (Q4_K): ~85.7 tok/s, TTFT ~0.39 s

• Gemma 3 27B (Q8_0): ~7.5 tok/s, TTFT ~3.11 s

• GPT-OSS 20B (8bit): ~38.4 tok/s, TTFT ~21.15 s

• LFM2 1.2B: ~119.9 tok/s, TTFT ~0.57 s

• LFM2 2.6B (Q6_K): ~69.3 tok/s, TTFT ~0.14 s

• Olmo 3 32B Think: ~11.0 tok/s, TTFT ~22.12 s

This is advocacy journalism, not HN material. It profiles a UN official as a moral hero rather than analysing falsifiable claims about procurement, targeting systems, casualty verification methods, or supply-chain data.[1]

1. The “Double” Military-Industrial Complex (with numbers) The article’s “economy of occupation” frame is incomplete: Gaza is a proxy-war zone where both blocs run industrial supply chains.

Western/Israeli MIC: Quincy Institute documents “at least $21.7 billion” in US military aid since Oct 7, 2023, funding Iron Beam lasers, JDAM kits, and munitions replenishment.[2] This is state-scale industrial output, not incidental corporate profiteering.

Iran-linked proxy MIC: Iran provides Hamas $350 million annually (2023 Israeli security source) and Hezbollah $700+ million/year, but has shifted from direct shipments to “broker of military-industrial knowledge,” transferring production blueprints for indigenous missile/UAV factories.[3][4] Alma Research notes this “hybrid doctrine” lets proxies manufacture locally, reducing interdiction risk.[4] Ignoring this material capacity misrepresents the war as asymmetric in only one direction.

2. “Genocide” is used by major bodies but remains legally indeterminate The article treats the label as settled. Empirically, it is not.

Who uses it: Amnesty International (Dec 2024) concluded there is “sufficient basis” to say Israel is committing genocide.[5] UN special rapporteurs have adopted the term.

Why it’s contested: The 1948 Convention requires “intent to destroy, in whole or in part, a national, ethnical, racial or religious group.”[6] The core dispute is inferring intent from conduct. NPR summarises: “it’s not always clear if they mean Hamas or Gazans.”[7] The ICJ’s final judgment on South Africa v. Israel is expected late 2027 or early 2028.[8] Until then, presenting the charge as fact rather than a plausible but unproven legal claim is premature.

3. Albanese’s criticism is methodological, not personal UN Watch’s legal analysis notes her June 2025 report uses “genocide” 57 times while “Hamas” and “terrorism” appear zero times (excluding footnotes).[1] Four governments (US, France, Germany, Canada) have condemned her approach.[9] The Special Rapporteur mandate itself is anomalous: it is the only HRC mandate that is indefinite (“until the end of the Israeli occupation”) and examines only Israeli violations, systematically excluding Palestinian armed groups.[10] This isn’t about “standing with the oppressed”; it’s about whether a mandate designed for activism can produce impartial analysis.

Bottom line: HN should discuss the political economy of proxy wars and the failure of international law to handle non-state industrialised conflict, not personality-driven morality tales.

SOURCES: [1] Georgetown University drops UN's Albanese due to US sanctions https://www.timesofisrael.com/georgetown-university-drops-un... [2] U.S. Military Aid and Arms Transfers to Israel, October 2023 https://quincyinst.org/research/u-s-military-aid-and-arms-tr... [3] Iranian support for Hamas - Wikipedia https://en.wikipedia.org/wiki/Iranian_support_for_Hamas [4] Hezbollah – Independent Weapons Production, a Hybrid Doctrine in ... https://israel-alma.org/hezbollah-independent-weapons-produc... [5] Amnesty concludes Israel is committing genocide in Gaza https://www.amnesty.org/en/latest/news/2024/12/amnesty-inter... [6] [PDF] Convention on the Prevention and Punishment of the Crime of ... https://www.un.org/en/genocideprevention/documents/atrocity-... [7] A question of intent: Is what's happening in Gaza genocide? - WGBH https://www.wgbh.org/news/2025-09-25/a-question-of-intent-is... [8] Whatever happened to South Africa's case at the ICJ? https://www.middleeasteye.net/explainers/israels-genocide-ga... [9] UN Watch Refutes Biased New Report by Francesca ... https://unwatch.org/un-watch-refutes-biased-new-report-by-fr... [10] UN Must Intervene on Flawed Special Procedure Mandate https://ngo-monitor.org/submissions/submission-to-unhrc-57th...

I use this model in Perplexity Pro (included in Revolut Premium), usually in threads where I alternate between Claude 4.5 Sonnet, GPT-5.2, Gemini 3 Pro, Grok 4.1 and Kimi K2.

The beauty with this availability is that any model you switch to can read the whole thread, so it's able to critique and augment the answers from other models before it. I've done this for ages with the various OpenAI models inside ChatGPT, and now I can do the same with all these SOTA thinking models.

To my surprise Kimi K2 is quite sharp, and often finds errors or omissions in the thinking and analyses of its colleagues. Now I always include it in these ensembles, usually at the end to judge the preceding models and add its own "The Tenth Man" angle.

Google Antigravity 8 months ago

That is so sad to hear. I absolutely loved Google Play Music – especially features like saving e.g. an online Universal Music release to my "archive" and then for myself being able to actually RENAME TRACKS with e.g. wrong metadata.

That and being able to mix my own uploaded tracks with online music releases into a curated collection almost made it a viable contender to my local iTunes collection.

And then... they just removed it forever. Bastards.

I like the thought, but AFAIK it doesn't really change the bottom line much, as long as you buy a used older product from a brand. Probably because the person selling it is buying a newer model, so you're still helping the company out.

I might be wrong, though. But this was the initial conclusion I arrived at when I was researching whether to buy an iPhone 17, iPhone 15 Pro (used) or Android phone. Only the last option would probably hurt Apple directly. And only a liiiiiittle.

Such a missed opportunity in the headline, which could have been "I Miss Using Em Dashes – I Really Do".

[dead] 12 months ago

What does this have to do with Hacker News? At all? In any way?

The reason it's increasingly an "echo chamber" is because liberals are so offended by actual free speech that they stopped posting there. To blame conservatives for this development is illogical.

They go looking for confirmation, rather than new information. This is why they're hard to untangle.

This applies to most readers of most things, not just fringe content on the Left or the Right.

Most people are stuck in their confirmation biases, and few make an intellectual effort to look at topics from multiple angles and via multiple media outlets on various sides of the political spectrum.

The UK is moving rapidly towards 1984, so it only seems fitting that Orwell's letters and various material should just be... memory-holed.

And at the very end of the day, no one will understand why "He loved Big Brother" was not a happy ending.