HN user

jiocrag

364 karma
Posts5
Comments81
View on HN
Grok 4.5 14 days ago

You’re joking, right? The model that says Elon Musk could beat LeBron James in a 1 on 1 isn’t biased?

same here. Paying for Pro ($200) but the "try it" link just leads to the Pro sign up page, where it says I'm already on Pro. Hyper intelligent coding agents, but can't make their website work.

Gemini 2.5 Flash 1 year ago

Not at all. The model weights and training data remain the same, it's just RAG'ing real-time twitter data into its context window when returning results. It's like a worse version of Perplexity.

Gemini 2.5 Flash 1 year ago

Excellent point. If they can figure out how to either remunerate or drive traffic to third parties in conjunction with this, it would be huge.

OpenAI O3-Mini 1 year ago

why is this impressive at all? It effectively amounts to correcting a typo.

If Bard is using PaLM 2, Google is in serious trouble. Here's its offering for "the simplest PostgreSQL query to get month-over-month volume and percentage change." Note that no actual calculations take place and the query generates a syntax error because it references a phantom column. GPT 3.5 and 4 handle this with ease.

SELECT month, volume, percentage_change FROM ( SELECT date_trunc('month', created_at) AS month, SUM(quantity) AS volume FROM orders GROUP BY date_trunc('month', created_at) ) AS monthly_orders ORDER BY month;

It’s literally addressed in his first bullet point…

“ Response: While there is evidence that some tasks that appear emergent under exact match have smoothly improving performance under another metric, I don’t think this rebuts the significance of emergence, since metrics like exact match are what we ultimately want to optimize for many tasks.”

How could we possibly ever test for "not feeling like an automaton?" The only possible test, it seems, is internal subjective experience. Even if something reported this, how would you ever verify it?

Aren't we inherently projecting feelings onto anything that isn't inside our own direct experience? There is no way to confirm any alleged sentience outside of your own "feelings" is not an automaton, including other humans.

AFAIK, we have no way of assessing whether “someone” has acquired new information other than that someone demonstrating their knowledge. We can’t MRI “acquisition of information”. So, as per this thought experiment, we have no way of telling if Mary gained new information by “seeing” red when she already “knew” what red was and could interpret it as such. That’s the whole point of the the thought experiment, which clearly went over your head.

"Simply devise a way to test for information predicated on trichromacy using bichromatic metamers." lol... "Simply..." this is literally the 'the explanatory gap' described in the article. We don't currently have a way to measure if the brain "acquired new information" when it knows about something but then sees it for the first time... which is the whole point. "experience a functional difference" -- what does that even mean?