HN user

alphabetting

5,792 karma
Posts204
Comments384
View on HN
stencil.so 7d ago

You only need the frontier model for one single edit

alphabetting
3pts0
www.platformer.news 6mo ago

Debunking the AI food delivery hoax that fooled Reddit

alphabetting
12pts1
abseil.io 7mo ago

Performance Hints

alphabetting
59pts1
blog.jcz.dev 7mo ago

Gemini 3 Pro vs. 2.5 Pro in Pokemon Crystal

alphabetting
315pts92
blog.vllm.ai 9mo ago

VLLM TPU: A New Unified Back End Supporting PyTorch and Jax on TPU

alphabetting
1pts0
jax-ml.github.io 11mo ago

How to Think About GPUs

alphabetting
395pts122
ca.finance.yahoo.com 1y ago

OpenAI taps Google in unprecedented cloud deal despite AI rivalry

alphabetting
4pts0
www.thetimes.com 1y ago

Apple helped China become America's biggest tech rival

alphabetting
2pts1
blog.google 1y ago

DolphinGemma: How Google AI is helping decode dolphin communication

alphabetting
324pts138
speechmap.ai 1y ago

SpeechMap: The Free Speech Dashboard for Al

alphabetting
4pts3
matharena.ai 1y ago

Gemini 2.5 gets 24.4% on MathArena USAMO beating previous top score of 4.7%

alphabetting
54pts10
www.philschmid.de 1y ago

Function Calling Guide: Gemini 2.0 Flash

alphabetting
3pts0
techcrunch.com 1y ago

Trump calls for creation of 'crypto strategic reserve'

alphabetting
11pts0
services.google.com 1y ago

Report on specifically how state-backed hackers are using Gemini [pdf]

alphabetting
4pts0
cloud.google.com 1y ago

Adversarial Misuse of Generative AI

alphabetting
3pts0
variety.com 1y ago

Delta Inks Exclusive Pact with YouTube for In-Flight Viewing

alphabetting
2pts0
www.plasticlist.org 1y ago

Plastic List

alphabetting
12pts1
computer.tldraw.com 1y ago

Tldraw Computer

alphabetting
6pts0
arcprize.org 1y ago

Arc Prize 2024 Winners and Technical Report

alphabetting
127pts56
www.wsj.com 1y ago

T-Mobile Hacked in Chinese Breach of Telecom Networks

alphabetting
5pts1
www.theverge.com 1y ago

More than a quarter of new code at Google is generated by AI

alphabetting
7pts0
blog.google 1y ago

NotebookLM launches feature to customize and guide audio overviews

alphabetting
335pts127
www.wsj.com 1y ago

Mystery Drones Swarmed a US Military Base for 17 Days. The Pentagon Is Stumped

alphabetting
7pts2
half-potato.gitlab.io 1y ago

Ever: Exact Volumetric Ellipsoid Rendering for Real-Time View Synthesis

alphabetting
80pts14
twitter.com 1y ago

Example of NotebookLM converting obscure PDF into realistic podcast audio

alphabetting
1pts0
arxiv.org 1y ago

The Vizier Gaussian Process Bandit Algorithm

alphabetting
2pts0
huggingface.co 1y ago

Everchanging Quest: Rogue-like game powered by LLMs

alphabetting
1pts0
twitter.com 1y ago

OpenAI co-founder John Schulman announces he is joining Anthropic

alphabetting
43pts8
www.volexity.com 1y ago

StormBamboo Compromises ISP to Abuse Insecure Software Update Mechanisms

alphabetting
3pts0
www.youtube.com 2y ago

Momentum Isn't Magic – The Revival of the Hot Hand [video]

alphabetting
1pts0
Notes on DeepSeek 1 month ago

The government doesn’t have to ask NYT to restrict opinions.

This 1988 model of the flow of information in free societies and their media gatekeepers was probably correct. Nearly 40 years later it is not. The digital content flows in free societies is so diverse today that widely read content extremely critical of whichever parties or power-holders you'd like to read about is everywhere and easy to find. Not the case in authoritarian systems.

Gemini 3.1 Pro 5 months ago

the agentic benchmarks for 3.1 indicate Gemini has caught up. the gains are big from 3.0 to 3.1.

For example the APEX-Agents benchmark for long time horizon investment banking, consulting and legal work:

1. Gemini 3.1 Pro - 33.2% 2. Opus 4.6 - 29.8% 3. GPT 5.2 Codex - 27.6% 4. Gemini Flash 3.0 - 24.0% 5. GPT 5.2 - 23.0% 6. Gemini 3.0 Pro - 18.0%

I found this prompt online and tweaking it for audio overviews works extremely well for me.

https://open.substack.com/pub/lawsen/p/notebooklm-podcasts-b...

Generate a deep technical briefing, not a light podcast overview. Focus on technical accuracy, comprehensive analysis, and extended duration, tailored for an expert listener. The listener has a technical background comparable to a research scientist on an AGI safety team at a leading AI lab. Use precise terminology found in the source materials. Aim for significant length and depth. Aspire to the comprehensiveness and duration of podcasts like 80,000 Hours, running for 2 hours or more.

Gemini 2.5 1 year ago

The elo jump and big benchmark gains could be justification

What's disturbing? In all likelihood the close timing was world labs rushing to get their demo out the door knowing this was coming because they wouldn't get nearly the hype they did if this came before.

How so? I think if a team is fine-tuning specifically to beat ARC that could be true but when you look at Sonnet and o1 getting 20%, I think a standalone frontier model beating it would mean we are close or already at AGI.

It’s a lot easier to pay off your would-be competitors than it is to innovate. I’m hesitant to say that antitrust is good for its subjects, but Google does make you wonder.

This line would ring more true if Google hadn't invented transformers and a host of other ML breakthroughs to bolster search. Also not search related, but inventing self-driving cars and solving protein folding are innovations that benefit consumers greatly.

I'd bet Ben himself has compared Google to Bell Labs in the past. To argue they don't innovate is insane.

Id wager SEO/AI slop is a much bigger problem for Google than ad load for their userbase. Only 20% of queries have ads so 4 out of 5 times ads won't even be seen. But on those 4 out of 5 do have the chance of turning users away from Google with awful SEO optimized sites and AI slop.