HN user

sanxiyn

15,137 karma

Seo Sanghyeon <sanxiyn@gmail.com>.

Posts79
Comments4,085
View on HN
blog.tilderesearch.com 2mo ago

Aurora: A Leverage-Aware Optimizer for Rectangular Matrices

sanxiyn
1pts0
www.spec.org 2mo ago

Spec CPU 2026

sanxiyn
2pts0
arcprize.org 3mo ago

Measuring Human Performance on ARC-AGI-3

sanxiyn
2pts0
www.biorxiv.org 10mo ago

Generative design of novel bacteriophages with genome language models

sanxiyn
3pts0
www.owlposting.com 1y ago

Plasmidsaurus: The unreasonable effectiveness of plasmid sequencing as a service

sanxiyn
4pts1
arxiv.org 1y ago

Can we improve the energy efficiency of EUV lithography?

sanxiyn
2pts0
arxiv.org 2y ago

Transcendence: Generative Models Can Outperform the Experts That Train Them

sanxiyn
4pts0
petapixel.com 2y ago

Instagram's Made with AI Tag Is Inaccurate, Misleading, and Needs to Go

sanxiyn
3pts0
www.science.org 2y ago

Why are all proteins left-handed? New theory could solve origin of life mystery

sanxiyn
4pts0
doi.org 2y ago

Global Room-Temperature Superconductivity in Graphite

sanxiyn
1pts0
www.nber.org 2y ago

Generative AI at Work

sanxiyn
1pts0
arxiv.org 2y ago

Predicting Transition Temperature of Superconductors with Graph Neural Networks

sanxiyn
1pts0
old.reddit.com 2y ago

Visit to the LK-99 basement lab (Chinese video with English subtitle)

sanxiyn
3pts0
ntia.gov 3y ago

AI Accountability Policy Request for Comment

sanxiyn
3pts0
predr.ag 3y ago

Falsehoods programmers believe about undefined behavior

sanxiyn
151pts226
github.com 3y ago

CodeGeeX: An Open Multilingual Code Generative Model

sanxiyn
2pts0
cen.acs.org 3y ago

The helium shortage that wasn't supposed to be

sanxiyn
2pts0
www.reuters.com 3y ago

SK Hynix says has developed its most advanced 238-layer storage chip

sanxiyn
5pts0
old.reddit.com 4y ago

GPT-3 2nd Anniversary

sanxiyn
2pts0
www.planetary.org 4y ago

Planetary Science Decadal Survey: After the Red Planet, an Ice Giant

sanxiyn
1pts0
github.com 5y ago

Mrustc Bootstraps Rustc 1.39.0

sanxiyn
3pts0
mediumworkersunion.org 5y ago

Medium Workers Union

sanxiyn
2pts0
www.cnet.com 5y ago

United Arab Emirates Hope Mars probe enters orbit and makes history

sanxiyn
2pts0
arxiv.org 5y ago

CPM: A Large-Scale Generative Chinese Pre-Trained Language Model

sanxiyn
1pts0
scite.ai 6y ago

Smart Citations

sanxiyn
3pts0
en.wikipedia.org 6y ago

Hreflang

sanxiyn
6pts1
gizmodo.com 6y ago

Facebook Must Delete Content Globally If It's Considered Defamatory in Europe

sanxiyn
1pts0
twitter.com 6y ago

In this house, we fire union organisers

sanxiyn
4pts0
www.nature.com 6y ago

A national experiment reveals where a growth mindset improves achievement

sanxiyn
2pts0
www.nature.com 6y ago

Eat less meat: UN climate change report calls for change to human diet

sanxiyn
457pts547

I know it is old, but the classic is a classic for reasons. It is criminal this survey does not include the venerable Secrets of the Glasgow Haskell Compiler inliner by Simon Peyton Jones et al.

A few words on DS4 2 months ago

There is a benchmark for performance work, and I think it is not being optimized by model vendors. The latest result from GSO is that both Opus 4.6 and 4.7 slightly outperforms GPT 5.5. This also matches my experience.

https://gso-bench.github.io/

The main point is that just as you can't ask for tiny nuclear explosion because nuclear physics just doesn't work that way, you also can't ask for factoring of 21 with Shor's algorithm. Quantum computing just doesn't work that way, sorry.

I will let Scott Aaronson speak. (See https://scottaaronson.blog/?p=9668)

Sometimes these days, I'll survey the spectacular recent progress in fault-tolerance, 2-qubit gate fidelities, programmable hundred-qubit systems, etc., only to be answered with a sneer: "What's the biggest number that Shor's algorithm has factored? Still 15 after all these years? Haha, apparently the emperor has no clothes!" I've commented that this is sort of like dismissing the Manhattan Project as hopelessly stalled in 1944, on the ground that so far it hasn't produced even a tiny nuclear explosion... If there's a reason why you think it can't work beyond a certain scale, say so. But don't fixate on one external benchmark and ignore everything happening under the hood, if the experts are telling you that under the hood is where all the action now is, and your preferred benchmark is only relevant later.

Last three CVEs are collections of bugs. CVE-2026-6784 is a collection of 55 bugs. CVE-2026-6785 is a collection of 154 bugs. CVE-2026-6786 is a collection of 107 bugs.

As for credits, I think bugs are ultimately credited to people, and this time Mozilla people used Mythos, as opposed to Anthropic people using Opus or Mythos.

The gap between formal and informal has been pointed out as an Achilles' heel of formal methods from the dawn of the field, so critique is not particularly new. The standard reference is Social processes and proofs of theorems and programs (1979), which is worth reading.

ARC-AGI-3 4 months ago

While I think all of your design choices are defensible, I do think you should release the full human baseline data. The second best action count is fine, but other choices are reasonable as well.

Claude Opus 4.6 6 months ago

You can disable this at Settings > Capabilities > Memory > Search and reference chats.