HN user

alvis

1,206 karma
Posts115
Comments185
View on HN
openai.com 7d ago

GPT‑Red: Unlocking Self-Improvement for Robustness

alvis
35pts1
support.claude.com 10d ago

Claude Code May–July 2026 weekly limits promotion

alvis
45pts66
twitter.com 22d ago

Claude Sonnet 5 is here

alvis
3pts0
www.notion.com 28d ago

Claude Agents in Notion

alvis
2pts0
www.bbc.co.uk 2mo ago

Rice cooker leads to £260k payout to sacked university cleaner

alvis
2pts0
community.openai.com 2mo ago

Codex Plugin for Claude Code

alvis
1pts0
help.openai.com 6mo ago

ChatGPT Atlas Default Browser Promotion

alvis
2pts0
assets.anthropic.com 8mo ago

System Card: Claude Opus 4.5 [pdf]

alvis
3pts0
www.independent.co.uk 8mo ago

Most Americans say 'Arabic numerals' should not be taught in school (2019)

alvis
26pts28
www.theverge.com 9mo ago

OpenAI is about to launch its new AI web browser, ChatGPT Atlas

alvis
4pts2
www.bbc.co.uk 9mo ago

Physics Nobel awarded to three scientists for work on quantum computing

alvis
4pts0
developers.openai.com 9mo ago

Apps SDK

alvis
468pts382
supabase.com 9mo ago

Supabase Series E

alvis
6pts0
www.nytimes.com 9mo ago

First Woman Appointed Archbishop of Canterbury

alvis
3pts0
www.perplexity.ai 9mo ago

The Internet Is Better on Comet

alvis
4pts11
twitter.com 9mo ago

Comet is now available to everyone

alvis
2pts0
www.youtube.com 9mo ago

Claude plays Catan: Managing agent context with Sonnet 4.5

alvis
1pts0
twitter.com 10mo ago

New in Notion: Agent instructions and memory

alvis
3pts1
platform.openai.com 1y ago

Use webhooks to receive real-time updates from the OpenAI API

alvis
5pts2
docs.anthropic.com 1y ago

Claude Code GitHub Action

alvis
2pts0
www.gov.uk 1y ago

AI Opportunities Action Plan: government response

alvis
2pts2
www.forbes.com 1y ago

Daisy, the 'AI Granny' Outwitting Scammers

alvis
20pts5
openai.com 2y ago

ChatGPT Team

alvis
13pts0
twitter.com 2y ago

Sam Altman is the CEO of a new AI group under Microsoft

alvis
55pts1
www.theverge.com 2y ago

Microsoft Hires Former OpenAI CEO Sam Altman

alvis
87pts3
www.theverge.com 2y ago

Turmoil at OpenAI: after firing Sam Altman, what's next for the home of ChatGPT?

alvis
1pts0
www.bbc.co.uk 2y ago

Sam Altman: The extraordinary firing of an AI superstar

alvis
4pts1
fortune.com 3y ago

Elon Musk needs payments to make Twitter an ‘everything app’

alvis
2pts0
qz.com 3y ago

The UK's top universities reached an agreement on how to deal with generative AI

alvis
3pts0
twitter.com 3y ago

OpenAI's ChatGPT Blocked in Italy

alvis
4pts2
Claude Sonnet 5 22 days ago

Ironically, the key message of today's release is that Sonnet 5 is far less capable than Opus 4.8 and Mythos 5. It's a funny development is the past few weeks

Claude Sonnet 5 22 days ago

What I starting to hate is that each model's effort level can mean completely different power.

Today sonnet 5's med level effort is equivalent to sonnet 4.6 low level effort :/

Claude Fable 5 1 month ago

Another thing to note: 30-day retention for all traffic on Mythos-class models

Is it good or bad? 30 days is a long time for anything bad to happen

Claude Fable 5 1 month ago

It’s too obvious that antropic need to find way to earn enough revenue before IPO. Claude subscription isn’t earning earning much money I bet

The problem is that it's testing claims (or some people would prefer calling them "truths") without much context.

Take just one random example: `Hostels in Kota, Rajasthan commonly use caged ceiling fans as a preventive measure against student suicides`

While `Hostels in Kota, Rajasthan commonly use caged ceiling fans` may be a verifiable facts (though I doubt if there are any statistics for verification but let's say there are), `a preventive measure against student suicides` is a claim that no one can prove that. It can just a believe at most.

Arh. Did Biden stole Thump 2nd term? Truth or fact or claim?

Claude Opus 4.7 3 months ago

I don't have much quality drop from 4.6. But I also notice that I use codex more often these days than claude code

Claude Opus 4.7 3 months ago

TL;DR; iPhone is getting better every year

The surprise: agentic search is significantly weaker somehow hmm...

Claude Opus 4.7 3 months ago

TL;DR; iPhone is getting better every year

The surprise: agentic search is significantly weaker somehow hmm...

Claude Opus 4.5 8 months ago

“For Max and Team Premium users, we’ve increased overall usage limits, meaning you’ll have roughly the same number of Opus tokens as you previously had with Sonnet.” — seems like anthropic has finally listened!

Claude Opus 4.5 8 months ago

What surprise me is that Opus 4.5 lost all reasoning scores to Gemini and GPT. I thought it’s the area the model will shine the most

ChatGPT Atlas 9 months ago

I’m not sure even if these browser automation brings much value other than financial analysts, at least for the moment

ChatGPT Atlas 9 months ago

Not another comet or Claude browser extension plz. I stopped using them only a few days after launching

Claude Skills 9 months ago

Well. I bet Notion simply forget some of APIs are private before. I started developing using Notion APIs on the first day it got released. They have constant updates and I have seen lots of improvement. There is just no reason why they intentionally want to make the duplicate page API on MCP but not api.

PS. Just want to say, Notion MCP is still very buggy. It can't handle code block, nor large page very well

Apps SDK 10 months ago

The question is, whether having UI in chatgpt a game changer, fundamentally?

Apps SDK 10 months ago

Given there is already a MCP-UI project, I’m not surprised it can be done. But even that I’m not very convinced that it’s the right approach. After all, it’s still far too slow for real usage…

Apps SDK 10 months ago

So it’s take 2 for Open AI’s App Store moment. But this time surfing Anthropic’s MCP wave. Smart interop.. or just chasing the cool kids?

GPT-5-Codex 10 months ago

It's interest to see this quote: `for the bottom 10% of user turns sorted by model-generated tokens (including hidden reasoning and final output), GPT‑5-Codex uses 93.7% fewer tokens than GPT‑5`

It sounds like it can make simple tasks much more correct. It's impressive to me. Today coding agent tends to pretend they're working hard by generating lots of unnecessary code. Hope it's true