HN user

mirzap

2,221 karma

Curious developer.

[ my public key: https://keybase.io/mirzap; my proof: https://keybase.io/mirzap/sigs/5G0Sj20ied-ISqZWBbkxcQQ3pm3o5_yWUSkV9aHeRSk ]

Posts87
Comments325
View on HN
news.ycombinator.com 5d ago

Just got an AWS billing alert projecting my monthly cost at $140B

mirzap
90pts5
repoprompt.com 1mo ago

Repoprompt is going Open Source

mirzap
3pts0
www.testingcatalog.com 2mo ago

SpaceXAI prepares Grok Build desktop app to rival OpenAI Codex

mirzap
1pts0
huggingface.co 2mo ago

Zyphra releases the ZAYA1-8B MoE model optimized for intelligence density

mirzap
7pts2
www.npmcharts.com 2mo ago

Spike in Codex Downloads

mirzap
2pts0
www.lost-pixel.com 2mo ago

Lost Pixel is joining Figma and sunsetting the OSS product

mirzap
3pts0
hypirion.com 3mo ago

Understanding Clojure's Persistent Vectors, pt. 1 (2013)

mirzap
114pts27
twitter.com 3mo ago

AWS has officially removed all EC2 instances in Bahrain from their docs

mirzap
58pts1
twitter.com 3mo ago

Qwen 3.6 Plus Preview Available for Free for a Limited Time

mirzap
5pts0
twitter.com 3mo ago

Chroma Context-1 a 20B Parameter Open Source Search Agent

mirzap
4pts0
twitter.com 3mo ago

OpenAI's latest repo has Claude as the third top contributor

mirzap
59pts25
twitter.com 4mo ago

Cursor Composer 2 is just Kimi K2.5 with RL

mirzap
276pts168
microsoft.ai 4mo ago

MAI-Image-2: for limitless creativity

mirzap
1pts0
elixir-language-tour.swmansion.com 4mo ago

New Version of the Elixir Language Tour

mirzap
6pts0
www.theverge.com 4mo ago

Meta is shutting down its VR metaverse on June 15th

mirzap
5pts1
old.reddit.com 4mo ago

15.03. 1999 (27 years ago) ICQ chat where the name Counter-Strike was decided

mirzap
3pts1
expo.dev 4mo ago

Jetpack Compose now available for React Native apps

mirzap
1pts0
electrek.co 4mo ago

Musk admits xAI "not built right" weeks after Tesla invested $2B

mirzap
8pts2
www.amd.com 4mo ago

Way to Run OpenClaw Locally on AMD Ryzen

mirzap
5pts1
twitter.com 4mo ago

Vite+ Is Now MIT

mirzap
13pts2
www.reuters.com 4mo ago

Reducing Europe's nuclear energy sector was 'strategic mistake', EU chief says

mirzap
6pts0
www.geophysical-forensics.ch 4mo ago

Independent Geophysical Forensic Analysis of the Nordstream Pipeline Sabotage

mirzap
4pts0
upstash.com 6mo ago

Context7 Without Context Bloat

mirzap
2pts1
docs.google.com 6mo ago

The Honey Files Expose Major Fraud

mirzap
19pts2
twitter.com 6mo ago

Z.ai is set for its IPO on Jan 8, 2026

mirzap
2pts0
twitter.com 7mo ago

NanoGPT is the first LLM to train and run inference in space

mirzap
2pts0
twitter.com 7mo ago

Google insider profited $1M in a single day betting on the Google search markets

mirzap
2pts0
tanstack.com 8mo ago

How we accidentally made route matching more performant

mirzap
2pts0
vercel.com 1y ago

NuxtLabs Joins Vercel

mirzap
20pts1
github.com 1y ago

FrankenPHP is now under PHP org

mirzap
6pts0

Yeah, I do that too. Essentially, the system I described begins working on a task that is small enough and clearly defined. Each “slice” in a milestone usually have 5-10 subtasks (for instance, Slice E1 has P1...P6 subtasks). The orchestrator then receives the prompt to implement E1-P1.

I dedicate a significant amount of time to defining the precise actions that agents should perform (PRD/ADR). I break down the feature sets into Milestones and slices (tasks). These tasks are small, well-defined, and scoped. I have a prompt template that the “architect” agent prepares whenever I want to initiate a new feature. This ensures that the prompt structure remains consistent and standardized over time. The generated prompt is then pasted to the “orchestrator,” which performs context discovery (using Repoprompt) and finalizes the plan then proceeds to launch subagents to do the work.

Based on the size and complexity of the task, as well as any inter-task dependencies, the orchestrator deploys one or more subagents (sometimes 5 or 6 subagents) to work on these mini tasks. Once all tasks are completed, the orchestrator initiates verification and launches a review workflow. This workflow uses the original prompt, acceptance criteria, repository internal guidelines, and relevant skills to conduct a thorough review of the agents’ work.

Typically, there are one or two review iterations, during which the review agent identifies any issues. Sometimes, I may also notice issues and have to "steer" the orchestrator. The time required for a slice to complete ranges from 30 minutes to 4 or 5 hours, depending on its size, complexity, and the number of subtasks it contains.

Only if I run about 3 such orchestration in parallel I can reach hourly limit.

DeepSeek v4 3 months ago

My thoughts exactly. I also believe that subscription services are profitable, and the talk about subsidies is just a way to extract higher profit margins from the API prices businesses pay.

Why would they have that feature in claude code cli if it goes against the ToS? You can use Claude Code programatically. This is not the issue. The issue is that Anthropic wants to lock you in within their dev ecosystem (like Apple does). Simple as that.

I’m not saying charging above marginal cost to fund R&D is weird. That’s how every R&D company works.

My point was simpler: they’re almost certainly not losing money on subscriptions because of inference. Inference is relatively cheap. And of course the big cost is training and ongoing R&D.

The real issue is the market they’re in. They’re competing with companies like Kimi and DeepSeek that also spend heavily on R&D but release strong models openly. That means anyone can run inference and customers can use it without paying for bundled research costs.

Training frontier models takes months, costs billions, and the model is outdated in six months. I just don’t see how a closed, subscription-only model reliably covers that in the long run, especially if you’re tightening ecosystem access at the same time.

They are not losing money on subscription plans. Inference is very cheap - just a few dollars per million tokens. What they’re trying to do is bundle R&D costs with inference so they can fund the training of the next generation of models.

Banning third-party tools has nothing to do with rate limits. They’re trying to position themselves as the Apple of AI companies -a walled garden. They may soon discover that screwing developers is not a good strategy.

They are not 10× better than Codex; on the contrary, in my opinion Codex produces much better code. Even Kimi K2.5 is a very capable model I find on par with Sonnet at least, very close to Opus. Forcing people to use ONLY a broken Claude Code UX with a subscription only ensures they loose advantage they had.

Why are you all obsessed with this question when it comes to Chinese models? Here are some of the questions you should be asking Western governments and models instead: Who protects the pedophiles at the top of Western governments and corporations? How many people have been convicted in relation to the Epstein files? Who protects powerful politicians and Western oligarchs from pedophilia charges? Who did Epstein work for, and why (hint: it’s not Russia or China)?

How so? Apple's subscription cancellation is one click away, and you don't get overcharged when canceling.

I think convenience inherently comes with bloat. This is true for many things in life. Most people use cars simply to drive to work and back, yet they don’t use 90% of the features sold to them by salespeople.