HN user

mixermachine

287 karma
Posts0
Comments150
View on HN
No posts found.

The OpenCode CLI does not work as well for me as the PI CLI. I'm a subscriber of OpenCode Go (the sub, good value for me really) but I had not great experiences with OpenCode CLI. It multiple times with different models deadlocked itself into listing endlessly to non ending processes (Android Debugging Bridge, COM serial log, ...). There was also a problem where the OpenCode CLI would crash after sometime with a Bun error.

I switched to the PI CLI and have no problems with hanging processes anymore. OpenCode Go allows for API access so I'm keeping this sub.

It is quite interesting how this is handled world wide. For me PII is very sensitive and I advice people to be very cautious. Every business in the EU (were I live) also has to be very careful with such data by law. Fines are now at a level were they can hurt the business significantly.

During vacation in an Asian country on the other side all of this was basically a no brainer for smaller to medium businesses. I once rented a scooter there and the business owner had all her documents organised in WhatsApp chats. Including now my passport plus drivers licence... The people in general in that country were also very relaxed when it came to giving out their contact details to random businesses.

I don't want to throw shade on them, thus no country name. Incredible friendly and welcoming people there.

Right, I did swap that. Still, you have to pay that 4k then every year and give out the code. I also assume that prices will go up as no AI company (but NVIDIA -> selling shovels) is currently making any money.

For some projects the giving out the code part might be ok (i use Codex there too) but for the core app at the company I'm working at there is currently a strict no-AI policy. A local GPU solves this.

With parallelism of 16 you can still get around 25 to 30 tokens per user when all 16 channels are running. Not everyone will use the model at the same time but it certainly will be tight, especially for agentic coding. For pure chat applications this should be quite fine.

The 5h quota of Codex Pro on GPT 5.4 Medium lasts me for around an hour and a half, maybe 2 hours. And this is already the "savy" setup. Enable GPT 5.5 High fast and you will be beached in 30 minutes with active development.

For continues all day work you definitely need a higher tier sub level.

I'm actually looking into deploying a GPU at my company because we can not give out our code. Qwen 3.6 looks good

RAM + GPU are getting more expensive but mostly for applications that require a lot of it like AI. The hardware cost for regular applications has not vastly increased (especially when factoring in inflation). Spending 2x development time on a problem often is not worth it (or only with large deployments).

UI development is an even more special case here. The customer buys the machine which runs the code, not the company. So sadly "good enough" is the standard.

One example for me here is the "switch product option" button on Amazon listings (e.g. switch green to blue color, smaller to larger model). On my phone this sometimes takes >5 seconds to properly load. Horribly optimised.

In my experience in software architecture, drawing a diagram often saves you >60 minutes of discussion and potentially multiple meetings. This works even with a badly drawn but truthful one.

Use an Ai agent + Mermaid.js for a quick scribble if you are in a remote meeting. Use white boards or pen + paper in a local meeting.

Diagrams are so much clearer then words, especially if the concept or logic in question is not trivial.

The cutting edge, max size models will likely stay in the GPU space for a long time. But these models are not needed for most general requests. With a fine tuned 30B quantisized model you can serve a large portion of requests with around 32GB of RAM. Free users will likely only get these kinds of models.

At some point we will get these models in hardware and the cost per token will be minimal.

I'm using the Codex Business subscription (about 30€) already for multiple months. Even there they cut back on the quota. A few months back it was hard for me to reach the limit. Now it is easier.

Still, in comparison with Claude Code, the quota of Codex is a much better deal. However, they should not make it worse...

Fully agree. I went to school in Germany and many of our textbooks were free there. Sometimes you would get a textbook that is already >= 10 years and out of shape but who cares? Especially the basic knowledge does not change often. Buying all these textbooks new every year feels like a scam to me as they are then only used for one year by the pupil.

Btw when you damaged a book beyond repair, you needed to pay the full price. Only the exercise books needed to be bought freshly as they were "used up" fully after the year. Still, they were often seen as optional.

This seems quite strange to claim. Basically every city in the developed world already has power plants on the outside and a lot of wires to get the electricity in

It exists and does degrade panels but the time horizon is pretty wide. Real world data shows something like 0.5% to 0.7% degregation per year on average. At the start the degregation is higher and but it slows down with age.

So a 20 year old panel might be at around 80% in the worst case. Often they are in much better shape. This seems like a pretty good deal to me.

Especially for hot and sunny areas solar is insane. At mid day, max heat, you get the peak production and can run your AC at full throttle. That enables you to efficiently work at nice temperatures.

The article states the same solar production numbers as your comment. I agree that the headline is overly positive but the ramp up of solar can't really be denied. Change at this scale is sadly slow in this rather conservative sector.

The biggest thing is truly that solar has now reached a price tag where it just makes sense to replace other sources. You don't need to think about the environment any more to prefer it.

Crazy thing is, the cost of these systems has gone down even further since then. You can get a 800w plug-in solar set with panels and an inverter for around 200€. Shipping might be 70€ or you can pick it up locally at the dealer. Add another 50-100€ for attachment material and you are good.

So for 250 to 400€ you can get a system that will break even latest after four years, likely earlier.

Nothing will come close to Opus 4.6 here. You will be able to fit a destilled 20B to 30B model on your GPU. Gpt-oss-20B is quite good in my testing locally on a Macbook Pro M2 Pro 32GB.

The bigger downside, when you compare it to Opus or any other hosted model, is the limited context. You might be able to achieve around 30k. Hosted models often have 128k or more. Opus 4.6 has 200k as its standard and 1M in api beta mode.

Vibecoding #2 6 months ago

Regarding the $200 subscription. For Claude Code with Opus (and also Sonnet) you need that, yes.

I had ChatGPT Codex GPT5.2 high reasoning running on my side project for multiple hours the last nights. It created a server deployment for QA and PROD + client builds. It waited for the builds to complete, got the logs from Github Actions and fixed problems. Only after 4 days of this (around 2-4 hours) active coding I reached the weekly limit for the ChatGPT Plus Plan (23€). Far better value so far.

To be fully honest, it fucked up one flyway script. I have to fix this now my self :D. Will write a note in the Agent.md to never alter existing scripts. But the work otherwise was quite solid and now my server is properly deployed. If I would switch between High reasoning for Planing and Middle reasoning for coding, I would get even more usage.

Don't get me wrong, but somebody has to operate an exit node and somehow there needs to be a consensus on the protocol + routing.

If the network is only earth bound fixed wireless, the distance might be small enough that the state comes for the operator itself... This raises the cost of running this network from just money to life threat.

Getting many open source satellites up in orbit might not be feasible.

Got to say, I like the current Android versions. In the early days I flashed my Motorola Defy every second month with some cool new ROM. Always rooted and Xposed, always enabling something new.

Now I run a S23 Ultra and after two years it still does everything I need. OneUI 8.0 and Android 16. For work (app de) I also have a Pixel 7a, always with the newest Android Beta. Also works well.

Even the entry level phones work OK to pretty good now. My Samsung A16 5G (also for work) functions surprisingly well for 150€.

Fully agree. ChatGPT is often very confident and tells me that X and Y is absolutely wrong in the code. It then answers with something worse... It also does rarely say "sorry, I was wrong" when the previous output was just plain lies. You really need to verify every answer because it is so confident.

I fully switched to Gemini 3 Pro. Looking into an Opus 4.5 subscription too.

My GF on the other side prefers ChatGPT for writing tasks quite a lot (school teacher classes 1-4).