HN user

captainkrtek

1,361 karma
Posts7
Comments344
View on HN

Amen to this.

I've bounced back and forth on my feelings for AI and have landed in the realm of: - there are certain things it is exceptional at that humans cannot replicate. - there are certain things I do not want to use it for.

And review falls squarely in that first category. Similarly, it is exceptional at working through "low hanging fruit" type problems such as spotting inefficiencies, analyzing a profile to find flaws in software, etc.

Claude Opus 4.7 3 months ago

I use Claude Opus 4.6 as an enterprise user, and have also noticed a lobotomization. In recent weeks it's been much more self-correcting even within singular responses ("This is the problem - no wait, we already proved it can't be this - but actually ...") I'm wary of 4.7 being a change in this pattern, it's frustrating to have such a substantial change in experience every few months.

Was talking about this with some colleagues who are from Ukraine, Russia, and other countries.

In the US, it seems corruption is only allowed at the top. If you tried to bribe your way out of a traffic ticket as a regular person, you'd get in big trouble, then meanwhile the president pardons wealthy fraudsters [1].

Meanwhile, in countries like Russia, everyone can get in on the action. A colleague of mine told me if he were to get drafted to the war, he knew exactly how much to pay and who to pay off locally to get his name off the list. It's equal opportunity corruption.

[1] https://techcrunch.com/2025/03/28/nikola-founder-trevor-milt...

On my team I've been adding additional linters and analyzers (some I've written with Claude) to run at CI or locally to prevent codified "bad patterns" from entering our systems. This has been nice as a backstop, as I can't enforce what everyone's Claude prompts and local workflows are, but we can agree what CI checks run before merging. Not a 100% solution, but it has been helpful so far.

I've been toying with this too.

I added a Claude skill (/gather-history) that consolidates the history of our session(s) specific to the change into a series of: decision log, involvement (how much did I write vs. AI, how many refinement iterations, reviews, etc.) that I can then include in the PR. So far this has been helpful for my colleagues to understand how I arrived at the change and how thoroughly it's been developed.

There seems to be so much value in planning, but in my organization, there is no artifact of the plan aside from the code produced and whatever PR description of the change summary exists. It makes it incredibly difficult to assess the change in isolation of its' plan/process.

The idea that Claude/Cursor are the new high level programming language for us to work in introduces the problem that we're not actually committing code in this "natural language", we're committing the "compiled" output of our prompting. Which leaves us reviewing the "compiled code" without seeing the inputs (eg: the plan, prompt history, rules, etc.)

One challenge with code review as an antidote to poor quality gen-AI code, is that we largely see only the code itself, not the process or inputs.

In the pre-gen-AI days, if an engineer put up a PR, it implied (somewhat) they wrote their code, reviewed it implicitly as they wrote it, and made choices (ie: why is this the best approach).

If Claude is just the new high level programming language, in terms of prompting in natural language, the challenge is that we're not reviewing the natural language, we're reviewing the machine code without knowing what the inputs were. I'm not sure of a solution to this, but something along the lines of knowing the history of the prompting that ultimately led to the PR, the time/tokens involved, etc. may inform the "quality" or "effort" spent in producing the PR. A one-shotted feature vs. a multi-iteration feature may produce the same lines of code and general shape, but one is likely to be higher "quality" in terms of minimal defects.

Along the same lines, when I review some gen-AI produced PR, it feels like I'm reading assembly and having to reverse how we got here. It may be code that runs and is perfectly fine, but I can't tell what the higher level inputs were that produced it, and if they were sufficient.

Will there be an interest in vision based wearables?

Google Glasses - dead

Apple Vision Pro - dead

FB/Meta x RayBan - dead soon(?)

It seems they can’t get over the social hurdle of having a camera strapped to your face, and the effects of that on people around you. I think the tech is neat, but not socially accepted as a concept to make it viable. My sister is big into tiktok and filming all the time, and it personally makes me hesitant to be nearby as I’m not comfortable being filmed all the time.

50+ weeks? so a year?

I've been in big tech for 12+ years now. The first handful of years are definitely a grind to earn your spot, get a couple promos. After that though, it can become quite a bit easier to coast if that's what you're looking for. People know you, know you're probably valuable cause you're "senior" or "staff" and still here, and likely leave you alone. But yeah, as a newer engineer these days, it still requires the initial commitment to earn the privilege of coasting in a big tech company.

My biggest problem with usage of an LLM in coding is that it removes engineers from understanding the true implementation of a system.

Over the years, I learned that a lot of one's value as an engineer can come from knowing how things actually work. I've been in many meetings with very senior engineers postulating how something works arguing back and forth, when quietly one engineer taps away on their laptop, then spins it around to say "no, this is the code here, this is how it actually works".

I'm no economist, but if (when?) the AI bubble bursts and demand collapses at the price point memory and other related components are at, wouldn't price recover?

not trying to argue, just curious.

I've lived in Seattle my whole life, and have worked in tech for 12+ years now as a SWE.

I think the SEA and SF tech scenes are hard to differentiate perfectly in a HN comment. However, I think any "Seattle hates AI" has to do more with the incessant pushing of AI into all the tech spaces.

It's being claimed as the next major evolution of computing, while also being cited as reasons for layoffs. Sounds like a positive for some (rich people) and a negative for many other people.

It's being forced into new features of existing products, while adoption of said features is low. This feels like cult-like behavior where you must be in favor of AI in your products, or else you're considered a luddite.

I think the confusing thing to me is that things which are successful don't typically need to be touted so aggressively. I'm on the younger side and generally positive to developments in tech, but the spending and the CEO group-think around "AI all the things" doesn't sit well as being aligned with a naturally successful development. Also, maybe I'm just burned out on ads in podcasts for "is your workforce using Agentic AI to optimize ..."

I think the obvious things are:

- Deviation in consistency/texture/color/etc.

- Obvious signs related to the above (eg: diarrhea, dehydration, blood in stool).

Ultimately though, you can get the same results by just looking down yourself and being curious if things look off...

tldr: this feels like literal internet-of-shit IoT stuff.

This has not been my experience at least at the more remote-friendly places I worked. However, I can see this at companies with different culture / pace / attitude.

My most recent role the entire company of ~200 was remote, and so there was rarely the expectation of immediacy in a response. If something was truly urgent you'd be paged.

This is excellent and aligns with my own experience.

During my day I try to minimize interruptions by batching them. I will largely ignore Slack, and as notifications come in I glance and determine quickly if it really is urgent or if it can wait. If it can wait, I will punt all of those messages to a "remind me later" of a few hours, and get back to my task. I think this keeps my "recovery time" small as I'm not looking too close at these messages. It's not perfect, but definitely helps over pausing my "real work" to fully dive into each notification or ask.

It's so sad how much money leaders will effortlessly pump into something like this, when we still have existential threats of climate change, incurable diseases, poverty, housing, and so on.

Meanwhile ungodly amounts of money are being used so some boomer can generate a AI video of a baby riding a puppy.

A company I worked at also did this, though there was no limits. Some folks would choose to spend the whole week working on a larger refactor, for example, I unified all of our redis usage to use a single modern library compared to the mess of 3 libraries of various ages across our codebase. This was relatively easy, but tedious, and required some new tests/etc.

Overall, I think this kind of thing is very positive for the health of building software, and morale to show that it is a priority to actually address these things.

I've struggled throughout my life with anxiety, ADHD, and bouts of depression.

I've done years of therapy, and use some medication to help with my ADHD. I will say though, singlehandedly, the best (and hardest) thing I've had to do is fight my own phone/internet/computer usage.

I grew up with computers and still work professionally with them day to day, but have made a serious effort in the last year to cut down my usage in an extreme manner:

- Using a 'brick' device to control which apps work on my phone, requiring I physically tap my phone to the brick to lock/unlock the restricted mode. This is always on.

- Blocking tons of sites via multiple means (iOS screen time, eero network profiles)

- Turning my iPhone into a very very basic phone: in "bricked" mode, which I will find myself using for continuous days at a time, I can only use: gmail, photos, notes, weather, maps, spotify, telephone, imessage. No news, no internet browsing, no social media.

- I've deleted all social media accounts (except LinkedIn, but this too is blocked on my phone).

From all of this, the initial realization was, "wow, I'm bored..", which is hard at first to sit with as a feeling, when normally my first instinct to that feeling was "let's open some app/youtube/etc." Then you slowly find positive things creeping in to occupy that boredom time: reading, calling friends/family, getting chores done, etc. And for anything "restricted" that I can't do on my phone, I largely can do on my computer (though I still block sites like reddit, youtube). But this is much healthier as I'm much less likely to pick up my laptop for hours on end, vs. opening my phone at every moment I'm bored.

It would be a good thing, if it would cause anything to change. It obviously won't.

I agree wholeheartedly. The only change is internal to these organizations (eg: CloudFlare, AWS) Improvements will be made to the relevant systems, and some teams internally will also audit for similar behavior, add tests, and fix some bugs.

However, nothing external will change. The cycle of pretending like you are going to implement multi-region fades after a week. And each company goes on continuing to leverage all these services to the Nth degree, waiting for the next outage.

Not advocating that organizations should/could do much, it's all pros/cons. But the collective blast radius is still impressive.

I'm Canadian and American, and have lived in both places and seen the stark differences myself. In the US, the police culture is certainly militarized and proud of it. Even in small towns you have days where the police roll out the biggest armored vehicles they have to show off, and that's their idea of a "community event", kids think its cool obviously, but it's really just "lets show off all of our high power toys".

There is a new doc on Netflix called "Country Doctor" which follows a doctor who works in a very rural area, it shows what happens when the hospital goes through another round of being sold off to a new firm and the difficult access of the county.

The fact that people are unimpressed that we can have a fluent conversation with a super smart AI that can generate any image/video is mindblowing to me.

It's gimmicky and it seems the only people "impressed" or "mindblown" with AI content are boomers on Facebook.

Any time I see some obviously AI generated content, whether it be some LinkedIn-influencer type or some blog spam that got well ranked on SEO, I'm immediately disgusted. The internet is eating itself.