HN user

SoMomentary

45 karma
Posts0
Comments36
View on HN
No posts found.

I thought the Gemini 3.5 Flash Lite response was quite telling myself. I personally like the Pelican SVG test, to me it is still a charming snapshot of model performance anecdata. No one would argue it's rigorous but I don't think it was ever intended to be.

I get people burning out on the pelican SVG test alongside the rest of the AI burnout, but I guess for myself I'm just choosing to keep enjoying it while I still can.

Huh, looked this up as Acetaminophen is one of the only OTC painkillers I can take. I thought NAC (N-Acetyl Cysteine) sounded familiar and low and behold I already have some I've purchased as a supplement.

Anyone can buy this stuff, no need to make it sound like it's some controlled substance. I will say if I take to much Acetaminophen I'll just head to the hospital instead of winging it!

Superpowers 6 22 days ago

I've loved Superpowers right along. I think a lot of what it does has been ingested into Claude Code proper now so I'll be interested to see if this release actually changes things up.

You could have your Coding agent use this to do it for you! Really though it's about using the right tool for the job, and this seems like a better choice to use a tool like this so you don't have your agent burn tons of tokens imperfectly working through all your documents at a slower pace.

Personally I'll be giving this a whirl as part of my note distillation process. I end up with hundreds of pages of PDFs and docx files that I'm sure would be easy to convert with a dedicated tool.

The speed was impressive when I tested it but unfortunately the accuracy left a lot to be desired. Be interesting to do the math on some of my normal workflows to see where the break even is between them, assuming the tasks you have can tolerate a couple of failures.

Just today I was thinking about threading the needle on that and making my own Steam Machine with an AMD BC-250 board. Maybe I still will, it'd be a 10th the price and I do love to tinker.

I actually feel the opposite of what most people are saying here. I thought your writing style was great. It felt like you respected my time as a reader and got to the fucking point.

I thought the lack of fluff was refreshing!

Not at all! My company has 100s of clients and we track time in 6 minute increments. I feed in my browser history, terminal logs, session scripts, calendar, git commits, etc etc into it and voila it produces a highly accurate timesheet in no time flat.

Automating it has been way better for me than the alternative of breaking my flow whenever I'm switching tasks to chart my time, or logging all my hours for the week in one sitting. Different strokes for different folks I suppose.

Local privacy respecting inference can be worth it. I use a local model to log everything I do all week to automate my timesheet. I also have it do a bunch of other data tasks. I won't say that larger SOTA models wouldn't do these tasks better than a local model but PII is a concern and my employer wouldn't approve of me just setting tokens on fire everyday to do what I could do myself.

Having seen all the AI interactions that you can get through clickstream data I have no doubt that $GOVERNMENT_AGENCY can see much much more.

Awesome, thank you so much for that! I'll have to dig in and get this working for my work machine where I still have to use Chrome. I literally quit using Chrome for personal use over this but I guess it was premature.

I thought these stopped working altogether with the release of Chrome 142? I know you could override it for awhile there, but I've been lead to believe that option is gone.

This is part of what forced my hand and made Firefox my daily driver, at least for personal use.

Yes, it's been a way better deal to go for a subscription than pay as you go for me in the past. I had a month where I burnt through ~3.8b tokens which was somewhere in the ballpark of $8k worth of savings.

Now though I don't dare use spend tokens for basic note taking with Sonnet because I'm hitting the limit over a couple million tokens on the 20x plan, so they've really tightened the purse strings since November.

I think some of this comes down to undeclared A/B testing. I've had the worst week of interactions I have ever had using Claude Code. The whole week whenever I have a session that isn't failing miserably I seem to get tapped for a session survey but on any that are out and out shitting the bed it never asks. It has felt a little surreal. I'd love to see a product wide stats graph for swearing, I would 100% believe that it is hitting an all time high but maybe I'm just a victim of a bad A/B round.

Not at all! My shop is having one of its best years ever. We've been switching to value based billing with our clients which has made them happier and us more profitable.

Using AI powered prototypes to sell clients on new features has gone really well for us. Many of our clients have presold themselves on adding AI features of some kind, which is nice because it generally means they will need support for these features going forwards.

Their latest update just introduced indexing for handwriting + text, so now you're able to search through all of your notes! This has been a huge QoL upgrade for me.

I think their point is that all of the content out there is turning in to AI Slop so it won't matter if search changes because the results themselves have already been changed.

Onlycats 1 year ago

They should also reject duplicates and probably rate limit uploads...

[dead] 1 year ago

Did he have to leave for the same 130 day rule that made Musk leave?