HN user

gsuuon

397 karma

https://github.com/gsuuon

Posts5
Comments313
View on HN
Gemini CLI 1 year ago

I will now write the corrected code.You are right. I have been making a mess of this. I will get it right this time.

It keeps putting import statements in the middle of files or duplicating edits at the wrong places, but it's hard to get frustrated since it seems pretty self-aware. I didn't say it was making a mess - it's just being hard on itself.

A short session ended up sending over 3 million tokens though - wonder how the economics of this work out for google?

Pretty sure OP was talking about OpenAI expiring their credits (just had mine expire).

Btw, unsurprisingly, the time for expiry appears to be in UTC in case anyone else is in the situation of trying to spend down their credits before they disappear.

This makes sense, effective post-mortems don't focus on assigning blame. They try to identify and fix the problem. The issue with AI is that it's such a black-box right now, this process is likely not feasible. If an AI makes a decision that turns out to be wrong, there's no reliable way to identify and fix the problem. You can prompt engineer or finetune or re-train the model, but in the end you can only hope that issue has been fixed. I think the black-box nature of AI makes it different from other systems we use.

Is it just me or does it seem like the Catholic church might have a better grasp on technology than the US government?

  46. While responsibility for the ethical use of AI systems starts with those who develop, produce, manage, and oversee such systems, it is also shared by those who use them. As Pope Francis noted, the machine “makes a technical choice among several possibilities based either on well-defined criteria or on statistical inferences. Human beings, however, not only choose, but in their hearts are capable of deciding.”[92] Those who use AI to accomplish a task and follow its results create a context in which they are ultimately responsible for the power they have delegated. Therefore, insofar as AI can assist humans in making decisions, the algorithms that govern it should be trustworthy, secure, robust enough to handle inconsistencies, and transparent in their operation to mitigate biases and unintended side effects.[93] Regulatory frameworks should ensure that all legal entities remain accountable for the use of AI and all its consequences, with appropriate safeguards for transparency, privacy, and accountability.[94] Moreover, those using AI should be careful not to become overly dependent on it for their decision-making, a trend that increases contemporary society’s already high reliance on technology.
That is, "an AI told me so" should never be a valid excuse for anything.

I also really liked:

  62. In light of the above, it is clear why misrepresenting AI as a person should always be avoided; doing so for fraudulent purposes is a grave ethical violation that could erode social trust. Similarly, using AI to deceive in other contexts—such as in education or in human relationships, including the sphere of sexuality—is also to be considered immoral and requires careful oversight to prevent harm, maintain transparency, and ensure the dignity of all people.[124]
I think it should be a legal requirement that AI identifies itself as such given certain key-phrases and that there's no way to prompt engineer it out.

Really interesting read overall, thanks for sharing.

I'm really hopeful for e-ink or low-fidelity devices to help ween us off media addiction. Hopefully Nothing pursues something in that space since it aligns with their mission. Would love to switch most of my work screens to e-ink and only have 'normal' screens for explicit recreation time.

I was impressed this is open source, then impressed it was done by one person, then impressed it only took 6 months, and eventually somehow impressed again that it was a _high schooler_. My mind is blown. Kudos for managing to do something insane like this. Very inspirational.

I'm a little suspicious of the Isaac Newton example. The values of the better answer are very close, I wonder if the ordering holds up against small rewordings of the prompt?

Another approach if you're working with a local model is to ask for a summary of one word and then work with the resulting logits (wish I could find the article/paper that introduced this). You could compare similarity by just seeing how many shared words are in the top 500 of two queries, for example.

This is generally a pro for me. I can skip to the bottom of a file and see what the main purpose of it is, knowing that it's the only place that could reference everything else in the file.

This was cool (spacelancers didn't work in Arc but did in Chrome, forest demo 403's) - audio and fullscreen fail because they aren't initiated by a user gesture. Impressive how quick it is to get into a game with this level of fidelity, normally you'd need several minutes of downloading.

I really like this idea and have (briefly) tried something similar with a 'personal' discord server. I think this sort of thing would be great to rebuild local community over the internet instead of mostly faceless or parasocial interaction.

I default to never answering now. Every time I feel like it might be an actual person on the other end and risk answering, it has turned out to be a scammer or spam call. Not sure why governments don't do more against them - scam and spam callers have destroyed phone communication.

I'm really confused by this policy - isn't security and enforcement the value prop of the app store in the first place? Why are they offloading the work of accountability to user mob-rule by basically saying "here's where they live, go get them"? Who does this actually benefit? No sane disgruntled user is ever physically showing up at a developers home.

I made a neovim plugin awhile back with that explicit purpose: being a toolbox to assemble your own AI stuff with. I struggled a lot to make AI useful and was hoping if the tools were there to make it easy to play around with, folks would figure out how to utilize LLMs effectively and share their results. Til now though, I'm not sure if anyone's even using the plugin beyond the starter prompts (which were only meant as examples). Maybe the API sucks.. idk.

I really like her content, especially the earthship videos. The 'home engineering' that goes into these buildings is awesome. Though I'll probably never get to build one, would happily play a game based on this (like some sort of physics based survival game? Is this a thing?)

Sometimes complex topics are really like rubric's cubes - some changes here break things over there and then you need to make a bunch of turns to fix things elsewhere. Thinking through writing is necessary for these, because they look much simpler until they aren't and all the gory details start tripping over each other. The unfortunate part is that it feels very difficult to 're-enter' the topic as if reading for the first time, so the writing can easily become difficult to understand for a fresh-reader since it was edited by someone who's read it dozens of times in various incarnations and orders.

1) that sounds like a Montessori school? 2) I feel like Walter White is one of the more memorable character names (w/ the alliteration, no?)