HN user

kthartic

200 karma
Posts1
Comments117
View on HN

Hmm I feel this might just be nostalgia at play. I've not played any Anno games, but comparing the two side-by-side and 1800 looks leaps & bounds better to me. 1602 looks like it was drawn in MS Paint.

Apple Invites 1 year ago

my key takeaway is that they should really read the room

I guarantee the pool of people who really care about those issues/are affected by them is tiny. For example, I use my MacBook daily and haven't noticed any Spotlight issues. I have an iPad too - what's wrong with it? This is coming from someone who works in tech, the average Joe isn't gonna care.

I for one am happy to see an app like this. Currently the only way to get my friends together is through a group chat, and it's always a mess.

I asked Gemini 2.0 Flash (with my voice) whether it natively understands audio or is converting my voice to text. It replied:

"That's an insightful question. My understanding of your speech involves a pipeline first. Your voice is converted to text and then I process the text to understand what you're saying. So I don't understand your voice directly but rather through a text representation of it."

Unsure if this is a hallucination, but is disappointing if true.

Edit: Looking at the video you linked, they say "native audio output", so I assume this means the input isn't native? :(

I would definitely agree with you there, but I also can't blame people for consuming such media. It's incredibly difficult to bury your head in the sand/avoid all media in today's ultra-connected world. Even if I were to ditch my phone, my friends would still be talking about it.

I hope though my perspective in my earlier comment might give you some insight into the general psyche of (at least some) younger folks.

As someone who you'd probably refer to as one of the "younger folks", I think part of it is a coping mechanism. The outside world looks incredibly bleak. Most of my friends and housemates appear burnt out or have a depressed outlook on the state of the world - politics, global warming, house/rent prices, crumbling healthcare, dating (which in today's world means swiping on an app), competitive work life, social pressures (exacerbated by instagram, tiktok, etc), gym, seemingly no time in the world to do anything.

Perhaps it's just because I live in London, but this a snapshot of my social circle right now. It's also no secret that the West is suffering from a mental health crisis.

With the weight of the world looming over me, a nice meal for lunch is about the only thing keeping me together (being slightly sarcastic here).

I do wonder, was it always this way? Genuine question - did you ever feel this way in your 20s/30s? Did your friends?

Google Down 2 years ago

If it works for you it must work for everyone. I guess the reports on the linked site and in this thread must be wrong

Games -> AI-powered characters that interact with you in realtime

Commercials/tutorials/corporate training videos -> Voiceover work

TV shows -> Dubbing in various languages

Fast food drive-throughs -> Taking customer orders

In the Ford case they hired an impersonator to sing one of her copyrighted songs, so it's clearly an impersonation.

In OpenAI's case the voice only sounds like her (although many disagree) but it isn't repeating some famous line of dialog from one of her movies etc, so you can't really definitively say it's impersonating SJ.

GPT-4o 2 years ago

GPT-4o voice isn't out yet, so you were likely chatting with the old/current tech (which is still really good).

From OpenAI: "We'll roll out a new version of Voice Mode with GPT-4o in alpha within ChatGPT Plus in the coming weeks"

GPT-4o 2 years ago

I think they were just interrupting on purpose to show off that as a feature (or they just wanted to keep the live presentation brief)

GPT-4o 2 years ago

The voice getting cut off was likely just a problem with their live presentation setup, not the ChatGPT app. It was flawless in the 2nd half of the presentation.

This is huge for indie game developers! They can voice every line of dialogue for every character themselves (or with just 1 professional voice actor).

Text-to-speech AI voice generators exist, but you don't have fine control over the emotion/expressiveness/intonation of the lines like you do with this approach.