HN user

heavyarms

134 karma
Posts4
Comments35
View on HN

The last time I checked (a few days ago) it only had an "Upload Image" option... and I have been playing with Gemini on and off for months and I have never been able to actually upload an image.

It's basically what I've come to expect from most Google products at this point: half-baked, buggy, confusing, not intuitive.

What mechanism would make it possible to enforce non-paywalled, non-authenticated access to public web pages? This is a classic "problem of the commons" type of issue.

The AI companies are signing deals with large media and publishing companies to get access to data without the threat of legal action. But nobody is going to voluntarily make deals with millions of personal blogs, vintage car forums, local book clubs, etc. and setup a micro payment system.

Any attempt to force some kind of micro payment or "prove you are not a robot" system will add a lot of friction for actual users and will be easily circumvented. If you are LinkedIn and you can devote a large portion of your R&D budget on this, you can maybe get it to work. But if you're running a blog on stamp collecting, you probably will not.

Whenever I see one of these posts, I click just to see if the proposed solution to testing the output of an LLM is to use the output of an LLM... and in almost all cases it is. It doesn't matter how many buzzwords and acronyms you use to describe what you're doing, at the end of the day it's turtles all the way down.

The issue is not the technology. When it comes to natural language (LLM responses that are sentences, prose, etc.) there is no actual standard by which you can even judge the output. There is no gold standard for natural language. Otherwise language would be boring. There is also no simple method for determining truth... philosophers have been discussing this for thousands of years and after all that effort we now know that... ¯\_(ツ)_/¯... and also, Earth is Flat and Birds Are Not Real.

Take, for example, the first sentence of my comment: "Whenever I see one of these posts, I click just to see if the proposed solution to testing the output of an LLM is to use the output of an LLM... and in almost all cases it is." This is absolutely true, in my own head, as my selective memory is choosing to remember that one time I clicked on a similar post on HN. But beyond the simple question of if it is true or not, even an army of human fact checkers and literature majors could probably not come up with a definitive and logical analysis regarding the quality and veracity of my prose. Is it even a grammatically correct sentence structure... with the run-on ellipsis and what not... ??? Is it meant to be funny? Or snarky? Who knows ¯\_(ツ)_/¯ WFT is that random pile of punctuation marks in the middle of that sentence... does the LLM even have a token for that?

There are lots of valid use cases for speech synthesis and text-to-speech technology, and there are like 1 or 2 valid/legal use cases for voice cloning that I can think of. Ignoring the moral and ethical questions, why would anybody devote time and resources building a company around a very niche solution... one in which your customer churn rate is partially dependent on users not ending up in prison.

edit: typo

This makes sense on a number of fronts.

1. If you have capital to invest, you could do worse than AI startups at the moment.

2. Nvidia's long-term threat is not just direct competitors (AMD, Intel), but the big cloud-players going to in-house chips. Supporting the next wave of your customers makes sense.

3. Using Nvidia is the path of least resistance right now. If you only invest in startups using your products (and you are an active investor), you give startups another reason to avoid taking a risk on the alternative.

edit: typo

Gemini AI 3 years ago

I assume if one of the names in the paper was O'Shaughnessy you would immediately think: "Irish immigrant!" Schmidt? German immigrant!

Ignoring the obvious issue that this whole anonymous story seems suspiciously perfect for selling a related product...

On the one hand... Companies spent the past couple of decades engaging in various SEO hacks to rank high on search results and OpenAI scraped the internet to train a language model. Theoretically, it seems possible that some of the SEO techniques at least partially colored the flavor of LLM-generated text, and an "AI detector" could pick that up. So if you do a great job writing SEO optimized text (wordy, structured, lots of repeated key words, etc.) you are more likely to be flagged.

But really.. "AI Detector" services are snake oil and will lead to the creation of "Anti AI Detector" services that offer protective spells against the original snake oil. See, we eliminate a bunch of jobs with AI but we create whole new disciplines of work that didn't exist before. "AI Generated Content Obfuscation Specialist - III - W2" coming to a job board near you soon.

I've been thinking along the same lines. The token window IMO should be a conceptual inverted pyramid, where there most recent tokens are retained verbatim but previous iterations are compressed/pooled more and more as the context grows. I'm sure there's some effort/research in this direction. It seems pretty obvious.

I think the claim is based on the public political statements made by leaders in Texas. The fact that there is a huge discrepancy in what they say publicly against the science of global warming and the utility of renewables versus what the investment numbers say is the really sad part. Basically it boils down to: I'm going to lie through my teeth to pander to the stupid people who vote for me, but I'm also going to create favorable conditions for my wealthy buddies to make a killing in renewables.

I used to buy into some of this JFK stuff when I was a X-Files watching teenager. What really burst the bubble for me was a documentary I watched where a team of snipers and forensic scientists re-created the exact shot with mannequins with bones and ballistic gel. They didn't even have to try that hard. Using the same rifle and ammo, the first shot they tried resulted in almost the same exact trajectory. I can't find a clip of that exact documentary (circa 2004-2006), but there are others who have done the same. You don't have to look hard to find very comprehensive and scientific explanations for the exact trajectory of that specific shot. But you do have to look very hard to find an actual explanation for why it is impossible that is beyond the level of "golly gee folks, I done shot lots of guns in my life and let me tell you, it ain't possible."

https://youtu.be/Q7ERXm9OwuE?t=250

I dabble in music production and know some of the people in the "Lofi" world, so I know for a fact that this is not true. It's just a formulaic sub-genre where people are trying to make similar instrumentals with the same vibe. It would be jarring to listen to a playlist while studying and each song had wildly different tempos, instruments, etc.

Also, the music doesn't sound "Lofi" because it's generated by algorithms. A lot of hard work and software goes into taking a clean, pitch-perfect digital signal and making it sound like something playing on a record player from the 70s.

First of all, I'd like to say that this looks like a great project and I wish you the best of luck. I've done a bit of work on building knowledge graphs from semi-structured data and I know that every aspect of it is challenging. Obviously there's the data pipelines, ETL, semantic matching/categorization, statistical models, etc. Just building a simple UI for presenting a large knowledge graph was more challenging than most front end work I've ever done.

Question: if the goal is to build a knowledge graph that can "explain how anything in the world is related to everything else" how do you measure progress toward that goal? And how do you measure the quality? Just having a bunch of topics and relationships is not a great metric in my opinion. Obviously this is still very early, but here's an example I found in about 30 seconds of clicking around:

"Evidence suggests that Heart Failure is related to Income and COVID-19." [https://www.system.com/view/topic/P0XELnR0PaK]

There are topics in System for "Obesity" and "Smoking", but those are not associated to Heart Failure.

Roland 50 Studio 4 years ago

If you inspect the network traffic you can see what the "hidden" additional devices behind the countdowns are.

SP404 2022-04-04T12:00:00.000Z

TR606 2022-06-06T12:00:00.000Z

TR707 2022-07-07T12:00:00.000Z

TR909 2022-09-09T12:00:00.000Z

Remember, Stephenson’s target audience consisted of “scientists, mathematicians, engineers, and entrepreneurs.” Given his choice to court private wealth, it’s no surprise that Project Hieroglyph was doomed from the start. After all, you can’t very well expect to succeed as a hero if you stop to ask the villains for their permission.

It's really, really hard to take somebody serious when their political frame of reference makes them see the world in such crisp black and white contrast they just assume, without any further explanation needed, that clearly everybody already agrees that entrepreneurs (or maybe private wealth? As in, non-government wealth?) are the real villains.

There's a good book by Kevin Poulsen called "The Kingpin: How one Hacker Took Over the Billion-Dollar Cybercrime Underground" that is a bit out of date at this point (2011), but it goes into great length on all of the dynamics of the early forums where all of carding/spam/botnet operators did business.

In a forum/marketplace like this, your reputation is worth a lot of money. And if you scam someone and get banned, sure, you can just join again under a new identity, but building your reputation up again means you will lose out on a lot of potential sales.

I've been thinking this since the Apple Watch came out, but other than practical limitations (battery life) there's always this problem: have you tried to hold your arm out in front of your face and stare at your watch for a while? It's not the most comfortable position. You might have to add some extra arm/shoulder days in your exercise routine.

I'm working on a project with a somewhat related concept (dynamically generated PWA built from configuration generated by a simple UI) and I can appreciate some of the complexities involved in building a a framework and composable component library.

From what is shown in the video, I think the biggest limitation is that this approach would only work if the data model is stored in a document store. Each new "itemId" just becomes a property in a big JSON file and the framework knows what to update and how to update it because the "schema" is just controlled by the shape of the component tree and the type of component. That probably works for some use cases, but it's not something that scales IMO.

In other news:

Having a $3 million mansion at the beach confers much greater happiness than an affordable house in the suburbs!

Having a new top-of-the-line Mercedes SUV confers much greater status than a used Honda minivan!

Having a beautiful face and 6-pack abs confers much greater attractiveness than looking like Shrek!

There are two mutually exclusive questions at play here... and I'm totally convinced the population at large has ability to discern the difference and not conflate one with the other.

1. What is the best/strongest/longest lasting type of immunity, natural or vaccine?

2. What is the safest type of immunity, natural or vaccine?

Is natural immunity better? There's some data to show it is. Should you risk your life to get? Probably not.

I highlighted almost the exact quote you have here and it's nice to see it at the top of the discussion.

I agree with your sentiment, but I also think it's worth thinking carefully about two of the main points that stuck out to me:

- Access to compute for large models

- Access to large datasets (in this case mostly taxpayer funded academic research)

Every company and/or research group has access to the data, but some have a huge advantage in terms of compute. If there's a question about commercializing research, the scales are tilted toward those with more compute.

In this specific case, I think the intention to make AlphaFold open source and available to the community is obviously the best solution. But my question is, what happens if a less altruistic for-profit entity uses its huge compute advantage to develop new techniques and insights, and then patents everything before it becomes available to the community?

I understand that is the basic mechanism for how medical/pharmaceutical research gets translated into life-saving treatments, but if we're approaching a generalized model that can pump out "patent-worthy" discoveries only bound by the amount of data and access to compute, there's an obvious opportunity for a winner-take-most scenario.

You are agreeing to a chain of comments about an article that reaches an exaggerated conclusion based mostly on opinion and not on facts and careful analysis, and your response is to claim that the entire news industry is guilty of it too based on a vague notion that there was a lot of "sources say" reporting that was (made up? inaccurate?) over a specific period.

So, in conclusion, human beings excel at pattern matching, even when the patterns aren't there, and sometimes use their pattern matching ability to validate the point of view which aligns most closely to their identity and psychological needs.

I'm so tired of hearing this argument. The amount of revenue generated by customers who use only Microsoft Teams without an Office365 subscription is exactly $0 [1]. MS gives you exactly 2 options to get teams:

1. Free

2. Included as part of Office365.

Teams is currently eatings Slacks lunch, but the lunch was paid for by a corporate IT guy who switched to Office365 so he could lay off some IT admins and save some money compared to managing an on-premise Exchange server.

[1](https://www.microsoft.com/en-us/microsoft-365/microsoft-team...)

I disagree. I think we need to put the paddle to the metal and take this baby for a spin. I do think the analogy needs a tune up. Web browsers are not the "automobiles of the information superhighway" because while they let you go forward or go back, they don't let you turn left or right. So clearly, web browsers must be the trains of the industrial revolution, riding on the coast-to-coast rail network which was built in Dot Com Boom and Bust fashion [1]. So I think Silicon Valley will chug along for many years to come. Of course, there's a chance this gravy train could be derailed. Owning all of the train stations has something to do with Monopoly. I'm running out of steam... need more analogies and puns to shovel into the furnace..

[1] https://en.wikipedia.org/wiki/Railway_Mania

Anybody who has coded Game of Life can draw a parallel to this. Simple rules can lead to arbitrarily complex systems. I'm all on board with this concept. Graph theory is amazing and useful in many ways we don't understand yet. I'm all on board with this concept as well.

But a graph has nodes and edges. Nodes, in this case, can be particles.. I guess? But what are the edges? When a "simple rule" is applied to a collection of particles, what is the force that connects them after the interaction? I read some of the material in detail and skimmed some of the rest, but there was a lot of setup and cool graph visualizations and not a lot speaking to this core question.

Disclaimer: I'm not a theoretical physicist but I have read "Quantum Physics for Babies" at least 50 times.