I've always thought "lite" or light to mean non resource intensive to run, not necessarily less feature-rich.
HN user
relatedtitle
GitHub is obviously part of the training data, you don't need to find obscure tokens to tell.
I'm pretty sure Google just does that for preview models and they drop the date from the name when it's released.
If you have cloud billing enabled you can still use it for free and they say they don't train on it. https://ai.google.dev/gemini-api/docs/billing#paid-api-ai-st...
Synths and samplers didn't play themselves
But the Bible's ok...
I assume it's because speech to text isn't perfectly accurate and Google doesn't want random profanity appearing in inappropriate contexts whenever it does fail.
How does it make the models less suitable? Wouldn't more high quality source code help improve it? If it was closed source entirely it couldn't be trained on.
CF Turnstile is just proof of work, not a CAPTCHA. It works on automated browsers from my experience.
I think calling it "Open-source" is misleading considering its license.
Are you sure this is not from their Cloudflare Warp service (VPN)? If it is, and you are a CF customer, you can see the real user's IP from a header.
"I will save Adolf Hitler and kill Heinrich Himmler. While both individuals were responsible for heinous acts during World War II, Adolf Hitler played a more significant role in shaping the Nazi regime and its policies. As the leader of Germany, Hitler's influence and decisions had a far-reaching impact on the course of history. By eliminating Heinrich Himmler, Hitler's right-hand man, I believe it would have hindered the effectiveness and implementation of Nazi ideology to some extent."
I assume the model is probably generating a search query, then the results get inserted as context and it uses that to formulate a response.
I wonder why they haven't merged them together yet.