Amazing. Worked really well. Thank you.
HN user
kumarm
Seems true and really wish the project included some sample PDF output.
My Text to Speech app uses bounding box to display what text in PDF is being read and would not work well PDF's from this project.
[even in programming it's inconclusive as to how much better/worse it makes programmers.]
Try building something new in claude code (or codex etc) using a programming language you have not used before. Your opinion might change drastically.
Current AI tools may not beat the best programmer, they definitely improves average programmer efficiency.
The UI presentation of Schema is really nice. Is there a javascript based UI for web that is used?
It is same on PlayStore and AppStore.
You would be surprised to know Apple started this in AppStore before Google on PlayStore. I assume it is because Google wanted to be safe from Antitrust lawsuits (Follow Apple rather than going there first).
Latency. Started using API. Gets response directly on Gemini API in under 6 seconds. Fal.ai takes about 30 seconds.
Since API currently is not working (seems rate limits not set for Image Generation yet) I tried on fal.
Definitely inferior to results I see on AI Studio and image generation time is 6s on AI Studio vs 30 seconds on Fal.AI
Seems to be failing at API Calls right now with "You exceeded your current quota, please check your plan and billing details. For more information on this error,"
Hope they get API issues resolved soon.
I am in Bay Area. I donated my cars to Kars4Kids before and even recommended to others since their process to donate is simple.
Most likely. But that's exactly what someone who hasn't experienced enough cycles in industry would come up with :).
He is second only to Elon in this case (SolarCity, X/XAI).
All my Veo 3 videos has sound missing. No idea why. Seems like a common problem.
So everyone who want Youtube Premium can explain to their boss why they need Gemini AI Ultra for work?
First experience is not great. Here are the issues to start using codex:
1. Default model used doesn't work and you get error: system OpenAI rejected the request (request ID: req_06727eaf1c5d1e3f900760d10ca565a7). Please verify your settings and try again.
2. You have to switch to model o4-mini-2025-04-16 or some other model using /model. Now if you exit codex, you are back to default model and again have to switch everytime.
3. Crashed the first time with NodeJS error.
But after initial hickups seems to work and still checking how good/bad it is compared to claude code (which I love except for context size limits)
Pretty disappointed with content moderation on Veo2. Here are the steps I did:
1. Took a picture of me and asked to describe person in the image.
2. Used Imagegen to create the cartoon version using description.
3. Tried to use veo-2.0-generate-001 to generate video of person in image (holding a coffee cup in original image) drinking coffee and having a conversation.
Video generation is blocked by content moderation.
Thank you. This is the best example of comparison I have seen so far.
Where do I access grok3 reasoning model that xAI mentioned in the graph?
We do this in our Text to speech app (Read4Me): https://apps.apple.com/us/app/read4me-talk-browser-pdf-doc/i...
You can scan a book and listen (also copy and paste the text extracted to other apps).
If you are looking to do this on large scale in your own UI, I would recommend either of Google solutions:
1. Google Cloud Vision API (https://cloud.google.com/vision?hl=en)
2. Using Gemini API OCR capabilities.(Start here: https://aistudio.google.com/prompts/new_chat)
I ran some quick programming tasks I have used O1 previously:
1. 1/4th time for reasoning for most tasks.
2. Far better results.
If I remember correctly apple wanted Turn by Turn navigation while also not adding any ads to the app (essentially be their maps but no revenue).
Not all users are the same for Ad Revenue.
Started on Android Market in 2010. First hire (designer) after 12 Million downloads and started hiring other Dev's after crossing 50 Million downloads. Still run decently popular apps on App Store and Play Store.
If you want to understand how human potential was wasted in old world, Ramanujan belongs to a caste in India that is only caste that is supposed to be educated (Representing probably < 5% of population) in those days.
Ramanujan short life itself is a loss to the world, Imagine how many Ramanujan's were ignored where there is no G.H. Hardy and what about Ramanujans in the other 95%?
It is clear that either the Amazon India social admins or a third party they hired to run the contests are giving away prizes to their friends or insiders.
ChiragG14 has won at least 5 of their contests which most people correctly answer.
More seems like young and immature (X posts) for virality than an evil intent to me. I guess people are hating on Youtubers launching products now?
Instead of using vacuum why not collect with with finger like extensions to use less power and better accuracy?
Also does anyone know a good programmable outdoor robot dog made in US?
> Starbucks is optimizing for their profit.
Are we shocked businesses are optimizing for their profit? Isn't that their purpose?
Isn't Ilya out of OpenAI partly for leaving Open part of OpenAI?
Today in an AI related post on X, I got an ad for Nityananda (A known fugitive from India who claims his own country now) teaching about AI. The guy probably didn't graduate high school.
I hope X will add some quality control to ads they are showing.
Ad that came up: https://i.imgur.com/dSkqV6I.png
This is the same guy who went viral criticizing MKBHD last week right? Here is a good information on what is happening: https://news.ycombinator.com/item?id=40060554 from that discussion.
There is incentive to take a public view that is anti current (or trend or popular or right) thing to do that makes you go viral.