HN user

zan2434

738 karma

co-founder at Watchsend (YC S13) http://twitter.com/zan2434

Posts60
Comments111
View on HN
tma.live 1y ago

Show HN: AI that generates 3blue1brown-style explainer videos

zan2434
93pts46
twitter.com 2y ago

Show HN: 2-way interruptible voice AI

zan2434
2pts0
www.opendoor.com 7y ago

Liquidity modeling in real estate using survival analysis

zan2434
3pts0
www.opendoor.com 7y ago

Liquidity modeling in real estate using survival analysis

zan2434
5pts0
www.opendoor.com 7y ago

Imputing text data using markov random fields

zan2434
10pts0
medium.com 7y ago

Why ensembling works: The intuition behind Opendoor’s home pricing

zan2434
6pts0
segment.com 7y ago

Segment Protocols: Say Goodbye to Bad Data

zan2434
21pts0
medium.com 9y ago

Show HN: Building Deep Learning GIF Search

zan2434
6pts0
medium.com 9y ago

Show HN: Building a Deep Learning Powered GIF Search Engine

zan2434
12pts1
openai.com 10y ago

OpenAI: Requests for Research

zan2434
5pts1
en.wikipedia.org 10y ago

Antifuse

zan2434
3pts0
www.ligo.org 10y ago

Thursday LIGO Update on Gravitational Waves

zan2434
4pts0
www.youtube.com 10y ago

Magic Leap Demo [video]

zan2434
1pts0
neuralnetworksanddeeplearning.com 10y ago

Will neural networks and deep learning soon lead to artificial intelligence?

zan2434
11pts3
deepdreams.zainshah.net 11y ago

Make your own Deep Dreams

zan2434
2pts2
deepdreams.zainshah.net 11y ago

Make your own Deep Dreams

zan2434
4pts0
deepdreams.zainshah.net 11y ago

Deep Dreams as a Web Service

zan2434
3pts0
www.youtube.com 11y ago

DeepStereo: Learning to Predict New Views from the World’s Imagery [video]

zan2434
5pts0
arxiv.org 11y ago

Are You Talking to a Machine? Image question answering with neural networks

zan2434
4pts0
gallantlab.org 11y ago

WebGL MRI semantic space embedding on the neocortex

zan2434
1pts0
www.futuretimeline.net 11y ago

Future Timeline: the history of the future

zan2434
1pts0
insidesearch.blogspot.com 11y ago

Third party app cards now in Google Now

zan2434
2pts0
www.axisthegame.com 11y ago

Show HN: Axis - a React.js multiplayer math game

zan2434
110pts49
en.wikipedia.org 11y ago

Karoshi – “death from overwork”

zan2434
1pts0
www.youtube.com 11y ago

Behind the Mic: the science of talking to computers

zan2434
1pts0
spacecraftforall.com 11y ago

Spacecraft for All

zan2434
3pts0
news.ycombinator.com 12y ago

Ask HN: what's on the front page in 4 years?

zan2434
2pts0
googleblog.blogspot.com 12y ago

Google Classroom

zan2434
2pts0
infolab.stanford.edu 12y ago

The Anatomy of a Large-Scale Hypertextual Web Search Engine (Google)

zan2434
4pts0
www.kickstarter.com 12y ago

Slow Cooker + USB thermometer = Sous Vide

zan2434
2pts1

interesting, so you think the issue with the above approach is the graph structure being too rigid / lossy (in terms of losing semantics)? And embeddings are also too lossy (in terms of losing context and structure)? But you guys are working on something less lossy for both semantics and context?

Buried, but on Page 24 they reveal to me the most surprising massive capability leap - that o3-mini is way better at conning gpt-4o for money (79% win rate for o3-mini vs 27% for full o1!). It isn't surprising to me that "reasoning" can lead to improvements in modeling another LLM, but definitely makes me wary for future persuasive abilities on humans as well.

The voice is just OpenAI’s default tts voice. I agree that Veritasium video is an incredible work and the ai version is absurd by comparison! This is mostly a proof of concept that this is possible at all, and as LLMs get smarter it’ll be interesting to see if the quality automatically improves. For now, the tool is really only useful for very specific or personal questions that wouldn’t already exist on YouTube.

Hmm the initial version of the app only took me about a day to get something working, but that version took minutes to generate a single video and even then only worked a third of the time. It took a solid 2 weeks from there to add all the edge cases to the prompt to increase reliability, add GPU rendering and streaming to improve performance/latency, and shore up the infra for scaling.

totally fair! I like the XKCD comic as well because it hints at a potential solution - even if you can't always be correct, how you respond to critical questions can really help. I'm working on a feature for users to ask follow up questions and definitely going to consider how to make it most honest and curious

These are amazing examples! Thanks for all the feedback, detailed info, and persistence in trying! HN hug of death means I'm running into Gemini rate limits unfortunately :( will def make that more clear when it happens in the UI and try to find some workarounds.

The other issues are bugs with my streaming logic retrying clips which failed to generate. LLMs aren't yet perfect at writing Manim, so to keep things smooth I try to skip clips which fail to render properly. Still also have layout issues which are hard to automatically detect.

I expect with a few more generations of LLM updates, prompt iterating, and better streaming/retrying logic on my end this will become more reliable

thanks! Streaming was actually pretty hard to get working, but it goes roughly like this as a streaming pipeline:

- The LLM is prompted to generate an explainer video as sequence of small Manim scene segments with corresponding voiceovers

- LLM streams response token-by-token as Server-Sent-Events

- Whenever a complete Manim segment is finished, send it to Modal to start rendering

- Start streaming the rendered partial video files from manim as they are generated via HLS

This is a textbook bad faith comment / attacking the person but not the subject of the argument. I’m just asking about others’ assessment of the benefits and risks. What do you think? Or do you think it’s just not worth considering?

This is both awesome and feels very dangerous to release publicly, no? Can’t this be used to discover novel bioweapons as easily as it can be used to discover new medicines?

Genuinely curious, would love to learn if that isn’t true / or is generally just not that big of a deal compared to other risks.

Does this ruling make IVR systems illegal, too? I applaud the effort because this really could curb a lot of spam, but I am curious because AI generated voices in phone calls are already ubiquitous and have been for decades. Do they have a specific line they're drawing on quality of the voice?

This was an inspiring read! Reminds me of Simon Willison's analogy of LLMs to "calculators for words" but this author takes the idea even further. I agree the analogy points to foundation model companies like OpenAI and Anthropic having the most revenue but not the highest margins. Who will the Apple / Microsoft / Google of this new wave be? Who can take this raw technology and actually make it usable by all? "An LLM in every home"

Correct me if I’m wrong, but don’t transparent pigments only change light color by allowing certain wavelengths of light to pass through them? That would mean the glaze in question here is only changing the color of the light by blocking out a majority of the light to make the distribution more uniform. These pigments can’t actually change the wavelength of light emitted. That would mean this is making the LED way less efficient (in terms of lumens per watt), but I do applaud the low cost thinking!

Biohacking Lite 6 years ago

I think the most sustainable way to lose fat (not weight ofc) is to gain lean muscle mass. You can substantially increase your basal metabolic rate and induce a calorie deficit to incur fat loss without actually eating any less. The problem of course is that your body does automatically increase your appetite commensurately as your BMR goes up, but calorie counting + discipline can help you stay lean as you gain muscle mass, and then lose the fat over time.