HN user

dchichkov

2,070 karma

engineer. researcher. founder.

... "What I can not create, I do not understand." ...

... "Know how to solve every problem that has been solved." ...

... "Four no's. Five clues." ...

... "Reality is that which, when you stop believing in it, doesn’t go away." ...

... "A human being is a part of the whole called by us universe, a part limited in time and space." ...

Posts25
Comments840
View on HN
en.wikipedia.org 1y ago

Decree 770

dchichkov
2pts2
en.wikipedia.org 1y ago

Decree 770

dchichkov
2pts0
wellknown.game 2y ago

Show HN: Play any Well Known text game with GPTs

dchichkov
1pts0
www.nytimes.com 7y ago

A Haven for Spare Parts Lives on in Silicon Valley (2009)

dchichkov
1pts1
github.com 7y ago

Show HN: The Zen of Autonomous Vehicles

dchichkov
3pts0
www.nature.com 8y ago

Stephen Hawking. So close to fixing one of his problems. And so far now. :'-(

dchichkov
1pts0
techcrunch.com 10y ago

Mobileye says Tesla auto braking tech wasn’t designed for scenario

dchichkov
2pts0
github.com 10y ago

Show HN: Curious Namespace Trick – Limited static polymorphism in C++

dchichkov
41pts18
arxiv.org 11y ago

Common Mistakes When Machine Learning to Stock Market Modelling (2012)

dchichkov
1pts0
www.kuka-timoboll.com 12y ago

KUKA Robot vs. Timo Boll: The Duel

dchichkov
5pts0
ramcloud.stanford.edu 12y ago

Redis vs. RAMCloud

dchichkov
2pts0
www.theregister.co.uk 12y ago

Woz: Google Glass will be so cool... just like bluetooth

dchichkov
1pts0
setiathome.berkeley.edu 12y ago

SETIhome Donation History Since 1 Nov 2011

dchichkov
1pts0
9to5google.com 12y ago

Google testing new homepage design, removes black bar.

dchichkov
2pts0
matplotlib.org 12y ago

What's new in Matplotlib 1.3

dchichkov
2pts0
speakerdeck.com 13y ago

Bare-Metal Multicore Performance in a General-Purpose Operating System

dchichkov
3pts0
www.youtube.com 13y ago

The H.P. Touch Computer (1983)

dchichkov
4pts0
web.mit.edu 13y ago

The robotic equivalent of a Swiss army knife

dchichkov
8pts1
discussions.apple.com 13y ago

Forced Apple ID Security Questions - Choice of 3

dchichkov
2pts2
www.disam.upm.es 13y ago

Halloween treat: BaTboT - a biologically inspired morphing-wing bat robot

dchichkov
3pts0
www.slashgear.com 13y ago

Apple reveals Lightning to microUSB adapter to pacify Europe

dchichkov
39pts65
www.nature.com 14y ago

After experiment seven

dchichkov
1pts0
boingboing.net 14y ago

Aerial robot performing a perching maneuver

dchichkov
2pts0
www.nature.com 14y ago

Self-powered cyborgs

dchichkov
1pts0
www.nature.com 14y ago

Waiting for Landauer, only two decades left

dchichkov
5pts0

> In the proposal, OpenAI also said the U.S. needs “a copyright strategy that promotes the freedom to learn” and on “preserving American AI models’ ability to learn from copyrighted material.”

Perhaps also symmetric "freedom to learn" from OpenAI models, with some provisions / naming convention? U.S. labs are limited in this way, while labs in China are not.

0 1 00 01 10 11 000 001 010 011 100 101 110 111 0000 0001 0010 0011 0100 0101 0110 0111 1000 1001 1010 1011 1100 1101 1110

And no, I don't think the knowledge of language is necessary. To give a concrete example, tokens from TinyStories dataset (the dataset size is ~1GB) are known to be sufficient to bootstrap basic language.

For long context sizes AGI is not useless without vast knowledge. You could always put a bootstrap sequence into the context (think Arecibo Message), followed by your prompt. A general enough reasoner with enough compute should be able to establish the context and reason about your prompt.

I agree, they are only starting the data flywheel there. And at the same time making users pay $200/month for it, while the competition is only charging $20/month.

And note, the system is now directly competing with "interns". Once the accuracy is competitive (is it already?) with an average "intern", there'd be fewer reasons to hire paid "interns" (more expensive than $200/month). Which is maybe a good thing? Fewer kids wasting their time/eyes looking at the computer screens?

[dead] 1 year ago

The approach of "cutting funding and then observing whether anything critical fails or is impacted" only works if outcomes follow a normal distribution.

This is far from the case — many areas are characterized by heavy-tailed loss distributions, where extreme negative consequences could really ruin the day and erase any efficiency gains.

I've suggested that long context should be included into the prompt.

In your particular case the prompt would look something like: <pubmed dump> what are the plants that aren't poisonous to most people?

A general reasoner would recover language and relevant world model from pubmed dump. And then would proceed to reason about it, to perform the task.

It doesn't look like a particularly efficient process.

If you look at the benchmarks of the DeepSeek-V3-Base, it is quite capable, even in 0-shot: https://huggingface.co/deepseek-ai/DeepSeek-V3-Base#base-mod... This is not from scratch. These benchmark numbers are an indication that the base model already had a large number of reasoning/LLM tokens in the pre-training set.

On the other hand, my take on it, the ability to do reasoning in a long context is a general capability. And my guess is that it can be bootstrapped from scratch, without having to do training on all of the internet or having to distill models trained on the internet.

MMMU is not particularly high. Janus-Pro-7B is 41.0, which is only 14 points better than random/frequent choice. I'm pretty sure, their base DeepSeek 7B LLM will get around 41.0 MMMU without access to images, this is a normal number for a roughly GPT4-level LLM base with no access to images.

Sorry, but this was ChatGPT/o1 with access to code execution (Python) and it used almost 4 minutes to do reasoning. It had done a few checks with smaller numbers, all of which had failed. And it proceeded to make a wrong conclusion (with high confidence).

I understand that it is mostly regulated at the state level. I'm not sure about other states, but The Computer Science Standards for California Public Schools (Kindergarten through Grade Twelve) also tend to be followed by private schools. So they can claim their programs meet state requirements.

This brings computers into the classroom, and once they’re available, it is a slippery slope. It is easier for teachers to have students use semi-gamified "educational" apps rather than engage themselves.

Example for K-2 - https://www.cde.ca.gov/be/st/ss/documents/csstandards.pdf:

  K-2.CS.1 Select and operate computing devices that perform a variety of tasks accurately and quickly based on user needs and preferences.

  K-2.CS.2 Explain the functions of common hardware and software components of computing systems.

  K-2.CS.3 Describe basic hardware and software problems using accurate terminology.

  K-2.NI.4 Model and describe how people connect to other people, places, information and ideas through a network.

  ...

  K–2 K-2.AP.12 Create programs with sequences of commands and simple loops, to express ideas or address a problem

  K-2.IC.20 Describe approaches and rationales for keeping login information private, and for logging off of devices appropriately

Another Gorilla is the schools, teachers and state-approved recommendations, that extend their reach even into private schools.

Imagine my frustration one day, when I've discovered that my kindergartner has full access to a brand-new, shiny iPad during class. Despite complaints from parents, the teacher refused to reduce iPad usage (or even activate Screen Distance and Screen Time controls on the iPad, or share usage statistics).

The only thing that I've learned, this is all in line with California’s state-approved computer literacy recommendations.

I wish that "Online Coupon Price Tags" in stores would also be banned. I'm talking about these yellow price tags that show lower than "Club" prices, which are only valid if you collect a coupon online.

Like FTC, I estimate that banning these would save U.S. consumers millions of hours they currently spend searching and clicking on pointless coupons on their phones before making purchases. It would also increase happiness, as it's extremely annoying to pay $20 extra, knowing that a lower price is available if only you spent ten minutes struggling with a store's website on your phone.

Whoever invented this is evil and is destroying happiness.

I wish that "Online Coupon Price Tags" in stores would also be banned. I'm talking about these yellow price tags that show lower than "Club" prices, which are only valid if you collect a coupon online.

Like FTC, I estimate that banning these would save U.S. consumers millions of hours they currently spend searching and clicking on pointless coupons on their phones before making purchases. It would also increase happiness, as it's extremely annoying to pay $20 extra, knowing that a lower price is available if only you spent ten minutes struggling with a store's website on your phone.

Whoever invented this is evil and is destroying happiness.

Yeah, I've also had difficulty finding something with enough I2S. It was a while back and I've used Sprocket carrier for Jetson TX2 - it had 6 lanes, so up to 96. It was for a SODAR application, so the sampling frequency was not that critical and to me it felt like the perfect trick to make an array with off-the-shelf hardware. So I was just curious, if this was something you've considered.

For something indoors, yes, I can see how low sampling frequency gets very limiting. And 192 microphones, that's really pushing it. Love it.

The $2/mic vs $0.5/mic argument is a fun one. You've obviously poured enormous amount of engineering in there, involving PCB design, FPGA and network programming, writing custom CUDA kernels, signal processing, PyTorch, the list goes on. And you've had 4090 plugged in your PC in 2023. Classic hobbit in a mithril vest ;)

I'm curious, why haven't you used TDM I2S microphones for your array and used PDM?

I understand that ICS-52000 is a relatively low cost ($2/100pcs) and there are even breakout boards available with 4 microphones, which can be chained to 8 or 16, like https://www.cdiweb.com/datasheets/notwired/ds-nw-aud-ics5200...

Then you can take Jetson (or any I2S capable hardware with DSP or GPU on it) and chain 16 microphones per I2S port. It would seem a lot easier to assemble and probgam, if comared to FPGA setup.

Decree 770 2 years ago

It is an interesting piece of history. It was mentioned in a middle of a rather long book: "Behave: The Biology of Humans at Our Best and Worst", and I was surprised that it is not on the surface and I didn't know about it before. Considering how relevant it is to the current happenings in the US.

Is there some tax data about the amount of Clean Vehicle/CA and Federal EV credits issued for Cybertruck?

The income cap on getting the clean vehicle rebates is $135k ($200k joint filers). And I'm not sure about the federal rebates. Tesla doesn't offer 0% financing, current Cybertruck APR deal is reported to be 5.29% for up to 72 months. So I don't see how someone with the income under the rebate cutoff can afford that $100k car or the financing option. The delta between the number of rebates (Federal EV vs Clean Vehicle/CA) may allow to estimate, how many of these are corporate (pre-income tax + rebate?) purchases.

And these "Cox Automotive estimates", are these reliable numbers that had been confirmed by Tesla earnings, or it is a "best guess by influencers" type of information?

I'm grateful to be getting a car from another manufacturer this year.

I'm curious, what is the alternative that you are considering? I've been delaying an upgrade to electric for some time. And now, a car manufacturer that is contributing to the making of another Jan 6th, 2021 is not an option, in my opinion.

As long as there's competition, it is fine. Boeing fits at least that role easily. Plus, they've built the vehicle with no drama and without purchasing Twitter in the middle. This is worth something.

We see similar situation in automotive. Other companies do allow to keep Tesla in check, so there's less opportunity to force "Cybertrucks" onto the market as the only option.

I agree, Open Weights are Open "Binary", not Open Source.

It's like taking an executable (.so module, firmware blob) and releasing it under permissive license, so anyone could disassemble, modify and hack it. And then disclosing what programming languages were used and pointing at a few libraries. And then saying that no, actual source code is not going to be released.