HN user

frozenport

3,002 karma

Been here for more than a decade.

Its been quiet a journey!

From wordpress to ai by way of hustling and a PhD!

Professional experience in hpc, publishing, wet lab biology, optical physics, startups, medicine, compilers and AI.

Posts48
Comments2,609
View on HN
www.eetimes.com 2y ago

Groq CEO: 'We No Longer Sell Hardware'

frozenport
204pts148
www.techdirt.com 2y ago

Does Elon Grok the Trademark Issues with 'Grok'? AI Chip Company Groq Does

frozenport
6pts1
godbolt.org 3y ago

Godbolt Runs CUDA

frozenport
2pts0
izzys.casa 8y ago

Millennials Are Killing the [C++] Modules TS

frozenport
4pts0
garfbert.com 9y ago

Garfbert by Jim Jadams

frozenport
1pts0
www.reddit.com 9y ago

Arab Israeli students call Syrian exiles 'traitors' in debate clash

frozenport
1pts0
www.npr.org 10y ago

Medical Errors Are No. 3 Cause of U.S Deaths, Researchers Say

frozenport
1pts0
www.reddit.com 10y ago

MSVC CRT Sneaks in Telemetry by Default?

frozenport
130pts33
github.com 10y ago

JSON Objects to add gender diversity into your website

frozenport
2pts2
www.youtube.com 11y ago

Hitler on C++17

frozenport
7pts1
www.doomworld.com 11y ago

The Doom Comic Revealed

frozenport
2pts1
www.fanfiction.net 11y ago

Minesweeper Fanfiction

frozenport
67pts32
armdevices.net 11y ago

$240 Macbook Air clones run Windows 8

frozenport
3pts1
springfieldpc.dyndns.org 11y ago

Explanation of Open Source Licenses

frozenport
2pts0
meta.stackoverflow.com 11y ago

[SO] This site is really strict like the Taliban

frozenport
4pts1
en.cppreference.com 12y ago

Std::get_money

frozenport
15pts4
www.dailydot.com 12y ago

DashCon 2014 descended into chaos

frozenport
2pts0
www.buzzfeed.com 12y ago

How Russia's Troll Army Hits America

frozenport
1pts0
chambana.craigslist.org 12y ago

iPhone looking for Android – 32 (Urbana)

frozenport
1pts0
www.nature.com 12y ago

Effects of Sexual Activity on Beard Growth in Man: letters to Nature (1970)

frozenport
22pts3
arxiv.org 12y ago

Division By Three (2006)

frozenport
42pts21
i.stack.imgur.com 12y ago

SO: How to add a number in Javscript?

frozenport
1pts0
www.westword.com 12y ago

4chan camgirl grows up

frozenport
2pts0
www.youtube.com 12y ago

Drone's eye view of Burning Man 2013

frozenport
1pts0
www.independent.co.uk 12y ago

Syria: Russia’s warning falls on deaf ears as Britain and US prepare to bomb

frozenport
10pts0
www.usatoday.com 12y ago

Halliburton shares up despite destroying evidence

frozenport
1pts0
www.zpub.com 13y ago

In Praise of Idleness [1932] By Bertrand Russell

frozenport
3pts0
twitter.com 13y ago

Elon Musk on TCP

frozenport
3pts0
www.nypost.com 13y ago

Edward Snowden and China

frozenport
5pts2
moodycamel.com 13y ago

A fast lock-free queue for C++

frozenport
14pts3

> if groq and cerebras combined

There isn't to be shared between the two techs, Groq's hardware is a like a railgun that installs all the weights into the optimal location before firing off an inference. Cerebras computer engineering more convention requiring the same data movement that GPUs struggle with optimizing.

Suspect Groq is complementary/superior to nvidia's GPUs, while it is unclear what Cerebras brings other then maybe some deals with TSMC.

    “exotic” signal representation without much practical utility despite the fact that they have been around the signal and image processing community for more than 30 years now. 
Maybe they aren't that good? Maxwell's equations got a lot better when they dumped them, same thing with the few uses in video game physics/camera tracing.

I think semi analysis commented that they have pipelines instead of batches[1].

So every clock cycle you're doing useful work rather than loading up people into batches. And thats why the arch will probably win for inference, for training you're basically competing with software eco system and silicon density. AKA NVIDIA can give TSMC more money to get more ALUs on the die.

I think other places have attempted dataflow (FPGA etc) but they all basically had buffers (due to non-determinism in networks stack and even ram). SambaNova seems indistinguishable from an FPGA with a few clock cycles difference. I think they blew their shot with a Series D ($600 million???) where they made more of the same old. Maybe Intel will buy them to augment Altera? Looks like chasing parity with existing strategies.

I buy the Groq hype because its something different, certainly the public demo helped. HN is about the future.

[1] https://www.semianalysis.com/p/groq-inference-tokenomics-spe...

Well I've been using the groq public api, and its approx. the rates claimed.

Economics and costs are hard to predict. For example, Groq is not using HBM chips. So probably the cards are a lot easier to source.

Its not clear what the capacity of these systems are in terms of total users, or even tokens per second. Then you factor in cost. Then you realize all vendors will match a competitors pricing. Then you realize Groq doesn't sell chips.

¯\_(ツ)_/¯

The only thing you have is the public API to benchmark against: https://artificialanalysis.ai/