HN user

treesciencebot

1,947 karma

python, hot silicon and anything in between.

Posts97
Comments140
View on HN
infinite-kanvas.com 1y ago

Show HN: Open-source image restylization canvas

treesciencebot
1pts0
madebyoll.in 1y ago

World Emulation via Neural Network

treesciencebot
250pts46
blog.fal.ai 1y ago

MiniMax (Hailuo) video-01-Live

treesciencebot
2pts0
cloud.google.com 1y ago

Powerful infrastructure innovations for your AI-first future

treesciencebot
1pts0
www.recraft.ai 1y ago

Recraft introduces a revolutionary AI model that thinks in design language

treesciencebot
1pts0
www.recraft.ai 1y ago

New SOTA text-to-image model by Recraft

treesciencebot
3pts1
www.osmo.ai 1y ago

Scent Teleportation Update: We Did It

treesciencebot
1pts0
venturebeat.com 1y ago

Moondream raises $4.5M to prove that smaller AI models can still pack a punch

treesciencebot
3pts0
www.genmo.ai 1y ago

Mochi 1: A new SOTA in open-source video generation models

treesciencebot
2pts0
factorio.com 1y ago

Friday Facts #432 – Aquilo

treesciencebot
1pts0
fal.ai 1y ago

Show HN: Fal.ai – generative media platform for developers

treesciencebot
3pts0
techcrunch.com 1y ago

Character.ai CEO Noam Shazeer Returns to Google

treesciencebot
116pts83
blackforestlabs.ai 1y ago

Black Forest Labs – FLUX.1 open weights SOTA text to image model

treesciencebot
28pts3
blog.fal.ai 2y ago

AuraFlow v0.1: a open source alternative to Stable Diffusion 3

treesciencebot
164pts35
blog.fal.ai 2y ago

AuraSR – An open reproduction of the GigaGAN Upscaler

treesciencebot
4pts0
twitter.com 2y ago

Stability announces release date for Stable Diffusion 3 2B (small variant)

treesciencebot
2pts1
chipsandcheese.com 2y ago

Comparing Crestmonts: No L3 Hurts

treesciencebot
20pts0
imgsys.org 2y ago

Generative Image Generation Arena – Imgsys.org

treesciencebot
3pts0
imgsys.org 2y ago

Imgsys.org: a generative image model arena

treesciencebot
3pts0
chipsandcheese.com 2y ago

Intel's Ambitious Meteor Lake iGPU

treesciencebot
59pts23
sifted.eu 2y ago

Stable Diffusion maker leaves Stability AI

treesciencebot
2pts0
github.com 2y ago

OpenSora Releases its first trained checkpoints (2-5 SEC, 512x512 T2V)

treesciencebot
24pts7
stability.ai 2y ago

TripoSR: Fast 3D Object Generation from Single Images

treesciencebot
5pts0
chipsandcheese.com 2y ago

Qualcomm's Adreno 530, a Small Mobile iGPU

treesciencebot
2pts0
fastsdxl.ai 2y ago

Show HN: Real-time image generation with SDXL Lightning

treesciencebot
444pts104
github.com 2y ago

Accelerators

treesciencebot
2pts0
www.jasper.ai 2y ago

Jasper Expands by Acquiring Image Platform Clipdrop from Stability AI

treesciencebot
1pts0
fastsdxl.ai 2y ago

Real-time text-to-image generation powered by SDXL Lightning

treesciencebot
5pts0
huggingface.co 2y ago

Stable Diffusion XL Lightning

treesciencebot
3pts0
chipsandcheese.com 2y ago

Examining AMD's RDNA 4 Changes in LLVM

treesciencebot
4pts0

fal | San Francisco, CA / Remote | Full-time

We are on a mission to build world’s first generative media platform for developers. We are running inference on tens of thousands of GPUs, and looking for people (in all functions) to help us scale it to hundreds of thousands.

Featured Roles:

- Distributed Systems Engineer, https://fal.ai/careers/4009192009

- Virtualization Engineer, https://fal.ai/careers/4146037009

- ML Performance & Systems Engineer, https://fal.ai/careers/4009191009

Remaining: https://fal.ai/careers

the main question is going to be software stack. NVIDIA is already shipping NVFP4 kernels and perf is looking good. It took a really long time after MI300X's that the FP8 kernels were OK (not even good, compared to almost perfect FP8 support in NVIDIA side of things).

I will doubt that they will be able to reach %60-70 of the FLOPs in majority of the workloads (unless they hand craft and tune a specific GEMM kernel for their benchmark shape). But would be happy to be proven wrong, and go buy a bunch of them

fal | Growth Engineer | San Francisco (on site 5 days/wk)

Help us scale generative‑media infra: hack demos in the AM, pitch partners over coffee.

You’ll build quick client libs & microsites, run data A/Bs, write content that drives sign‑ups, and hand‑hold new devs.

Need: Python, JS/React/Next.js, SQL; speed, ownership, love for gen‑AI. Get: strong salary + equity, platinum health, unlimited “build‑something” stipend and most importantly a seat at a rocketship.

Shoot a link to something you’ve built to careers@fal.ai

Apple M3 Ultra 1 year ago

GH200 is nowhere near $343,000 number. You can get a single server order around 45k (with inception discount). If you are buying bulk, it goes down to sub-30k ish. This comes with a H100's performance and insane amount of high bandwith memory.

For traditional LLMs this might be true (especially large MoEs at bs=1) but I highly disagree with "multi-modal models" phrase since most of the models that output in other modalities are generally compute bound. Which means less flops will make the experience so much worse (imagine waiting a couple minutes for an image and hours for videos).

Sora is here 2 years ago

Hunyuan at other providers like fal.ai is cheaper than SORA for the same resolution (720p 5 seconds gets you ~15 videos for $20 vs almost 50 videos at fal). It is slower than SORA (~3 minutes for a 720p video) but faster than replicate's hunyuan (by 6-7x for the same settings).

https://fal.ai/models/fal-ai/hunyuan-video

all high end "gaming" rigs are either using ~16 real cores or 8:24 performance/efficiency cores these days. threadripper/other HEDT options are not particularly good at gaming due to (relatively) lower clock speed / inter-CCD latencies.

Zed AI 2 years ago

running a much worse model at a higher latency (since local GPU power is limited) is a worse experience for Zed.

backlight is now the main bottleneck for consumption heavy uses. I wonder what are the main advancements that are happening there to optimize the wattage.

comparing it to lmsys chatbot arena, what sort of an option would you expect? the prompts essentially come from public HF datasets like parti prompts where they test a bunch of stuff (prompt adherence, attention mapping [something in front of something else etc], aesthetics, photo-realism, etc.) so it is hard to ask about each category.