HN user

helloericsf

676 karma
Posts25
Comments63
View on HN
manus.im 10mo ago

Context Engineering for AI Agents: Lessons

helloericsf
120pts4
manus.im 1y ago

Context Engineering for AI Agents: Lessons

helloericsf
3pts0
huggingface.co 1y ago

Better than DeepSeek R1? MiniMax-M1:open-weight hybrid-attention reasoning model

helloericsf
6pts0
github.com 1y ago

kit - Code Intelligence Toolkit

helloericsf
1pts0
github.com 1y ago

DeepSeek Open Source Optimized Parallelism Strategies, 3 repos

helloericsf
103pts8
twitter.com 1y ago

DeepSeek Open Source DeepGEMM – FP8 GEMM Library(300 lines for 1350+ FP8 TFLOPS)

helloericsf
4pts1
twitter.com 1y ago

Alibaba Open Source Large-Scale Video Generative Models: Wan2.1

helloericsf
8pts2
github.com 1y ago

DeepSeek open source DeepEP – library for MoE training and Inference

helloericsf
536pts71
github.com 1y ago

DeepSeek Open Source FlashMLA – MLA Decoding Kernel for Hopper GPUs

helloericsf
441pts108
twitter.com 1y ago

New Qwen2.5-Max Outperforms DeepSeek V3 in Benchmarks

helloericsf
3pts2
github.com 1y ago

Longest context up to 4M, MiniMax-01 hybrid 456B Open source model

helloericsf
19pts1
huggingface.co 1y ago

DeepSeek v3 beats Claude sonnet 3.5 and way cheaper

helloericsf
48pts9
twitter.com 1y ago

NeurIPS and Dr. Picard released statement for singling out Chinese scholars

helloericsf
2pts2
github.com 1y ago

Tencent Hunyuan-Large

helloericsf
148pts103
huggingface.co 1y ago

Chinese AI Community: open-source Heatmap

helloericsf
1pts1
techcrunch.com 2y ago

Poolside is raising $400M+ at a $2B valuation to build a coding co-pilot

helloericsf
3pts1
bentoml.com 2y ago

Is LMDeploy the Ultimate Solution? Why It Outshines VLLM, TRT-LLM, TGI, and MLC

helloericsf
16pts8
arxiv.org 2y ago

21.2× faster than llama.cpp? plus 40% memory usage reduction

helloericsf
43pts14
www.datagravity.dev 2y ago

Databricks acquires Tabular, Snowflake fork Iceberg?

helloericsf
2pts3
huggingface.co 2y ago

New Yi 1.5 models under Apache 2.0

helloericsf
2pts0
www.snowflake.com 2y ago

Snowflake Arctic

helloericsf
1pts1
arxiv.org 2y ago

Training-Free Long-Context Scaling of Large Language Models

helloericsf
2pts0
www.deeplearning.ai 2y ago

Multi-agent collaboration design patterns

helloericsf
2pts1
blog.eleuther.ai 2y ago

Yi-34B, Llama 2, and common practices in LLM training

helloericsf
41pts3
developers.cloudflare.com 2y ago

Cloudflare Calls – Build real-time serverless video, audio and data applications

helloericsf
8pts1
Qwen3-VL 10 months ago

If you're in SF, you don't want to miss this. The Qwen team is making their first public appearance in the United States, with the VP of Qwen Lab speaking at the meetup below during SF teach week. https://partiful.com/e/P7E418jd6Ti6hA40H6Qm Rare opportunity to directly engage with the Qwen team members.

HF:https://huggingface.co/Wan-AI/Wan2.1-T2V-14B Github:https://github.com/Wan-Video/Wan2.1

Wan2.1, a comprehensive and open suite of video foundation models that pushes the boundaries of video generation. Wan2.1 offers these key features:

- SOTA Performance: Wan2.1 consistently outperforms existing open-source models and state-of-the-art commercial solutions across multiple benchmarks. - Supports Consumer-grade GPUs: The T2V-1.3B model requires only 8.19 GB VRAM, making it compatible with almost all consumer-grade GPUs. It can generate a 5-second 480P video on an RTX 4090 in about 4 minutes (without optimization techniques like quantization). Its performance is even comparable to some closed-source models. - Multiple Tasks: Wan2.1 excels in Text-to-Video, Image-to-Video, Video Editing, Text-to-Image, and Video-to-Audio, advancing the field of video generation. - Visual Text Generation: Wan2.1 is the first video model capable of generating both Chinese and English text, featuring robust text generation that enhances its practical applications. - Powerful Video VAE: Wan-VAE delivers exceptional efficiency and performance, encoding and decoding 1080P videos of any length while preserving temporal information, making it an ideal foundation for video and image generation.

- 389 billion parameters and 52 billion activation parameters, capable of handling up to 256K tokens. - outperforms LLama3.1-70B and exhibits comparable performance when compared to the significantly larger LLama3.1-405B model.

OpenCL was discussed more frequently in classes about a decade ago. However, I haven't heard it mentioned in the last five years or so.