HN user

martianlantern

88 karma
Posts21
Comments30
View on HN
darshanmakwana412.github.io 3mo ago

Gumbel Max trick for LLM sampling

martianlantern
3pts0
martianlantern.github.io 5mo ago

Strassen's Matmul with AVX 512

martianlantern
1pts0
bytes.usc.edu 6mo ago

System Design Interview: An insider's guide (Alex Xu) [pdf]

martianlantern
1pts0
martianlantern.github.io 6mo ago

3 Months of Using Neovim

martianlantern
5pts1
martianlantern.github.io 6mo ago

Fail Faster on Your Idea

martianlantern
2pts1
martianlantern.github.io 6mo ago

Fail Faster on Your Ideas

martianlantern
1pts0
martianlantern.github.io 7mo ago

Distilling persona vectors into LLM weights

martianlantern
1pts0
martianlantern.github.io 7mo ago

Fast Median Filter over arbitrary datatypes

martianlantern
38pts2
martianlantern.github.io 7mo ago

Updating My Bash Prompt

martianlantern
2pts3
martianlantern.github.io 7mo ago

How to Establish SSH over Tor

martianlantern
2pts0
martianlantern.github.io 7mo ago

Median Filter over Arbitrary Datatypes

martianlantern
1pts0
github.com 9mo ago

Show HN: Python Asyncio Puzzles

martianlantern
1pts0
martianlantern.github.io 10mo ago

LLM Routing Strategies

martianlantern
2pts0
martianlantern.github.io 10mo ago

Paged Attention Performance Analysis

martianlantern
1pts0
martianlantern.github.io 10mo ago

Paged Attention Performance Analysis

martianlantern
3pts0
github.com 11mo ago

ThinkMesh: A Python lib for parallel thinking in LLMs

martianlantern
73pts4
github.com 11mo ago

Train a small GPT to solve mazes

martianlantern
1pts0
github.com 11mo ago

Reasoning Cot Injection

martianlantern
1pts0
github.com 11mo ago

Injecting doubts in the CoT of reasoning models

martianlantern
2pts0
github.com 11mo ago

Injecting doubts in CoT of reasoning models

martianlantern
4pts0
kubernetes.io 11mo ago

Show HN: Kubernetes Tutorial

martianlantern
1pts0

Cool project, but just a side thought I was having about how do people have resources and the money to make things like this and make it avl for public, I mean it's fair to say they have their own GPUs or if they are using api keys for gpt or Gemini with enterprise subsidized inference

But still coming from a frugal background I still cannot wrap my head around this

I liked the new approach but I don't like the pseudo security framing. I don't see how an LLM rewritten query is secure than just sending my raw query, in both cases I don't feel secure at all

DeepSeek-v3.1-Base 11 months ago

Is there any benchmarks and comparisons compared to gpt-oss? I believe it far exceeds gpt oss or even gpt5 otherwise they wounldn't have released it

Hey, really cool work love the idea of focusing on key decision points. I was curious though since confidence can be non monotonic during CoT[1], how does binary search handle cases where there are multiple ups and downs in confidence? It seems like there might be more than one "pivotal" token, so I wonder if there's a plan to support multi-token pivots or use a different approach than binary search?

[1] https://arxiv.org/abs/2505.14489

I’m not familiar with Zig and would appreciate an explanation of how this works. My understanding is that cache behavior is managed by the CPU, and programmers only influence it indirectly through the sequence of instructions (i.e., access patterns). Is that accurate? Also, is this approach specific to Zig, or could it be achieved in C or Rust as well? Thanks

Very insightful post, this may work in the IMO setting because mathematical problems are inherently binary if we ignore somethings like the incompleteness theorem. In contrast, subjective tasks, such as evaluating a painting or rating a poem, lack absolute truth. How would such reasoners estimate confidence in these cases, and to what extent could RL techniques effective in the IMO transfer to real world problems?

Living with LLMs 12 months ago

Laymen used LLMs for all kinds of things as well. In an interview a prominent Finnish actor admitted to (paraphrasing) “pasting a full script into an AI app for content analysis and structuring”

I think it is now soon that hollywood movies or netflix dramas will be using generative AI. First it will be for blending frames, or increasing aesthetics, then some frames being generated to reduce production cost, then an entire shoot to be done in latent world and slowly AI will get adopted in everything that we see