HN user

rbitar

118 karma

Founder/CEO @ Frontend.co. AI-powered app development platform.

Previously Founder @ SkillHire and CTO @ Goop. Ex-Google and YouTube PM. Stanford, UCSB Cornell.

React, TypeScript, Next.js, Tailwind, Ruby/Rails developer. Based in Miami, FL.

Email: rami+hn@frontend.co

Posts15
Comments55
View on HN
Claude Haiku 4.5 9 months ago

I regularly use @ key to add files to context for tasks I know require edits or patterns I want claude to follow, adds a few extra key strokes but in most cases the quality improvement is worth it

Claude Haiku 4.5 9 months ago

Interesting and if they are using speculative decoding that variance would make sense. Also your numbers line up with what openrouter is now publishing at 169.1tps [1]

Anthropic mentioned this model is more then twice as fast as claude sonnet 4 [2], which OpenRouter averaged at 61.72 tps for sonnet 4 [3]. If these numbers hold we're really looking at an almost 3x improvement in throughput and less then half the initial latency.

[1] https://openrouter.ai/anthropic/claude-haiku-4.5 [2] https://www.anthropic.com/news/claude-haiku-4-5 [3] https://openrouter.ai/anthropic/claude-sonnet-4

Congrats to the team, I'm surprised the industry hasn't been as impressed with their benchmarks on token throughput. We're using the Qwen 3 Coder 480b model and seeing ~2000 tokens/second, which is easily 10-20x faster then most LLM models on the market. Even some of the fastest models still only achieve 100-150 tokens / second (see OpenRouter stats by provider). I do feel after around 300-400 tokens/second the gains in speed feel more incremental, so if there was a model at 300+ tokens/second, I would consider that a very competitive alternative.

Really excited for this product, the industry needs alternatives to WebContainers which has become more restrictive around licensing. Also great to see that non-node runtimes (ruby / python) will be supported. Having said that, really wish this was open-source, even if that meant the OSS version had more limited features then the commercial alternative.

Cerebras Code 12 months ago

This token throughput is incredible and going to set a new bar in the industry. The main issue with the cerebras code plan is that number of requests/minute is throttled, and with agentic coding systems each tool call is treated as new "message" so you can easily hit the api limits (10 messages/minute).

One workaround we're doing now that seems to work is use claude for all tasks but delegate specific tools with cerebras/qwen-3-coder-480b model to generate files or other token heavy tasks to avoid spiking the total number of requests. This has cost and latency consequences (and adds complexity to the code), but until those throttle limits are lifted seems to be a good combo. I also find that claude has better quality with tool selection when the number of tools required is > 15 which our current setup has.

Frontend.co | REMOTE | Full-time & Part-time | https://www.frontend.co

Frontend is building an AI-powered Shopify development platform. We use AI to generate full-stack application built using Next.js / Tailwind for Shopify-connected storefronts.

Experience: 7-10+ years as a full-stack experience with recent experience using TypeScript, Next.js, Tailwindcss, Postgres, and Ruby on Rails. E-commerce development using Shopify GraphQL APIs is a plus. Our tech stack is:

- Next.js - Supabase / Postgres - TypeScript - Tailwind - Ruby on Rails - Shopify Storefront + Admin GraphQL APIs

Send us a note at: info[plus]hn@frontend.co

Excited to try this out, it will solve two problems we’ve had: applying a code diff reliably and selecting which files from a large codebase to use for context.

We quickly discovered that RAG using a similarity search over embedded vectors can easily miss relevant files, unless we cast a very wide net during retrieval.

We’ve also had trouble getting any LLM to generate a diff format (such as universal diff) reliably so your approach to applying a patch is exciting.

This looks great, glad to see this project and congrats on the launch. Having said that, how does this project fit in with the Shopify Hydrogen effort using Remix / React? There seems to be an ever growing number of ways to build a shopify storefront these days (ie, native templates, remix/hydrogen, web components, Shopify JS Buy SDK, etc.) so it's not clear what technology to "bet on" from a developer perspective.

Separately, nice touch adding the refined LLM instructions, this looks like a nice pattern for other UI frameworks to follow.

RubyLLM has been a joy to work with so nice to see it’s being used here. This project is also great and will make it easier to build an agent that can fetch data outside of the codebase for context and/or experiment with different system prompts. I’ve been a personal fan of claude code but this will be fun to work with

SEEKING WORK | New York, Florida | Remote Only

- Website: www.skillhire.com

- Email: info+hn@skillhire.com

SkillHire is a community of vetted, Freelance software developers for hire. We work primarily with startups to scale their engineering teams with a remote dev team.

Common tech stacks we work with include React, React Native (iOS), Node, and Ruby on Rails.

Drop us a line if we can help you build or augment your eng team.

We help senior remote developers around the world contract with companies here in the US, mostly Startups, for freelance work. Feel free to check us out at Skillhire.com

You’ll have to go through a technical interview process but most jobs are seeking JS developers, especially Node, React-Redux and React Native devs.

Hopefully we’re one more resource to help devs like yourself who want to travel and work.

SEEKING FREELANCER - Remote OK

Type: Freelance Ruby on Rails or React engineer

Location: Based in LA / NYC

Interested in a strong Ruby / Rails developer, preferably full-stack (we use HAML and React) or someone who is comfortable as a tech lead. Ideal if you have experience with Spree or Solidus (eCommerce). We're building a content / commerce site using React + Rails on the backend. If you have some DevOps experience using Docker on EC2, even better.

Contact: rami@goop.com

I'm the developer/founder of the site. I'm pivoting Cofoundr from its initial purpose as a social community to a flash sales site for founders. I am planning on sending offers I would actually find useful myself (more focus on quality rather than frequency). Any feedback you guys have would be helpful.

[dead] 16 years ago

Yes, there are. Follow the link and notice the redirect embeds params for the affiliate link.