HN user

porridgeraisin

1,787 karma
Posts48
Comments1,196
View on HN
aviationweek.com 17h ago

GE Aerospace Marks Hybrid-Electric First with Saab 340B

porridgeraisin
1pts0
chatgpt.com 19h ago

Terry Tao's ChatGPT Session about the Jacobian Conjecture

porridgeraisin
7pts0
twitter.com 4d ago

Twitter user investigating potential Fable distillation in DeepSeek V4 Pro

porridgeraisin
1pts2
www.space.com 4d ago

Vikram-1, India's first private orbital rocket, aces debut launch

porridgeraisin
7pts0
www.lightreading.com 5d ago

Nvidia has a new AI-RAN plan – a 6G radio unit chip

porridgeraisin
4pts1
www.lightreading.com 5d ago

Nokia says long-term 6G is not doable without Nvidia

porridgeraisin
3pts0
www.tomshardware.com 7d ago

Starlink V5

porridgeraisin
6pts0
renewablesnow.com 7d ago

Reflect Orbital gets FCC nod for inaugural solar reflector

porridgeraisin
4pts0
www.xda-developers.com 17d ago

WSL Keeps Getting Better

porridgeraisin
5pts1
www2.eecs.berkeley.edu 19d ago

Understanding Latency Hiding on GPUs [pdf]

porridgeraisin
1pts0
wtop.com 20d ago

Germany unveils reform push: Tax cuts, pension overhaul and new sick leave rules

porridgeraisin
4pts0
wasm.chdb.io 21d ago

A complete ClickHouse OLAP engine, compiled to WebAssembly

porridgeraisin
65pts15
news.ycombinator.com 27d ago

Ask HN: Haven't posts about web front end frameworks completely stopped?

porridgeraisin
4pts3
newsroom.ibm.com 27d ago

IBM debuts sub-1 nanometer chip technology

porridgeraisin
392pts205
eccc.weizmann.ac.il 1mo ago

Bipartite Matching Is in NC

porridgeraisin
2pts0
twitter.com 1mo ago

Malware devs added nuclear and bioweapons text to trigger LLM safety refusals

porridgeraisin
3pts1
twitter.com 1mo ago

A new YC tool promises "your code never leaves your machine." It does

porridgeraisin
6pts0
x.ai 2mo ago

Use Grok in OpenCode

porridgeraisin
5pts0
cloud.google.com 2mo ago

AI-aided code migration: Google got 6x faster migration from TensorFlow to Jax

porridgeraisin
3pts0
www.fortuneindia.com 2mo ago

IIT Madras establishes Menlo Park Centre to help Indian startups scale globally

porridgeraisin
3pts1
www.bloomberg.com 3mo ago

Microsoft Takes over Norway Stargate Data Center from OpenAI

porridgeraisin
3pts0
iisc.ac.in 5mo ago

Redefining GAN power devices for adoption in EVs and data centres

porridgeraisin
1pts0
thejaggi.blogspot.com 5mo ago

Beyond WaPo angst: Why journalists need to abandon hubris and look within

porridgeraisin
4pts0
www.reuters.com 5mo ago

Apple acquires Israeli audio AI startup Q.ai

porridgeraisin
2pts1
physicsworld.com 8mo ago

The Forgotten Pioneers of Computational Physics

porridgeraisin
3pts0
fabiensanglard.net 8mo ago

Floating Point Visually Explained (2017)

porridgeraisin
2pts0
huggingface.co 8mo ago

Maya1: Open-source 3B Voice Model

porridgeraisin
4pts0
arxiv.org 8mo ago

Language Models Are Injective and Hence Invertible

porridgeraisin
1pts0
semiengineering.com 9mo ago

SRAM Scaling Issues, and What Comes Next (2024)

porridgeraisin
3pts0
www.government.nl 9mo ago

Minister of Economic Affairs Invokes Goods Availability Act for Nexperia

porridgeraisin
2pts0

From my skim of their paper, they have gotten robodojo 14% versus previous sota 9%. So its an improvement, but doesnt seem like a reliable model yet, which is fine. I also notice the model becomes worse with more data on the printer refilling task, so thats either a bad run or its a data point against the robustness. The transfer learning results are meaningful. 75% on new complex tasks from <10 hours of demo is really good. I believe this is a 2x improvement over the SOTA. 10 trials per task is standard for these papers but is pretty weak statistically. Memory tasks its weaker at but they have acknowledged it, and I cant see why this arch would have problems w that in the future.

Yes. The dhammapada uses a lot of Anushthubh chandas, which is an option in the website. However, the language differences between pali and sanskrit are enough to make TTS trip up. E.g only one "s" variant, and karma -> kamma, etc,. If you finetuned on pali a little bit it should work fine.

GRPO is policy gradient/PPO with your value function baseline monte carlo estimated using k rollouts. The only new thing is finding out it works well with binary rewards and LLM policies.

https://x.com/eliebakouch/status/2077428860639973471

gpqa diamond alone makes up 20% of their "capability index" and is responsible for ~70% of the gap between their model and nemotron nano (10 point difference!)

so what about contamination? i looked at the mixture and they train on a dataset "AIML-TUDA/QA-base" (for 10 epochs lol) that literally is a very very light rephrasing of the gpqa diamond test set

tl;dr: they trained on 10 epochs of a "very light rephrasing" of EVERY gpqa diamond EVAL data :)

Yep, I gave up on them for now. One of my older laptops also needs a charger that has the third ground pin, lest the trackpad become laggy due to electrical joise. I couldn't find a compact gan charger with a three pin plug and the third actually being connected to ground.

But this particular instance isn't much related to that. In my school too, we did all that, and extending to beach cleanups. It's not like the same effect was there.

Besides, most garbage in most Indian cities/towns is not consumer litter in the first place. Side note: this is why 'guarded' areas such as metro stations are spotless.

This whole thread is 100% correct. Zoning, rezoning, related points are all correct. One minor correction though, the megacities where the comment mentions zoning being better is in most cases only true definitionally. For example, there are portions of chennai which are by the book, not part of chennai, but for all practical purposes, are. e.g, the new major city-wide bus stand is definitionally outside the city, but practically it is the bus stand to go to if you want to travel to the rest of the state, is not at all zoned properly and faces the same issues.

The only exception where rezoning has actually been done properly is hyderabad.

Here is the reason for this:

- The 74th amendment in india only calls for a _planning commitee_ (reminiscent of india's past) for any area with a population of over 10 lakh. But this MPC is completely distinct from the organisation that does the actual service delivery. In Chennai's example, the CMDA are the planning guys, and the GCC (greater chennai corporation) are the execution guys who actually employ garbage collection folks, buy cleaning machines, fix storm water drains, etc etc,.

CMDA covers ~6000 sq. km. and all that is practically chennai, and they plan bus stops and FSI and roads and all that across this region. However, the GCC is _still_ limited _to this day_ to just 500 sq. km and this is the "chennai" that appears in reports. So all your sanitation workers and clean roads, tree lined roads, good storm water drains, etc, that you see in say besant nagar, are completely absent in the remaining 5500 sq. km. Instead, you have multiple different corporations, Tambaram, Avadi, Kanchipuram, etc, and the coordination problems here are just insane. It also absolutely does not help that political fracturing happens due to dominant castes differing across these regions. You also lose the simple concept of rich people cross subsiding poor people - which is essential in any public service. The taxes paid by boat club road residents, remain with the GCC, and does not contribute even a little bit towards better roads on land under the Tambaram corporation, ironically leading to worse infrastructure for even the boat club road resident making a trip to Trichy. So you get nice parks and heart-shaped traffic lights in nungambakkam before even a road is paved in thoraipakkam.

In hyderabad, they did a very nice thing. They merged 27 ULBs directly into the GHMC, tripling its area from 650 to 2000 sq. km. So its the same thing CMDA did but through a merger. Those 27 bodies - with their staff, their ward councillors, their budgets - are now part of GHMC. The sanitation workers who used to report to, say, Medipally Municipality now report to GHMC. The property tax collected there now goes into GHMC's budget. The roads there are now GHMC's responsibility to maintain.

This unified thing devoid of coordination problems and with great cross subsidisation seems to work much better. It will be a good model for other cities to follow.

I don't know about amdavad and pune but I am willing to bet they are also good for similar reasons.

In the classic FDE model palantir pioneered, one of the main features was that you would use your learnings from one customer and integrate that back into your in house product, so that similar customers are serviced for cheaper later on.

With AI coding agents FDEs are now everywhere. One because they can demand a higher salary due to simply doing more due to AI. And two, because AI really accelerates the whole bespoke solution implementation thing. However, from what I hear, none of the actual "integrate that back into your in house platform" stuff is actually happening. So it's a tiny bit of a farce.

Yes this. Anytime a SaaS supports UPI Autopay, I'm 10-100x more likely to subscribe.

For context: it's basically a recurring payment but you open the Autopay section in Google pay or whichever app and it shows you all your subscriptions right there (which is a major +), the transaction history, and cancellation happens in the UPI app itself I don't have to faff around in the SaaS and deal with potentially some dark patterns.

Funnily enough, I have my agent use vim sometimes.

I prefer commits to be granular to the extent that I manually edit the patch in git add -p, albeit rarely. So I have a pty plugin that my agent literally sends control codes to. This is also useful in general for LLMs driving any interactive terminal program. Long live text, I guess.

Grok 4.5 14 days ago

There is one thing to note here.

The fact that it is more token efficient will itself lead it to be smarter since the context will be smaller for the same task. However, in opus models, you'll have built internal correlations like "if it did X it will usually do Y" which may not be true here, since grok 4.5 may have done X purely due to the smaller context size, but can't do Y cos it wasn't RLd on that pattern enough. So it will be a unique experience as far as opus tier models go.

It would have been in a different tone had it been human written. In fact, if it had not been written with AI, there is no chance SVD gets called the fancy "engine behind image compression", the blog would have said exactly what you said in the comment above: "Its one of the simple forms of lossy compression, here's a simple before/after on a sample image" or whatever.