HN user

rldjbpin

290 karma
Posts11
Comments712
View on HN

to me this behaviour is aligned with other industries China's been involved in. take solar or EVs as examples.

the dumping and involution involved in these industries is apparent and has been disrupting for the external markets affected. simultaneously both have helped make it economical to go green.

if you stop looking at LLMs as nukes, this becomes more convincing. China clearly does not want to rely on the West for what it considers core tech. they have done it for search and social; this is a natural next frontier. while they encourage a deluge of options for local use, others can make hay while the sun shine. albeit with the usual caveats.

autoregressive decoding is not an optimal paradigm for local or decentralized inference. amidst the mania and shortage, make the most of what is made available, like with solar, than crying over spilt milk.

hard to say the company was working the same post 7 series. the headlines is more like lexus leaving the US (ironic) than the parent company (Oppo).

those who still love their phones can still get most of the way there by using an equivalent oppo phone and tweak some software.

I think that's the most expensive pelican I've rendered through a Chinese model so far.

quite insane that it costs as much as 5.6 Terra [1], and twice the European counterpart (albeit dated for today's standards?) [2].

to be fair, the pelicans from Terra were quite weird all things considered. also, given the limited TPS from the first-party, it has to be pushing the limits of inference capabilities.

[1] https://openrouter.ai/openai/gpt-5.6-terra

[2] https://openrouter.ai/mistralai/mistral-medium-3-5

this. the problem with a lot of "think of the children" argument is that we often rule out the failure in whatever policing is needed for the general public.

the adults of today, albeit the younger ones, were kids growing up with post-IM world. the ones before that could speak with their friends from their homes instead of only meeting in person (due to telephones). we all are products of the environment we were raised in. but despite those learnings, it seems like we are not self-regulating as a society.

charity begins at home, and before bringing children in a conversation, we should think if we all are alright.

literally don't use any of the three parties involved in the drama (judge me all you want!).

however, i am glad this blog post exists, even though this is coming from reading the edited version. it might be down to ignorance, but the writeup contradicts the clarification in the bottom. it does seem like the series of events have had a personal impact to the writer. i am happy people can express themselves at their own leisure, and choose how to do so.

to those virtue signalling on language features or personal beliefs, go on a hike! the vast majority of projects that touch our lives are run by opinionated people, whether community-driven or from a for-profit. the more that is shared here about this, the more this seems like a byproduct of this reality.

Muse Spark 1.1 13 days ago

keep spending a few billion dollars developing frontier models, release them as open weights, and turn coding models into a commodity

the muse family is no longer open-source. it is still priced very "cheaply", but a significant shift in approach.

the title is doing the heavy lifting here. the paper discusses how they apply an existing standard (CXL), to provide an intermediate memory tier in a niche use case:

Each MemServer combines 768 GB of DDR5 memory alongside 256 GB of DDR4 connected through Vistara ASICs.

while the ratio above would be nice the other way around, this approach surely adds enough latency to make it comparable to intel optane than system memory at face value.

this should not take away ddr4 supply for those who would like to run it at home.

GPT-5.6 13 days ago

terra is just weird. in this nothingburger test, time nor higher costs seem to not strongly correlate with the aesthetics.

GPT-5.6 13 days ago

I really wish there was just an easy guide on when to use Sol vs Terra vs Luna

Their dev guide has the following:

Use gpt-5.6-sol for frontier capability, gpt-5.6-terra for a balance of intelligence and cost, or gpt-5.6-luna for efficient, high-volume workloads. The gpt-5.6 alias routes requests to gpt-5.6-sol

https://developers.openai.com/api/docs/guides/latest-model#u...

we used to believe lead could be turned into gold, and used it in fuel to paint not too long ago.

hindsight 20/20, maybe i will be proven ignorant about aluminium in the future.

it is amazing that meta continues to burn money and keep afloat. ads DO really make money, especially when they are going to overtake google in this [1].

to all the critics, i would suggest letting him/them cook. the snowden-era privacy concerns was exactly around ai getting trained or using personal data, and they have a treasure trove of it.

their money dump do sometimes also spend money on r&d unrelated to the current mindhive topics. we never know what the scatterbrain approach may give birth to tomorrow.

do be critical and vocal on how they squeeze money out of other areas or how they increase their revenue by ruining our lives instead.

[1] https://news.ycombinator.com/item?id=47758569

Resetting Xbox 16 days ago

xbox is not a startup and neither m$ a vc. the board is not in love with gaming to continue sponsoring it. but if the solution is to make it even worse to game in their platform, then i hope this crashes and burns for good.

this product really deserves the "halo" in the name. very difficult to gauge its fit in the current market.

if you want inference, go mac with much higher memory bandwidth. given the price premium here, or the little there is, you might as well.

if you want to finetune and experiment, cuda still has the moat and the kit is not much cheaper, if at all, than dgx spark.

from personal experience, i had access to amd developer cloud with a fair bit of credits. however, even doing inference outside of their supported use cases (which are often dated btw) using vllm was a pain. in the end, despite great computing potential on paper, i decided to not spend more time than its worth on it. if their enterprise cloud continue to have these grievances, i am not optimistic about this kit.

it might be down to skill issue on my end. perhaps if these sell and it gives amd enough motivation to add more software staff in-house, more power to them.

otherwise, good article from labs as usual. nice to know that other kits based on this soc are more or less the same (unsurprisingly).

The Vespa at 80 17 days ago

i lived in an interesting time and place combo where i'd see more apes around than vespas. primarily because we have local alternatives for vespas, which were also levied a lot of import duties to begin with.

while vespas came earlier and had the advantage of interchangeably calling scooters "vespas". however, did not figure out that they both shared the same parent company.

being true or not is irrelevant for decisions such as this. has been done on both sides, whether at software level or "hardware".

from Teslas not allowed parked around sensitive areas in the city, to blocking a (very famous and quite well-made) Russian antivirus, or Huawei communication stack.

Anthropic is not doing itself any favours with their recent (?) antics [1], so it is completely well-founded to do this imho. regardless, harness as a moat is not quite as established as the underlying models to the same extent.

[1] https://news.ycombinator.com/item?id=48734373

no matter your luck with hardware or your sysadmin skills, doing local inference for just yourself and/or to emulate typical usage (e.g. your coding workflow and deep research, etc.) is just very inefficient in current model architecture.

to me, this is a "truck" approach to city driving as a single person who does not do furniture hauling every weekend. the sense of privacy and freedom is nice but online inference is more "economical" as multi-user load is more effectively served than going solo.

maybe new architectures would make it effective to do text inference locally [1], till then great on you if you can spend car money on your setup. hope it is a great learning experience as well.

[1] https://deepmind.google/models/gemma/diffusiongemma/

all this talk, and the student pack's access to copilot (and i can imagine the free tier) is completely neutered.

you still cannot choose any models nor have much credits to work with (200 per month, around 2 usd). this was luckily the nudge i really needed personally to get out of the vs code paradigm, and i am lowkey glad for it.

overall, we have come full circle where vs code is using a model that one of its fork (cursor) has adapted as the core underlying model for its product.

time to short gamestop?

jokes aside, as someone who has never bought physical games, the landscape for physical ownership in this space was already shrinking.

there is a major chunk of playerbase who only plays online multiplayers, which are tied to always-online mode and constant updates. even if you could buy these titles on a disc, practically speaking they carried very little value of physical ownership.

we are just boiling the frog with little more intensity now. if it is possible to repurpose the bluray reader of ps5 to connect to pc, it might be a nice way to invest in one before they stop selling too.

the list of hostnames and words they compare the base url values with is just a nice advertising for these providers for me.

regardless, while you are not logged in and using a non-anthropic model (which is now fortunately feasible), there is nothing that affects your day-to-day.

the rest is just lame cat-and-mouse shenanigans to keep an eye out for.

the "physical" part is stressed too much. in today's landscape, it is no longer the only feasible way to address the different grievances made by the op.

if the content is going to be digital* anyway, despite fitting in on physical media, buying media online without drm is the way to go for most people. the bandcamp, gog of the lot that is. then your responsibility is usually just to download and find a way to self-preserve your collection.

physical media is easy on the latter part, but mostly painful on the former. in some cases, it is no longer feasible anyway. the new gta game will be digital-only, and if the nail is not on the coffin yet, this would be the "apple" moment to kill of any major publication ever supporting the medium again.

*i.e. all except vinyls which is still the "best" medium available for music

quite interesting to find hackable hardware in commodity looking smart devices.

advice for the op: the images in the page took half a minute to load (on multigigabit internet not being stressed). might be a routing issue between the server and my isp, but the images could use some optimization.

this industry has made people way too much money. so it attracts those who wants to make a quick buck with whichever means possible.

it took me a decade from falling in love with pc building to finally build one for myself. the crypto bros ruined it earlier, then now the crankers. both underlying tech is amazing. but any interesting idea or "breakthrough" trigger a gold rush mania.

same with the people working in certain parts of the globe. we were compensated disproportionately compared to other industries, thanks to the above tailwind and ZIRP.

i wonder how many people in the industry would be there if it wasn't the "future", at least in putting zeroes in our bank balance. personally would have not been here if not for the money, but would still have love for computers themselves and what we can do with thme.

while i understand the sentiment to switch out, i find one thing quite intriguing. the same crowd (read the website title) is happy to create workarounds to make their workflow adapt to the new os. however, very minimal mention of the various efforts to bypass and disable annoyances m$ is throwing at us. even the article mentions the OOBE as a major point.

personally have been able to use windows without ever logging in since windows 8.1. the iot (embedded) version is the path of least resistance for most. it even contains old paint and calculator from windows 7! might be a skill issue, but i've found it easier to debloat once than to figure out fractional scaling in linux.

not much info to personally go with tbh.

for those commenting on relying on a competitor for a "major" component of their product, look no further to how they source their displays from samsung all this time. even their silicon at one point! [1]

while it is now all about harness and "holding it right", their approach seem no different from other instances on hardware side. in terms of why not a more "leading" provider, look how the openai integration panned out so far. moreover, they have little to no initiative nor incentive to provide edge models or a hybrid solution so far. and anthropic's moat is not in audio.

[1] https://news.ycombinator.com/item?id=6419506