If you mean CUDA specific then yes. The biggest benefit of these machines over the others is the CUDA ecosystem and tools like cuDF, cuGraph etc
HN user
xs83
TTFT on a Mac is terrible and only increases as the context increases, thats why many are selling their M3 Ultra 512GB
Now this looks much more interesting! Is the top one input tokens and the second one output tokens?
So 38.54 t/s on 120B? Have you tested filling the context too?
I've been on both sides of the software cycle as a stream lead and as a developer and ultimately I think this is it. Agile gets bastardised as soon as specific reporting structures and deadlines come into place "What do we need to plan for this increment?" etc
A team with a good work ethic and some milestone goals to go to and a "Take from the top" mentality, Kanban is perfect for what needs to be done.
Can't take from the top "because of reasons"? Thats a conversation with the people that particular problem involves and things are re-prioritised.
I'll never understand removing a whole day of productivity for ceremonies to plan being more productive.
Async standups via Slack using Geekbot are absolutely fine, A daily simple async retro for teams on a different timezone also help massively for synchronisation, discipline to update tickets with progress and a push regularly mentality all helps with the sync problem if across timezones!
I like the interface - I will say that it doesnt work when there is an ad-blocker installed (the workspace wont open).
When I am into the app it seems the database wasnt generated for items like the authentication system (which I would expect as it is included and mandatory).
Still more work needed but overall it looks great and I can see the advanced workings you have put in behind the scenes to make it do what it does.
https://aws.amazon.com/solutions/implementations/serverless-...
AWS just released this too which looks like a much more expensive version of their previous Lambda@Edge functionality.
Number one reason is probably * gestures wildly at everything *
I will die on the hill of squash-merge only. It doesnt make the history that confusing having branching visible with some run of the mill standardisation in process.
My biggest issue is that if you have to delay a feature then the sheer amount of conflicts you have to resolve (many of which you had nothing to do with) becomes prohibitive.
I have no problems with rebasing master or a published branch - thats the release managers problem to deal with everything on but I really dont know where this "rebase everything" comes from
How is its laziness? I found Opus to be very quick to curtail output and default into "Add the rest of your code here" type things.
I am a lazy data engineer - I want to prompt it into something I can basically copy and paste
Awesome news - we will eradicate HIV within the a single generation at this rate!
Nice!
I have been a long term Rectangle user on Mac and a WinSplitRevolution for decades on Windows, my biggest preference as a developer is for it to be keyboard centric.
I love the use of a numpad as a positional reference for the screen.
Hotkey + NumPad 6 = right side of screen full Hotkey + NumPad 5 = full screen Hotkey + Numpad 1 = Bottom left Quarter
Repeated presses of these cycle between Half, 3rd and Quarter
This reminds me of the Crysis UI which I still love as a weapon / armor switcher :)
I think the differences is that it isnt weaponised as it is in the US but it is very much valued and expected.
Yeah thats ridiculously expensive for a home lab IMO, the thinkcenters can be picked up for a couple of hundred bucks used and a T400 runs about the same. So $400 for an AI capable home bench would run $1200 for a 3 node cluster - I can live with that.
My 5 node Rpi5 Cluster ran $1500 with NVMe Hats and a PoE hat and cant do any GPU work :(
I think the Thinkstation M920 can house a single slot short GPU in it, something like the T400 (https://www.nvidia.com/content/dam/en-zz/Solutions/design-vi...) would give you the capability to run modern GPU workloads.
It wouldnt touch a data centre but the simple ability to run CUDA (Rapids.ai for example) would mean you could prototype things pretty efficiently all on a local setup and not have to pay for GPU costs.
UK and Australia.
The pros and cons are the same pros and cons you get with any system that has some form of centralised control.
The main one comes down to triage, if you need / want a procedure that isnt life threatening or helps your quality of life you can expect a long wait or not being able to get it. For example - any kind of cosmetic surgery isnt going to be done for free unless it affects your quality of life (e.g. Rhinoplasty for deviated septum)
If your situation is an emergency you get seen pretty quickly - otherwise it can take a while - as an example my mother needed a knee replacement - it took almost a year from when the doctor recognised the need for her to actually have it done. But it got done and didnt cost a penny.
She could have gone private, had it done the following week and paid 5-6 figures for it - it just wasnt that urgent and she couldnt afford it.
I have played rough sports and been hospitalised from it and I was treated immediately (emergency) - again this was free.
I think this concept is lost on people - not everything medical needs to done immediately and its sometimes fine to wait (even if its not preferred).
I think dismantling an existing system that is based on capitalism and is protected by the people rich enough to keep lobbying and getting kick backs will always be difficult.
It would need to be started again completely.
I dont think the US even has the concept of a public hospital does it?
He cant because thats exactly what it is.
The UK NHS system is on its knees because the conservative government has started selling it out from under the people to guess - US companies.
But all the gammons are spouting exactly the same as he is that its the "immigrants" causing it (even though Brexit was going to solve that problem - right?)
I am from 2 countries with "Free" healthcare system, I know that in either country if I really required it I would be seen, treated and discharged with no more than a few payments needed for medically necessary prescription medicines if I have a job, If I dont those are also covered.
In both countries - I can CHOOSE to go private for whatever reasons (for example, one of my friends has just had a baby and she chose to go private so she could get the OBGYN that she wanted and have a planned caesarian).
But it is not necessary and for those who cant afford private healthcare (or those who dont want to take it out) that safety net exists.
The foolproof solution (if such a thing could ever exist) - is to not have a healthcare system that is built for-profit. That way the government is the largest bargainer, you dont have companies colluding to drive up prices, everyone know what they are going to get and can't try and sway it that way.
By it coming from taxes the prices are managed better than if you have a bunch of cough self regulating companies running the racket.
I understand the pros and cons very well - and even with my private healthcare I get the benefits of the service I pay for with my taxes so my bill would never even be 10% of that 800k you somehow have paid.
nvm - I just saw OpenFaas's licence since I last looked at it
I believe that no one should have to pay for medical care at the point of receiving it. This is what Taxes and government planning are for
So how is this different to OpenFaas?
Perhaps its because the average house (e.g. the Simpsons house) is no longer achievable for the majority of young people, smaller footprints mean things have to become multipurpose.
Digression but Medical debt should not exist in the first place - especially not in a supposedly modern country.
I did a conversion of 500GB of data using dask_cudf on a GTX 1060 with 6GB of VRAM and was able to do it faster than a 20 node m3.xlarge Cluster.
What you can do on even consumer GPU's is mind blowing.
This and Rapids.ai is the single reason that NVIDIA is the leader in AI.
They made GPU processing at scale accessible to everyone, I have been a long term user of Rapids and found that even as a data engineer I can do things on an old consumer GPU that would otherwise require a 20+ node cluster to do in the same time.
It requires NVIDIA
I remember when mobile phones started to become big, I still had a pager / beeper at the time and realised that you could SMS directly to the beeper number for them to appear - that was a pretty cool few months until everything went the way of mobile and 2 way texting!
Training AI on AI generated data produces some increasingly weird outputs, I am sure we are already seeing the results of this in some models but the level of Hallucination is only going to increase unless some kind of checks and balances are implemented
I find extra long grain basmati rice doesnt do well in a rice cooker, much prefer the soak and huge pan of water method so I can get consistent results for inch long rice grains!
I measure for all other rices in my rice cooker depending on the rice