crazy how they only show benchmark results against their own models
HN user
up6w6
I am very suspicious of the results. A few months ago they published a LLM benchmark, calling it "perfect" while it actually contained like only 50 inputs (academic benchmark datasets usually contain tens of thousands of inputs).
Very audacious to call it "almost perfect" when it has only what appears to be 50 questions. For comparison, MMLU contains 57 tasks and more than 100k questions.
I remember that you can detect the "curl | bash" server side and serve a different script than what the user would get by downloading using other methods[1]. But yeah, the binary itself already has enough attack surface.
Not only that, but Servo is only receiving less 3000 USD/month in donation, less than their goal of 10k/month, and much less than what they deserve if consider Ladybird is receiving millions.
https://servo.org/blog/2024/06/28/input-text-emoji-devtools/
Does anyone knows what is the current status of Apple silicon hardware emulation in qemu?
Related news: Servo Web Engine Continues Advancing But Seeing Just $1.6k In Monthly Donations
Lack of speed and ghosting felt like it made traditional Eink impossible to do most computing tasks. So we focused on making the most Paperlike epaper display that has no ghosting and high refresh rate - 60 to 120fps. We started working on this in 2018.
The website mentions 60hz, will it also support 120hz?
The Opus model that seems to perform better than GPT4 is unfortunately much more expensive than the OpenAI model.
Pricing (input/output per million tokens):
GPT4-turbo: $10/$30
Claude 3 Opus: $15/$75
Chatgpt-3.5 price reduction seems to be a direct response to Mixtral, which was cheaper (~0.0019 vs 0.0020 for 1K tokens) and better (https://arena.lmsys.org/) until now.
Sadly we will still have to wait for application. Firefox doesn't have HDR support even on Windows right now.
I think the medium is trying to compete with Anthropic's Claude than Openai's products
https://www-files.anthropic.com/production/images/model_pric...
For shell commands, there is a tool from Github that does it quite well (and it's included on the copilot plan)[1].
It's perfect to run fast bash commands that I forgot but are simple enough that the first Google search result would solve it anyways.
[1] https://www.npmjs.com/package/@githubnext/github-copilot-cli
Love how they want to keep the focus to the personal usage instead of selling themselves into building features for business.
Does anyone knows what is the current state of the voice commands? I always wanted a self hosted version of Google Assistant. I bet it's possible to use a fine tuned version of llama and open source models like whisper to have a local assistant smarter than any of these smart speakers available in the market.
True. The results from codex are actually from code-cushman-001 (Chen et al. 2021), which is an older model that Copilot was based on.
Even the 7B model of code llama seems to be competitive with Codex, the model behind copilot
https://ai.meta.com/blog/code-llama-large-language-model-cod...
usually can have their TDP modified in BIOS, via Ryzen Master, or third party tools like RyzenAdj (Phoenix support: https://github.com/FlyGoat/RyzenAdj/pull/256) to perform pretty closely.
I didn't know about that. Do you have citations or benchmarks?
Usernames is a pending feature[1], something discussed in their "secret" non-indexed forum.
[1] https://community.signalusers.org/t/usernames-in-signal/9157
They also don't have any publicly defined approach to combat criminal activity on their network, like they can't give the IP address of personal information, but they can still delete the groups and related accounts[1] - just create an account from these free temporary phone numbers from the internet, you will be able to see the past groups accessed (and non exited) by the account. To give a contrast, Matrix has a clear approach to moderation[2].
* they have a forum[3] that is not even indexed by search engines[4], which is not community friendly at all.
[1] https://community.signalusers.org/t/could-signal-become-the-...
[2] https://matrix.org/docs/communities/moderation/
[3] https://community.signalusers.org
[4] https://community.signalusers.org/t/google-site-search-doesn...
I guess it isn't related to grep.app, which is an excellent code search engine, right?
I think because Find My relies solely on the GPS report of surround Apple devices, it may happen in the situation where the only Ithings around the stolen devices are the ones in this family.
It's actually a mix of GPT3 (the one with an API available) and web results.
the difference is that matrix isn't trying to become the One True Standard, but just glue the others together. @xkcdComic
https://twitter.com/matrixdotorg/status/841424770025545730/p...
Most bridges work by running a program that will emulate a client. For example, with Telegram/Whatsapp/Signal you will authenticate the bridge bot using a qr-code just like if you were authenticating on a computer.
Also see [1], they have every bridge's features well documented.
We know there's been a lot of chatter recently about runtime speed.
For reference: https://news.ycombinator.com/item?id=32457587
Btw, I like where it's going but I find it quite sad that there is no official linux ARM64 support yet - which means I can't try to use it on AWS lambda for example.
A comparison between the mentioned Ampere Altra Q80-30 and AWS Graviton2:
https://blog.cloudflare.com/arms-race-ampere-altra-takes-on-...
A similar instance to the RX220 at Equinix (bare metal) would cost around $1800/month [1]. (8x Hetzner's price)
[1] https://metal.equinix.com/product/servers/c3-large-arm64/
https://news.gandi.net/en/2022/03/for-all-the-people-and-one...
Cutting off Russians and Belarusians would only encourage the creation of different closed worlds and digital networks. We have chosen to hold out our hand to these people. We are not at war with them. Only their leaders, and their madness, need to be stopped. We will of course react quickly against war propaganda of any kind.
Few points that made me choose them (though I would probably take Cloudflare if they supported the TLD of my domains):
I'm using Oracle's ARM servers and I thought it was some weird patch they did to the kernel, the bug only disappeared when I force upgraded it to 22.04. Ubuntu/Canonical itself would be the last place I would have thought to be the source of a problem like that.
There is also Kagi[1]