HN user

roflmaostc

411 karma
Posts6
Comments139
View on HN

I found once super old books in our lab (like hundreds of years) and was wondering what they were used for.

Apparently they did CT scans of closed books and read the content. Polevoy, Dmitry V., et al. "From tomographic reconstruction to automatic text recognition: the next frontier task for the artificial intelligence." Fifteenth International Conference on Machine Vision (ICMV 2022). Vol. 12701. SPIE, 2023. https://iris.unive.it/bitstream/10278/3687069/1/Albertin_et-...

So yeah, but lottery companies probably make it harder by engineering against it.

Many of the issues could potentially be solved by modern LLMs?

Reading, analyzing and assembling documentation could be probably done by LLMs.

And by including old code and snippets into the training set, the LLM could be fairly proficient in writing this code probably too?

Maybe someone knows more about the use/not-use of LLMs in this context?

Good initiative!

The problem is: publication is based on reputation. Reputation takes time and effort from the entire community.

I feel like modern infrastructure (Google Scholar, AI research, LinkedIn, etc) helped to decrease the importance of high-impact journals such as Nature, etc. Researchers don't rely on highly curated printed journals in their physical mailbox to get informed what's happening. You can just use tools to scrape content much faster.

But still: It can be career decisive if a reseachers lands a publication in a for-profit journal such as Nature.

The CS community has a much nicer publishing pipeline where most top journals/proceedings are attached to non-profit conferences and the fee is 0 (beside a conference fee).

I wish more fields would work like this: you publish with a conference proceeding and talk on the conference about your paper.

Researchers are themselves responsible for typesetting, advertising, etc. This and removing for-profit stakeholders can reduce the costs a lot.

Partially agree. However, this problem has existed with scam e-mails since the 90s.

For me the solution is in signed e-mails and signed documents. If the person invites me to a online meeting with a signed e-mail, I trust that person that it's really them.

Same for footage of wars, etc. The journalist taking it basically signs the videos and verifies it's authenticity. It is AI generated, then we would loose trust in that person and wouldn't use their material anymore.

It doesn't surprise me it happens within the Elsevier ecosystem. Elsevier has a long tradition of scientific misconduct and scientifically immoral behavior (see Wikipedia).

The operating margin of Elsevier is around 40% which is huge! At the end mostly paid by tax-payer money.

Personally, I never review or publish with Elsevier.

Prism 6 months ago

I am not so skeptical about AI usage for paper writing as the paper will be often public days after anyways (pre-print servers such as arXiv).

So yes, you use it to write the paper but soon it is public knowledge anyway.

I am not sure if there is much to learn from the draft of the authors.

Eat Real Food 7 months ago

Beef (red meat) is classified as a probable carcinogen, while chicken (white meat) is safe according to current research.

try to calculate 12312312.123213 * 123123.3123123

A computer uses orders of magnitude less energy than a human.

It's all about the task, humans are specialized too.

EDIT: maybe add a logarithm or other non-linear functions to make the gap even bigger.

Koralm Railway 7 months ago

Haha, check who updated this article. Only afterwards I realized we're not past the 14th yet...

Interesting. Personally I find it questionable to squat so many domains for ads. But they pay for it and it is within the legal framework.

Can you provide more details please?

The FFT is still easy to use, and it you want a higher frequency resolution (not higher max frequency), you can zero pad your signal and get higher frequency resolution.

Q: What if I need matrix dimensions (M, N, K) not found in your configurations? A: 1. You can find the nearest neighbor configuration (larger than yours) and pad with zeros. 2. Feel free to post your dimensions on GitHub issues. We are happy to release kernels for your configuration.

Lol, this will be potentially much slower than using the general matmul kernel.

However, I like this kind of research because it really exploits specific hardware configurations and makes it measurable faster (unlike some theoretical matmul improvements). Code specialization is cheap, and if it saves in the order of a few %, it quickly reimburses its price, especially for important things like matmul.

This is good news.

At EPFL we observe worrying trends that all services are moved to Microsoft (e-mails, cloud).

What happened to universities to host elemental services themselves?

EPFL also partnered up recently with Omnissa Work Space One to strengthen security of IT on campus. Mandatory (American) software which EPFL IT office wants to install on machines...