HN user

mdlincoln

927 karma

https://matthewlincoln.net

https://www.linkedin.com/in/mdlincoln

https://github.com/mdlincoln

Posts62
Comments18
View on HN
whyy.org 2y ago

William Noel, groundbreaking librarian and open data advocate, has died

mdlincoln
1pts0
about.jstor.org 2y ago

JSTOR is Now Available in 1k Prisons

mdlincoln
140pts96
arxiv.org 4y ago

Misogyny, pornography, and malignant stereotypes in LAION-400M image dataset

mdlincoln
2pts0
www.theartnewspaper.com 6y ago

Ghent Altarpiece: latest phase of restoration unmasks 16th century overpainting

mdlincoln
11pts4
www.lespetitescases.net 6y ago

I don’t use Semantic Web technologies anymore, though they still influence me

mdlincoln
119pts66
www.loc.gov 7y ago

Digital Strategy for the Library of Congress

mdlincoln
32pts6
blogs.msdn.microsoft.com 8y ago

Azure Government and ICE

mdlincoln
32pts34
blog.cmog.org 8y ago

Photographing glass: Lighting techniques for transparent glass objects

mdlincoln
137pts9
en.wikipedia.org 8y ago

"MS Fnd in a Lbry" (1961)

mdlincoln
2pts0
ryancordell.org 8y ago

Humorless Man Yells at English Major Jokes

mdlincoln
4pts2
www.crummy.com 8y ago

Tool Safety: Beautiful Soup and Software Ethics

mdlincoln
15pts0
www.theguardian.com 8y ago

Black and Latino representation in Silicon Valley has declined, study shows

mdlincoln
55pts64
www.theverge.com 8y ago

Transgender YouTubers had their videos used train facial recognition software

mdlincoln
2pts0
quamproxime.com 8y ago

The Temporality of Artificial Intelligence

mdlincoln
7pts0
techvibes.com 9y ago

It’s Time to Take ‘Diversity Debt’ Seriously

mdlincoln
1pts0
arxiv.org 9y ago

Generating “Art” by Learning About Styles and Deviating from Style Norms

mdlincoln
2pts0
www.dancohen.org 9y ago

Irrationality and Human Computer Interaction

mdlincoln
1pts0
mobile.nytimes.com 9y ago

Finland Works, Quietly, to Bury Its Nuclear Reactor Waste

mdlincoln
3pts0
adcontrarian.blogspot.com 9y ago

Display Ads: My 3¢ Worth

mdlincoln
1pts1
llamasandmystegosaurus.blogspot.com 9y ago

A Translation of Genesis I using word2vec

mdlincoln
1pts1
aeon.co 9y ago

Consciousness is not a thing, but a process of inference

mdlincoln
30pts2
medium.com 9y ago

Banning exploration in my infovis class

mdlincoln
165pts32
www.jacobinmag.com 9y ago

Duke Nukem’s Dystopian Fantasies

mdlincoln
2pts0
www.huffingtonpost.com 9y ago

Practicing Ada's “Poetical Science”

mdlincoln
1pts0
www.microsoft.com 9y ago

Inclusive Design at Microsoft

mdlincoln
2pts0
txt.fyi 9y ago

Txt.fyi

mdlincoln
655pts183
www.cnn.com 9y ago

Intel chiefs presented Trump with claims of Russian efforts to compromise him

mdlincoln
5pts0
boingboing.net 9y ago

Librarians to create fake patrons to fight automated book-culling software

mdlincoln
2pts0
opentranscripts.org 9y ago

Programming Is Forgetting: Toward a New Hacker Ethic

mdlincoln
10pts1
people.csail.mit.edu 9y ago

Purposes, Concepts, Misfits, and a Redesign of Git [pdf]

mdlincoln
6pts1

Prolific | Senior Software Engineer | Hybrid ONSIDE 1-2 days/wk Bay Area | $200k-$250k

Prolific is not just another player in the AI space – we are the architects of the human data infrastructure that's reshaping the landscape of AI development. In a world where foundational AI technologies are increasingly commoditized, it's the quality and diversity of human-generated data that truly differentiates products and models.

We’re looking for impact-focused Software Generalists to join our specialized team focused on serving frontier model creators and enterprise AI application developers. As a full-stack engineer, you will work across Prolific’s domains to solve customer and product problems.

This is an exciting opportunity to work directly with frontier AI companies, making critical technical decisions that balance scrappy startup execution with scalable, reliable engineering, as Prolific revolutionizes research for the AI community. You'll will have regular in person collaboration with customers and our US team, as well as collaborate closely with our UK-based tech teams.

Unfortunately we don't sponsor US visas at this time.

Apply: https://job-boards.eu.greenhouse.io/prolific/jobs/4767348101

context-dependent, or "reified" assertions are a pain point for sure. I come from the perspective of cultural heritage data, where context is king. Which expert made this attribution for this painting? Who owned it _when_? According to which archival document? etc.

Almost all the engineering problems cited in the original post are still basically there, but graphical models are still the least painful way of doing this, particularly when trying to share data between institutions. Example: https://linked.art/model/assertion/

I find that a fascinating reaction given how rapidly %>% have been taken up across a large segment of the R universe, to great excitement! Personally, I find it far MORE legible than endlessly-nested function calls.

It results in code that more closely resembles executed order of operations (e.g. filter -> mutate -> group -> summarize). Context is also key: it's most often used for data processing pipelines in specific analytical scripts or literate-code documents - less so used when defining generalizable/testable functions in packages (again, just a personal perspective - YMMV of course)

A related aside: while forgeries - deliberate imitations to mislead and deceive - are exciting, they only represent a very tiny portion of art attribution questions. In reality, these tend to deal more with discerning between artists working in the same period, rather than those attempting to fool the eye at several centuries' remove.

For example, the Rembrandt Research Project infamously set out to identify genuine vs. fake Rembrandt paintings in his corpus of known works under the false assumption that there would be a lot of 18th/19th/20th century forgeries. In fact, most of the "non-Rembrandt" cases they found were not later imitations, but instead works done by his own students or contemporaries - or works co-produced by Rembrandt and another. The result - deconstructing the project's original false assumption - proved revolutionary for our understanding of artistic studio practice from the period, but failed to locate many "forgeries" as such.

A review (paywalled, sorry!): http://www.sciencedirect.com/science/article/pii/02604779899...

And a Met exhibition: http://www.metmuseum.org/art/metpublications/Rembrandt_Not_R...

It's not mentioned in this guide, but Hadley Wickham's tidyr is a more streamlined version of the reshape2 package for fitting your data into a "tidy" format necessary for ideal faceting.

Yes, I am guilty of writing the post for an audience already largely familiar with the context.

I should probably add that the types of "explanations" I put forward in this post are actually not of central concern to me - certainly not explanations derived solely from parsing quantitative results. I'm far more interested in the descriptive evidence this kind of measurement can provide. It can give wider context to what tends to be a very case-study-centric discipline (e.g. oh, this guy happened to work a lot with Italian publishers in this period? We didn't realize it before just looking at 5-10 artists per article/monograph, but actually that is quite exceptional/normal for this period...)

Then again, proposing these kinds of explanations is also something of a disciplinary norm, for better or worse.

Righted Museum 11 years ago

http://www.washingtonpost.com/news/the-intersect/wp/2015/03/...

A project apparently exploring how copyright claims result in selective censoring in the Street-View-esque images of museum collections produced by the Google Art Project.

The WaPo article actually conflates copyright and reproduction rights (I work in a museum curatorial dept. FYI) Copyright would apply to works where artists, or their estates, can still make copyright claims over their artworks (although the role of fair use in reproducing images of art is evolving: http://www.collegeart.org/fair-use/)

But why can older artworks that are now in the public domain still have their images blurred out? Although the museum may have agreed to openly release representations of the public domain works that they own, it is often the case that museums may temporarily hang works on loan from private collectors in their galleries. In cases like these, museums and the lenders work out loan terms that frequently include provisions about photography. These loan agreements supersede copyright issues. Whether or not museums should agree to such terms is, of course, a good question.