HN user

acabal

9,146 karma

I run Scribophile, at www.scribophile.com, and Writerfolio, at writerfolio.com.

My company is called Turkey Sandwich Industries, at turkeysandwichindustries.com.

I'm also the Editor-in-Chief of Standard Ebooks, at standardebooks.org. Standard Ebooks is a volunteer-driven project dedicated to producing commercial-quality public domain ebooks edited to strict typography and coding standards for free and libre distribution.

My website is at alexcabal.com.

Posts25
Comments986
View on HN
alexanderpetros.com 2mo ago

Support Put, Patch, and Delete in HTML Forms

acabal
1pts0
standardebooks.org 1y ago

Public Domain Day in Literature

acabal
7pts2
www.newyorker.com 1y ago

A Controversial Rare-Book Dealer Tries to Rewrite His Own Ending

acabal
81pts34
www.berfrois.com 3y ago

Virginia Woolf on Not Knowing Greek

acabal
2pts0
www.vice.com 3y ago

Thank You for Your Feedback

acabal
1pts1
alexcabal.com 4y ago

Standard Ebooks serves millions of requests per month with a 2GB VPS

acabal
34pts6
nymag.com 4y ago

Fears of a Clown (2015)

acabal
1pts0
harpers.org 5y ago

The Anxiety of Influencers

acabal
206pts185
harpers.org 5y ago

The Anxiety of Influencers

acabal
3pts0
aeon.co 5y ago

Where Loneliness Can Lead

acabal
1pts0
www.polygon.com 7y ago

Morrowind: An Oral History

acabal
3pts0
www.artic.edu 7y ago

Art Institute of Chicago marks 52400 public domain works as CC0 in site redesign

acabal
2pts0
www.techdirt.com 8y ago

Appeals Court Says It's Okay to Copyright an Entire Style of Music

acabal
171pts66
www.nytimes.com 8y ago

Here’s How Far the World Is from Meeting Its Climate Goals

acabal
2pts0
motherboard.vice.com 8y ago

Mathematical Formula Predicts Global Mass Extinction Event in 2100

acabal
2pts1
aeon.co 8y ago

Step by step, Americans are sacrificing the right to walk

acabal
4pts0
defectivebydesign.org 9y ago

Tim Berners-Lee approves Web DRM, but W3C members have two weeks to appeal

acabal
249pts215
www.wired.com 9y ago

India’s Silicon Valley Is Dying of Thirst

acabal
3pts0
www.newyorker.com 9y ago

Greenland is Melting

acabal
1pts0
www.slate.com 9y ago

The uncertain future of copyright in the on-demand age

acabal
1pts0
alexcabal.com 13y ago

The results from our pay-what-you-want ebook pricing experiment

acabal
2pts0
standardebooks.com 13y ago

Show HN: Pay-what-you-want for DRM-free, beautifully-illustrated ebook classics

acabal
1pts1
phpbestpractices.org 13y ago

Show HN: PHP Best Practices, a short guide for common and confusing PHP tasks

acabal
16pts12
alexcabal.com 14y ago

How I made $1000 in 5 hours with an idea, a sales page, and a tribe

acabal
2pts2
sigildev.blogspot.com 15y ago

An Analysis of Epub3: "Website in a box," and why that's bad

acabal
3pts1

Because if HTTP is the language of the web, then HTML forms are how humans speak that language to computers. Right now we humans can only speak GET and POST.

In other words, right now if a human wants to DELETE a widget, the human has click on an HTML form to `POST /widgets/123/delete` - i.e. use an incorrect verb on an incorrect URL/object - or use some other workaround like smuggling a special `_method=DELETE` variable. This is unnatural and semantically incorrect, resulting in ugly hacks that break HTTP-level expectations like idempotency; and it also requires additional app-level logic to process.

Meanwhile a machine is allowed to simply `DELETE /widgets/123` because their interface to HTTP is not clicking on HTML forms.

We humans could converse with websites in semantically correct HTTP, have clean URLs in which both REST APIs and human-facing URLs are identical without hacks, and require no extra app/framework logic, if HTML forms simply allowed all (human-relevant) HTTP verbs.

Home folder litter is one of my top pet peeves in computing. In fact it's the only reason why I refuse to use snaps on Ubuntu. I don't even care about whatever technical stuff everyone argues about - but snaps create a permanent `~/snap/` directory and Ubuntu devs don't care. There's been a bug report on Launchpad for over a decade[1] and it's the second highest voted bug in Ubuntu history, but no, Ubuntu devs think littering the home folder with highly visible system-level machinery is totally unavoidable.

It's like putting your car's engine in the passenger seat - rude, intolerable, and plain stupid. What if Grandma was browsing her home folder and deleted `~/snap/` because she has no idea what it is?

[1] https://bugs.launchpad.net/ubuntu/+source/snapd/+bug/1575053

The gem in this post is Pure, which I haven't heard of until now. I also have my prompt show the git status, and for large repos `git status` can take 10+ seconds to load and cache.

I had no idea that you could do that asynchronously, and then have ZSH update the already printed prompt with the status later! That blows my mind!

SE editor in chief here. What you describe is incorrect. The only thing we do is very light sound-alike spelling modernization, like "to-night" -> "tonight". We do not do things like change from en-GB to en-US, replace old words with different modern words, or change text for "American readers", whatever that means. I have no idea where you got that impression.

I personally worked on the Forsyte saga. If you think something was done in error, please let us know and we'll be happy to fix it.

This article is rediscovering the same phenomenon that happened when the steam-powered machinery was invented, leading to the Luddite movement.

Machinery at the dawn of the industrial revolution was supposed to be a time-saving miracle that freed capitalists from having to deal with workers, and also freed workers from backbreaking labor, letting them spend their hours in the pursuit of leisure.

Of course, the opposite happened. Machinery meant workers could produce more output in the same amount of time, so they didn't work less, they worked at least the same and eventually even more to keep up with competition and the demands of consumers. It took decades of unrest and bloody conflict to give us the 8-hour workday.

This article is rediscovering that same history, but for a different class. AI is to white-collar knowledge workers what steam-powered machinery was to the rough-handed working class of the 1800s. It promises capitalists freedom from having to deal with highly-paid knowledge workers, and it promises highly-paid knowledge workers freedom from their labor so they can spend their time in the pursuit of leisure.

Look to history to see how that worked out.

I've always told people, Kindles are ereaders seeming designed by people who hate books.

The renderer is atrocious and is holding back the entire industry, much like IE6's crappy renderer and monopoly on users held the entire web back a decade. Browsers (and thus ebooks, which are just HTML/CSS) can now do pretty decent typography, but Amazon inexplicably refuses to get on board with epub.

Their file formats are equally garbage. Mobi, a format that has hardly changed since circa the year 2005, was still in active use until just recently. Their other proprietary formats are confusing in feature set and are opaque to create. The official tool to create Amazon ebooks only runs on Windows![1]

Kindles still can't natively read epubs, but since they accept epubs via email, their customers get confused and email me about it. (Epubs sent via email are quietly convert to Amazon's propriety format, meaning all bets are off on the result. Good luck, publisher!)

I always tell people, buy literally any other ereader.

[1] Calibre can also create them but it's reverse-engineering and not the official implementation.

The reading ease algorithm we use is the Flesh-Kincaid algorithm, which works pretty well for regular prose books but clearly fails very badly on avant-garde prose like Ulysses or As I Lay Dying.

The lost art of XML 6 months ago

XML lost because 1) the existence of attributes means a document cannot be automatically mapped to a basic language data structure like an array of strings, and 2) namespaces are an unmitigated hell to work with. Even just declaring a default namespace and doing nothing else immediately makes your day 10x harder.

These items make XML deeply tedious and annoying to ingest and manipulate. Plus, some major XML libraries, like lxml in Python, are extremely unintuitive in their implementation of DOM structures and manipulation. If ingesting and manipulating your markup language feels like an endless trudge through a fiery wasteland then don't be surprised when a simpler, more ergonomic alternative wins, even if its feature set is strictly inferior. And that's exactly what happened.

I say this having spent the last 10 years struggling with lxml specifically, and my entire 25 year career dealing with XML in some shape or form. I still routinely throw up my hands in frustration when having to use Python tooling to do what feels like what should be even the most basic XML task.

Though xpath is nice.

No, none have reached out yet. I've had some brief, high-level discussion along those lines with some people in the library industry, and the conclusion I drew is that public libraries in the US are highly fragmented in terms of technological capability. Instead of partnering with individual local library systems, it would make the most sense to - as you mentioned - partner with Overdrive. But there's been no movement in that direction. If anyone from Overdrive is reading, get in touch :)

I know you griped about this in a different thread, but we won't be doing that, sorry. You can uniquely identify an ebook and its version by using dc:identifier in combination with dcterms:modified in the metadata file. If you desperately need a filesystem-safe string then concatenate those two and sha it.

As Robin mentioned the typical style is "fine art oil painting", with some wiggle room allowed for exceptionally difficult cases (like Asian-themed books, as there just wasn't much fine art on that subject pre-1930).

We also require that the art have some kind of connection to the book itself, so it's not just some random fine art. Sometimes the connection is a little fuzzy, but we do the best we can given that art must be pre-1930 and also must have been previously published.

(My personal favorite artwork selection of the books I worked on is The Communist Manifesto[1]. That painting was actually made specifically for a different book by Willa Cather[2], but I thought the peasant laborer, holding a sickle in one hand, with a faraway look in her eyes as the red sun rises behind her was just too good to pass up for Marx!)

1920ish was when it started becoming much more common for books to have illustrated dust jackets, so now that more books from that era and onwards are entering the public domain, we opt to use the first edition dust jacket if it's in the appropriate style. Fortunately for us, that era also happens to be the so-called Golden Age of Illustration so it's not hard finding beautiful art to use!

[1] https://standardebooks.org/ebooks/karl-marx_friedrich-engels...

[2] https://standardebooks.org/ebooks/willa-cather/the-song-of-t...

The ebooks we produce are entirely in the US public domain, including metadata and any other files. Unfortunately there are basically no good fonts released under the CC0 license. (Most open fonts are released under the OFL license, which is not the same.) Therefore we don't embed any font files, except for Standard Blackletter[1] when necessary, which is a font we developed especially for our use based on public domain specimens, and released via the CC0 license.

[1] https://github.com/standardebooks/standard-blackletter

I'm shocked and saddened to hear this. Greg was a deep source of knowledge and support as I started and shepherded Standard Ebooks. He was generous with his time and experience, and unbelievably patient with me, some guy he had never heard of or met before who was just another cold-email in what must have been an endless stream in his inbox. We should all aspire to his high spirit of camaraderie, charity, and kindness. The world has lost a champion of both literature and the free web.

"Don't like it? Here is a full refund and you are free to read some other version."

That is not at all what I said.

You can't claim to care about preserving the works while changing them, and that is changing them.

We do not and have never made that claim. We are creating our own editions of these public domain books, not engaging in historical preservation.

If you want to read classic books in their original spelling, then you must locate first editions. Editors and publishers have updated both spelling and punctuation as a matter of course for centuries. Just look at any three editions of any Jane Austen novel - and you could never read an edition of Shakespeare more recent than 1800.

Spelling varies widely across the eras our ebooks were published in. Therefore we attempt to standardize spelling to what a modern reader might be familiar with. We only make sound-alike changes, like to-morrow -> tomorrow.

This is a common practice that editors and publishers have quietly engaged in for centuries. For example, today you are not reading Shakespeare in the way it was spelled in its first printing.

In addition to what Robin mentioned below, some of these placeholders are for books on our Wanted list. I also think it's useful to show readers that particular books are looking for volunteers to produce, and also to show that some books they might want are locked away by copyright for possibly decades. In that sense it's partly a political message.