This is a really good point. My wife and I produce a history podcast where, right now, a lot of it draws on primary sources such as letters, diaries, etc. from the late 1800's to early 1900's. The ones in German she translates into English, but it feels "wrong" to change someone's words that are already in English.
The net result is that it's easier (for me) to understand a letter written by someone in a translated foreign language than one written by someone in English from the same period.
It's hard to pick a favorite - everyone brings their own perspective to their time. One of my early favorites was 'A Jew Who Stood Guard for Hitler'[1], the account of a kid who joined the Nazis for a while - till they found out he was Jewish and kicked him out. Ah, the naivety of youth.
Yeah? And here I've been a happy Transkribus customer for some time now. If there are better models or interfaces out there for analyzing historical handwriting, I'll definitely take a look.
In the movie Sneakers, a whole scene is taken up sending some guy on a date with Mary McDonnell so she could record clips of his voice. Today she'd just need a phone call or his Instagram. It's getting harder to keep up with who _people_ are online, much less organizations and domain names.
I don't understand what your point is. Do you disagree with any of the concrete suggestions in the blog post about what should have been done differently, or do you think they're hard to follow?
Yes, we've been using Transkribus for this extensively. My wife is a historian who spends quite a bit of time sorting through old letters and diaries, and it has been a considerable quality of life improvement.
Even if you are able to read someone's scratches, having a model to do the bulk lifting saves your eyes a lot of squinting. One thing that makes Transkribus useful for research vs a chat interface is that it can line up its interpretation alongside the original image so you can examine its work directly.
Whereas if we're talking about lossy compression (as is the person to whom you replied) we certainly can compress arbitrary data - almost as much as we want.
The hard question, then, is how much the decompressed output looks like the original.
reply