<?xml version="1.0" encoding="utf-8"?>
<feed xmlns="http://www.w3.org/2005/Atom">

  <title>weltliteratur.net</title>
  <subtitle>A Black Market for the Digital Humanities</subtitle>
  <icon>http://weltliteratur.net/favicon.ico</icon>
  <link href="https://weltliteratur.net/atom.xml" rel="self"/>
  <link href="https://weltliteratur.net/"/>
  <updated>2025-07-14T09:45:29+00:00</updated>
  <id>https://weltliteratur.net</id>

  
  <entry>
    <title>DraCor Platform Update (July 2025)</title>
    <link href="https://weltliteratur.net/dracor-platform-update-2025/"/>
    <updated>2025-07-14T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/dracor-platform-update-2025</id>
    <content type="html">&lt;ul&gt;
  &lt;li&gt;&lt;strong&gt;tl;dr:&lt;/strong&gt; Enhanced API, Frontend and Schema, Comprehensive Documentation, and Upcoming Summit in Berlin.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;After bigger updates in &lt;a href=&quot;https://weltliteratur.net/dracor-summer-update-2021/&quot;&gt;August 2021&lt;/a&gt; and &lt;a href=&quot;https://weltliteratur.net/streamlining-the-dracor-api/&quot;&gt;December 2023&lt;/a&gt;, we are excited to share that DraCor just got another upgrade!&lt;/p&gt;

&lt;p&gt;This latest release introduces a wide range of improvements – from enhancements to the API and frontend to refinements in the underlying TEI/XML data model. It is the result of months of collaborative work by our team, all in the spirit of keeping DraCor open, transparent, and easy to hack, build on, and explore. The new features come without any breaking changes.&lt;/p&gt;

&lt;p&gt;DraCor was built as an open digital infrastructure for the computational study of European drama – spanning from ancient Greek tragedy to 20th-century theatre. Since our last major technical update, the platform has grown not only in functionality but also in content, thanks in large part to our global community of researchers and enthusiasts who have contributed new corpora. Today, DraCor hosts more than 4,000 plays across nearly 30 corpora, representing over 20 languages – and the platform continues to grow.&lt;/p&gt;

&lt;p&gt;Before diving into the roundup of changes, we would like to give credit where it is due. The lion’s share of this update was made possible thanks to the tireless work of our technical lead, Carsten Milling, and co-lead, Ingo Börner. Development was supported through funding from &lt;a href=&quot;https://cordis.europa.eu/project/id/101004984&quot;&gt;CLS INFRA&lt;/a&gt; and &lt;a href=&quot;https://oscars-project.eu/projects/dracoros-fostering-open-science-digital-humanities-connecting-dracor-ecosystem-eosc&quot;&gt;DraCorOS&lt;/a&gt;.&lt;/p&gt;

&lt;h2 id=&quot;-a-smarter-cleaner-api&quot;&gt;💡 A Smarter, Cleaner API&lt;/h2&gt;

&lt;p&gt;The DraCor API has seen its most comprehensive revision since the project began. Among the many changes are the following:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;&lt;strong&gt;DTS Support:&lt;/strong&gt; We have implemented a &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;/dts&lt;/code&gt; endpoint to align the API in line with the &lt;a href=&quot;https://distributed-text-services.github.io/specifications/&quot;&gt;Distributed Text Services (DTS) specification&lt;/a&gt;. A DraCor Notebook documenting the new endpoint is available &lt;a href=&quot;https://github.com/dracor-org/dracor-notebooks/blob/main/dts/dts.ipynb&quot;&gt;here&lt;/a&gt;.&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Plaintext Endpoint:&lt;/strong&gt; Users can now fetch a plain text version of any play via a dedicated endpoint – making it easier to use DraCor data in NLP pipelines and other tooling.&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Corpus Metadata Enhancements:&lt;/strong&gt; We now track Git commit hashes from the corpus repositories, adding another layer of provenance and traceability to our data.&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Improved Filtering:&lt;/strong&gt; The &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;spoken-text&lt;/code&gt; endpoint now features refined relation filter parameters, offering more precise ways to query dramatic dialogue.&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Inclusive Metadata:&lt;/strong&gt; The API now supports both &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;sex&lt;/code&gt; and &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;gender attributes&lt;/code&gt;, enhancing its ability to reflect diverse encodings of character identity.&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Better Developer Experience:&lt;/strong&gt; We have extended and refined our OpenAPI specification, added more comprehensive tests, and introduced a robust CI workflow for building Docker images.&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Upgrades:&lt;/strong&gt; Under the hood, we have moved to &lt;strong&gt;&lt;a href=&quot;https://exist-db.org/&quot;&gt;eXist-db&lt;/a&gt; 6.4.0&lt;/strong&gt; and &lt;strong&gt;dracor-metrics 1.5.1&lt;/strong&gt;, improving both performance and maintainability.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Full API release notes here: &lt;a href=&quot;https://github.com/dracor-org/dracor-api/releases&quot;&gt;dracor-api changelog&lt;/a&gt;.&lt;/p&gt;

&lt;h2 id=&quot;️-a-refined-frontend-for-exploration&quot;&gt;🖥️ A Refined Frontend for Exploration&lt;/h2&gt;

&lt;p&gt;The DraCor frontend has also received a series of upgrades that improve usability and extend functionality:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;&lt;strong&gt;New “Tools” Tab:&lt;/strong&gt; Data from each play can be sent directly to third-party analysis tools, such as Voyant Tools, Gephi Lite or the CLARIN Switchboard.&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Corpus Status Overview:&lt;/strong&gt; We have added a &lt;a href=&quot;https://dracor.org/doc/corpora&quot;&gt;status page&lt;/a&gt; that provides an overview of the current state, maintenance responsibilities, and licencing of all DraCor corpora – useful for developers, curators, and researchers alike.&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Rendering Fixes &amp;amp; UI Improvements:&lt;/strong&gt; A variety of bugs were squashed, including issues with rendering segment lists. The toolchain has been updated, and we have migrated to &lt;strong&gt;pnpm&lt;/strong&gt; for improved dependency management.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Full frontend release notes here: &lt;a href=&quot;https://github.com/dracor-org/dracor-frontend/releases&quot;&gt;dracor-frontend changelog&lt;/a&gt;.&lt;/p&gt;

&lt;h2 id=&quot;-tei-schema-10-a-solid-foundation&quot;&gt;🧾 TEI Schema 1.0: A Solid Foundation&lt;/h2&gt;

&lt;p&gt;One of the most substantial and long-awaited developments is the release of the &lt;strong&gt;first stable version of the DraCor TEI schema&lt;/strong&gt;. It all began during a coffee break at the &lt;strong&gt;&lt;a href=&quot;https://jcls.io/site/ccls2024/&quot;&gt;CCLS 2024&lt;/a&gt;&lt;/strong&gt; in Vienna – and ended in a full overhaul of how we define and validate DraCor corpora. Highlights include:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;&lt;strong&gt;Based on tei_drama (TEI 4.9.0):&lt;/strong&gt; The schema builds on the official TEI customisation for drama, ensuring compatibility with broader TEI ecosystems.&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;More Flexibility &amp;amp; Clarity:&lt;/strong&gt; We have revised the structure to make it easier for contributors to encode new plays and for the schema to accommodate varied metadata.&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;New Elements &amp;amp; Better Validation:&lt;/strong&gt; We now support elements such as &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;&amp;lt;region&amp;gt;&lt;/code&gt;, &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;&amp;lt;orgName&amp;gt;&lt;/code&gt;, and &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;&amp;lt;addName&amp;gt;&lt;/code&gt;, and introduced Schematron rules tailored to DraCor’s specific needs.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Full schema release notes here: &lt;a href=&quot;https://github.com/dracor-org/dracor-schema/releases&quot;&gt;dracor-schema changelog&lt;/a&gt;.&lt;/p&gt;

&lt;h2 id=&quot;-comprehensive-documentation&quot;&gt;📜 Comprehensive Documentation&lt;/h2&gt;

&lt;p&gt;Our contribution to the EU project &lt;a href=&quot;https://clsinfra.io/&quot;&gt;CLS INFRA&lt;/a&gt; is documented in several detailed reports that provide comprehensive insights into DraCor:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;&lt;strong&gt;D7.1:&lt;/strong&gt; The report “On Programmable Corpora” presents DraCor as a prototype for an infrastructural ecosystem of &lt;em&gt;Programmable Corpora&lt;/em&gt;, defined as corpora that expose an open, transparently documented, and a research-driven API, enabling machine-actionable access to texts. This concept builds on on Aaron Swartz’s vision of “A Programmable Web” (2013), in which applications form an interoperable, dynamic ecosystem through the “natural growth” of APIs from the resources they make accessible. (&lt;a href=&quot;https://doi.org/10.5281/zenodo.7664964&quot;&gt;doi:10.5281/zenodo.7664964&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;D7.2:&lt;/strong&gt; The “Report on API Libraries for R and Python for the Programmable Corpora Prototype DraCor” introduces the DraCor libraries, (&lt;a href=&quot;https://cran.r-project.org/web/packages/rdracor/index.html&quot;&gt;rdracor&lt;/a&gt; for R and &lt;a href=&quot;https://pypi.org/project/pydracor/&quot;&gt;pydracor&lt;/a&gt; for Python), which provide researchers with direct programmatic access to DraCor’s corpora in their preferred programming language. (&lt;a href=&quot;https://doi.org/10.5281/zenodo.15302236&quot;&gt;doi:10.5281/zenodo.15302236&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;D7.3:&lt;/strong&gt; The report “On Versioning Living and Programmable Corpora” addresses the challenge of ensuring research reproducibility with &lt;em&gt;living corpora&lt;/em&gt; by recommending Git-based versioning for dynamic corpora and containerisation for complex programmable corpora infrastructures. (&lt;a href=&quot;https://doi.org/10.5281/zenodo.11081934&quot;&gt;doi:10.5281/zenodo.11081934&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;D7.4:&lt;/strong&gt; The final report of our CLS INFRA work package, “Report on the Implementation of Programmable Corpora”, documents the evolution of DraCor into a standards-compliant, extensively documented service. It now features a stable API (version 1.x), Distributed Text Services (DTS) endpoints, and new tools such as &lt;a href=&quot;https://github.com/dracor-org/ezdrama&quot;&gt;EzDrama&lt;/a&gt; and &lt;a href=&quot;https://github.com/dracor-org/epdracor-whois&quot;&gt;Who-is-@who&lt;/a&gt;, all of which contribute to transforming the experimental prototype into a sustainable piece of European infrastructure for computational literary studies. (&lt;a href=&quot;https://doi.org/10.5281/zenodo.15301341&quot;&gt;doi:10.5281/zenodo.15301341&lt;/a&gt;)&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;-save-the-date-dracor-summit-2025&quot;&gt;📅 Save the Date: DraCor Summit 2025&lt;/h2&gt;

&lt;p&gt;&lt;strong&gt;From 1 to 5 September 2025, the &lt;a href=&quot;https://summit.dracor.org/&quot;&gt;DraCor Summit&lt;/a&gt; will take place in Berlin&lt;/strong&gt; – hosted by the Freie Universität Berlin and the University of Potsdam. It will be five days of workshops, talks, a corpora conference, a computational drama analysis workshop, a barcamp, and even co-working time. The Summit is open to anyone working in or curious about computational literary studies, corpus development, or cultural analytics. There is no registration fee. General registration opened on 6 June – we would love to see you there!&lt;/p&gt;

&lt;h2 id=&quot;-follow-the-drama&quot;&gt;🧭 Follow the Drama&lt;/h2&gt;

&lt;p&gt;To stay up to date with DraCor news, project updates, and community events, subscribe to &lt;a href=&quot;https://www.listserv.dfn.de/sympa/info/dracormailinglist&quot;&gt;our mailing list&lt;/a&gt; and follow the official &lt;a href=&quot;https://bsky.app/profile/dracor.org&quot;&gt;DraCor account on Bluesky&lt;/a&gt;.&lt;/p&gt;

&lt;center&gt;&lt;p style=&quot;font-size:16px;&quot;&gt;(Portions of this blog post were composed with assistance from Claude Opus 4.)&lt;/p&gt;&lt;/center&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer, 
	
          
          Carsten Milling, 
	
          
          Ingo Börner, 
	
          
          Julia Jennifer Beine, 
	
          
          Peer Trilcke
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>How to Identify Key Passages in Literary Works? Using Algorithms and Machine Learning!</title>
    <link href="https://weltliteratur.net/Key-Passages/"/>
    <updated>2023-12-19T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/Key-Passages</id>
    <content type="html">&lt;h1 id=&quot;context&quot;&gt;Context&lt;/h1&gt;

&lt;p&gt;In the
&lt;a href=&quot;https://gepris.dfg.de/gepris/projekt/424207720?language=en&quot;&gt;DFG-funded&lt;/a&gt;
project &lt;a href=&quot;https://www.projekte.hu-berlin.de/en/schluesselstellen/what-matters-key-passages-in-literary-works&quot;&gt;What matters? Key Passages in Literary
Works&lt;/a&gt;
(which is part of the special priority programm &lt;a href=&quot;https://dfg-spp-cls.github.io/&quot;&gt;Computational
Literary Studies&lt;/a&gt;) we set out to
identify and characterize key passages in literary works.&lt;/p&gt;

&lt;p&gt;We understand key passages as passages that are particularly important
to expert readers when interpreting texts. In a mixed-methods
approach, we investigate empirically which textual characteristics of
literary genres can be revealed through patterns of citation and
quotation.&lt;/p&gt;

&lt;h1 id=&quot;corpus&quot;&gt;Corpus&lt;/h1&gt;

&lt;p&gt;Our main corpus consists of two literary works &lt;a href=&quot;https://en.wikipedia.org/wiki/Die_Judenbuche&quot;&gt;Die
Judenbuche&lt;/a&gt; by Annette
von Droste-Hülshoff and &lt;a href=&quot;https://en.wikipedia.org/wiki/Michael_Kohlhaas&quot;&gt;Michael
Kohlhaas&lt;/a&gt; by Heinrich
von Kleist with 44 and 49 scholarly articles,
respectively. Fortunately, we could build on the previous work of the
&lt;a href=&quot;https://gepris.dfg.de/gepris/projekt/372804438?language=en&quot;&gt;ArguLIT
project&lt;/a&gt;
and their annotation of all direct quotations.&lt;/p&gt;

&lt;h1 id=&quot;automatic-identification-of-quotations&quot;&gt;Automatic Identification of Quotations&lt;/h1&gt;

&lt;p&gt;Scholarly texts contain different types of quotations. For example,
verbatim quotes of single words to longer quotations spanning multiple
sentences, and indirect quotations in the form of summarizations or
re-narrations. In the first phase of the project, we focused on the
automatic identification and linking of direct quotations starting
with quotations of a length of five or more words. In &lt;a href=&quot;https://aclanthology.org/2021.nlp4dh-1.7.pdf&quot;&gt;Lotte and
Annette: A Framework for Finding and Exploring Key Passages in
Literary Works&lt;/a&gt;&lt;sup id=&quot;fnref:1&quot; role=&quot;doc-noteref&quot;&gt;&lt;a href=&quot;#fn:1&quot; class=&quot;footnote&quot; rel=&quot;footnote&quot;&gt;1&lt;/a&gt;&lt;/sup&gt;, we
outline the current landscape for text reuse detection and the
development of our tool &lt;a href=&quot;https://hu.berlin/quid&quot;&gt;Quid&lt;/a&gt;. Although there
are a number of existing tools, we found that all had limitations for
our specific use case. We evaluated Quid and compared it to the
existing tools.&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th align=&quot;center&quot; rowspan=&quot;3&quot;&gt;Approach&lt;/th&gt;
      &lt;th align=&quot;center&quot; colspan=&quot;3&quot;&gt;Die Judenbuche&lt;/th&gt;
      &lt;th align=&quot;center&quot; colspan=&quot;3&quot;&gt;Michael Kohlhaas&lt;/th&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;th align=&quot;left&quot;&gt;Precision&lt;/th&gt;
      &lt;th align=&quot;left&quot;&gt;Recall&lt;/th&gt;
      &lt;th align=&quot;left&quot;&gt;F&lt;sub&gt;1&lt;/sub&gt;&lt;/th&gt;
      &lt;th align=&quot;left&quot;&gt;Precision&lt;/th&gt;
      &lt;th align=&quot;left&quot;&gt;Recall&lt;/th&gt;
      &lt;th align=&quot;left&quot;&gt;F&lt;sub&gt;1&lt;/sub&gt;&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td align=&quot;left&quot;&gt;&lt;a href=&quot;https://www.sciencedirect.com/science/article/abs/pii/S0022283605803602?via%3Dihub&quot; target=&quot;_blank&quot;&gt;BLAST&lt;/a&gt;&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.59&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.61&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.60&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.37&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.59&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.45&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td align=&quot;left&quot;&gt;&lt;a href=&quot;https://plagiarism.bloomfieldmedia.com/software/copyfind/&quot; target=&quot;_blank&quot;&gt;Copyfind&lt;/a&gt;&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.85&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.75&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.79&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.76&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.79&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.78&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td align=&quot;left&quot;&gt;&lt;a href=&quot;https://people.f4.htw-berlin.de/~weberwu/simtexter/app.html&quot; target=&quot;_blank&quot;&gt;SimilarityTexter&lt;/a&gt;&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;&lt;strong&gt;0.91&lt;/strong&gt;&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.64&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.76&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;&lt;strong&gt;0.83&lt;/strong&gt;&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.74&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;&lt;strong&gt;0.79&lt;/strong&gt;&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td align=&quot;left&quot;&gt;&lt;a href=&quot;https://github.com/JonathanReeve/text-matcher&quot; target=&quot;_blank&quot;&gt;Textmatcher&lt;/a&gt;&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.69&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.37&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.48&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.68&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.42&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.52&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td align=&quot;left&quot;&gt;Quid&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.82&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;&lt;strong&gt;0.90&lt;/strong&gt;&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;&lt;strong&gt;0.86&lt;/strong&gt;&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.70&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;&lt;strong&gt;0.90&lt;/strong&gt;&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.78&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;
&lt;center&gt;&lt;p style=&quot;font-size: 16px; line-height: 24px;&quot;&gt;&lt;b&gt;Table 1.&lt;/b&gt; Comparison of different approaches for text reuse detection with an evaluation on our corpus.&lt;/p&gt;&lt;/center&gt;

&lt;p&gt;Considerably more difficult to identify are quotations which are
shorter than 5 words. In &lt;em&gt;A Novel Approach for Identification and
Linking of Short Quotations in Scholarly Texts and Literary
Works&lt;/em&gt;&lt;sup id=&quot;fnref:2&quot; role=&quot;doc-noteref&quot;&gt;&lt;a href=&quot;#fn:2&quot; class=&quot;footnote&quot; rel=&quot;footnote&quot;&gt;2&lt;/a&gt;&lt;/sup&gt;, we develop and compare two approaches to tackle this
challenge, &lt;em&gt;ProQuo&lt;/em&gt; and &lt;em&gt;ProQuoLM&lt;/em&gt;.&lt;/p&gt;

&lt;p&gt;For ProQuo, we use the (page) references for long quotations as
examples to tell apart (page) references for short quotations from
other text in parenthesis. This includes references like those to the
Bible or other literary works. We then relate short quotes to their
source in the literary work by figuring out the relationships between
the quotes and references. We also use the positions of long quotes as
guides to link short quotations to the correct passage of the literary
work.&lt;/p&gt;

&lt;p&gt;For our second approach, ProQuoLM, we fine-tune a &lt;a href=&quot;https://huggingface.co/dbmdz/bert-base-german-uncased&quot;&gt;German BERT&lt;/a&gt;
for classification. First, we identify potential short
quotes, and then use the fine-tuned model to filter them.&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th align=&quot;center&quot; rowspan=&quot;2&quot;&gt;Approach&lt;/th&gt;
      &lt;th align=&quot;center&quot; colspan=&quot;3&quot;&gt;Die Judenbuche&lt;/th&gt;
      &lt;th align=&quot;center&quot; colspan=&quot;3&quot;&gt;Michael Kohlhaas&lt;/th&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;th align=&quot;left&quot;&gt;Precision&lt;/th&gt;
      &lt;th align=&quot;left&quot;&gt;Recall&lt;/th&gt;
      &lt;th align=&quot;left&quot;&gt;F&lt;sub&gt;1&lt;/sub&gt;&lt;/th&gt;
      &lt;th align=&quot;left&quot;&gt;Precision&lt;/th&gt;
      &lt;th align=&quot;left&quot;&gt;Recall&lt;/th&gt;
      &lt;th align=&quot;left&quot;&gt;F&lt;sub&gt;1&lt;/sub&gt;&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td align=&quot;left&quot;&gt;Baseline&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.65&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;&lt;strong&gt;0.78&lt;/strong&gt;&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.71&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.59&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;&lt;strong&gt;0.75&lt;/strong&gt;&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.66&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td align=&quot;left&quot;&gt;ProQuo&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.87&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.72&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.79&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;&lt;strong&gt;0.87&lt;/strong&gt;&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.66&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.75&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td align=&quot;left&quot;&gt;ProQuoML&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;&lt;strong&gt;0.88&lt;/strong&gt;&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.75&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;&lt;strong&gt;0.81&lt;/strong&gt;&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;&lt;strong&gt;0.87&lt;/strong&gt;&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.69&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;&lt;strong&gt;0.77&lt;/strong&gt;&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;
&lt;center&gt;&lt;p style=&quot;font-size: 16px; line-height: 24px;&quot;&gt;&lt;b&gt;Table 2.&lt;/b&gt; Evaluation results of our two approaches compared to a baseline which always links a quotation from the scholarly work to the first matching instance in the literary work.&lt;/p&gt;&lt;/center&gt;

&lt;h1 id=&quot;quidex--visualization-and-exploration&quot;&gt;QuidEx – Visualization and Exploration&lt;/h1&gt;

&lt;p&gt;We created &lt;a href=&quot;https://hu.berlin/quidex&quot;&gt;QuidEx&lt;/a&gt;, a website for
visualization and exploration of the results, which is shown in this
screenshot:&lt;/p&gt;

&lt;figure style=&quot;text-align:center;&quot;&gt;
  &lt;img src=&quot;/images/key-passages-website.jpg&quot; alt=&quot;Key passages, website&quot; style=&quot;width:900px; border: 1px solid transparent; border-color: black;&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;On the left, there’s a heatmap that displays the distribution of
quoted passages in the entire literary work. The darker the area, the
more frequently it has been quoted, suggesting its significance. Right
beside the heatmap is the literary work itself. The grayscale
indicates how many scholarly works quote any part of a crucial
passage. This means the level of gray remains constant for the entire
key passage. The font size is adjusted based on how often a minimal
segment is quoted. At the bottom, alongside the literary text, there’s
a list of all scholarly works.&lt;/p&gt;

&lt;p&gt;The source code of the website is available as a &lt;a href=&quot;https://scm.cms.hu-berlin.de/schluesselstellen/quidex-wh&quot;&gt;white-label version&lt;/a&gt; which facilitates adoption by others, for example, for the &lt;a href=&quot;https://clarin09.ims.uni-stuttgart.de/sdc4lit/lotte/index.html&quot;&gt;comparison of intertextual relations of literary texts&lt;/a&gt;.&lt;/p&gt;

&lt;figure style=&quot;text-align:center;&quot;&gt;
  &lt;img src=&quot;/images/key-passages-banner.jpg&quot; alt=&quot;Key passages, banner&quot; style=&quot;height:700px; border: 1px solid transparent; border-color: black;&quot; /&gt;
  &lt;img src=&quot;/images/key-passages-banner-info.jpg&quot; alt=&quot;Key passages, banner&quot; style=&quot;height:700px; border: 1px solid transparent; border-color: black;&quot; /&gt;
&lt;/figure&gt;

&lt;h1 id=&quot;summary-and-outlook&quot;&gt;Summary and Outlook&lt;/h1&gt;

&lt;p&gt;In the first phase of the project, we developed tools to identify,
link, visualize, and explore direct quotations of all lengths.  In
August 2023, the project went into its second phase, titled &lt;a href=&quot;https://www.projekte.hu-berlin.de/en/schluesselstellen/&quot;&gt;Is Expert
Knowledge Key? Scholarly Interpretations as Resource for the Analysis
of Literary Texts in Computational Literary
Studies&lt;/a&gt;.
One important task we are currently working on, is the identification
and linking of indirect quotations, that is, summarizations and
re-narrations.&lt;/p&gt;

&lt;p&gt;For daily key passages, follow us on
&lt;a href=&quot;https://bsky.app/profile/fredr0id.bsky.social&quot;&gt;Bluesky&lt;/a&gt; or try Quid
online with our &lt;a href=&quot;https://hu.berlin/quidweb&quot;&gt;web interface&lt;/a&gt;.&lt;/p&gt;
&lt;div class=&quot;footnotes&quot; role=&quot;doc-endnotes&quot;&gt;
  &lt;ol&gt;
    &lt;li id=&quot;fn:1&quot; role=&quot;doc-endnote&quot;&gt;
      &lt;p&gt;Lotte and Annette have since been renamed to Quid and QuidEx, respectively. &lt;a href=&quot;#fnref:1&quot; class=&quot;reversefootnote&quot; role=&quot;doc-backlink&quot;&gt;&amp;#8617;&lt;/a&gt;&lt;/p&gt;
    &lt;/li&gt;
    &lt;li id=&quot;fn:2&quot; role=&quot;doc-endnote&quot;&gt;
      &lt;p&gt;Accepted at JCLS 2023 and soon to be published. &lt;a href=&quot;#fnref:2&quot; class=&quot;reversefootnote&quot; role=&quot;doc-backlink&quot;&gt;&amp;#8617;&lt;/a&gt;&lt;/p&gt;
    &lt;/li&gt;
  &lt;/ol&gt;
&lt;/div&gt;
</content>
    <author>
      <name>
	
          
          Frederik Arnold, 
	
          
          Robert Jäschke
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Streamlining the DraCor API</title>
    <link href="https://weltliteratur.net/streamlining-the-dracor-api/"/>
    <updated>2023-12-01T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/streamlining-the-dracor-api</id>
    <content type="html">&lt;h2 id=&quot;again-what-is-dracor&quot;&gt;Again, What Is DraCor?&lt;/h2&gt;

&lt;p&gt;DraCor, the Drama Corpora platform, is an infrastructure built to facilitate the digital research on European drama. We have currently 25 (!) TEI-encoded DraCor corpora in 19 (!) languages/dialects:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;Alsatian, Bashkir, Czech, Dutch, English, French, German, (Ancient) Greek, Hebrew, Hungarian, Italian, Latin, Polish, Russian, Spanish, Swedish, Tatar, Ukrainian, Yiddish&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;16 of these corpora are already running on our production server (&lt;a href=&quot;https://dracor.org/&quot;&gt;dracor.org&lt;/a&gt;), the others are still in the development phase and will be released when ready. Some corpora (like the Shakespeare corpus) have a stable number of plays. Depending on the corpus strategies of their maintainers, other corpora can generally grow and be updated with new plays. We call them ‘living corpora’.&lt;/p&gt;

&lt;h2 id=&quot;dracor-api-10&quot;&gt;DraCor API 1.0&lt;/h2&gt;

&lt;p&gt;The TEI-encoded plays are stored in an &lt;a href=&quot;http://exist-db.org/&quot;&gt;eXist database&lt;/a&gt;. Various encoding layers are extracted and made available via our API: customised data for individual research purposes. You can easily access the full text or text slices (a list of characters, spoken texts, stage directions) or co-occurrence networks. All API endpoints &lt;a href=&quot;https://dracor.org/doc/api&quot;&gt;are documented&lt;/a&gt; following the OpenAPI specification.&lt;/p&gt;

&lt;p&gt;A little bit of history. The DraCor API has grown organically ever since we introduced it sometime in August 2017. Yet the first publicly documented release on GitHub was &lt;a href=&quot;https://github.com/dracor-org/dracor-api/releases/tag/v0.72.0&quot;&gt;0.72.0&lt;/a&gt; in September 2020.&lt;/p&gt;

&lt;p&gt;New endpoints were added on a frequent basis and made DraCor what it is today, a research-prone platform serving data for computational literary studies. We always tried to use meaningful names and adhere to naming patterns, but over the years some inconsistencies have slipped in.&lt;/p&gt;

&lt;p&gt;Henny Sluyter-Gäthje revisited the consistency of names when working on her pydracor library, a wrapper for DraCor API endpoints (&lt;a href=&quot;https://github.com/dracor-org/dracor-api/issues/186&quot;&gt;see GitHub ticket&lt;/a&gt;). Building on her work, we revised the DraCor API, making it more consistent and sustainable.&lt;/p&gt;

&lt;p&gt;This work took a lot of time and after the streamlined API has worked well on our staging server for a while now, we release today the first version of the DraCor API that we consider stable, &lt;a href=&quot;https://github.com/dracor-org/dracor-api/releases&quot;&gt;our 1.0.0&lt;/a&gt;.&lt;/p&gt;

&lt;figure style=&quot;text-align:center;&quot;&gt;
  &lt;img src=&quot;/images/dracor/dracor-mockup-2023.jpg&quot; alt=&quot;DraCor Mockup 2023&quot; style=&quot;width:600px; border: 1px solid transparent; border-color: black;}&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;p style=&quot;font-size: 16px; line-height: 24px;&quot;&gt;&lt;b&gt;Fig. 1.&lt;/b&gt; This mockup is part of the new &lt;b&gt;&lt;a href=&quot;https://dracor.org/doc/media-kit&quot;&gt;DraCor Media Kit&lt;/a&gt;&lt;/b&gt; by Mark Schwindt, released along with the new API version.&lt;/p&gt;&lt;/center&gt;

&lt;h2 id=&quot;legacy-api&quot;&gt;Legacy API&lt;/h2&gt;

&lt;p&gt;From here on out, we would like that everyone who works with the DraCor API to use version 1.x. The URL paths on dracor.org now have a &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;/v1&lt;/code&gt; prefix (e.g., &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;https://dracor.org/api/v1/info&lt;/code&gt;).&lt;/p&gt;

&lt;p&gt;However, the last pre-release version of the API will remain to be available under URLs featuring a &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;/v0&lt;/code&gt; prefix (e.g., &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;https://dracor.org/api/v0/info&lt;/code&gt;). Old URLs without a version prefix will be redirected to the versioned ones. For instance, &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;https://dracor.org/api/info&lt;/code&gt; now redirects to &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;https://dracor.org/api/v0/info&lt;/code&gt;. This should allow old scripts to function as long as they follow the redirect.&lt;/p&gt;

&lt;p&gt;We plan to keep the API version 0.x available for at least a year. After this period, we may phase it out so as not to have to maintain multiple versions of the API indefinitely. Also see &lt;a href=&quot;https://dracor.org/doc/faq&quot;&gt;dracor.org/doc/faq&lt;/a&gt;.&lt;/p&gt;

&lt;h2 id=&quot;pydracor-and-rdracor&quot;&gt;pydracor and rdracor&lt;/h2&gt;

&lt;p&gt;Our API wrappers will be updated to work with our API version 1.0.&lt;/p&gt;

&lt;p&gt;pydracor 2.0.0 is already available &lt;a href=&quot;https://pypi.org/project/pydracor/&quot;&gt;via PyPI&lt;/a&gt;, the updated rdracor will be available &lt;a href=&quot;https://cran.r-project.org/web/packages/rdracor/index.html&quot;&gt;via CRAN&lt;/a&gt;.&lt;/p&gt;

&lt;h2 id=&quot;other-updates&quot;&gt;Other Updates&lt;/h2&gt;

&lt;p&gt;We introduced a new &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;/wikidata/mixnmatch&lt;/code&gt; endpoint to automatise the integration of DraCor permalinks into Wikidata. A corresponding &lt;a href=&quot;https://www.wikidata.org/wiki/Wikidata:Property_proposal/DraCor_ID&quot;&gt;DraCor property&lt;/a&gt; was also proposed and we wait for approval.&lt;/p&gt;

&lt;p&gt;We retired the &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;/corpora/{corpus}/plays/{play}/segmentation&lt;/code&gt; endpoint so as to reduce redundancies (the same information can be obtained via the &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;/corpora/{corpus}/plays/{play}&lt;/code&gt; endpoint).&lt;/p&gt;

&lt;p&gt;For now, the RDF endpoint has been removed from the OpenAPI specification as it needs more work. This is something we prioritise and plan to release rather sooner than later, probably with a 1.1 release of our API.&lt;/p&gt;

&lt;p&gt;With this release, we also updated our eXist-db from 6.0.1 to 6.2.0.&lt;/p&gt;

&lt;h2 id=&quot;the-team-behind-the-update-&quot;&gt;The Team Behind the Update …&lt;/h2&gt;

&lt;p&gt;… includes Carsten Milling, Henny Sluyter-Gäthje, Ingo Börner, Peer Trilcke, Daniil Skorinkin (all University of Potsdam), Frank Fischer, Viktor J. Illmer, Mark Schwindt, Julia Beine, Heinz-Alexander Fütterer (all Freie Universität Berlin) and Ivan Pozdniakov. The implementation was largely carried out as part of the “CLS INFRA” project (&lt;a href=&quot;https://cordis.europa.eu/project/id/101004984&quot;&gt;funded through EU’s Horizon 2020 programme&lt;/a&gt;) and the Cluster of Excellence “EXC2020 Temporal Communities” (&lt;a href=&quot;https://gepris.dfg.de/gepris/projekt/390608380&quot;&gt;funded through DFG&lt;/a&gt;).&lt;/p&gt;

&lt;p&gt;DraCor was founded as a community project and, as always, numerous friends and colleagues have contributed feedback and ideas while working on their corpora or conducting research based on data from DraCor.&lt;/p&gt;

&lt;p&gt;The platform, &lt;a href=&quot;https://web.archive.org/web/20200119005611/https://twitter.com/eumanismo/status/1218066125969412096&quot;&gt;once praised&lt;/a&gt; for its original guerilla strategy – although this is debatable 😊 – has become a well-frequented part of the international research infrastructure. We trust that this API 1.0 release will contribute to making digital research on European drama (and beyond) more reliable and sustainable.&lt;/p&gt;

&lt;p&gt;Goodbye for now, may your requests be valid, your responses swift, and your status codes always successful!&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer, 
	
          
          Carsten Milling, 
	
          
          Mark Schwindt
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Reverse Engineering "Frembdsch", A Fantasy-Language Radio Play by Dagmara Kraus and Marc Matter</title>
    <link href="https://weltliteratur.net/reverse-engineering-frembdsch/"/>
    <updated>2023-02-09T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/reverse-engineering-frembdsch</id>
    <content type="html">&lt;p&gt;First aired on the German radio station SWR2 &lt;a href=&quot;https://www.swr.de/swr2/hoerspiel/frembdsch-swr2-ohne-limit-2022-12-17-100.html&quot;&gt;on December 17, 2022&lt;/a&gt;, “Frembdsch” (which could be translated into English as &lt;em&gt;Fornlish&lt;/em&gt; or &lt;em&gt;Foreysh&lt;/em&gt;) is an audio play composed by author/translator Dagmara Kraus and media artist/sound poet Marc Matter. In the performance the composers are joined by the voices of Susanne Reuter, François Cavaillès, and Uwe-Peter Spinner. For the most part, the dialogue of the play is in an invented language, while the stage directions, which were derived using the German Drama Corpus at &lt;a href=&quot;https://dracor.org/&quot;&gt;dracor.org&lt;/a&gt;, are spoken in relatively understandable, mostly standard High German. &lt;a href=&quot;https://www.hoerspielundfeature.de/frembdsch-100.html&quot;&gt;Hörspiel und Feature&lt;/a&gt; describes the piece as being based on the concept of &lt;a href=&quot;http://stella.atilf.fr/scripts/fantomes.exe&quot;&gt;“mots fantômes,”&lt;/a&gt; incorrectly read and written phantom words that have found their way into dictionaries and then into the mouths of speakers. What is produced is a broad polysemy with very indefinite semantics.&lt;/p&gt;

&lt;p&gt;What sort of language is Frembdsch? A multilingual speaker of a Romance, Germanic, or Slavic language might be reminded of Hungarian for its opaque semantics and familar phonology. Common word endings include -os and -ć which are reminiscent of the Greek noun and the Polish verb. Some words are repeated, among “gudem,” “finestre,” and “mazgrab,” which, as a German speaker, sound like “good,” “gloomy” or “window,” and “mass grave” respectively. The R is in most cases an alveolar trill, as in Italian or in Slavic languages. Phonologically it is clearly European, although which language it most closely resembles differs based on the speaker. Its least Indo-European feature is rather prevalent reduplication (as in some creole languages), whose function seems to be primarily narrative. The only period in which I detected a consonant unused in Indo-European languages was between ‘20:00 and ‘21:00, where an actor produces what sounds like a dental click (k͡ǀ), a sound only found in Africa and in the Damin ritual jargon of Australia. To this, the person saying the stage directions says “Befremdet” (tr. disconcerted or estranged).&lt;/p&gt;

&lt;p&gt;As stated above, the stage directions were produced with the help of the &lt;a href=&quot;https://dracor.org/ger&quot;&gt;German Drama Corpus&lt;/a&gt;. Many of these stage directions appear word for word in the corpus (some many times), other have been distorted and denatured, so as to make their identification impossible. Some were simply invented by the authors. I did a search in order to see how many times each stage direction used in Frembdsch appears in the German Drama corpus. The top ten are:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th style=&quot;text-align: left&quot;&gt;Stage direction&lt;/th&gt;
      &lt;th style=&quot;text-align: right&quot;&gt;Count&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Für sich. (tr. to oneself)&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;3015&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Heftig. (tr. severely)&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1537&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Verwirrt (tr. confused)&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;234&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Zuckend. (tr. twitching/flinching)&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;216&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Geheimnisvoll. (tr. mysteriously)&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;199&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Spöttisch. (tr. mockingly)&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;189&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Von innen. (tr. from within)&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;187&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Taumelt. (tr. staggering)&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;162&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Neugierig. (tr. curiously)&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;146&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Befremdet. (tr. disconcerted/estranged)&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;99&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;To oneself, severely, confused, flinching: these are the bread and butter of German stage directions.&lt;/p&gt;

&lt;p&gt;I also compiled a list of the works in which the stage directions from Frembdsch most often appear. The top ten:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th style=&quot;text-align: left&quot;&gt;Play&lt;/th&gt;
      &lt;th style=&quot;text-align: right&quot;&gt;Count&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Ignorabimus by Arno Holz&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;173&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Das Konzert by Hermann Bahr&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;103&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Das Haus der Temperamente by Johann Nestroy&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;82&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Sonnenfinsternis by Arno Holz&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;80&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Die Industrie-Ausstellung by Friedrich Kaiser&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;75&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Der Vampyr by Wilhelm August Wohlbrück&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;74&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Einen Jux will er sich machen by Johann Nestroy&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;74&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Der Talisman by Johann Nestroy&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;65&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Zar und Zimmermann by Albert Lortzing&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;65&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Faust by August Klingemann&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;61&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;These are generally the works in which the above most common stage directions most often appear. In Arno Holz’s Ignorabimus, for example, the word “heftig” appears no less than 93 times. “Für sich” appears 101 times in Nestroy’s Das Haus der Temperamante.&lt;/p&gt;

&lt;p&gt;I wanted to see which of the stage directions could be traced back to single sources, that is to say, which of them are distinctively from a single work. The 31 stage directions with a distinct, exact match in the German Drama Corpus were:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th style=&quot;text-align: right&quot;&gt;Time&lt;/th&gt;
      &lt;th style=&quot;text-align: left&quot;&gt;Stage direction&lt;/th&gt;
      &lt;th style=&quot;text-align: left&quot;&gt;Play&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;00:43&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Kein Licht. (tr. No light)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Ödipus und die Sphinx by Hugo von Hoffmansthal&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;02:56&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Noch halb im Fass. (tr. Still half in the barrel.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Eulenspiegel by Johann Nestroy&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;04:00&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Der Wind stöhnt. (tr. The wind moans.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Draußen vor der Tür by Wolfgang Borchert&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;05:00&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Springt mit allen Zeichen des Entsetzens zurück. (tr. Springs back with every indication of distress.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Der Sohn by Walter Hasenclever&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;05:37&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Von großer Erregung übermannt. (tr. Overwhelmed by great agitation.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Der Sohn by Walter Hasenclever&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;07:32&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Tod geht ab. (tr. Exit death.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Draußen vor der Tür by Wolfgang Borchert&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;08:32&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Übertrieben elegant. (tr. With exaggerated elegance.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;XYZ by Klabund&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;08:41&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Immer mehr schwatzhaft und schnodderig. (tr. Ever more loquacious and flippant.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Draußen vor der Tür by Wolfgang Borchert&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;08:47&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Absichtlich gemein. (tr. Purposefully base.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Draußen vor der Tür by Wolfgang Borchert&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;10:27&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Abwechselnd lauschend und sprechend. (tr. Alternately listening and speaking.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Meister Oelze by Johannes Schlaf&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;11:38&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Nimmt sich das Rätsel vor. (tr. Gets to work on the riddle.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;XYZ by Klabund&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;12:41&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Bricht disharmonisch ab. (tr. Discontinues dissonantly.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;XYZ by Klabund&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;13:01&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Kugelfüsse. (tr. Spherical feet.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Der Wetterfürst by Paul Scheerbart&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;13:35&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Wird immer leiser. (tr. Grows ever quieter.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Draußen vor der Tür by Wolfgang Borchert&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;13:45&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Mit strengem Ton. (tr. With a harsh tone.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Almansor by Heinrich Heine&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;14:10&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;In größter Ungeduld. (tr. Extremely impatient.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Eulenspiegel by Johann Nestroy&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;14:14&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Tut pikiert. (tr. Acts piqued.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Meister Oelze by Johannes Schlaf&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;14:26&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Robust, um nicht zu heulen. (tr. Stolidly so as not to sob.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Draußen vor der Tür by Wolfgang Borchert&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;15:45&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Perplex zurück. (tr. Back, perplexed)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;XYZ by Klabund&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;16:05&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Mit komischer Gravität. (tr. With comic gravity.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Zar und Zimmermann by Albert Lortzing&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;16:36&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Mit ungeschickter Ängstlichkeit. (tr. With clumsy anxiety.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Eulenspiegel by Johann Nestroy&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;18:14&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Mit den Bewegungen einer Marionette. (tr. With the movements of a marionette.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Eulenspiegel by Johann Nestroy&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;18:23&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Alle fahren erschrocken herum. (tr. All wheel about, shocked.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Der Sohn by Walter Hasenclever&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;19:36&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Auf beide los. (tr. Coming at both.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Zar und Zimmermann by Albert Lortzing&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;21:29&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Und ziehen sich zurück. (tr. And draw back.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Hidalla oder Sein und Haben by Frank Wedekind&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;22:13&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Mit affektiertem Schmerz. (tr. With affected anguish.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Almansor by Heinrich Heine&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;22:43&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Sich etwas erhebend. (tr. Rising somewhat.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Aslauga by Friedrich de La Motte Fouqué&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;23:58&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Alle in größter Verwirrung ab. (tr. All exit in great confusion.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Eulenspiegel by Johann Nestroy&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;28:09&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Fällt rauschend ein. (tr. Collapses noisily.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Zar und Zimmermann by Albert Lortzing&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;29:04&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Buntes Gewühl. (tr. Pied turmoil.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Almansor by Heinrich Heine&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;29:52&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Burleskes Ballett. (tr. Burlesque ballet.)&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Almansor by Heinrich Heine&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;I then wanted to see which plays have more than one distinctive line in Frembdsch. They were:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th style=&quot;text-align: left&quot;&gt;Play&lt;/th&gt;
      &lt;th style=&quot;text-align: right&quot;&gt;Count&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Draußen vor der Tür by Wolfgang Borchert&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;6&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Eulenspiegel by Johann Nestroy&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;5&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;XYZ by Klabund&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;4&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Almansor by Heinrich Heine&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;4&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Zar und Zimmermann by Albert Lortzing&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;3&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Der Sohn by Walter Hasenclever&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;3&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Meister Oelze by Johannes Schlaf&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;It seems Borchert, Nestroy, Klabund, Heine, Lortzing, Hasenclever, and Schlaf, entered Frembdsch relatively unharmed. Note that only one of these plays (Lortzing’s Zar und Zimmermann) is present in the previous list of works in which stage directions from Frembdsch most often appear. Only two authors are represented in both lists: Lortzing and Nestroy. Nestroy also appeared three separate times in the previous list, suggesting a general stylistic affinity between his stage directions and those of Frembdsch.&lt;/p&gt;

&lt;p&gt;A closer look into Nestroy’s Eulenspiegel, the play with the second-highest number of distinctive lines in Frembdsch, reveals narrative borrowings that would otherwise not be captured by this strategy of exact matching. Reminiscent of Beckett’s Endgame, the narrative, as it is, of Frembdsch begins and ends with stage directions indicating a character exiting and then again entering a barrel. For example:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th style=&quot;text-align: right&quot;&gt;Time&lt;/th&gt;
      &lt;th style=&quot;text-align: left&quot;&gt;Stage direction&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;01:05&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;In der Ecke steht ein Fass. (tr. A barrel stands in the corner.)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;01:10&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Lauscht ein Kopf heraus. (tr. A head listens out. [sic!])&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;02:56&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Noch halb im Fass. (tr. Still half in the barrel.)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;03:28&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Aus dem Fass steigend. (tr. Climbing out of the barrel.)&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;Then near the end:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th style=&quot;text-align: right&quot;&gt;Time&lt;/th&gt;
      &lt;th style=&quot;text-align: left&quot;&gt;Stage direction&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;31:52&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Lauscht in das Fass. (tr. Listens into the barrel.)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;31:56&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Steigt ins Fass. (tr. Climbs into the barrel.)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;31:59&lt;/td&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Bückt sich, schlägt Deckel zu. (tr. Bends, closes the lid.)&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;The only exact match with the German Drama Corpus is “Noch halb im Fass.” This comes from Nestroy’s Eulenspiegel, a play with a great deal of business with people in barrels. A closer look into Eulenspiegel reveals two near misses: instead of “Steigt ins Fass,” Nestroy’s text has “Steigt ins Faß,” and instead of “aus dem Fass steigend,” Nestroy’s text has “aus dem Fasse steigend.” It seems that Nestroy’s work, and Eulenspiegel in particular, is disproportionately represented in Frembdsch.&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Jonah Lubin
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Vossian Antonomasia Extraction Using Pre-Trained Language Models</title>
    <link href="https://weltliteratur.net/vossian-antonomasia-extraction/"/>
    <updated>2022-11-23T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/vossian-antonomasia-extraction</id>
    <content type="html">&lt;p&gt;Hi there, let’s get back to one of our favorite topics: &lt;a href=&quot;https://vossanto.weltliteratur.net/&quot;&gt;Vossian
Antonomasia&lt;/a&gt; (VA).  In &lt;a href=&quot;https://weltliteratur.net/vossian-antonomasia-next-level/&quot;&gt;our 2019
EMNLP-IJCNLP
paper&lt;/a&gt;, we
tried to detect VA automatically using rule-based methods and a simple
neural network approach. The latter showed very promising results, so
we continued along this line and brought in some heavier machinery –
the Michael Jordans in the field of natural language processing (NLP):
pre-trained language models (PLMs).&lt;/p&gt;

&lt;p&gt;Neural networks and especially PLMs like
&lt;a href=&quot;https://aclanthology.org/N19-1423.pdf&quot;&gt;BERT&lt;/a&gt; have shown that they can
improve a wide range of NLP tasks, especially those for which large
labeled datasets are not available.&lt;/p&gt;

&lt;p&gt;The advantage of language models is the pre-training, which is
conducted with large amounts of unlabeled text data. For example, BERT
was trained on the English Wikipedia and
&lt;a href=&quot;https://arxiv.org/pdf/1506.06724.pdf&quot;&gt;BooksCorpus&lt;/a&gt;. In the
pre-training phase, the model learns a basic understanding of language
that can be used further for downstream tasks, for example, named
entity recognition (NER), sentiment analysis, or part-of-speech
tagging.&lt;/p&gt;

&lt;p&gt;We decided to try this on Vossian Antonomasia for &lt;a href=&quot;https://doi.org/10.3389/frai.2022.868249&quot;&gt;our latest paper
published in Frontiers in Artificial
Intelligence&lt;/a&gt;.  Instead of
classifying complete sentences, that is, deciding whether a sentence
contains a VA expression or not, we reformulated the task. Now, the
machine learning model is trained to identify all parts of a VA within
a sentence, that is, the source, target and modifier, and distinguish
them from one another. This is called &lt;em&gt;sequence tagging&lt;/em&gt;.&lt;/p&gt;

&lt;table&gt;
  &lt;tr&gt;
    &lt;th&gt;Words:&lt;/th&gt;
    &lt;td&gt;A&lt;/td&gt;
    &lt;td&gt;Spice&lt;/td&gt;
    &lt;td&gt;Girls&lt;/td&gt;
    &lt;td&gt;of&lt;/td&gt;
    &lt;td&gt;hip-hop&lt;/td&gt;
    &lt;td&gt;,&lt;/td&gt;
    &lt;td&gt;the&lt;/td&gt;
    &lt;td&gt;Wu-Tang&lt;/td&gt;
    &lt;td&gt;Clan&lt;/td&gt;
    &lt;td&gt;offers&lt;/td&gt;
    &lt;td&gt;something&lt;/td&gt;
    &lt;td&gt;for&lt;/td&gt;
    &lt;td&gt;every&lt;/td&gt;
    &lt;td&gt;kind&lt;/td&gt;
    &lt;td&gt;of&lt;/td&gt;
    &lt;td&gt;rap&lt;/td&gt;
    &lt;td&gt;fan&lt;/td&gt;
  &lt;/tr&gt;
  &lt;tr&gt;
    &lt;th&gt;Tags:&lt;/th&gt;
    &lt;td&gt;-&lt;/td&gt;
    &lt;td&gt;B-SRC&lt;/td&gt;
    &lt;td&gt;I-SRC&lt;/td&gt;
    &lt;td&gt;-&lt;/td&gt;
    &lt;td&gt;B-MOD&lt;/td&gt;
    &lt;td&gt;-&lt;/td&gt;
    &lt;td&gt;B-TRG&lt;/td&gt;
    &lt;td&gt;I-TRG&lt;/td&gt;
    &lt;td&gt;I-TRG&lt;/td&gt;
    &lt;td&gt;-&lt;/td&gt;
    &lt;td&gt;-&lt;/td&gt;
    &lt;td&gt;-&lt;/td&gt;
    &lt;td&gt;-&lt;/td&gt;
    &lt;td&gt;-&lt;/td&gt;
    &lt;td&gt;-&lt;/td&gt;
    &lt;td&gt;-&lt;/td&gt;
    &lt;td&gt;-&lt;/td&gt;
  &lt;/tr&gt;
&lt;/table&gt;

&lt;p&gt;For the training of neural networks, we &lt;a href=&quot;https://github.com/weltliteratur/vossanto/tree/master/frontiers&quot;&gt;annotated our VA
dataset&lt;/a&gt;
on the word-level, that is we marked for each word in 3,066 sentences
whether it is part of a source, target, or modifier or does not belong
to a VA at all.&lt;/p&gt;

&lt;p&gt;We then used the BERT base model and fine-tuned it with the annotated
data. In particular, we added an additional layer on top of the BERT
model that computes a tag for each word of the input sentence. Then
the parameters were re-computed based on the input data.&lt;/p&gt;

&lt;p&gt;In addition, we also trained a neural network model from scratch using
a concatenation of &lt;a href=&quot;https://allenai.org/allennlp/software/elmo&quot;&gt;ELMo&lt;/a&gt;
and &lt;a href=&quot;https://nlp.stanford.edu/projects/glove/&quot;&gt;GloVe&lt;/a&gt; embeddings and a
bidirectional &lt;a href=&quot;https://en.wikipedia.org/wiki/Long_short-term_memory&quot;&gt;long short-term
memory&lt;/a&gt; (BLSTM)
neural network with a &lt;a href=&quot;https://en.wikipedia.org/wiki/Conditional_random_field&quot;&gt;conditional random
field&lt;/a&gt; (CRF)
on top that also tags each word.&lt;/p&gt;

&lt;!--
| Task               | Approach    |   Precision |   Recall |      F1 |
| :----------------: | :---------: | :---------: | :------: | :-----: |
|                    | Baseline    |       0.876 |    0.880 |   0.878 |
| Classification     | BLSTM-ATT   |       0.921 |    0.074 |   0.947 |
|                    | BERT-CLF    |       0.971 |    0.977 |   0.974 |
|                    |             |             |          |         |
|                    | BASELINE    |       0.765 |    0.616 |   0.682 |
| Sequence-Tagging   | BLSTM-CRF   |       0.908 |    0.907 |   0.907 |
|                    | BERT-SEQ    |       0.908 |    0.944 |   0.926 |
--&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th align=&quot;center&quot;&gt;Task&lt;/th&gt;
      &lt;th align=&quot;center&quot;&gt;Approach&lt;/th&gt;
      &lt;th align=&quot;center&quot;&gt;Precision&lt;/th&gt;
      &lt;th align=&quot;center&quot;&gt;Recall&lt;/th&gt;
      &lt;th align=&quot;center&quot;&gt;F1&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;th align=&quot;left&quot; rowspan=&quot;3&quot;&gt;Classification&lt;/th&gt;
      &lt;td align=&quot;left&quot;&gt;Baseline&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.876&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.880&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.878&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td align=&quot;left&quot;&gt;BLSTM-ATT&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.921&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.074&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.947&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td align=&quot;left&quot;&gt;BERT-CLF&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.971&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.977&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.974&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;th align=&quot;left&quot; rowspan=&quot;3&quot;&gt;Sequence-Tagging&lt;/th&gt;
      &lt;td align=&quot;left&quot;&gt;BASELINE&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.765&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.616&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.682&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td align=&quot;left&quot;&gt;BLSTM-CRF&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.908&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.907&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.907&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td align=&quot;left&quot;&gt;BERT-SEQ&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.908&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.944&lt;/td&gt;
      &lt;td align=&quot;right&quot;&gt;0.926&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;First, we could improve the sentence classification task by almost 0.1
points in F1 score.  Also, we could achieve strong results (0.93 in F1
score) on the new sequence tagging task, where BERT outperforms the
BLSTM-CRF model.&lt;/p&gt;

&lt;p&gt;In addition to the evaluation on our annotated dataset, we conducted a
robustness study on real-world newspaper data. We also studied the
ability of the model to predict new types of VA focussing on new types
of source entities (e.g., organizations, locations, fictional
characters) and on new syntactic variations around the source (e.g.,
“a SOURCE on”, “of SOURCE of”). In total, the model identified around
10,000 VA candidates in the NYT corpus that our previous models were
not able to find. The following table shows the most frequently
predicted source candidates. Due to limited capacity, we only
evaluated samples of these candidates.&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th style=&quot;text-align: left&quot;&gt;Source Candidates&lt;/th&gt;
      &lt;th style=&quot;text-align: right&quot;&gt;Count&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;&lt;a href=&quot;https://www.wikidata.org/wiki/Q162808&quot;&gt;Holy Grail&lt;/a&gt;&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;116&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;&lt;a href=&quot;https://www.wikidata.org/wiki/Q27436&quot;&gt;Cadillac&lt;/a&gt;&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;88&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;&lt;a href=&quot;https://www.wikidata.org/wiki/Q106880435&quot;&gt;Pied Piper&lt;/a&gt;&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;85&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;&lt;a href=&quot;https://www.wikidata.org/wiki/Q243278&quot;&gt;RollsRoyce&lt;/a&gt;&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;71&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;&lt;a href=&quot;https://www.wikidata.org/wiki/Q90&quot;&gt;Paris&lt;/a&gt;&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;60&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;&lt;a href=&quot;https://www.wikidata.org/wiki/Q13371&quot;&gt;Harvard&lt;/a&gt;&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;58&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;&lt;a href=&quot;https://www.wikidata.org/wiki/Q2283&quot;&gt;Microsoft&lt;/a&gt;&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;43&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;&lt;a href=&quot;https://www.wikidata.org/wiki/Q641&quot;&gt;Venice&lt;/a&gt;&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;42&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;Demon Barber&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;39&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;&lt;a href=&quot;https://www.wikidata.org/wiki/Q116&quot;&gt;King&lt;/a&gt;&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;37&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;&lt;a href=&quot;https://www.wikidata.org/wiki/Q39&quot;&gt;Switzerland&lt;/a&gt;&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;37&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;&lt;a href=&quot;https://www.wikidata.org/wiki/Q38076&quot;&gt;McDonalds&lt;/a&gt;&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;35&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;&lt;a href=&quot;https://www.wikidata.org/wiki/Q12206942&quot;&gt;Darth Vader&lt;/a&gt;&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;34&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;&lt;a href=&quot;https://www.wikidata.org/wiki/Q14947899&quot;&gt;Wild West&lt;/a&gt;&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;33&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;&lt;a href=&quot;https://www.wikidata.org/wiki/Q11841&quot;&gt;Cinderella&lt;/a&gt;&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;32&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;&lt;a href=&quot;https://www.wikidata.org/wiki/Q192785&quot;&gt;Goliath&lt;/a&gt;&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;29&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td style=&quot;text-align: left&quot;&gt;&lt;a href=&quot;https://www.wikidata.org/wiki/Q164815&quot;&gt;Woodstock&lt;/a&gt;&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;29&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;In summary, in our newest paper, we developed new models for
extracting VAs on the word-level, that is, the models tag all words
that belong to a VA expression in a sentence.  In addition to the high
evaluation scores on our annotated dataset, we showed in multiple
robustness studies that the best model is able to predict new versions
of VAs regarding syntactic variations and also types of named
entities.&lt;/p&gt;

&lt;p&gt;The full annotation and deeper analytics of the predicted candidates
is one of our next projects. If you are interested in participation,
feel free to contact us. Stay tuned!&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Michel Schwab, 
	
          
          Robert Jäschke, 
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Automatic Extraction of Network Data From Amazon Prime Videos (Using "1917" As an Example)</title>
    <link href="https://weltliteratur.net/extracting-network-data-from-amazon-prime-videos/"/>
    <updated>2022-02-16T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/extracting-network-data-from-amazon-prime-videos</id>
    <content type="html">&lt;p&gt;Social networks extracted from movies or series are interesting research material. You can either download them (via pages like &lt;a href=&quot;https://moviegalaxies.com/&quot;&gt;Moviegalaxies&lt;/a&gt;) or create them yourself (e.g. using our tool &lt;a href=&quot;https://ezlinavis.dracor.org/&quot;&gt;ezlinavis&lt;/a&gt;). They are also well suited as a basis for tool-based introductory courses in network analysis.&lt;/p&gt;

&lt;p&gt;Often, however, we don’t know exactly how the data came about (for example, there is still the problem that &lt;a href=&quot;https://moviegalaxies.com/movies/view/512/the-lord-of-the-rings-the-fellowship-of-the-ring/&quot;&gt;Frodo appears twice&lt;/a&gt; in the network extracted from the first of the three “Lord of the Rings” movies on Moviegalaxies – how did that happen?). Another problem is that sometimes it just takes too long to extract your own network data in a scene-by-scene fashion using &lt;a href=&quot;https://ezlinavis.dracor.org/&quot;&gt;ezlinavis&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;So it might come in handy if we introduce another way to retrieve network data.&lt;/p&gt;

&lt;p&gt;If you happen to have an Amazon Prime Video account (or know someone who has 😊), you may have noticed the &lt;a href=&quot;https://www.amazon.com/primeinsider/video/pv-xray-tips.html&quot;&gt;&lt;strong&gt;X-Ray feature&lt;/strong&gt;&lt;/a&gt;. When you press the ⏸️ button, you see background information about the current production, including information about the characters/actors currently on screen.&lt;/p&gt;

&lt;p&gt;Turns out that this information is stored in a JSON file that you can download for each movie/series in a roundabout way. Once you have this file, you can do some analysis with it, like visualising the screentime of characters or extracting co-presence networks, which is really neat. (One advantage is that you can very easily compile datasets on your own favourite movies or series, so long as these productions are available on Prime Video.)&lt;/p&gt;

&lt;p&gt;To start with a visual argument, here is the result for Sam Mendes’ movie &lt;a href=&quot;https://en.wikipedia.org/wiki/1917_(2019_film)&quot;&gt;“1917”&lt;/a&gt;, followed by a description of how we did it:&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/1917-network.png&quot; alt=&quot;1917, movie, network&quot; style=&quot;width:1024px;&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;We are far from being the first to discover this possibility. In fact, there’s &lt;a href=&quot;https://www.curiousgnu.com/movie-character-screen-time&quot;&gt;a blog post from 2016 on this topic&lt;/a&gt; – and although things usually change all the time on the internet, the method described still works almost 6 years later.&lt;/p&gt;

&lt;p&gt;The first step is to &lt;a href=&quot;https://www.curiousgnu.com/movie-character-screen-time#how-to-get-x-ray-data&quot;&gt;download the JSON file&lt;/a&gt; using your browser’s built-in developer tools.&lt;/p&gt;

&lt;p&gt;The resulting JSON file is a bit convoluted, but it is relatively easy to find your way around it. Amazon’s X-Ray data divides the entire movie into 14 scenes (&lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;&quot;id&quot;:&quot;/xray/scene/1&quot;&lt;/code&gt; etc.):&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;Scene&lt;/th&gt;
      &lt;th&gt;Starts at&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;00:00:00&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;2&lt;/td&gt;
      &lt;td&gt;00:00:48&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;3&lt;/td&gt;
      &lt;td&gt;00:01:01&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;4&lt;/td&gt;
      &lt;td&gt;00:11:21&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;5&lt;/td&gt;
      &lt;td&gt;00:16:29&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;6&lt;/td&gt;
      &lt;td&gt;00:25:45&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;7&lt;/td&gt;
      &lt;td&gt;00:30:50&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;8&lt;/td&gt;
      &lt;td&gt;00:44:37&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;9&lt;/td&gt;
      &lt;td&gt;00:54:57&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;10&lt;/td&gt;
      &lt;td&gt;01:06:33&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;11&lt;/td&gt;
      &lt;td&gt;01:19:56&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;12&lt;/td&gt;
      &lt;td&gt;01:22:59&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;13&lt;/td&gt;
      &lt;td&gt;01:38:20&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;14&lt;/td&gt;
      &lt;td&gt;01:49:32&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;These scene divisions are, of course, contingent and can also lead to rather poor results for some movies. But for “1917”, a station drama, it works quite well.&lt;/p&gt;

&lt;p&gt;Let’s look at Scene 6, which features only two characters, Schofield and Blake:&lt;/p&gt;

&lt;div class=&quot;language-json highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;p&quot;&gt;{&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
  &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;changesCollection&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:[&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
    &lt;/span&gt;&lt;span class=&quot;p&quot;&gt;{&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
      &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;changeType&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;s2&quot;&gt;&quot;AddItem&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
      &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;itemId&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;s2&quot;&gt;&quot;/name/nm1126657/Lance Corporal Schofield&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
      &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;timePosition&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;mi&quot;&gt;1548000&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
    &lt;/span&gt;&lt;span class=&quot;p&quot;&gt;},&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
    &lt;/span&gt;&lt;span class=&quot;p&quot;&gt;{&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
      &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;changeType&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;s2&quot;&gt;&quot;AddItem&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
      &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;itemId&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;s2&quot;&gt;&quot;/name/nm2835616/Lance Corporal Blake&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
      &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;timePosition&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;mi&quot;&gt;1560000&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
    &lt;/span&gt;&lt;span class=&quot;p&quot;&gt;}&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
  &lt;/span&gt;&lt;span class=&quot;p&quot;&gt;],&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
  &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;initialItemIds&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:[],&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
  &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;timeRange&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:{&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
    &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;endTime&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;mi&quot;&gt;1850000&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
    &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;startTime&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;mi&quot;&gt;1545000&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
  &lt;/span&gt;&lt;span class=&quot;p&quot;&gt;}&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;}&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;p&gt;As you can see, in addition to the mere co-presence in a scene, there is also a temporal marker for the appearance of a character. But if you think this would make it easy to measure the overlap of the presence of two characters, alas, no. Because, unfortunately, there are only markers for their entrance, not their exit, so a character is always “present” until the end of a scene.&lt;/p&gt;

&lt;p&gt;To extract the co-presence network data, we wrote a little Python script (based on the example &lt;a href=&quot;https://www.curiousgnu.com/movie-character-screen-time&quot;&gt;here&lt;/a&gt;). We also included shared scenes as edge labels and fed all this into Gephi, the result of which you have seen above.&lt;/p&gt;

&lt;h2 id=&quot;conclusion&quot;&gt;Conclusion&lt;/h2&gt;

&lt;p&gt;We don’t know how Amazon Prime’s X-Ray data is encoded in practice, so we can’t assume that all JSON files were created in a similar process. But the fact that we can read this JSON data relatively easily, despite the lack of documentation, allows us to do some meaningful things with it.&lt;/p&gt;

&lt;p&gt;And there’s a lot more we could do with the X-Ray data. For example, the JSON also contains &lt;a href=&quot;https://www.imdb.com/&quot;&gt;IMDb&lt;/a&gt; IDs of actors (e.g. &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;&quot;id&quot;:&quot;/name/nm1126657/Lance Corporal Schofield&quot;&lt;/code&gt; → &lt;a href=&quot;https://www.imdb.com/name/nm1126657/&quot;&gt;imdb.com/name/nm1126657/&lt;/a&gt;). But for now, as Ed Wood used to say so enthusiastically, that’s a wrap!&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Lisa Poggel, 
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Quantification of Scholarly Articles on German Drama</title>
    <link href="https://weltliteratur.net/scholarly-articles-on-german-drama/"/>
    <updated>2021-10-12T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/scholarly-articles-on-german-drama</id>
    <content type="html">&lt;h1 id=&quot;context&quot;&gt;Context&lt;/h1&gt;

&lt;p&gt;In our project &lt;a href=&quot;https://www.projekte.hu-berlin.de/en/schluesselstellen/&quot;&gt;“What matters? Key passages in literary
works”&lt;/a&gt; we
analyse how scholarly articles cite literary works and whether these
citations can be used to identify key passages. We started with a
corpus of 100 scholarly articles dealing with the interpretion of one
of two novellas:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://en.wikipedia.org/wiki/Michael_Kohlhaas&quot;&gt;&lt;em&gt;Michael Kohlhaas&lt;/em&gt;&lt;/a&gt; by Heinrich von Kleist (1808)&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://en.wikipedia.org/wiki/Die_Judenbuche&quot;&gt;&lt;em&gt;Die Judenbuche&lt;/em&gt;&lt;/a&gt; by Annette von Droste-Hülshoff (1842)&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;In the next phase of our project, we want to focus on German-language
drama and are therefore building a corpus of scholarly articles dealing
with German plays. For our distant-reading approach,
this corpus should be substantial, that is, for each play we would like
to have a number of scholarly articles dealing with it. So one of our
questions for operationalisation was: “Which plays should we pick so
that we find a reasonable number of scholarly works per play?” Or
more generally: “Which German plays have been interpreted most
frequently by researchers?”&lt;/p&gt;

&lt;h1 id=&quot;data-sources&quot;&gt;Data Sources&lt;/h1&gt;

&lt;p&gt;To answer this question, we utilise two data sources:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://dracor.org/ger&quot;&gt;GerDraCor&lt;/a&gt;, the German Drama Corpus, which
is part of the larger &lt;a href=&quot;https://dracor.org/&quot;&gt;DraCor&lt;/a&gt; project and
currently comprises 545 German plays from the period
1657&lt;a href=&quot;https://dracor.org/id/ger000538&quot;&gt;¹&lt;/a&gt; to
1947&lt;a href=&quot;https://dracor.org/id/ger000476&quot;&gt;²&lt;/a&gt;. DraCor provides the full
texts of all plays (which is crucial for our project) as well as
detailed metadata such as title, author and year of publication.&lt;/li&gt;
  &lt;li&gt;The &lt;a href=&quot;http://www.bdsl-online.de/&quot;&gt;BDSL online catalogue&lt;/a&gt;, short for
&lt;a href=&quot;https://www.ub.uni-frankfurt.de/bdsl/&quot;&gt;Bibliographie der deutschen Sprach- und Literaturwissenschaft&lt;/a&gt;
(Bibliography of German Linguistics and Literature), a comprehensive
bibliography of more than 300,000 scholarly works on German language
and literature published between 1985 and 2010. It offers broad
search options, in particular it is possible to search for articles
that refer to a specific work (by title) or an author (by name).&lt;/li&gt;
&lt;/ul&gt;

&lt;h1 id=&quot;implementation&quot;&gt;Implementation&lt;/h1&gt;

&lt;p&gt;&lt;strong&gt;1.&lt;/strong&gt; We downloaded a GerDraCor snapshot on August 31, 2021, by cloning
its git repository (&lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;git clone git@github.com:dracor-org/gerdracor.git&lt;/code&gt;).&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;2.&lt;/strong&gt; Using &lt;a href=&quot;https://scm.cms.hu-berlin.de/jaeschkr/interpretatorisch/-/blob/master/bdsl/crawltitles.py&quot;&gt;a small Python script&lt;/a&gt;, we extracted titles
(XPath: &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;tei:teiHeader/tei:fileDesc/tei:titleStmt/tei:title&lt;/code&gt;)
authors (XPath:
&lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;tei:teiHeader/tei:fileDesc/tei:titleStmt/tei:author&lt;/code&gt;) of each play.
The resulting TSV file contains 545 plays. The name of authors is stored
in the format “Surname, Forename”, titles of plays separated from it by a tab character.&lt;/p&gt;

&lt;p&gt;Here the first entries of the resulting list:&lt;/p&gt;
&lt;div class=&quot;language-plaintext highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;Alberti, Konrad   Brot!
Alberti, Konrad   Im Suff
André, Johann   Der Comödienfeind
Anzengruber, Ludwig   Das vierte Gebot
Anzengruber, Ludwig   Der Gwissenswurm
[...]
&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;p&gt;&lt;strong&gt;3.&lt;/strong&gt; Looping over the list of plays, we queried BDSL on the same
day, August 31, 2021. It is possible to
search for “Behandeltes Werk” (work treated), which allows us to
restrict the search to scholarly articles whose subject is a specific
work (using the title of plays). As some titles are ambiguous,
initial tests have shown that we need to further restrict the
search to “Behandelte Person” (person treated), which names the
author of the work.&lt;/p&gt;

&lt;p&gt;⚠️ There are a few limitations to our approach: We do limit our search
to certain fields, but ultimately rely on the string matching
of BDSL’s search engine, so we miss some hits. If, for example,
we search for Wagner’s “Tannhäuser” using its short rather than
its full title (“Tannhäuser und Der Sängerkrieg auf Wartburg”),
we would arrive at 16 hits instead of just one. BDSL also employs
keywords for works, using the keyword “Wagner, Richard / Tannhäuser und der
Sängerkrieg auf Wartburg” we even get 44 hits. All in all, working with
the BDSL interface is a bit opaque, interoperable
IDs for both authors and works (aligned with the
&lt;a href=&quot;https://en.wikipedia.org/wiki/Integrated_Authority_File&quot;&gt;Gemeinsame Normdatei GND&lt;/a&gt;)
would be a crucial improvement here.&lt;/p&gt;

&lt;h1 id=&quot;results-per-play&quot;&gt;Results per Play&lt;/h1&gt;

&lt;p&gt;For the following 172 plays we were able to find an article
(that is, for 373 we could not):&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;Play&lt;/th&gt;
      &lt;th&gt;Author&lt;/th&gt;
      &lt;th style=&quot;text-align: right&quot;&gt;Articles&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;Faust&lt;/td&gt;
      &lt;td&gt;Johann Wolfgang Goethe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2037&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Dantons Tod&lt;/td&gt;
      &lt;td&gt;Georg Büchner&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;309&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Nathan der Weise&lt;/td&gt;
      &lt;td&gt;Gotthold Ephraim Lessing&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;307&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Penthesilea&lt;/td&gt;
      &lt;td&gt;Heinrich von Kleist&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;304&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Iphigenie auf Tauris&lt;/td&gt;
      &lt;td&gt;Johann Wolfgang Goethe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;243&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Emilia Galotti&lt;/td&gt;
      &lt;td&gt;Gotthold Ephraim Lessing&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;193&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Prinz Friedrich von Homburg&lt;/td&gt;
      &lt;td&gt;Heinrich von Kleist&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;172&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Räuber&lt;/td&gt;
      &lt;td&gt;Friedrich Schiller&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;171&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Wilhelm Tell&lt;/td&gt;
      &lt;td&gt;Friedrich Schiller&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;147&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der zerbrochne Krug&lt;/td&gt;
      &lt;td&gt;Heinrich von Kleist&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;143&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Maria Stuart&lt;/td&gt;
      &lt;td&gt;Friedrich Schiller&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;123&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Leonce und Lena&lt;/td&gt;
      &lt;td&gt;Georg Büchner&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;116&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Jungfrau von Orleans&lt;/td&gt;
      &lt;td&gt;Friedrich Schiller&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;116&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Hermannsschlacht&lt;/td&gt;
      &lt;td&gt;Heinrich von Kleist&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;114&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Elektra&lt;/td&gt;
      &lt;td&gt;Hugo von Hofmannsthal&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;111&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Torquato Tasso&lt;/td&gt;
      &lt;td&gt;Johann Wolfgang Goethe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;109&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Kabale und Liebe&lt;/td&gt;
      &lt;td&gt;Friedrich Schiller&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;102&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die letzten Tage der Menschheit&lt;/td&gt;
      &lt;td&gt;Karl Kraus&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;100&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Amphitryon&lt;/td&gt;
      &lt;td&gt;Heinrich von Kleist&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;97&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Reigen&lt;/td&gt;
      &lt;td&gt;Arthur Schnitzler&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;85&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Schwierige&lt;/td&gt;
      &lt;td&gt;Hugo von Hofmannsthal&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;77&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Zauberflöte&lt;/td&gt;
      &lt;td&gt;Emanuel Schikaneder&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;71&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Miß Sara Sampson&lt;/td&gt;
      &lt;td&gt;Gotthold Ephraim Lessing&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;70&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Egmont&lt;/td&gt;
      &lt;td&gt;Johann Wolfgang Goethe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;66&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Familie Schroffenstein&lt;/td&gt;
      &lt;td&gt;Heinrich von Kleist&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;57&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Soldaten&lt;/td&gt;
      &lt;td&gt;Jakob Michael Reinhold Lenz&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;50&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Rosenkavalier&lt;/td&gt;
      &lt;td&gt;Hugo von Hofmannsthal&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;47&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Frühlings Erwachen&lt;/td&gt;
      &lt;td&gt;Frank Wedekind&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;47&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Draußen vor der Tür&lt;/td&gt;
      &lt;td&gt;Wolfgang Borchert&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;45&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die natürliche Tochter&lt;/td&gt;
      &lt;td&gt;Johann Wolfgang Goethe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;45&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Ariadne auf Naxos&lt;/td&gt;
      &lt;td&gt;Hugo von Hofmannsthal&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;44&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Judith&lt;/td&gt;
      &lt;td&gt;Friedrich Hebbel&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;43&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Maria Magdalene&lt;/td&gt;
      &lt;td&gt;Friedrich Hebbel&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;43&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;König Ottokars Glück und Ende&lt;/td&gt;
      &lt;td&gt;Franz Grillparzer&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;42&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Verschwörung des Fiesco zu Genua&lt;/td&gt;
      &lt;td&gt;Friedrich Schiller&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;41&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Jüdin von Toledo&lt;/td&gt;
      &lt;td&gt;Franz Grillparzer&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;40&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Turm&lt;/td&gt;
      &lt;td&gt;Hugo von Hofmannsthal&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;40&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Philotas&lt;/td&gt;
      &lt;td&gt;Gotthold Ephraim Lessing&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;39&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Libussa&lt;/td&gt;
      &lt;td&gt;Franz Grillparzer&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;37&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Jedermann&lt;/td&gt;
      &lt;td&gt;Hugo von Hofmannsthal&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;36&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Stella&lt;/td&gt;
      &lt;td&gt;Johann Wolfgang Goethe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;35&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Frau ohne Schatten&lt;/td&gt;
      &lt;td&gt;Hugo von Hofmannsthal&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;34&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Geschichten aus dem Wiener Wald&lt;/td&gt;
      &lt;td&gt;Ödön von Horváth&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;34&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Vor Sonnenaufgang&lt;/td&gt;
      &lt;td&gt;Gerhart Hauptmann&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;31&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Büchse der Pandora&lt;/td&gt;
      &lt;td&gt;Frank Wedekind&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;31&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Juden&lt;/td&gt;
      &lt;td&gt;Gotthold Ephraim Lessing&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;30&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Ein Bruderzwist in Habsburg&lt;/td&gt;
      &lt;td&gt;Franz Grillparzer&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;29&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Hermannsschlacht&lt;/td&gt;
      &lt;td&gt;Christian Dietrich Grabbe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;28&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Herzog Theodor von Gothland&lt;/td&gt;
      &lt;td&gt;Christian Dietrich Grabbe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;28&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Schwärmer&lt;/td&gt;
      &lt;td&gt;Robert Musil&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;27&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Anatol&lt;/td&gt;
      &lt;td&gt;Arthur Schnitzler&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;26&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der grüne Kakadu&lt;/td&gt;
      &lt;td&gt;Arthur Schnitzler&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;26&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Ratten&lt;/td&gt;
      &lt;td&gt;Gerhart Hauptmann&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;25&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Agnes Bernauer&lt;/td&gt;
      &lt;td&gt;Friedrich Hebbel&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;25&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Liebelei&lt;/td&gt;
      &lt;td&gt;Arthur Schnitzler&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;23&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der gestiefelte Kater&lt;/td&gt;
      &lt;td&gt;Ludwig Tieck&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;23&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Don Juan und Faust&lt;/td&gt;
      &lt;td&gt;Christian Dietrich Grabbe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;21&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Arabella&lt;/td&gt;
      &lt;td&gt;Hugo von Hofmannsthal&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;21&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Unbestechliche&lt;/td&gt;
      &lt;td&gt;Hugo von Hofmannsthal&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;21&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Kasimir und Karoline&lt;/td&gt;
      &lt;td&gt;Ödön von Horváth&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;21&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Professor Bernhardi&lt;/td&gt;
      &lt;td&gt;Arthur Schnitzler&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;21&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Clavigo&lt;/td&gt;
      &lt;td&gt;Johann Wolfgang Goethe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;20&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Proserpina&lt;/td&gt;
      &lt;td&gt;Johann Wolfgang Goethe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;20&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Gyges und sein Ring&lt;/td&gt;
      &lt;td&gt;Friedrich Hebbel&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;20&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Marquis von Keith&lt;/td&gt;
      &lt;td&gt;Frank Wedekind&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;20&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Scherz, Satire, Ironie und tiefere Bedeutung&lt;/td&gt;
      &lt;td&gt;Christian Dietrich Grabbe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;19&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Traum ein Leben&lt;/td&gt;
      &lt;td&gt;Franz Grillparzer&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;19&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Weh dem, der lügt!&lt;/td&gt;
      &lt;td&gt;Franz Grillparzer&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;19&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Cardenio und Celinde oder Unglücklich Verliebete&lt;/td&gt;
      &lt;td&gt;Andreas Gryphius&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;19&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Hermanns Schlacht&lt;/td&gt;
      &lt;td&gt;Friedrich Gottlieb Klopstock&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;18&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Kindermörderin&lt;/td&gt;
      &lt;td&gt;Heinrich Leopold Wagner&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;18&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Tod des Tizian&lt;/td&gt;
      &lt;td&gt;Hugo von Hofmannsthal&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;17&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Talisman&lt;/td&gt;
      &lt;td&gt;Johann Nestroy&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;17&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Das weite Land&lt;/td&gt;
      &lt;td&gt;Arthur Schnitzler&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;17&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Des Meeres und der Liebe Wellen&lt;/td&gt;
      &lt;td&gt;Franz Grillparzer&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;15&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Herodes und Mariamne&lt;/td&gt;
      &lt;td&gt;Friedrich Hebbel&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;15&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Wupper&lt;/td&gt;
      &lt;td&gt;Else Lasker-Schüler&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;15&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Gründung Prags&lt;/td&gt;
      &lt;td&gt;Clemens Brentano&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;14&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Ödipus und die Sphinx&lt;/td&gt;
      &lt;td&gt;Hugo von Hofmannsthal&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;14&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Das Liebeskonzil&lt;/td&gt;
      &lt;td&gt;Oskar Panizza&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;14&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Ahnfrau&lt;/td&gt;
      &lt;td&gt;Franz Grillparzer&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;13&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Das Salzburger große Welttheater&lt;/td&gt;
      &lt;td&gt;Hugo von Hofmannsthal&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;13&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Zwillinge&lt;/td&gt;
      &lt;td&gt;Friedrich Maximilian Klinger&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;13&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Zerrissene&lt;/td&gt;
      &lt;td&gt;Johann Nestroy&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;13&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Triumph der Empfindsamkeit&lt;/td&gt;
      &lt;td&gt;Johann Wolfgang Goethe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Hannibal&lt;/td&gt;
      &lt;td&gt;Christian Dietrich Grabbe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Medea&lt;/td&gt;
      &lt;td&gt;Franz Grillparzer&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Judith und Holofernes&lt;/td&gt;
      &lt;td&gt;Johann Nestroy&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Erdgeist&lt;/td&gt;
      &lt;td&gt;Frank Wedekind&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Ponce de Leon&lt;/td&gt;
      &lt;td&gt;Clemens Brentano&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;11&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Ein treuer Diener seines Herrn&lt;/td&gt;
      &lt;td&gt;Franz Grillparzer&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;11&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Sohn&lt;/td&gt;
      &lt;td&gt;Walter Hasenclever&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;11&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Genoveva&lt;/td&gt;
      &lt;td&gt;Friedrich Hebbel&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;11&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Julius von Tarent&lt;/td&gt;
      &lt;td&gt;Johann Anton Leisewitz&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;11&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der einsame Weg&lt;/td&gt;
      &lt;td&gt;Arthur Schnitzler&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;11&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Journalisten&lt;/td&gt;
      &lt;td&gt;Gustav Freytag&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;10&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Sappho&lt;/td&gt;
      &lt;td&gt;Franz Grillparzer&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;10&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Blunt oder der Gast&lt;/td&gt;
      &lt;td&gt;Karl Philipp Moritz&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;10&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Alceste&lt;/td&gt;
      &lt;td&gt;Christoph Martin Wieland&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;10&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der letzte Held von Marienburg&lt;/td&gt;
      &lt;td&gt;Joseph von Eichendorff&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;9&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Almansor&lt;/td&gt;
      &lt;td&gt;Heinrich Heine&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;9&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Familie Selicke&lt;/td&gt;
      &lt;td&gt;Arno Holz&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;9&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Zriny&lt;/td&gt;
      &lt;td&gt;Theodor Körner&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;9&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Komödie der Verführung&lt;/td&gt;
      &lt;td&gt;Arthur Schnitzler&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;9&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Leben und Tod der heiligen Genoveva&lt;/td&gt;
      &lt;td&gt;Ludwig Tieck&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;9&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Florian Geyer&lt;/td&gt;
      &lt;td&gt;Gerhart Hauptmann&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;8&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Sturm und Drang&lt;/td&gt;
      &lt;td&gt;Friedrich Maximilian Klinger&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;8&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der junge Gelehrte&lt;/td&gt;
      &lt;td&gt;Gotthold Ephraim Lessing&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;8&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Einen Jux will er sich machen&lt;/td&gt;
      &lt;td&gt;Johann Nestroy&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;8&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Datterich&lt;/td&gt;
      &lt;td&gt;Ernst Elias Niebergall&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;8&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Alpenkönig und der Menschenfeind&lt;/td&gt;
      &lt;td&gt;Ferdinand Raimund&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;8&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Wallensteins Lager&lt;/td&gt;
      &lt;td&gt;Friedrich Schiller&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;8&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Alarcos&lt;/td&gt;
      &lt;td&gt;Friedrich Schlegel&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;8&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die verkehrte Welt&lt;/td&gt;
      &lt;td&gt;Ludwig Tieck&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;8&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Franziska&lt;/td&gt;
      &lt;td&gt;Frank Wedekind&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;8&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Ugolino&lt;/td&gt;
      &lt;td&gt;Heinrich Wilhelm von Gerstenberg&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;7&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Engländer&lt;/td&gt;
      &lt;td&gt;Jakob Michael Reinhold Lenz&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;7&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Wallensteins Tod&lt;/td&gt;
      &lt;td&gt;Friedrich Schiller&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;7&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Canut&lt;/td&gt;
      &lt;td&gt;Johann Elias Schlegel&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;7&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Halle&lt;/td&gt;
      &lt;td&gt;Achim von Arnim&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;6&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Jerusalem&lt;/td&gt;
      &lt;td&gt;Achim von Arnim&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;6&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Freier&lt;/td&gt;
      &lt;td&gt;Joseph von Eichendorff&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;6&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Des Epimenides Erwachen&lt;/td&gt;
      &lt;td&gt;Johann Wolfgang Goethe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;6&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Demetrius&lt;/td&gt;
      &lt;td&gt;Friedrich Hebbel&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;6&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Freiheit in Krähwinkel&lt;/td&gt;
      &lt;td&gt;Johann Nestroy&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;6&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Moral&lt;/td&gt;
      &lt;td&gt;Ludwig Thoma&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;6&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Faust&lt;/td&gt;
      &lt;td&gt;Friedrich Theodor Vischer&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;6&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Perdu! oder Dichter, Verleger und Blaustrümpfe&lt;/td&gt;
      &lt;td&gt;Annette von Droste-Hülshoff&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;5&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Erwin und Elmire&lt;/td&gt;
      &lt;td&gt;Johann Wolfgang Goethe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;5&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Babel und Bibel&lt;/td&gt;
      &lt;td&gt;Karl May&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;5&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Magdalena&lt;/td&gt;
      &lt;td&gt;Ludwig Thoma&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;5&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Laune des Verliebten&lt;/td&gt;
      &lt;td&gt;Johann Wolfgang Goethe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;4&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Mitschuldigen&lt;/td&gt;
      &lt;td&gt;Johann Wolfgang Goethe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;4&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Satyros oder der vergötterte Waldteufel&lt;/td&gt;
      &lt;td&gt;Johann Wolfgang Goethe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;4&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Tod Adams&lt;/td&gt;
      &lt;td&gt;Friedrich Gottlieb Klopstock&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;4&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Diamant des Geisterkönigs&lt;/td&gt;
      &lt;td&gt;Ferdinand Raimund&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;4&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Verschwender&lt;/td&gt;
      &lt;td&gt;Ferdinand Raimund&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;4&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Kammersänger&lt;/td&gt;
      &lt;td&gt;Frank Wedekind&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;4&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Hidalla oder Sein und Haben&lt;/td&gt;
      &lt;td&gt;Frank Wedekind&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;4&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Codrus&lt;/td&gt;
      &lt;td&gt;Johann Friedrich von Cronegk&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;3&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Merlin&lt;/td&gt;
      &lt;td&gt;Karl Immermann&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;3&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Tristan und Isolde&lt;/td&gt;
      &lt;td&gt;Richard Wagner&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;3&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Zensur&lt;/td&gt;
      &lt;td&gt;Frank Wedekind&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;3&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Kaiser Friedrich Barbarossa&lt;/td&gt;
      &lt;td&gt;Christian Dietrich Grabbe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Uriel Acosta&lt;/td&gt;
      &lt;td&gt;Karl Gutzkow&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Gabriel Schillings Flucht&lt;/td&gt;
      &lt;td&gt;Gerhart Hauptmann&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Colberg&lt;/td&gt;
      &lt;td&gt;Paul Heyse&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Alkestis&lt;/td&gt;
      &lt;td&gt;Hugo von Hofmannsthal&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Ein Trauerspiel in Berlin&lt;/td&gt;
      &lt;td&gt;Karl von Holtei&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Koralle&lt;/td&gt;
      &lt;td&gt;Georg Kaiser&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Kreidekreis&lt;/td&gt;
      &lt;td&gt;Klabund&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Erbförster&lt;/td&gt;
      &lt;td&gt;Otto Ludwig&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Das Haus der Temperamente&lt;/td&gt;
      &lt;td&gt;Johann Nestroy&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Barometermacher auf der Zauberinsel&lt;/td&gt;
      &lt;td&gt;Ferdinand Raimund&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Piccolomini&lt;/td&gt;
      &lt;td&gt;Friedrich Schiller&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Bettler&lt;/td&gt;
      &lt;td&gt;Reinhard Sorge&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Heimat&lt;/td&gt;
      &lt;td&gt;Hermann Sudermann&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Musik&lt;/td&gt;
      &lt;td&gt;Frank Wedekind&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Johann Faust&lt;/td&gt;
      &lt;td&gt;Paul Weidmann&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Nase des Michelangelo&lt;/td&gt;
      &lt;td&gt;Hugo Ball&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der blaue Boll&lt;/td&gt;
      &lt;td&gt;Ernst Barlach&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Narziß&lt;/td&gt;
      &lt;td&gt;Albert Emil Brachvogel&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Brutus&lt;/td&gt;
      &lt;td&gt;Joachim Wilhelm von Brawe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Frau im Fenster&lt;/td&gt;
      &lt;td&gt;Hugo von Hofmannsthal&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Simsone Grisaldo&lt;/td&gt;
      &lt;td&gt;Friedrich Maximilian Klinger&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Franz von Sickingen&lt;/td&gt;
      &lt;td&gt;Ferdinand Lassalle&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Die Makkabäer&lt;/td&gt;
      &lt;td&gt;Otto Ludwig&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Golo und Genovefa&lt;/td&gt;
      &lt;td&gt;Maler Müller&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Der Weibsteufel&lt;/td&gt;
      &lt;td&gt;Karl Schönherr&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Ernst Herzog von Schwaben&lt;/td&gt;
      &lt;td&gt;Ludwig Uhland&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Tannhäuser und Der Sängerkrieg auf Wartburg&lt;/td&gt;
      &lt;td&gt;Richard Wagner&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Faust&lt;/td&gt;
      &lt;td&gt;Hermann Ludwig Wolfram&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;Although it is no surprise to find Goethe’s “Faust” with all its variants
in first place, it is surprising that the second-placed play, Büchner’s
“Danton’s Death”, lags behind by an order of magnitude in terms of the
articles in which the play is discussed.&lt;/p&gt;

&lt;h1 id=&quot;results-per-author&quot;&gt;Results per Author&lt;/h1&gt;

&lt;p&gt;To see which author was written about most often, we can group and sort the
result by author:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;Author&lt;/th&gt;
      &lt;th style=&quot;text-align: right&quot;&gt;Plays&lt;/th&gt;
      &lt;th style=&quot;text-align: right&quot;&gt;Articles&lt;/th&gt;
      &lt;th style=&quot;text-align: right&quot;&gt;Mean Articles per Play&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;Johann Wolfgang Goethe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;22&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2610&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;118&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Heinrich von Kleist&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;7&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;887&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;126&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Friedrich Schiller&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;11&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;717&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;65&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Gotthold Ephraim Lessing&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;12&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;647&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;53&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Hugo von Hofmannsthal&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;17&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;478&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;28&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Georg Büchner&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;425&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;212&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Franz Grillparzer&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;13&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;247&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;19&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Arthur Schnitzler&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;13&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;218&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;16&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Friedrich Hebbel&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;13&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;163&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Frank Wedekind&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;12&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;131&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;10&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;We can also sort this table by the average number of articles per play
per author, although the result should be treated with caution for several
reasons:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;Author&lt;/th&gt;
      &lt;th style=&quot;text-align: right&quot;&gt;Plays&lt;/th&gt;
      &lt;th style=&quot;text-align: right&quot;&gt;Articles&lt;/th&gt;
      &lt;th style=&quot;text-align: right&quot;&gt;Mean Articles per Play&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;Georg Büchner&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;425&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;212&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Heinrich von Kleist&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;7&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;887&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;126&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Johann Wolfgang Goethe&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;22&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2610&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;118&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Emanuel Schikaneder&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;71&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;71&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Friedrich Schiller&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;11&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;717&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;65&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Gotthold Ephraim Lessing&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;12&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;647&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;53&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Karl Kraus&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;2&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;100&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;50&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Wolfgang Borchert&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;45&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;45&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Hugo von Hofmannsthal&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;17&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;478&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;28&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Robert Musil&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;1&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;27&lt;/td&gt;
      &lt;td style=&quot;text-align: right&quot;&gt;27&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;On the one hand, authors who have become known primarily for just one or a few
plays appear at the top (Karl Kraus, Wolfgang Borchert, Georg Büchner). However,
we only looked for plays that are already represented in GerDraCor. In addition
to Büchner’s two plays “Leonce and Lena” and “Danton’s Death”, his famous
fragment “Woyzeck” is missing, which limits the usefulness of the above table.
All the results presented only make statements about the plays and authors
currently contained in GerDraCor.&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Robert Jäschke, 
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>DraCor Summer Update (August 2021)</title>
    <link href="https://weltliteratur.net/dracor-summer-update-2021/"/>
    <updated>2021-08-05T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/dracor-summer-update-2021</id>
    <content type="html">&lt;p&gt;&lt;img src=&quot;/images/dracor/dracor-2021-mobile-mockup-05.jpg&quot; alt=&quot;DraCor 2021&quot; /&gt;&lt;/p&gt;

&lt;p&gt;DraCor, the &lt;strong&gt;Dra&lt;/strong&gt;ma &lt;strong&gt;Cor&lt;/strong&gt;pora platform, can be accessed at &lt;a href=&quot;https://dracor.org/&quot;&gt;dracor.org&lt;/a&gt;. Our aim is to provide an infrastructure component for computational literary studies in order to facilitate access to relevant research data and to ensure compliance with the &lt;a href=&quot;https://www.go-fair.org/fair-principles/&quot;&gt;FAIR Principles&lt;/a&gt; when dealing with literary corpora. In our case, the focus is of course on drama and theatre, but we plan to extend our concept of &lt;a href=&quot;https://doi.org/10.5281/zenodo.4284002&quot;&gt;Programmable Corpora&lt;/a&gt; to novels as part of the H2020-funded &lt;a href=&quot;https://cordis.europa.eu/project/id/101004984&quot;&gt;CLS INFRA&lt;/a&gt; project that started earlier this year.&lt;/p&gt;

&lt;p&gt;Today, we are pleased to introduce the next iteration of our DraCor website, along with new API functions that make it much easier to access our data for research. Our front-end, i.e. web design, is now at &lt;a href=&quot;https://github.com/dracor-org/dracor-frontend/releases/&quot;&gt;version 1.1.0&lt;/a&gt;, the DraCor-API at &lt;a href=&quot;https://github.com/dracor-org/dracor-api/releases/&quot;&gt;version 0.82.0&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;We have worked on all levels of the website, the homepage, the corpus overviews and pages of individual plays:&lt;/p&gt;

&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/dracor/dracor-2021-mobile-mockup-01.jpg&quot; alt=&quot;&quot; style=&quot;width:49%;&quot; /&gt;
  &lt;img src=&quot;/images/dracor/dracor-2021-mobile-mockup-03.jpg&quot; alt=&quot;&quot; style=&quot;width:49%;&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;As before, the website should work well on all types of devices, large screens, laptops, tablets and mobile phones. There are many small details to discover in the new design (our favourites are the author images that we automatically retrieve from Wikimedia Commons via Wikidata QIDs, just one small example of the benefits of our highly interoperable approach).&lt;/p&gt;

&lt;p&gt;The overall design is even better aligned with the TEI encoding on which all our corpora are based to increase the visibility of our rich annotations.&lt;/p&gt;

&lt;p&gt;Here’s a non-exhaustive summary of the new functions added since December 2020:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;infoboxes are now context-sensitive (their content differs depending on which tab is activated: “Network”, “Relations”, “Speech distribution”, “Full text”, or “Downloads”)&lt;/li&gt;
  &lt;li&gt;more explanations on the spot, e.g. the “Network” tab is accompanied by a short description of what the network graph actually shows&lt;/li&gt;
  &lt;li&gt;we also introduced an ℹ️ symbol that leads to longer explanations in our newly introduced (and still somewhat empty) FAQ section&lt;/li&gt;
  &lt;li&gt;author information and images sourced live from Wikimedia Commons via Wikidata identifiers&lt;/li&gt;
  &lt;li&gt;improved readability of network graphs (annotated information on gender is now integrated, colour contrast has been increased)&lt;/li&gt;
  &lt;li&gt;plays with historical or mythological characters now display a Wikidata QID if available, a new feature provided by our API (not all corpora are yet encoded with this information, but once they are, it will be much easier to compare all plays with e.g. a Faust character or to do a contrastive analysis of the vocabulary of mythological and historical characters in different authors and even between corpora in different languages)&lt;/li&gt;
  &lt;li&gt;the “Full text” tab now shows the print source in the context box, along with an overview of speaking characters per segment&lt;/li&gt;
  &lt;li&gt;corpus metadata is now much more accessible, you can download it directly from the corpus overviews (try the CSV version to create some quick charts in MS Excel or LibreOffice Calc)&lt;/li&gt;
  &lt;li&gt;added a button to copy DraCor IDs to the clipboard&lt;/li&gt;
  &lt;li&gt;new corpus added: &lt;a href=&quot;https://dracor.org/bash&quot;&gt;BashDraCor&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;new plays were added to our in-house corpora, &lt;a href=&quot;https://dracor.org/ger&quot;&gt;GerDraCor&lt;/a&gt; and &lt;a href=&quot;https://dracor.org/rus&quot;&gt;RusDraCor&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://twitter.com/Migabaj&quot;&gt;Michael Sonkin&lt;/a&gt; added numerous character annotations to &lt;a href=&quot;https://dracor.org/ita&quot;&gt;ItaDraCor&lt;/a&gt;, significantly improving the corpus which we originally inherited from Biblioteca Italiana&lt;/li&gt;
  &lt;li&gt;introduction of TypeScript (to gradually improve the quality of our codebase)&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Ok, we’ll stop here. As always the complete history of changes can be traced on GitHub.&lt;/p&gt;

&lt;h1 id=&quot;how-did-we-get-here&quot;&gt;How Did We Get Here?&lt;/h1&gt;

&lt;p&gt;Our &lt;a href=&quot;https://reactjs.org/&quot;&gt;React&lt;/a&gt;-powered website and the DraCor API have been up and running since 2017. We started with a German and a Russian drama corpus, both encoded in TEI, but later added many more corpora, a handful of which were provided by the community. At the moment we host 12 active drama corpora, and number 13, the rather large French Drama Corpus (more than 1,500 plays so far), is just around the corner. Many colleagues have approached us asking if they can add their corpus in the future, and we very much welcome their contribution. (If you have or want to create or maintain a TEI-encoded drama corpus and add it to DraCor, please drop us a line.)&lt;/p&gt;

&lt;p&gt;Looking back, we find it hard to believe that we didn’t initially intend to run a fancy website at all, our focus was on functionality. It wasn’t until we had developed the concept of &lt;a href=&quot;https://doi.org/10.5281/zenodo.4284002&quot;&gt;Programmable Corpora&lt;/a&gt; and the website was functional and running stably that we eventually decided we wanted a fancy website too. &lt;a href=&quot;https://twitter.com/umblaetterer/status/1334080571979542528&quot;&gt;So in December 2020&lt;/a&gt;, we released version 1.0.0 of the DraCor front-end, and since then the project has received much more attention. In the past it wasn’t a big problem when DraCor went offline on a Friday afternoon because of an unexpected server error, we just fixed it the next week. Well, we can’t wait that long anymore. 😊 Now we try to fix a problem as soon as it arises. If this isn’t possible, we collect more complex issues (and also feature suggestions) in the issue tracker.&lt;/p&gt;

&lt;p&gt;Another way to measure our impact is through the increasing number of third-party research conducted with our data. The most recent example is &lt;a href=&quot;https://escholarship.org/uc/item/9rr5k9p7&quot;&gt;Inna Wendell’s PhD thesis at UCLA&lt;/a&gt; on Russian five-act comedy in verse. A growing list of DraCor-enabled research can be found &lt;a href=&quot;https://dracor.org/doc/research&quot;&gt;on our website&lt;/a&gt;. We also just set up a &lt;a href=&quot;https://www.zotero.org/groups/4357494/clsinfra-wp7/collections/46CJJQM8/items/ZLP2NLW8/collection&quot;&gt;Zotero collection&lt;/a&gt; to keep track.&lt;/p&gt;

&lt;h1 id=&quot;api&quot;&gt;API&lt;/h1&gt;

&lt;p&gt;DraCor is of course more than a website, as we never tire of pointing out. The crucial element is its API. All currently active functions are &lt;a href=&quot;https://dracor.org/doc/api&quot;&gt;documented via Swagger&lt;/a&gt;, but this is of course not enough, especially if you are not yet that used to dealing with APIs. We are planning some workshops in the near future to train how to use the API so that the TEI-encoded DraCor corpora can unleash their full power to answer research questions. Stay tuned.&lt;/p&gt;

&lt;h1 id=&quot;ways-of-interaction&quot;&gt;Ways of Interaction&lt;/h1&gt;

&lt;p&gt;I’m sure we’ve forgotten this or that detail in this blogpost, but hope you enjoy exploring &lt;a href=&quot;https://dracor.org/&quot;&gt;the new DraCor&lt;/a&gt;. If you run into an issue, have a question, or want to contribute, you can always check our &lt;a href=&quot;https://dracor.org/credits&quot;&gt;Credits page&lt;/a&gt; to find the right contact person. We also use the GitHub Discussions feature for the &lt;a href=&quot;https://github.com/dracor-org/dracor-frontend/discussions&quot;&gt;dracor-frontend&lt;/a&gt; repo (all things website) and &lt;a href=&quot;https://github.com/dracor-org/dracor-api/discussions&quot;&gt;dracor-api&lt;/a&gt; (all things API functionality).&lt;/p&gt;

&lt;p&gt;That’s it for today. It’s been great to do some hacking together over the last months, even if the infamous in-person Potsdam/Moscow hackathon spirit is &lt;a href=&quot;https://dlina.github.io/Potsdam-Hackathon-2017/&quot;&gt;still unmatched&lt;/a&gt;. See you soon.&lt;/p&gt;

&lt;p&gt;🌞   🌴   🐬   ⛱️&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer, 
	
          
          Carsten Milling, 
	
          
          Mark Schwindt
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Ted Underwood's "Distant Horizons": Reviews and Replications</title>
    <link href="https://weltliteratur.net/underwood-distant-horizons/"/>
    <updated>2021-05-18T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/underwood-distant-horizons</id>
    <content type="html">&lt;p&gt;Ted Underwood’s most recent book – 206 pages, published in February 2019 – was widely acclaimed and much discussed. Based on the assumption that digital methods and corpora show their strengths especially in diachronic observations over long periods of time, he has provided aspects of a digital literary historiography.&lt;/p&gt;

&lt;p&gt;But this “computational monograph” – to quote Maciej Maryl – is not just something to read. Data and code have been published and can be replicated and questioned. This is not always easy and is not yet a living part of our practice in the Digital Humanities.&lt;/p&gt;

&lt;p&gt;That’s why this blogpost attempts to bring together the &lt;a href=&quot;#replications-of-code-andor-data&quot;&gt;&lt;strong&gt;REPLICATIONS&lt;/strong&gt;&lt;/a&gt; made so far to put Underwood’s technical setup to the test. Such efforts to verify scientific claims should be encouraged more and can be exemplary for similar undertakings in the future.&lt;/p&gt;

&lt;p&gt;I’m also gathering all &lt;a href=&quot;#reviews&quot;&gt;&lt;strong&gt;REVIEWS&lt;/strong&gt;&lt;/a&gt; of the book I found (as neither the publisher’s website nor the author’s personal page offers such an overview). If you have any additions, please lemmino (you can find me &lt;a href=&quot;https://twitter.com/umblaetterer&quot;&gt;on Twitter&lt;/a&gt;).&lt;/p&gt;

&lt;figure style=&quot;text-align:center;&quot;&gt;
  &lt;img src=&quot;/images/underwood-distant-horizons-cover.jpg&quot; alt=&quot;Underwood, Distant Horizons, cover&quot; style=&quot;width:300px; border: 1px solid transparent; border-color: black;}&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;p style=&quot;font-size: 16px; line-height: 24px;&quot;&gt;&lt;b&gt;Fig. 1.&lt;/b&gt; Ted Underwood: &lt;b&gt;Distant Horizons. Digital Evidence and Literary Change.&lt;/b&gt; Chicago and &lt;br /&gt;London: The University of Chicago Press 2019. (&lt;a href=&quot;https://doi.org/10.7208/chicago/9780226612973.001.0001&quot;&gt;doi:10.7208/chicago/9780226612973.001.0001&lt;/a&gt;).&lt;/p&gt;&lt;/center&gt;

&lt;p&gt;Underwood’s book has a somewhat mediating character. The machine-learning methods used for the analyses are not bleeding edge, he mostly relies on regularised logistic regression on several thousand features. Underwood refutes the claim that machine-learning models are always inscrutable black boxes. He spends a lot of time looking into these black boxes and interpreting the word frequencies and their distribution in order to craft narratives relevant to literary studies.&lt;/p&gt;

&lt;p&gt;But we are still at the beginning, as Underwood makes clear: “Statistical models can help us understand the social forces shaping a long arc of literary change. But turning those models into fully satisfying stories &lt;strong&gt;could take several more decades.&lt;/strong&gt;” (p. 109)&lt;/p&gt;

&lt;p&gt;I originally wanted to write a review myself, but never found the time to finish it (and finally got discouraged late last year when I encountered Maciej Maryl’s excellent and comprehensive review 😊). The following list of reviews I have compiled may serve to better sort through the rich discussions surrounding “Distant Horizons” and find good entry points.&lt;/p&gt;

&lt;h2 id=&quot;reviews&quot;&gt;REVIEWS&lt;/h2&gt;

&lt;ul&gt;
  &lt;li&gt;&lt;strong&gt;Dan Sinykin&lt;/strong&gt;, &lt;a href=&quot;https://post45.org/2019/05/distant-reading-and-literary-knowledge/&quot;&gt;&lt;em&gt;Post45&lt;/em&gt; (6 May 2019)&lt;/a&gt;
    &lt;ul&gt;
      &lt;li&gt;&lt;strong&gt;Tess McNulty&lt;/strong&gt; &lt;a href=&quot;https://post45.org/2019/05/seeing-double-a-response-to-dan-sinykin/&quot;&gt;responds to his review, &lt;em&gt;Post45&lt;/em&gt; (6 May 2019)&lt;/a&gt;&lt;/li&gt;
    &lt;/ul&gt;
  &lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Mareike Schumacher&lt;/strong&gt;, &lt;a href=&quot;https://lebelieberliterarisch.de/ted-underwoods-distant-horizons/&quot;&gt;&lt;em&gt;lebelieberliterarisch.de&lt;/em&gt; (25 Jul 2019)&lt;/a&gt; 🇩🇪 – followed by &lt;a href=&quot;https://lebelieberliterarisch.de/ted-underwoods-distant-horizons/#comment-580&quot;&gt;a comment from &lt;strong&gt;Fotis Jannidis&lt;/strong&gt;&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Laura Jiménez Ríos&lt;/strong&gt;, &lt;a href=&quot;https://doi.org/10.5944/rhd.vol.4.2019.24217&quot;&gt;&lt;em&gt;Revista de Humanidades Digitales&lt;/em&gt;, vol. 4 (2019), pp. 212–215&lt;/a&gt; 🇪🇸&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Daniel Rosenberg&lt;/strong&gt;, &lt;a href=&quot;https://doi.org/10.1086/707111&quot;&gt;&lt;em&gt;Modern Philology&lt;/em&gt;, vol. 117, no. 3 (Feb 2020),  publ. online 22 Nov 2019&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Katherine Bode&lt;/strong&gt;, &lt;a href=&quot;https://doi.org/10.1215/00267929-7933102&quot;&gt;&lt;em&gt;Modern Language Quarterly&lt;/em&gt;, vol. 81, issue 1 (1 Mar 2020), pp. 95–124&lt;/a&gt; (&lt;a href=&quot;https://katherinebode.files.wordpress.com/2019/10/mlq2019_preprint.pdf&quot;&gt;preprint&lt;/a&gt;) – a more critical take&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Alison Booth&lt;/strong&gt;, &lt;a href=&quot;https://muse.jhu.edu/article/771272/summary&quot;&gt;&lt;em&gt;Victorian Studies&lt;/em&gt;, vol. 62, no. 3 (Spring 2020), pp. 547–550&lt;/a&gt; – also somewhat critical&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Marie-Christine Boucher&lt;/strong&gt;, &lt;a href=&quot;https://doi.org/10.22029/ko.2020.1025&quot;&gt;&lt;em&gt;KULT_online&lt;/em&gt;, no. 61 (April 2020)&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Maciej Maryl&lt;/strong&gt;, &lt;a href=&quot;http://www.jltonline.de/index.php/reviews/article/view/1090/2504&quot;&gt;&lt;em&gt;JLTonline&lt;/em&gt; (17 Oct 2020)&lt;/a&gt; – assigns a new genre to the book, calling it a “computational monograph”&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;replications-of-code-andor-data&quot;&gt;REPLICATIONS OF CODE AND/OR DATA&lt;/h2&gt;

&lt;p&gt;When Underwood submitted his book to the publisher, he published data and code in a ZIP package on Zenodo on 25 March 2018 (&lt;a href=&quot;https://doi.org/10.5281/zenodo.1207277&quot;&gt;doi:10.5281/zenodo.1207277&lt;/a&gt;). There’s also a (slightly updated) &lt;a href=&quot;https://github.com/tedunderwood/horizon&quot;&gt;Git Hub repo&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;Portions of chapters 2, 3 and 4 are reprints of prior collaboratory work. This is also the reason why most of the listed replications could take place before the book even went to press.&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;&lt;strong&gt;Andrew Goldstone&lt;/strong&gt;, &lt;a href=&quot;https://www.andrewgoldstone.com/blog/2016/01/04/standards/&quot;&gt;“Of Literary Standards and Logistic Regression: A Reproduction”, &lt;em&gt;andrewgoldstone.com&lt;/em&gt;, 4 Jan 2016&lt;/a&gt;
    &lt;ul&gt;
      &lt;li&gt;&lt;strong&gt;Sarah Allison&lt;/strong&gt; wrote a short piece reflecting on Goldstone’s replication: &lt;a href=&quot;https://doi.org/10.22148/001c.11822&quot;&gt;“Other people’s data: humanities edition”, &lt;em&gt;Journal of Cultural Analytics&lt;/em&gt;, vol. 1, issue 1 (18 Dec 2016)&lt;/a&gt;&lt;/li&gt;
    &lt;/ul&gt;
  &lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Jonathan Goodwin&lt;/strong&gt;, &lt;a href=&quot;https://www.jgoodwin.net/blog/more-suvin/&quot;&gt;“Darko Suvin’s Genres of Victorian SF Revisited”, &lt;em&gt;jgoodwin.net&lt;/em&gt;, 17 Oct 2016&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Nan Z. Da&lt;/strong&gt;, &lt;a href=&quot;https://doi.org/10.1086/702594&quot;&gt;“The Computational Case against Computational Literary Studies”, vol. 45, no. 3 (Spring 2019)&lt;/a&gt; – among a total of 14 CLS papers, she also discusses Underwood’s “The Life Cycle of Genres”, which expanded into chapter 2 of “Distant Horizons”
    &lt;ul&gt;
      &lt;li&gt;followed by &lt;strong&gt;Underwood&lt;/strong&gt;’s reaction in CI’s response forum, &lt;a href=&quot;https://critinq.wordpress.com/2019/04/01/computational-literary-studies-participant-forum-responses-8/&quot;&gt;&lt;em&gt;critinq.wordpress.com&lt;/em&gt;, 1 Apr 2019&lt;/a&gt;&lt;/li&gt;
      &lt;li&gt;and &lt;strong&gt;Da&lt;/strong&gt;’s follow-up, &lt;a href=&quot;https://critinq.wordpress.com/2019/04/02/computational-literary-studies-participant-forum-responses-day-2/&quot;&gt;“Errors”, &lt;em&gt;critinq.wordpress.com&lt;/em&gt;, 2 Apr 2019&lt;/a&gt;&lt;/li&gt;
    &lt;/ul&gt;
  &lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;Ted Underwood&lt;/strong&gt;’s reply to Da, &lt;a href=&quot;https://doi.org/10.1086/709229&quot;&gt;“The Theoretical Divide Driving Debates about Computation”, &lt;em&gt;Critical Inquiry&lt;/em&gt;, vol. 46, no. 4 (Summer 2020)&lt;/a&gt;, which re-analyses material in chapter 2 using different methods&lt;/li&gt;
&lt;/ul&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>The First of May in German Literature</title>
    <link href="https://weltliteratur.net/the-first-of-may-in-german-literature/"/>
    <updated>2021-05-04T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/the-first-of-may-in-german-literature</id>
    <content type="html">&lt;center&gt;[ &lt;b&gt;tl;dr:&lt;/b&gt; This blogpost introduces a small dataset comprising 249 mentions
&lt;br /&gt;of the &quot;1st of May&quot; extracted from a corpus of German-language fiction. ]&lt;/center&gt;

&lt;p&gt;Last Friday, German radio station Bayern 2 asked me for an interview about our research on &lt;strong&gt;the role of the month of May in literature&lt;/strong&gt; (and about the practice of Digital Humanities in general, for that matter). The announcement is &lt;strong&gt;&lt;a href=&quot;https://web.archive.org/web/20210618231138/https://www.br.de/radio/bayern2/programmkalender/ausstrahlung-2460628.html&quot;&gt;here&lt;/a&gt;&lt;/strong&gt; – unfortunately, they didn’t put the interview online, but if anyone wants to listen to it, I can send them an mp3 copy.&lt;/p&gt;

&lt;p&gt;I took this opportunity to revisit my earlier research on date extractions from literature, which I started with Jannik Strötgen right after we met at his poster at the 2013 Herrenhausen Conference &lt;a href=&quot;https://web.archive.org/web/20221201065112/https://www.volkswagenstiftung.de/en/events/event-reports/documentation-herrenhausen-conference-digital-humanities-revisited&quot;&gt;“(Digital) Humanities Revisited”&lt;/a&gt;. Although our work was mostly finished by 2015, we occasionally returned to this topic due to public interest, and it was always great to come back to it since we could never exhaust all the possibilities our research data offered us because we quickly got busy with other things.&lt;/p&gt;

&lt;p&gt;The journalists at Bayern 2 probably learned about our research through an &lt;a href=&quot;https://archive.org/download/als-einst-am-ersten-mai-2019-05-04/als-einst-am-ersten-mai-2019-05-04.pdf&quot;&gt;article I wrote for Die Literarische Welt&lt;/a&gt; two years ago:&lt;/p&gt;

&lt;figure style=&quot;text-align:center;&quot;&gt;
  &lt;img src=&quot;/images/first-of-may/als-einst-am-ersten-mai-2019-05-04.png&quot; alt=&quot;newspaper article&quot; style=&quot;width:350px; border: 3px solid transparent; border-color: coral;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 1.&lt;/b&gt; “Als einst am 1. Mai die Welt begann” (publ. &lt;a href=&quot;https://archive.org/download/als-einst-am-ersten-mai-2019-05-04/als-einst-am-ersten-mai-2019-05-04.pdf&quot;&gt;4 May 2019&lt;/a&gt;).&lt;/center&gt;

&lt;p&gt;This article draws on the main results of our study originally presented at the &lt;a href=&quot;https://doi.org/10.5281/zenodo.4623384&quot;&gt;DHd2015 (in Graz)&lt;/a&gt; and &lt;a href=&quot;https://dbs.ifi.uni-heidelberg.de/files/Team/jannik/publications/fischer-stroetgen_temporal-expressions-in-literary-corpora_dh2015_final_2015-03-01.pdf&quot;&gt;DH2015 (in Sydney)&lt;/a&gt; conferences.&lt;/p&gt;

&lt;p&gt;An interesting by-product of our work was the “TIWOLI” app (“Today in World Literature”) with iOS and Android versions (&lt;a href=&quot;/Introducing-TIWOLI/&quot;&gt;which we introduced in this blog in 2016&lt;/a&gt;). And, of course, there’s &lt;a href=&quot;https://twitter.com/tiwolichirp&quot;&gt;the TIWOLI Twitter bot&lt;/a&gt;, run by our colleagues at University of Cologne.&lt;/p&gt;

&lt;p&gt;In 2017, we finally managed to release the corpus on which our research is based, the &lt;strong&gt;“Corpus of German-Language Fiction (txt)”&lt;/strong&gt; (&lt;a href=&quot;https://doi.org/10.6084/m9.figshare.4524680&quot;&gt;doi:10.6084/m9.figshare.4524680&lt;/a&gt;). The corpus contains 2,735 German-language prose works (mainly novels and short stories) by 549 authors, ranging from about 1510 to the 1940s (the bulk of the texts dates from 1840–1930). [On a sidenote, our corpus was also used by Severin Simmler &lt;a href=&quot;https://huggingface.co/severinsimmler/literary-german-bert&quot;&gt;to fine-tune a German BERT model to adapt it to the literary domain&lt;/a&gt;.]&lt;/p&gt;

&lt;p&gt;To extract the mentioned month names and concrete days, we used Jannik’s cutting-edge &lt;a href=&quot;https://github.com/HeidelTime/heideltime&quot;&gt;HeidelTime&lt;/a&gt; engine. As you can see in the flower-ringed histogram featured on the newspaper page, the month of May is by far the most frequently mentioned month in the corpus. In addition, the 1st of May is, also by far, the most frequently mentioned day of the year, with 249 mentions in the corpus.&lt;/p&gt;

&lt;p&gt;Here is an overview with the frequency of all days as we extracted them from the corpus (‘1’ means 0–9 occurrences, ‘2’ means 10–19 occurrences, etc., ‘+’ means 90 or more occurrences):&lt;/p&gt;

&lt;figure style=&quot;text-align:center;&quot;&gt;
  &lt;img src=&quot;/images/first-of-may/heatmap-all-days.png&quot; alt=&quot;heatmap&quot; style=&quot;width:550px; border: none;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 2.&lt;/b&gt; Green fields: days mentioned 50+ times in the corpus &lt;br /&gt;(slide taken from &lt;a href=&quot;https://dbs.ifi.uni-heidelberg.de/files/Team/jannik/dhd2015-fischer-stroetgen-deutsche-literatur-slides.pdf#page=21&quot;&gt;our DHd2015 slideset&lt;/a&gt;).&lt;/center&gt;

&lt;p&gt;HeidelTime can extract various forms of dates, so here it has managed to extract “1. Mai”, “1. May” as well as “erste” / “ersten Mai”. Some example sentences:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;“Ich telephonierte Montag nachmittag (also am &lt;u&gt;1. Mai&lt;/u&gt;) dem Witschi nach Gerzenstein, er möge mich in meinem Bureau aufsuchen.” (Glauser: &lt;a href=&quot;https://www.projekt-gutenberg.org/glauser/studer/chap15.html&quot;&gt;Wachtmeister Studer&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;“[…]; den &lt;u&gt;1. May&lt;/u&gt; 1772 reiseten wir aus Gondar ab.” (Knigge: &lt;a href=&quot;https://www.projekt-gutenberg.org/knigge/noldmann/noldm120.html&quot;&gt;Benjamin Noldmanns Geschichte der Aufklärung in Abyssinien&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;“Du sollst doch auch wissen, daß heute der &lt;u&gt;erste Mai&lt;/u&gt; ist, du Einsiedler!” (Dahn: &lt;a href=&quot;https://www.projekt-gutenberg.org/dahn/herzen/chap003.html&quot;&gt;Ernst und Frank&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;“Am &lt;u&gt;ersten Mai&lt;/u&gt; anno domini 1835 war zu Haimbach ein großes Frühstück.” (Stifter: &lt;a href=&quot;https://www.projekt-gutenberg.org/stifter/feldblum/feldb032.html&quot;&gt;Feldblumen&lt;/a&gt;)&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Now we’ve always wanted to dig a little deeper into our dataset to go beyond mere quantification, and we’ve demonstrated this idea by qualifying all mentions of the 10th of August:&lt;/p&gt;

&lt;figure style=&quot;text-align:center;&quot;&gt;
  &lt;img src=&quot;/images/first-of-may/pie-chart-10th-of-august.png&quot; alt=&quot;pie chart&quot; style=&quot;width:460px; border: none;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 3.&lt;/b&gt; Semantics of the 10th of August in our corpus &lt;br /&gt;(slide with German labelling taken from &lt;a href=&quot;https://dbs.ifi.uni-heidelberg.de/files/Team/jannik/dhd2015-fischer-stroetgen-deutsche-literatur-slides.pdf#page=22&quot;&gt;our DHd2015 slideset&lt;/a&gt;).&lt;/center&gt;

&lt;p&gt;About three quarters of the mentions were plot-related (i.e., something happened on a 10th of August, including date references in epistolary novels and fictitious diary entries). But more than 20% were non-plot-related date mentions, instead a reference to the 10th of August 1792, a decisive event of the French Revolution, when armed revolutionaries stormed the Tuileries Palace in Paris. This date is a constant reference point in our corpus of German fiction and makes for a high number of non-plot-related date mentions.&lt;/p&gt;

&lt;p&gt;For our focal date, the 1st of May, this problem is even more complex. Many feasts and customs are associated with this date, such as Walpurgis Night and the Maypole Dance, and since 1889 also International Workers’ Day. All these meanings can overlap when a 1st of May is mentioned and it is very difficult to decide whether a mention of the day refers to a plot-related event or not.&lt;/p&gt;

&lt;p&gt;To start somewhere, however, let us first measure the precision of the 249 mentions counted and say something about the reliability of the data we generated and evaluated at the time. There are only 2 false positives, the rest is a problem of corpus composition. It is meant to be a corpus with works of fiction (hence the name), but some non-fiction works have slipped in (as already detailed &lt;a href=&quot;https://doi.org/10.6084/m9.figshare.4524680&quot;&gt;in the corpus description on Figshare&lt;/a&gt;). Some mentions are duplicates, and one work (by Strindberg) made it into the corpus of German originals, although it is a translation). Here’s an overview:&lt;/p&gt;

&lt;figure style=&quot;text-align:center;&quot;&gt;
  &lt;img src=&quot;/images/first-of-may/pie-chart-1st-of-may.png&quot; alt=&quot;theatre on Mars, poster, small version&quot; style=&quot;width:450px; border: none;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 4.&lt;/b&gt; Precision of ‘1st of May’ mentions extracted from our corpus: &lt;br /&gt;correctly extracted from fiction: 226 – from non-fiction: 11 – &lt;br /&gt;duplicates: 8 – false positive: 2 – from translations: 1.&lt;/center&gt;

&lt;p&gt;Although we cannot say anything about recall in our exploratory study, the precision of HeidelTime is really good. We knew that before, of course, but never really detailed this. So this blogpost confirms the special role that May plays in German prose literature, even if the means we used at the time are far from being able to provide an exhaustive answer to the questions of “when literature takes place”. But it was a start.&lt;/p&gt;

&lt;p&gt;You can explore our findings yourself in an enhanced subset of our data, which we never formally published before. So here is a list of 1st of Mays mentioned in German fiction (in the ‘timex value’ column, HeidelTime automatically tried to assign years from the surrounding context of a mentioned date, which was not always successful, but we left this column in, even though we never used this data). The file format is CSV:&lt;/p&gt;

&lt;center&gt;🌿 &lt;a href=&quot;/data/sentences-containing-first-of-may.csv&quot;&gt;&lt;b&gt;sentences-containing-first-of-may.csv&lt;/b&gt;&lt;/a&gt; 🌱&lt;/center&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Vossian Antonomasia, Next Level</title>
    <link href="https://weltliteratur.net/vossian-antonomasia-next-level/"/>
    <updated>2020-11-09T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/vossian-antonomasia-next-level</id>
    <content type="html">&lt;p&gt;Our &lt;a href=&quot;/vossian-antonomasia-new-york-times/&quot;&gt;last blogpost on Vossian
Antonomasia&lt;/a&gt; (VA, or, vossanto) appeared
almost two years ago. Since then, we were busy improving our methods, writing
papers, and updating our &lt;a href=&quot;https://vossanto.weltliteratur.net/&quot;&gt;web page&lt;/a&gt; with
some more data to explore. In this article we summarise the latest
developments.&lt;/p&gt;

&lt;h2 id=&quot;interactive-timeline-for-data-exploration&quot;&gt;Interactive Timeline for Data Exploration&lt;/h2&gt;

&lt;p&gt;In the past year, &lt;a href=&quot;https://www.ibi.hu-berlin.de/de/forschung/info_processing_analytics&quot;&gt;our research group at
IBI&lt;/a&gt;
conducted a small code sprint to create an interface for the visual
exploration of Vossian Antonomasia extracted from our &lt;em&gt;New York Times&lt;/em&gt; working
corpus (covering the years 1987–2007). Thanks to &lt;a href=&quot;https://github.com/sjaakp&quot;&gt;Sjaak
Priester&lt;/a&gt; and his excellent &lt;a href=&quot;https://github.com/sjaakp/dateline&quot;&gt;JavaScript Dateline
widget&lt;/a&gt; we were able to quickly set up an
&lt;a href=&quot;https://vossanto.weltliteratur.net/timeline/&quot;&gt;interactive timeline&lt;/a&gt;:&lt;/p&gt;

&lt;p&gt;&lt;img src=&quot;../images/va_timeline_mj.png&quot; alt=&quot;Timeline Example: Michael Jordan&quot; /&gt;&lt;/p&gt;

&lt;p&gt;It assembles all Vossian Antonomasia extracted with the method described in
our &lt;a href=&quot;https://doi.org/10.18653/v1/D19-1647&quot;&gt;2019 EMNLP-IJCNLP paper&lt;/a&gt;. Each VA
is represented by a colour dot and the named entity providing the source
(whose properties or qualities are transferred to another named entity by the
magic of Vossian Antonomasia). The colour of the dot indicates the &lt;em&gt;New York
Times&lt;/em&gt; desk responsible for the corresponding article. Clicking on an entry
will display more information, as shown in the above screenshot:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;a photo automatically retrieved from
&lt;a href=&quot;https://commons.wikimedia.org/&quot;&gt;Wikimedia Commons&lt;/a&gt; (if available),&lt;/li&gt;
  &lt;li&gt;the sentence containing the VA expression with &lt;strong&gt;source&lt;/strong&gt; and &lt;em&gt;modifier&lt;/em&gt;
highlighted,&lt;/li&gt;
  &lt;li&gt;the article ID with a direct link to its full text,&lt;/li&gt;
  &lt;li&gt;the name of the author,&lt;/li&gt;
  &lt;li&gt;the name of the desk,&lt;/li&gt;
  &lt;li&gt;source and licence for the photo,&lt;/li&gt;
  &lt;li&gt;a permanent link to the corresponding VA expression.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Feel free to scroll through the
&lt;a href=&quot;https://vossanto.weltliteratur.net/timeline/&quot;&gt;timeline&lt;/a&gt; or just follow &lt;a href=&quot;https://vossanto.weltliteratur.net/timeline/#1077956_3&quot;&gt;this
example&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;In the upper left corner you can find a box for full-text search. After typing
in a few letters it shows all matching VA expressions. Clicking on a match
scrolls to the corresponding place in the timeline and shows the details, a
very convenient way to find interesting Vossian Antonomasia. For example, type
in “literature” to acquaint yourself with &lt;a href=&quot;https://vossanto.weltliteratur.net/timeline/#1587118_0&quot;&gt;“the Tupac Shakur of American
literature”&lt;/a&gt; or &lt;a href=&quot;https://vossanto.weltliteratur.net/timeline/#1097313_0&quot;&gt;“the
Madonna of Cuban
literature”&lt;/a&gt;.&lt;/p&gt;

&lt;h2 id=&quot;automatic-detection-of-vossian-antonomasia&quot;&gt;Automatic Detection of Vossian Antonomasia&lt;/h2&gt;

&lt;p&gt;Our &lt;a href=&quot;https://doi.org/10.1093/llc/fqy087&quot;&gt;first approach to VA extraction (published in &lt;em&gt;Digital
Scholarship in the Humanities&lt;/em&gt; 35.1, April
2020)&lt;/a&gt; was semi-automated,
sentence-based, and strictly centered around humans. We used regular
expressions to extract all sentences featuring our VA source pattern
“the ENTITY of”. Then our core idea was to use Wikidata as knowledge
base for distant supervision. We only kept candidates that included
the exact name or alias of a Wikidata entity which had to be an
&lt;a href=&quot;https://www.wikidata.org/wiki/Property:P31&quot;&gt;instance of&lt;/a&gt; the class
&lt;a href=&quot;https://www.wikidata.org/wiki/Q5&quot;&gt;human&lt;/a&gt;. We used a manually created
&lt;a href=&quot;https://github.com/weltliteratur/vossanto/blob/master/theof/blacklist.tsv&quot;&gt;blacklist&lt;/a&gt;
to exclude candidates like “the House of” (for ‘House’ being, among
other things, a documented alias of botanist &lt;a href=&quot;https://www.wikidata.org/wiki/Q3139666&quot;&gt;Homer Doliver
House&lt;/a&gt;).&lt;/p&gt;

&lt;p&gt;In our &lt;a href=&quot;https://doi.org/10.18653/v1/D19-1647&quot;&gt;2019 EMNLP-IJCNLP paper&lt;/a&gt; we
automated the process of extracting VA from large newspaper corpora and
compared &lt;strong&gt;three new approaches&lt;/strong&gt; to our initial approach:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;The second approach is an extension of our previous method. We replaced the
manually curated blacklist with a popularity measure to identify candidates
that could be removed after linking an entity to a Wikidata entry. We compared
the ‘popularity’ (for which we took the number of Wikidata sitelinks per
entity) of a human entity to another entity featuring the same label. For
example, the human entity ‘House’ (&lt;a href=&quot;https://www.wikidata.org/wiki/Q3139666&quot;&gt;the
botanist&lt;/a&gt;) only has 9 sitelinks while
the entity ‘House’ (&lt;a href=&quot;https://www.wikidata.org/wiki/Q3947&quot;&gt;the building&lt;/a&gt;) has
178 sitelinks. We removed such candidates since it was unlikely that the label
was linked to the correct entity. We also removed all candidates where the
source (e.g., ‘Prince’) together with multiple subsequent words (e.g., ‘of
Wales’) matched the name or alias of another Wikidata entity. This allowed us
to get rid of frequent false positives like &lt;a href=&quot;https://www.wikidata.org/wiki/Q43274&quot;&gt;the Prince of
Wales&lt;/a&gt; (who, grammatically, could also
be an aspiring Welsh singer reminding of the Artist Formerly Known as
Prince).&lt;/li&gt;
  &lt;li&gt;Since we focus on people as VA sources, our third approach is based on
named-entity recognition (NER). Instead of using Wikidata to detect sentence
candidates, we now tried the &lt;a href=&quot;https://nlp.stanford.edu/ner/&quot;&gt;Stanford NER
tool&lt;/a&gt; to detect entities. We also applied the
last step from the first approach to detect false positives.&lt;/li&gt;
  &lt;li&gt;In our fourth approach we leveraged the annotations from our initial
approach to train a neural network. First, we transformed each word of a
sentence into a word vector using pre-trained word embeddings
(&lt;a href=&quot;https://nlp.stanford.edu/projects/glove/&quot;&gt;GloVe&lt;/a&gt;). Then we fed the vectors
into a neural network – a bi-directional long short-term memory layer (BLSTM)
with a feed-forward layer – to let it learn to distinguish sentences that do
contain a VA expression from those that don’t.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;In our case, evaluating the different approaches is especially tricky due to
the rarity of the phenomenon. It is unfeasible to determine the recall for the
1.85 million articles in our NYT corpus. Therefore we calculate precision,
recall and f1 based on the labelled corpus. The BLSTM performs best, boosting
the precision to 87%:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;approach&lt;/th&gt;
      &lt;th&gt;precision&lt;/th&gt;
      &lt;th&gt;recall&lt;/th&gt;
      &lt;th&gt;f1&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;1st: semi-automated&lt;/td&gt;
      &lt;td&gt;49.8%&lt;/td&gt;
      &lt;td&gt;–&lt;/td&gt;
      &lt;td&gt;–&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;2nd: Wikidata&lt;/td&gt;
      &lt;td&gt;67.3%&lt;/td&gt;
      &lt;td&gt;93.0%&lt;/td&gt;
      &lt;td&gt;78.1%&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;3rd: NER&lt;/td&gt;
      &lt;td&gt;71.8%&lt;/td&gt;
      &lt;td&gt;81.3%&lt;/td&gt;
      &lt;td&gt;76.2%&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;4th: BLSTM&lt;/td&gt;
      &lt;td&gt;86.9%&lt;/td&gt;
      &lt;td&gt;85.3%&lt;/td&gt;
      &lt;td&gt;86.1%&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;We are currently working on various approaches to detect all three parts
(source, target and modifier) of a VA expression in a sentence. As a
by-product, we will have an enriched corpus where all parts of a VA expression
are tagged in each positively labelled sentence.&lt;/p&gt;

&lt;h2 id=&quot;modifiers&quot;&gt;Modifiers&lt;/h2&gt;

&lt;p&gt;Since we need labelled data to train and evaluate our machine-learning
approaches, we examined more than 3,000 VA expressions to annotate their
targets and modifiers. As a result, we can now present more reliable
statistics on the most common modifiers in our corpus. The ten most frequent
modifiers are:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;count&lt;/th&gt;
      &lt;th&gt;modifier&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;56&lt;/td&gt;
      &lt;td&gt;his day&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;34&lt;/td&gt;
      &lt;td&gt;his time&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;29&lt;/td&gt;
      &lt;td&gt;Japan&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;17&lt;/td&gt;
      &lt;td&gt;China&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;16&lt;/td&gt;
      &lt;td&gt;tennis&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;16&lt;/td&gt;
      &lt;td&gt;his generation&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;16&lt;/td&gt;
      &lt;td&gt;baseball&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;14&lt;/td&gt;
      &lt;td&gt;her time&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;13&lt;/td&gt;
      &lt;td&gt;our time&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;13&lt;/td&gt;
      &lt;td&gt;her day&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;A &lt;a href=&quot;https://vossanto.weltliteratur.net/emnlp-ijcnlp2019/statistics.html#modifiers&quot;&gt;longer
list&lt;/a&gt;
with many more results can be found on our most recent &lt;a href=&quot;https://vossanto.weltliteratur.net/emnlp-ijcnlp2019/statistics.html&quot;&gt;statistics
page&lt;/a&gt;. As
described in our first paper, this list illustrates an important function of
antonomasia, which is the “inculturation” of lesser known phenomena (cf. also
&lt;a href=&quot;https://doi.org/10.1515/9783110230215.373&quot;&gt;Holmqvist/Płuciennik 2010, p.
379&lt;/a&gt;).&lt;/p&gt;

&lt;p&gt;By filtering the modifiers we can create rankings for specific categories; for
example,
&lt;a href=&quot;https://vossanto.weltliteratur.net/emnlp-ijcnlp2019/statistics.html#country&quot;&gt;countries&lt;/a&gt;
…&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;count&lt;/th&gt;
      &lt;th&gt;country&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;29&lt;/td&gt;
      &lt;td&gt;Japan&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;17&lt;/td&gt;
      &lt;td&gt;China&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;10&lt;/td&gt;
      &lt;td&gt;Brazil&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;8&lt;/td&gt;
      &lt;td&gt;Iran&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;7&lt;/td&gt;
      &lt;td&gt;Mexico&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;7&lt;/td&gt;
      &lt;td&gt;Israel&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;7&lt;/td&gt;
      &lt;td&gt;India&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;4&lt;/td&gt;
      &lt;td&gt;South Africa&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;4&lt;/td&gt;
      &lt;td&gt;Poland&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;3&lt;/td&gt;
      &lt;td&gt;Spain&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;… or &lt;a href=&quot;https://vossanto.weltliteratur.net/emnlp-ijcnlp2019/statistics.html#sports&quot;&gt;sports&lt;/a&gt;:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;count&lt;/th&gt;
      &lt;th&gt;sports&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;16&lt;/td&gt;
      &lt;td&gt;tennis&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;16&lt;/td&gt;
      &lt;td&gt;baseball&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;11&lt;/td&gt;
      &lt;td&gt;hockey&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;10&lt;/td&gt;
      &lt;td&gt;basketball&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;8&lt;/td&gt;
      &lt;td&gt;golf&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;8&lt;/td&gt;
      &lt;td&gt;football&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;6&lt;/td&gt;
      &lt;td&gt;soccer&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;6&lt;/td&gt;
      &lt;td&gt;racing&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;3&lt;/td&gt;
      &lt;td&gt;women’s basketball&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;3&lt;/td&gt;
      &lt;td&gt;sailing&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;3&lt;/td&gt;
      &lt;td&gt;auto racing&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;2&lt;/td&gt;
      &lt;td&gt;pro football&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;2&lt;/td&gt;
      &lt;td&gt;New York baseball&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;Yale football fame&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;women’s hockey&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;women’s college soccer&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;this year’s national collegiate basketball tournament&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;the tennis tour&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;the tennis field&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;the soccer set&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;the racing world&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;the Olympic hockey tournament&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;stock-car racing&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;Rotisserie baseball&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;pro football owners&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;professional basketball coaches&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;professional basketball&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;motocross racing in the 1980’s&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;micro golfers&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;major league baseball&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;Laser sailing&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;Japanese baseball&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;Iraqi soccer&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;horse racing&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;hockey in the former Soviet Union&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;hockey commentary&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;high school baseball in New York&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;harness racing&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;golf criticism&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;football teams&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;football owners&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;football announcers&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;European hockey&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;country-club golf&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;college football underclassmen&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;college football these days&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;college football&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;college basketball&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;Chinese baseball&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;Brazilian basketball for the past 20 years&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;BMX racing&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;biddy basketball&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;basketball announcers&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;basketball analysts&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;basketball analysis&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;baseball’s new era&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;baseball managers&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;baseball executives&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;baseball collections&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;baseball cards&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;Most modifiers are short and consist of only one to three words, as this plot
of the distribution of the length (number of words) of modifiers shows:&lt;/p&gt;

&lt;p&gt;&lt;img src=&quot;https://raw.githubusercontent.com/weltliteratur/vossanto/master/emnlp-ijcnlp2019/nyt_vossantos_modifier_length.svg&quot; alt=&quot;modifier length distribution&quot; /&gt;&lt;/p&gt;

&lt;p&gt;The longest modifier we have found so far contains 25 words:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;“And while he modestly demurs, Mr. Barker is widely regarded as the
&lt;strong&gt;Bob Fosse&lt;/strong&gt; of &lt;em&gt;the carefully choreographed event that consumes
Midtown Manhattan with tin whistles, step dancers and some two million
spectators on that invariably brisk March 17 morning&lt;/em&gt;.” (source:
&lt;a href=&quot;http://query.nytimes.com/gst/fullpage.html?res=9B04E3DE1E3BF934A35750C0A9679C8B63&quot;&gt;2001/03/07/1276052&lt;/a&gt;)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;h2 id=&quot;json-data-dump&quot;&gt;JSON Data Dump&lt;/h2&gt;

&lt;p&gt;As additional outcome of the timeline, our &lt;a href=&quot;https://github.com/weltliteratur/vossanto/blob/master/timeline/vossantos.json&quot;&gt;dataset is now also available in
JSON
format&lt;/a&gt;.
The enriched JSON data dump contains links to
&lt;a href=&quot;https://www.wikidata.org/&quot;&gt;Wikidata&lt;/a&gt; and to images of VA sources on
&lt;a href=&quot;https://commons.wikimedia.org/&quot;&gt;Wikimedia Commons&lt;/a&gt; (including licence
information), so that they can easily be reused for further analysis or demos.
The following excerpt shows a JSON entry for an exemplary VA expression:&lt;/p&gt;

&lt;div class=&quot;language-json highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;p&quot;&gt;{&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
   &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;id&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;w&quot;&gt; &lt;/span&gt;&lt;span class=&quot;s2&quot;&gt;&quot;0039183_0&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
   &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;date&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;w&quot;&gt; &lt;/span&gt;&lt;span class=&quot;s2&quot;&gt;&quot;1987-05-10&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
   &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;sourceId&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;w&quot;&gt; &lt;/span&gt;&lt;span class=&quot;s2&quot;&gt;&quot;Q9696&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
   &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;sourceLabel&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;w&quot;&gt; &lt;/span&gt;&lt;span class=&quot;s2&quot;&gt;&quot;John F. Kennedy&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
   &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;sourceImId&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;w&quot;&gt; &lt;/span&gt;&lt;span class=&quot;s2&quot;&gt;&quot;John_F._Kennedy,_White_House_color_photo_portrait.jpg&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
   &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;sourceImThumb&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;w&quot;&gt; &lt;/span&gt;&lt;span class=&quot;s2&quot;&gt;&quot;thumb/c/c3/John_F._Kennedy%2C_White_House_color_photo_portrait.jpg/180px-John_F._Kennedy%2C_White_House_color_photo_portrait.jpg&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
   &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;sourceImLicense&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;w&quot;&gt; &lt;/span&gt;&lt;span class=&quot;s2&quot;&gt;&quot;pd&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
   &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;fId&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;w&quot;&gt; &lt;/span&gt;&lt;span class=&quot;s2&quot;&gt;&quot;1987/05/10/0039183&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
   &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;aUrlId&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;w&quot;&gt; &lt;/span&gt;&lt;span class=&quot;s2&quot;&gt;&quot;9B0DE3DD1E3EF933A25756C0A961948260&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
   &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;text&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
   &lt;/span&gt;&lt;span class=&quot;s2&quot;&gt;&quot;In the 60&apos;s Mr. Bernstein looked like *the John F. Kennedy of* /culture/.&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
   &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;author&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;w&quot;&gt; &lt;/span&gt;&lt;span class=&quot;s2&quot;&gt;&quot;Botstein, Leon&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt;&lt;span class=&quot;w&quot;&gt; &lt;/span&gt;&lt;span class=&quot;nl&quot;&gt;&quot;desk&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;w&quot;&gt; &lt;/span&gt;&lt;span class=&quot;s2&quot;&gt;&quot;Book Review Desk&quot;&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;}&lt;/span&gt;&lt;span class=&quot;w&quot;&gt;
&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;p&gt;The fields contain the following information:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;key&lt;/th&gt;
      &lt;th&gt;content&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;id&lt;/td&gt;
      &lt;td&gt;A unique (within the dataset) identifier for a VA expression.&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;date&lt;/td&gt;
      &lt;td&gt;Publication date of the corresponding NYT article.&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;sourceId&lt;/td&gt;
      &lt;td&gt;The Wikidata ID of the VA source.&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;sourceLabel&lt;/td&gt;
      &lt;td&gt;The (English) Wikidata label of the source.&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;sourceImId&lt;/td&gt;
      &lt;td&gt;The name of the source’s image on Wikimedia Commons.&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;sourceImThumb&lt;/td&gt;
      &lt;td&gt;The path to the source’s image on Wikimedia Commons.&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;sourceImLicense&lt;/td&gt;
      &lt;td&gt;The licence of the source’s image.&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;fId&lt;/td&gt;
      &lt;td&gt;The ID of the article’s file in the NYT dataset.&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;aUrlId&lt;/td&gt;
      &lt;td&gt;The ID of the article as part of its URL (&lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;http://query.nytimes.com/gst/fullpage.html?res=&amp;lt;HERE&amp;gt;&lt;/code&gt;)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;text&lt;/td&gt;
      &lt;td&gt;The sentence containing the VA expression (including org-mode markup).&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;h2 id=&quot;future-endeavours&quot;&gt;Future Endeavours&lt;/h2&gt;

&lt;p&gt;We are very interested in analysing Vossian Antonomasia in languages other
than English (especially German) and are looking for suitable corpora. In
general, our aim is to better understand Vossian Antonomasia and its usage,
distribution and variety.&lt;/p&gt;

&lt;p&gt;As always, all details can be found on our project website
&lt;a href=&quot;https://vossanto.weltliteratur.net/&quot;&gt;https://vossanto.weltliteratur.net/&lt;/a&gt;.&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Robert Jäschke, 
	
          
          Michel Schwab, 
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Tools Mentioned in DH2020 Abstracts</title>
    <link href="https://weltliteratur.net/tools-mentioned-in-dh2020-abstracts/"/>
    <updated>2020-07-23T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/tools-mentioned-in-dh2020-abstracts</id>
    <content type="html">&lt;p&gt;Building on our &lt;a href=&quot;/dh-tools-used-in-research/&quot;&gt;previous attempts&lt;/a&gt; to extract digital tools &lt;strong&gt;mentioned&lt;/strong&gt; in abstracts of Digital Humanities conferences, we conducted a similar exercise for this week’s virtual &lt;a href=&quot;https://dh2020.adho.org/&quot;&gt;DH2020&lt;/a&gt;. We again used our simple string-matching programme &lt;a href=&quot;https://github.com/lehkost/ToolXtractor&quot;&gt;ToolXtractor&lt;/a&gt; and an updated list of tools in the &lt;a href=&quot;http://tapor.ca/&quot;&gt;TAPoR&lt;/a&gt; database.&lt;/p&gt;

&lt;p&gt;So here is an overview of tools mentioned in more than one abstract (out of &lt;a href=&quot;https://dh2020.adho.org/abstracts/&quot;&gt;a total of 475 abstracts&lt;/a&gt; for all conference formats), including links to the actual texts – followed by some observations &lt;a href=&quot;#some-acknowledgements-and-observations&quot;&gt;at the bottom&lt;/a&gt; of this blog post:&lt;/p&gt;

&lt;h2 id=&quot;spacy-5&quot;&gt;spaCy (5)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/217_AnnotatingspatialentitiesinRomanianNovels.html&quot;&gt;Annotating spatial entities in Romanian Novels&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/289_TheSemanticsofStructureinLargeHistoricalCorpora.html&quot;&gt;The Semantics of Structure in Large Historical Corpora&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/358_AnnotatingReaderAbsorption.html&quot;&gt;Annotating Reader Absorption&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/463_PoetryLabAnOpenSourceToolkitfortheAnalysisofSpanishPoetryCorpora.html&quot;&gt;PoetryLab. An Open Source Toolkit for the Analysis of Spanish Poetry Corpora&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/612_NetworkAnalysisandSpatialStylometryinAmericanDramaStudiesNASSA.html&quot;&gt;“Network Analysis and Spatial Stylometry in American Drama Studies” (NASSA)&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;word2vec-5&quot;&gt;word2vec (5)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/135_ComputationandRhetoricalInventionFindingThingsToSayWithword2vec.html&quot;&gt;Computation and Rhetorical Invention: Finding Things To Say With word2vec&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/375_WhenClassicalChineseMeetsMachineLearningExplainingtheRelativePerformancesofWordandSentenceSegmentationTasks.html&quot;&gt;When Classical Chinese Meets Machine Learning: Explaining the Relative Performances of Word and Sentence Segmentation Tasks&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/422_ThesemanticsoftheUSstreetmap.html&quot;&gt;The semantics of the U.S. street map&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/495_TopicsbucketsandpsychiatryOnthecollectivecreationofacorpusexplorationtool.html&quot;&gt;Topics, buckets, and psychiatry. On the collective creation of a corpus exploration tool&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/681_HumanCenteredComputingforHumanistsCaseStudiesfromtheComputationalThinkingandLearningInitiativeatVanderbiltUniversity.html&quot;&gt;Human-Centered Computing for Humanists: Case Studies from the Computational Thinking and Learning Initiative at Vanderbilt University&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;gephi-4&quot;&gt;Gephi (4)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/174_IdentifyingrelationsbetweencharactersinAfrikaansTshivenaandXitsongabook.html&quot;&gt;Identifying relations between characters in Afrikaans, Tshivenḓa, and Xitsonga book&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/419_TeachingattheIntersectionofDigitalHumanitiesandVisualization.html&quot;&gt;Teaching at the Intersection of Digital Humanities and Visualization&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/571_CultureAnalyticsWorkshopNetworks.html&quot;&gt;Culture Analytics Workshop: Networks&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/577_ExtractingaSocialNetworkofMusicologists.html&quot;&gt;Extracting a Social Network of Musicologists&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;natural-language-toolkit-4&quot;&gt;Natural Language Toolkit (4)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/252_StopWords.html&quot;&gt;Stop Words&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/612_NetworkAnalysisandSpatialStylometryinAmericanDramaStudiesNASSA.html&quot;&gt;“Network Analysis and Spatial Stylometry in American Drama Studies” (NASSA)&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/725_Acomparativestudyofsentimentandtopicsinmigrationrelatedtweets.html&quot;&gt;A comparative study of sentiment and topics in migration related tweets&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/727_StudyingGeographicalPatternsAcrossJohnMiltonsGenres.html&quot;&gt;Studying Geographical Patterns Across John Milton’s Genres&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;voyant-tools-4&quot;&gt;Voyant Tools (4)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/261_TransHispanicNetworksofFeministSolidarityTheRiseandSpreadof8M.html&quot;&gt;Trans-Hispanic Networks of Feminist Solidarity: The Rise and Spread of #8M&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/288_FromdatatovisualisationDantesDivineComedyasacasestudy1.html&quot;&gt;From data to visualisation: Dante’s Divine Comedy as a case study&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/419_TeachingattheIntersectionofDigitalHumanitiesandVisualization.html&quot;&gt;Teaching at the Intersection of Digital Humanities and Visualization&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/681_HumanCenteredComputingforHumanistsCaseStudiesfromtheComputationalThinkingandLearningInitiativeatVanderbiltUniversity.html&quot;&gt;Human-Centered Computing for Humanists: Case Studies from the Computational Thinking and Learning Initiative at Vanderbilt University&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;drupal-3&quot;&gt;Drupal (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/479_TenyearsrecoveringthememoryofrepublicanexilewithcitizencollaborationTheresultsofExiliadsProjectaperspectivefromtheDigitalHumanitiesandtheDigitalPublicHistory.html&quot;&gt;Ten years recovering the memory of republican exile with citizen collaboration. The results of E-xiliad@s Project: a perspective from the Digital Humanities and the Digital Public History.&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/552_TheHistoricGravescrowdsourcingproject10yearstranscribingIrishhistory.html&quot;&gt;The Historic Graves crowdsourcing project: 10 years transcribing Irish history&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/718_EnhancingCommunitythroughOpenDHWebsiteDesign.html&quot;&gt;Enhancing Community through Open DH Website Design&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;google-maps-3&quot;&gt;Google Maps (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/140_TheGeoNewsMinerAninteractivespatialhumanitiestooltovisualizegeographicalreferencesinhistoricalnewspapers.html&quot;&gt;The GeoNewsMiner: An interactive spatial humanities tool to visualize geographical references in historical newspapers&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/278_TimeInformationSystemHuTimeAVisualizationandAnalysisToolforChronologicalInformationofHumanities.html&quot;&gt;Time Information System, HuTime — A Visualization and Analysis Tool for Chronological Information of Humanities&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/468_LinkingTimeSpaceandStatementsinOneGISSystemAUseCaseofStudyingIndividualsBiographies.html&quot;&gt;Linking Time, Space, and Statements in One GIS System: A Use Case of Studying Individuals’ Biographies&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;tesseract-3&quot;&gt;Tesseract (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/293_ANeuralOCREngineforNorthSaami.html&quot;&gt;A Neural OCR Engine for North Saami&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/435_FastSearchwithPoorOCR.html&quot;&gt;Fast Search with Poor OCR&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/541_ServingthecityanautomaticinformationextractionformappingAmsterdamnightlife18201940.html&quot;&gt;Serving the city: an automatic information extraction for mapping Amsterdam nightlife (1820-1940)&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;arcgis-2&quot;&gt;ArcGIS (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/153_ProjectTwitterLiteratureScrapingAnalyzingandArchivingTwitterDatainLiteraryResearch.html&quot;&gt;Project Twitter Literature: Scraping, Analyzing, and Archiving Twitter Data in Literary Research&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/299_TheFightforNationalLanguageRightsintheUSSR.html&quot;&gt;The Fight for National Language Rights in the USSR&lt;/a&gt; + &lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/382_EmpoweredMinoritiesDifferentialOutcomesForMinoritiesEnjoyingKremlinSupport.html&quot;&gt;“Empowered Minorities: Differential Outcomes For Minorities Enjoying Kremlin Support”&lt;/a&gt; (lightning + poster)&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;cytoscape-2&quot;&gt;Cytoscape (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/376_TellingbiggerstoriesFormalontologicalmodellingofscholarlyargumentation.html&quot;&gt;“Telling bigger stories”: Formal ontological modelling of scholarly argumentation&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/734_WebsofViolenceinBeowulf.html&quot;&gt;“Webs of Violence in Beowulf”&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;mallet-2&quot;&gt;MALLET (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/265_HANDLEGetaGriponMALLET.html&quot;&gt;HANDLE: Get a Grip on MALLET&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/495_TopicsbucketsandpsychiatryOnthecollectivecreationofacorpusexplorationtool.html&quot;&gt;Topics, buckets, and psychiatry. On the collective creation of a corpus exploration tool&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;neo4j-2&quot;&gt;Neo4j (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/679_VisualizingaTranslationalQueerPoetics.html&quot;&gt;Visualizing a Translational Queer Poetics&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/729_AVisualizationAssistedReadingSystemforaNeoConfucianCanon.html&quot;&gt;A Visualization-Assisted Reading Systemfor a Neo-Confucian Canon&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;omeka-2&quot;&gt;Omeka (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/299_TheFightforNationalLanguageRightsintheUSSR.html&quot;&gt;The Fight for National Language Rights in the USSR&lt;/a&gt; + &lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/382_EmpoweredMinoritiesDifferentialOutcomesForMinoritiesEnjoyingKremlinSupport.html&quot;&gt;“Empowered Minorities: Differential Outcomes For Minorities Enjoying Kremlin Support”&lt;/a&gt; (lightning + poster)&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/566_CreatingDigitalCollectionswithMinimalInfrastructureHandsOnwithCollectionBuilderforTeachingandExhibits.html&quot;&gt;Creating Digital Collections with Minimal Infrastructure: Hands On with CollectionBuilder for Teaching and Exhibits&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;openstreetmap-2&quot;&gt;OpenStreetMap (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/612_NetworkAnalysisandSpatialStylometryinAmericanDramaStudiesNASSA.html&quot;&gt;“Network Analysis and Spatial Stylometry in American Drama Studies” (NASSA)&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/730_DynamicSystemsforHumanitiesAudioCollectionsTheTheoryandRationaleofSwallow.html&quot;&gt;Dynamic Systems for Humanities Audio Collections: The Theory and Rationale of Swallow&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;h1 id=&quot;some-acknowledgements-and-observations&quot;&gt;Some Acknowledgements and Observations&lt;/h1&gt;

&lt;ul&gt;
  &lt;li&gt;Programming languages were excluded in this list (for a reality check: Python is mentioned in 26 abstracts, JavaScript in 11, MySQL in 4). Also, no full-text, data or code repositories like HathiTrust, Google Books, GitHub, Dataverse or Islandora were included.&lt;/li&gt;
  &lt;li&gt;Disambiguation was needed: The &lt;a href=&quot;https://sites.google.com/site/computationalstylistics/stylo&quot;&gt;stylo R package&lt;/a&gt; didn’t make it into the list, although the term ‘stylo’ is mentioned in two abstracts. But the second time it did not refer to the well-known stylometry library, but to &lt;a href=&quot;https://dh2020.adho.org/wp-content/uploads/2020/07/504_StyloasemanticwritingtoolforscientificpublishinginHumanSciences.html&quot;&gt;another tool&lt;/a&gt;.&lt;/li&gt;
  &lt;li&gt;So, hm, only 14 tools that were explicitely mentioned more than once. Our guess is still that not all tools that contributed to a research project or workshop were mentioned, which makes it more difficult to understand how things were done. One reason for this could be the limited writing space for conference abstracts, but there is definitely room for improvement in terms of specifically mentioning the tools used.&lt;/li&gt;
  &lt;li&gt;The TAPoR list of tools, even if updated, still isn’t (and never will be) exhaustive, so as usual, take this with a grain of salt.&lt;/li&gt;
&lt;/ul&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer, 
	
          
          Yoann Moranville
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Network Modelling "The Last Days of Mankind"</title>
    <link href="https://weltliteratur.net/theatre-on-mars/"/>
    <updated>2020-05-07T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/theatre-on-mars</id>
    <content type="html">&lt;p&gt;Last level. Final boss. Arguably the longest play ever written: &lt;strong&gt;“The Last Days of Mankind”&lt;/strong&gt;, brought to light between 1915 and 1922 by your favourite Austrian crank: writer and journalist &lt;strong&gt;&lt;a href=&quot;https://en.wikipedia.org/wiki/Karl_Kraus_(writer)&quot;&gt;Karl Kraus&lt;/a&gt;&lt;/strong&gt; (1874–1936).&lt;/p&gt;

&lt;p&gt;More than 600 pages in the German original, five acts, prologue and epilogue, 220 scenes altogether documenting the apocalypse of World War I. A first complete translation to English was published &lt;a href=&quot;https://yalebooks.yale.edu/book/9780300207675/last-days-mankind&quot;&gt;at Yale University Press&lt;/a&gt; in 2015 (done by Fred Bridgham and Edward Timms).&lt;/p&gt;

&lt;p&gt;As Kraus notes in the preface:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;“The performance of this drama, which would take some 10 evenings in terrestrial time, is intended for a theatre on Mars. Theatre-goers on planet earth would find it unendurable.”&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Now what did we do? Dialogues in the play are distributed among a total of &lt;strong&gt;925 speaking instances&lt;/strong&gt; (individual characters, voices, groups, choirs). We extracted the social network graph (based on co-occurrences of characters per scene) from &lt;a href=&quot;https://dracor.org/api/corpora/ger/play/kraus-die-letzten-tage-der-menschheit/tei&quot;&gt;the TEI version&lt;/a&gt; of the play. And then we took Kraus at his word and projected the network graph onto planet Mars. Resulting in our poster contribution for &lt;a href=&quot;https://dhd2020.de/&quot;&gt;DHd2020&lt;/a&gt; held in Paderborn in early March (one of the last conferences before coronavirus lockdown):&lt;/p&gt;

&lt;figure style=&quot;text-align:center;&quot;&gt;
  &lt;img src=&quot;/images/theatre-on-mars-poster-small.jpg&quot; alt=&quot;theatre on Mars, poster, small version&quot; style=&quot;width:463px;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 1.&lt;/b&gt; &quot;Visit to the Mars Theatre.&quot; (DOI:&lt;a href=&quot;https://doi.org/10.6084/m9.figshare.11917902&quot;&gt;10.6084/m9.figshare.11917902&lt;/a&gt;)&lt;/center&gt;

&lt;p&gt;Our visit to the Mars theatre won us &lt;a href=&quot;https://dig-hum.de/dhd-awards&quot;&gt;the Best Poster Award&lt;/a&gt;, which we were very happy about. Feel free to download the &lt;a href=&quot;https://doi.org/10.6084/m9.figshare.11917902&quot;&gt;hi-res version&lt;/a&gt; from Figshare (beware, it’s heavy). There’s also a lenghty &lt;a href=&quot;https://twitter.com/umblaetterer/status/1235556225128886277&quot;&gt;Twitter thread&lt;/a&gt; about our work (in German).&lt;/p&gt;

&lt;p&gt;Some properties of the network:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;Number of nodes: 925&lt;/li&gt;
  &lt;li&gt;Number of edges: 12.425&lt;/li&gt;
  &lt;li&gt;Network density: 0,029 (very low in comparison to other plays)&lt;/li&gt;
  &lt;li&gt;Number of independent clusters: 120&lt;/li&gt;
  &lt;li&gt;Clustering coefficient: 0,976 (extremely high, corresponds to the value of other station dramas)&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;If you want to toy around with the network data yourself, you can download it in CSV, GEXF and GraphML format &lt;a href=&quot;https://dracor.org/ger/kraus-die-letzten-tage-der-menschheit&quot;&gt;from &lt;strong&gt;DraCor&lt;/strong&gt;&lt;/a&gt;.&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Measuring the Impact of DH Conference Abstracts: The Case of DH2016</title>
    <link href="https://weltliteratur.net/measuring-impact-dh-conference-proceedings/"/>
    <updated>2020-02-17T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/measuring-impact-dh-conference-proceedings</id>
    <content type="html">&lt;p&gt;We are still toying around with the proceedings of &lt;a href=&quot;https://adho.org/conference&quot;&gt;ADHO’s Digital Humanities conferences&lt;/a&gt;. This time we are interested in the measurability of the scientific impact of published abstracts.&lt;/p&gt;

&lt;p&gt;Long story short, since no other meaningful metrics are at hand (or are there?), we have chosen &lt;strong&gt;Google Scholar&lt;/strong&gt;’s citation counts. As a test case we decided to go with the &lt;strong&gt;abstracts from DH2016&lt;/strong&gt; in Kraków, because 1.) they are well covered in Google Scholar and 2.) it has been three and a half years since the conference, so that enough time has passed to actually have an impact and attract citations.&lt;/p&gt;

&lt;p&gt;Our starting point was the fabulous &lt;strong&gt;&lt;a href=&quot;https://en.wikipedia.org/wiki/DBLP&quot;&gt;DBLP&lt;/a&gt;&lt;/strong&gt;, a giant computer science bibliography founded at University of Trier. Fortunately they also cover ADHO conferences since some time, so here are all contributions to DH2016 in a concise and interoperable form: &lt;a href=&quot;https://dblp1.uni-trier.de/db/conf/dihu/dh2016.html&quot;&gt;https://dblp1.uni-trier.de/db/conf/dihu/dh2016.html&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;Before we dive into the data, some known issues:&lt;/p&gt;
&lt;ul&gt;
  &lt;li&gt;Google Scholar is &lt;a href=&quot;https://www.nature.com/articles/d41586-018-04190-5&quot;&gt;“notoriously hard to mine”&lt;/a&gt;, there’s no API and access is limited (if you don’t want to deal with CAPTCHAs and “We’re sorry, but your computer network may be sending automated queries” you will have to apply some IP address magic).&lt;/li&gt;
  &lt;li&gt;Google Scholar works pretty well, but there might be some incorrectly assigned citations – since their search algorithm is proprietary, we can’t really know why things go wrong if they do.&lt;/li&gt;
  &lt;li&gt;It is also unclear what exactly is monitored by Google Scholar, so there are definitely missing citations.&lt;/li&gt;
  &lt;li&gt;When counting and ranking citations we don’t check if they are mainly self-citations.&lt;/li&gt;
  &lt;li&gt;Given the above points, please take the following with a grain of salt.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Out of 435 accepted abstracts at DH2016 (plenary lectures, panels, long and short papers, posters, pre-conference workshops and tutorials), 158 are cited at least once according to Google Scholar, &lt;strong&gt;as of 15 February 2020&lt;/strong&gt; (= 3,5 years after the conference).&lt;/p&gt;

&lt;p&gt;Here are our resulting data sets in case you wanna explore them yourself (CSV format):&lt;/p&gt;
&lt;ol&gt;
  &lt;li&gt;&lt;a href=&quot;/data/impact-dh2016-abstracts/dh2016-citations-google-scholar.csv&quot;&gt;DH2016 citations according to Google Scholar&lt;/a&gt; (158 abstracts)&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;/data/impact-dh2016-abstracts/dh2016-dsh-special-issue-citations-google-scholar.csv&quot;&gt;DSH Special DH2016 Issue citations according to Google Scholar&lt;/a&gt; (15 papers)&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;/data/impact-dh2016-abstracts/dh2016-papers-published-elsewhere-citations-google-scholar.csv&quot;&gt;DH2016 follow-up papers published elsewhere (selection)&lt;/a&gt; (3 papers)&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;What follows are some rankings based on these data sets (plus some general thoughts at the bottom of this blog post):&lt;/p&gt;

&lt;h2 id=&quot;1-most-cited-abstracts-from-dh2016-as-of-15-february-2020&quot;&gt;1. Most cited abstracts from DH2016 (as of 15 February 2020)&lt;/h2&gt;

&lt;p&gt;There are 28 abstracts with at least 4 citations collected by Google Scholar:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;Eetu Mäkelä, Thea Lindquist, Eero Hyvönen: &lt;br /&gt;CORE - A Contextual Reader based on Linked Data – Long Paper – &lt;strong&gt;15&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/MakelaLH16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=7822185147793572398&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Esko Ikkala, Jouni Tuominen, Eero Hyvönen: &lt;br /&gt;Contextualizing Historical Places in a Gazetteer by Using Historical Maps and Linked Data – Short Paper – &lt;strong&gt;11&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/IkkalaTH16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=15673204855311481131&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Terhi Nurmikko-Fuller, Jacob Jett, Timothy W. Cole, Chris Maden, Kevin R. Page, J. Stephen Downie: &lt;br /&gt;A Comparative Analysis of Bibliographic Ontologies: Implications for Digital Humanities – Short Paper – &lt;strong&gt;10&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/Nurmikko-Fuller16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=8048646481335919612&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Peer Trilcke, Frank Fischer, Mathias Göbel, Dario Kampkaspar: &lt;br /&gt;Theatre Plays as ‘Small Worlds’? Network Data on the History and Typology of German Drama, 1730-1930 – Long Paper – &lt;strong&gt;8&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/TrilckeFGK16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=3941180453510624475&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Aris Xanthos, Isaac Pante, Yannick Rochat, Martin Grandjean: &lt;br /&gt;Visualising the Dynamics of Character Networks – Long Paper – &lt;strong&gt;8&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/XanthosPRG16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=14174769823231486276&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Manuel Burghardt, Lukas Lamm, David Lechler, Matthias Schneider, Tobias Semmelmann: &lt;br /&gt;Tool-based Identification of Melodic Patterns in MusicXML Documents – Short Paper – &lt;strong&gt;8&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/BurghardtLLSS16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=14025672038740466450&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Fotis Jannidis, Isabella Reger, Markus Krug, Lukas Weimer, Luisa Macharowsky, Frank Puppe: &lt;br /&gt;Comparison of Methods for the Identification of Main Characters in German Novels – Short Paper – &lt;strong&gt;8&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/JannidisRKWMP16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=1541757619442963302&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Roman Klinger, Surayya Samat Suliya, Nils Reiter: &lt;br /&gt;Automatic Emotion Detection for Quantitative Literary Studies – Poster – &lt;strong&gt;8&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/KlingerSR16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=3544129917977527749&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Maciej Maryl, Maciej Piasecki, Ksenia Mlynarczyk: &lt;br /&gt;Where Close and Distant Readings Meet: Text Clustering Methods in Literary Analysis of Weblog Genres – Long Paper – &lt;strong&gt;7&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/MarylPM16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=16480967516874896595,8343475867642195584&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Manuel Burghardt, Michael Kao, Christian Wolff: &lt;br /&gt;Beyond Shot Lengths - Using Language Data and Color Information as Additional Parameters for Quantitative Movie Analysis – Poster – &lt;strong&gt;7&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/BurghardtKW16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=11783808221396105186&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Christof Schöch, Daniel Schlör, Stefanie Popp, Annelen Brunner, Ulrike Henny, José Calvo Tello: &lt;br /&gt;Straight Talk! Automatic Recognition of Direct Speech in Nineteenth-Century French Novels – Long Paper – &lt;strong&gt;6&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/SchochSPBHT16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=5739210873641308458&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Johannes Hellrich, Udo Hahn: &lt;br /&gt;Measuring the Dynamics of Lexico-Semantic Change Since the German Romantic Period – Short Paper – &lt;strong&gt;6&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/HellrichH16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=431654030027025291&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Fahad Khan, Francesca Frontini, Federico Boschetti, Monica Monachini: &lt;br /&gt;Converting the Liddell Scott Greek-English Lexicon into Linked Open Data using lemon – Short Paper – &lt;strong&gt;6&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/KhanFBM16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=17882901537315793857&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Federico Nanni, Pablo Ruiz Fabo: &lt;br /&gt;Entities as topic labels: improving topic interpretability and evaluability combining Entity Linking and Labeled LDA – Short Paper – &lt;strong&gt;6&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/NanniF16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=6365131429834412760&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Susan Brown, Tanya E. Clement, Laura Mandell, Deb Verhoeven, Jacque Wernimont: &lt;br /&gt;Creating Feminist Infrastructure in the Digital Humanities – Panel – &lt;strong&gt;5&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/BrownCMVW16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=1060823707180841281&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Alexander Czmiel: &lt;br /&gt;Sustainable publishing - Standardization possibilities for Digital Scholarly Edition technology – Long Paper – &lt;strong&gt;5&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/Czmiel16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=1358527880657138486&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Isabella di Lenardo, Benoit Seguin, Frédéric Kaplan: &lt;br /&gt;Visual Patterns Discovery in Large Databases of Paintings – Long Paper – &lt;strong&gt;5&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/LenardoSK16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=8651297274583880104&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Francesca Frontini, Carmen Brando, Jean-Gabriel Ganascia: &lt;br /&gt;REDEN ONLINE: Disambiguation, Linking and Visualisation of References in TEI Digital Editions – Long Paper – &lt;strong&gt;5&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/FrontiniBG16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=11539777205701010405&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Kim Jautze, Andreas van Cranenburgh, Corina Koolen: &lt;br /&gt;Topic Modeling Literary Quality – Long Paper – &lt;strong&gt;5&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/JautzeCK16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=17014469150281250445&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Mikko Tolonen, Niko Ilomäki, Hege Roivainen, Leo Lahti: &lt;br /&gt;Printing in a Periphery: a Quantitative Study of Finnish Knowledge Production, 1640-1828 – Long Paper – &lt;strong&gt;5&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/TolonenIRL16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=2912387019136768846&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Peter Robinson, Barbara Bordalejo: &lt;br /&gt;Textual Communities – Poster – &lt;strong&gt;5&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/0003B16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=5322581207885764479&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Alexander Dunst, Rita Hartel, Sven Hohenstein, Jochen Laubrock: &lt;br /&gt;Corpus Analyses of Multimodal Narrative: The Example of Graphic Novels – Long Paper – &lt;strong&gt;4&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/DunstHHL16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=13172547380672638477&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Maciej Eder, Jan Rybicki: &lt;br /&gt;Go Set A Watchman while we Kill the Mockingbird in Cold Blood, with Cats and Other People – Long Paper – &lt;strong&gt;4&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/EderR16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=5596591169426158931&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Chao-Lin Liu: &lt;br /&gt;Quantitative Analyses of Chinese Poetry of Tang and Song Dynasties: Using Changing Colors and Innovative Terms as Examples – Long Paper – &lt;strong&gt;4&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/Liu16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=12111986076156693804&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Alexandre Rigal, Dario Rodighiero, Loup Cellard: &lt;br /&gt;The Trajectories Tool: Amplifying Network Visualization Complexity – Long Paper – &lt;strong&gt;4&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/RigalRC16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=9763356290816258798&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Valérie Beaudouin, Zeynep Pehlivan: &lt;br /&gt;The Great War on the Web: the Making of Citing and Referencing by Amateurs – Short Paper – &lt;strong&gt;4&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/BeaudouinP16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=5116065516112527925&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Angelo Mario Del Grosso, Davide Albanesi, Emiliano Giovannetti, Simone Marchi: &lt;br /&gt;Defining the Core Entities of an Environment for Textual Processing in Literary Computing – Poster – &lt;strong&gt;4&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/GrossoAGM16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=17440349976870797986&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Martin Reckziegel, Stefan Jänicke, Gerik Scheuermann: &lt;br /&gt;CTRaCE: Canonical Text Reader and Citation Exporter – Poster – &lt;strong&gt;4&lt;/strong&gt; citations &lt;br /&gt;(&lt;a href=&quot;https://dblp.org/rec/conf/dihu/ReckziegelJS16&quot;&gt;DBLP&lt;/a&gt;) (&lt;a href=&quot;https://scholar.google.com/scholar?cluster=1700136619339017477&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;2-citation-counts-for-follow-up-articles-published-in-the-dsh-special-conference-issue&quot;&gt;2. Citation counts for follow-up articles published in the DSH special conference issue&lt;/h2&gt;

&lt;p&gt;EADH’s journal “Digital Scholarship in the Humanities” (DSH) published a special edition with 15 selected articles from the DH2016 conference (&lt;a href=&quot;https://academic.oup.com/dsh/issue/32/suppl_2&quot;&gt;Volume 32, Issue suppl_2, December 2017&lt;/a&gt;). Here are their citation counts (titles sometimes differ from the original abstracts):&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;Stefan Evert, Thomas Proisl, Fotis Jannidis, Isabella Reger, Steffen Pielström, Christof Schöch, Thorsten Vitt: &lt;br /&gt;&lt;a href=&quot;https://doi.org/10.1093/llc/fqx023&quot;&gt;Understanding and explaining Delta measures for authorship attribution&lt;/a&gt; – &lt;strong&gt;22&lt;/strong&gt; citations (original abstract: 2) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=5709367929062932294&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Heather Richards-Rissetto: &lt;br /&gt;&lt;a href=&quot;https://doi.org/10.1093/llc/fqx014&quot;&gt;An iterative 3D GIS analysis of the role of visibility in ancient Maya landscapes: A case study from Copan, Honduras&lt;/a&gt; – &lt;strong&gt;10&lt;/strong&gt; (2) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=2176558170048823529&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Stefan Jänicke, David Joseph Wrisley: &lt;br /&gt;&lt;a href=&quot;https://doi.org/10.1093/llc/fqx033&quot;&gt;Visualizing Mouvance: Toward a visual analysis of variant medieval text traditions&lt;/a&gt; – &lt;strong&gt;9&lt;/strong&gt; (3) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=7973381259497772230&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Grace Muzny, Mark Algee-Hewitt, Dan Jurafsky: &lt;br /&gt;&lt;a href=&quot;https://doi.org/10.1093/llc/fqx031&quot;&gt;Dialogism in the novel: A computational model of the dialogic nature of narration and quotations&lt;/a&gt; – &lt;strong&gt;8&lt;/strong&gt; (2) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=5269169457931415184&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Elena González-Blanco, Clara Martínez Cantón, Gimena del Rio Riande, Salvador Ros, Rafael Pastor, Antonio Robles-Gómez, Agustín Caminero, María Luisa Díez Platas, Álvaro del Olmo, Miguel Urízar: &lt;br /&gt;&lt;a href=&quot;https://doi.org/10.1093/llc/fqx025&quot;&gt;EVI-LINHD, a virtual research environment for the Spanish-speaking community&lt;/a&gt; – &lt;strong&gt;5&lt;/strong&gt; (2) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=4672347431723511486&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Arianna Ciula: &lt;br /&gt;&lt;a href=&quot;https://doi.org/10.1093/llc/fqx042&quot;&gt;Digital palaeography: What is digital about it?&lt;/a&gt; – &lt;strong&gt;4&lt;/strong&gt; (0) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=1251668675601027055&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Katarzyna Bazarnik, Jakub Wróblewski: &lt;br /&gt;&lt;a href=&quot;https://doi.org/10.1093/llc/fqx027&quot;&gt;First We Feel Then We Fall: James Joyce’s Finnegans Wake as an interactive video application&lt;/a&gt; – &lt;strong&gt;4&lt;/strong&gt; (0) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=3232051797453798596&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;David L Hoover: &lt;br /&gt;&lt;a href=&quot;https://doi.org/10.1093/llc/fqx022&quot;&gt;The microanalysis of style variation&lt;/a&gt; – &lt;strong&gt;3&lt;/strong&gt; (0) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=6780436031879293349&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;James O Gawley, A Caitlin Diddams: &lt;br /&gt;&lt;a href=&quot;https://doi.org/10.1093/llc/fqx038&quot;&gt;Comparing the intertextuality of multiple authors using Tesserae: A new technique for normalization&lt;/a&gt; – &lt;strong&gt;3&lt;/strong&gt; (0) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=13830734924028487386&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Marine Riguet, Suzanne Mpouli: &lt;br /&gt;&lt;a href=&quot;https://doi.org/10.1093/llc/fqx026&quot;&gt;At the crossroads between the scientific and the literary discourse: Comparison as a figure of dialogism&lt;/a&gt; – &lt;strong&gt;3&lt;/strong&gt; (3) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=4270286382363276284&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Martijn Kleppe, Marco Otte: &lt;br /&gt;&lt;a href=&quot;https://doi.org/10.1093/llc/fqx030&quot;&gt;Analysing and understanding news consumption patterns by tracking online user behaviour with a multimodal research design&lt;/a&gt; – &lt;strong&gt;3&lt;/strong&gt; (0) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=3515976435227019090&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Claire Warwick: &lt;br /&gt;&lt;a href=&quot;https://doi.org/10.1093/llc/fqx036&quot;&gt;Beauty is truth: Multi-sensory input and the challenge of designing aesthetically pleasing digital resources&lt;/a&gt; – &lt;strong&gt;1&lt;/strong&gt; (0) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=7794905959483244893&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Taylor Arnold, Peter Leonard, Lauren Tilton: &lt;br /&gt;&lt;a href=&quot;https://doi.org/10.1093/llc/fqx035&quot;&gt;Knowledge creation through recommender systems&lt;/a&gt; – &lt;strong&gt;1&lt;/strong&gt; (0) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=18225641954619045804&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Rui Hu, Carlos Pallán Gayol, Jean-Marc Odobez, Daniel Gatica-Perez: &lt;br /&gt;&lt;a href=&quot;https://doi.org/10.1093/llc/fqx028&quot;&gt;Analyzing and visualizing ancient Maya hieroglyphics using shape: From computer vision to Digital Humanities&lt;/a&gt; – &lt;strong&gt;1&lt;/strong&gt; (2) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=16064761551397513000&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Joris J van Zundert, Tara L Andrews: &lt;br /&gt;&lt;a href=&quot;https://doi.org/10.1093/llc/fqx039&quot;&gt;Qu’est-ce qu’un texte numérique?—A new rationale for the digital representation of text&lt;/a&gt; – &lt;strong&gt;0&lt;/strong&gt; (0) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=17718895804126846238&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;3-articles-published-elsewhere&quot;&gt;3. Articles published elsewhere&lt;/h2&gt;

&lt;p&gt;As it is not always possible to say whether a full paper is really the elaborated version of a conference abstract, we can’t provide an exhaustive list. We only list here the three papers that go by the same titles as the original abstracts:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;Melissa Terras, James Baker, James Hetherington, David Beavan, Martin Zaltz Austwick, Anne Welsh, Helen O’Neill, Will Finley, Oliver Duke-Williams, Adam Farquhar: &lt;br /&gt;&lt;a href=&quot;https://doi.org/10.1093/llc/fqx020&quot;&gt;Enabling complex analysis of large-scale digital collections: humanities research, high-performance computing, and transforming access to British Library digital collections&lt;/a&gt; (DSH 33.2, June 2018) – &lt;strong&gt;12&lt;/strong&gt; citations (original abstract: 0) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=11682747324791540734&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Yuta Hashimoto, Yoichi Iikura, Yukio Hisada, SungKook Kang, Tomoyo Arisawa, Daniel Kobayashi-Better: &lt;br /&gt;&lt;a href=&quot;http://www.digitalhumanities.org/dhq/vol/11/1/000281/000281.html&quot;&gt;The Kuzushiji Project: Developing a Mobile Learning Application for Reading Early Modern Japanese Books&lt;/a&gt; (DHQ 11.1, 2017) – &lt;strong&gt;8&lt;/strong&gt; (0) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=13178070444385478730&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Pelle Snickars, Roger Mähler: &lt;br /&gt;&lt;a href=&quot;http://www.digitalhumanities.org/dhq/vol/12/1/000373/000373.html&quot;&gt;SpotiBot-Turing testing Spotify&lt;/a&gt; (DHQ 12.1, 2018) – &lt;strong&gt;7&lt;/strong&gt; (0) &lt;br /&gt;(&lt;a href=&quot;https://scholar.google.com/scholar?cluster=6199034850415843370&quot;&gt;Google Scholar&lt;/a&gt;)&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;some-thoughts&quot;&gt;Some Thoughts&lt;/h2&gt;

&lt;ul&gt;
  &lt;li&gt;ADHO conference abstracts do have a measurable scientific impact in the community.&lt;/li&gt;
  &lt;li&gt;Publishing a reworked conference abstract in a journal can (but doesn’t have to) increase your impact.&lt;/li&gt;
  &lt;li&gt;It would be nice to have a more stable, sustainable, citable format for ADHO conference proceedings (think DOIs maybe?).&lt;/li&gt;
  &lt;li&gt;Case in point: proceedings of DH2015 in Sydney seem to have vanished into the internet ether, there are only a bunch of unrendered XML files left somewhere, not really citable.&lt;/li&gt;
  &lt;li&gt;DOIs for books of abstracts are already a step in the right direction (see, e.g., DOI:&lt;a href=&quot;https://doi.org/10.5281/zenodo.2596095&quot;&gt;10.5281/zenodo.2596095&lt;/a&gt; for DHd2019).&lt;/li&gt;
  &lt;li&gt;Working with Google Scholar is still tedious. There are some Python libraries that are supposed to make it easier to work with it, but they don’t cover the full range of Google Scholar functions, and the proprietary mechanism can change at any time without warning and render these libraries unusable.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Anyway, so much for this little experiment. It’s a beautiful day in Moscow, time to visit the ice rink in Gorky Park! ❄️&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>DH Tools Mentioned in "The Programming Historian"</title>
    <link href="https://weltliteratur.net/dh-tools-programming-historian/"/>
    <updated>2020-01-17T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/dh-tools-programming-historian</id>
    <content type="html">&lt;p&gt;This is a little follow-up to our blog post &lt;a href=&quot;/dh-tools-used-in-research/&quot;&gt;&lt;strong&gt;“Which DH Tools Are Actually Used in Research?”&lt;/strong&gt;&lt;/a&gt; – The idea is to get a general overview of tool mentionings in the fabulous lectures of &lt;a href=&quot;https://programminghistorian.org/&quot;&gt;&lt;strong&gt;“The Programming Historian”&lt;/strong&gt;&lt;/a&gt; (PH). We again used our &lt;a href=&quot;https://github.com/lehkost/ToolXtractor/&quot;&gt;ToolXtractor&lt;/a&gt; to extract tool names from PH’s 87 (English-language) lectures to date (including 7 retired ones), directly from &lt;a href=&quot;https://github.com/programminghistorian/jekyll/tree/gh-pages/en/lessons&quot;&gt;their Markdown source files&lt;/a&gt;. Like last time, we used &lt;a href=&quot;http://tapor.ca/home&quot;&gt;TAPoR’s&lt;/a&gt; database as positive list (for ToolXtractor relies on simple string matching). Any tool we failed to extract was not in TAPoR (yet), e.g., “Unity”, “Notepad++”, or “Atom” – an obvious shortcoming of this approach.&lt;/p&gt;

&lt;p&gt;We also created a co-occurence graph (tools mentioned in the same lesson):&lt;/p&gt;

&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/proghist-co-occurrence-graph.png&quot; alt=&quot;co-occurrence graph for tools mentioned in ProgHist lessons&quot; style=&quot;width:1024px;&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;You can spot Python, R, Excel, GitHub and Zotero at the core of teaching in the Digital Humanities. Publishing tools are grouped on the left, network tools on the upper right side. GIS tools are on the bottom right-hand side, text analysis tools are located right above the centre, etc. (The graph was generated in a heartbeat with our tool &lt;a href=&quot;https://ezlinavis.dracor.org&quot;&gt;ezlinavis&lt;/a&gt; and embellished with Gephi.)&lt;/p&gt;

&lt;p&gt;No further analysis, just this overview:&lt;/p&gt;

&lt;h2 id=&quot;jupyter-notebooks-11&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/jupyter-notebooks&quot;&gt;jupyter-notebooks&lt;/a&gt; (11)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;GitHub&lt;/li&gt;
  &lt;li&gt;Google Drive&lt;/li&gt;
  &lt;li&gt;JavaScript&lt;/li&gt;
  &lt;li&gt;Omeka&lt;/li&gt;
  &lt;li&gt;OpenRefine&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;R&lt;/li&gt;
  &lt;li&gt;Ruby&lt;/li&gt;
  &lt;li&gt;Twitter&lt;/li&gt;
  &lt;li&gt;Voyant Tools&lt;/li&gt;
  &lt;li&gt;Zenodo&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;building-static-sites-with-jekyll-github-pages-9&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/building-static-sites-with-jekyll-github-pages&quot;&gt;building-static-sites-with-jekyll-github-pages&lt;/a&gt; (9)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Blogger&lt;/li&gt;
  &lt;li&gt;Drupal&lt;/li&gt;
  &lt;li&gt;Jekyll&lt;/li&gt;
  &lt;li&gt;Omeka&lt;/li&gt;
  &lt;li&gt;Ruby&lt;/li&gt;
  &lt;li&gt;Slack&lt;/li&gt;
  &lt;li&gt;Tumblr&lt;/li&gt;
  &lt;li&gt;Twitter&lt;/li&gt;
  &lt;li&gt;WordPress&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;mapping-with-python-leaflet-9&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/mapping-with-python-leaflet&quot;&gt;mapping-with-python-leaflet&lt;/a&gt; (9)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;ArcGIS&lt;/li&gt;
  &lt;li&gt;C&lt;/li&gt;
  &lt;li&gt;GeoNames&lt;/li&gt;
  &lt;li&gt;GitHub&lt;/li&gt;
  &lt;li&gt;Google Maps&lt;/li&gt;
  &lt;li&gt;Leaflet&lt;/li&gt;
  &lt;li&gt;Nominatim&lt;/li&gt;
  &lt;li&gt;OpenStreetMap&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;using-javascript-to-create-maps-9&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/using-javascript-to-create-maps&quot;&gt;using-javascript-to-create-maps&lt;/a&gt; (9)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;Gephi&lt;/li&gt;
  &lt;li&gt;Google Maps&lt;/li&gt;
  &lt;li&gt;GPS Visualizer&lt;/li&gt;
  &lt;li&gt;JavaScript&lt;/li&gt;
  &lt;li&gt;Leaflet&lt;/li&gt;
  &lt;li&gt;MapBox&lt;/li&gt;
  &lt;li&gt;Palladio&lt;/li&gt;
  &lt;li&gt;Zotero&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;corpus-analysis-with-antconc-8&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/corpus-analysis-with-antconc&quot;&gt;corpus-analysis-with-antconc&lt;/a&gt; (8)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;AntConc&lt;/li&gt;
  &lt;li&gt;Beautiful Soup&lt;/li&gt;
  &lt;li&gt;KWIC&lt;/li&gt;
  &lt;li&gt;Natural Language Toolkit · NLTK&lt;/li&gt;
  &lt;li&gt;OpenRefine&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;R&lt;/li&gt;
  &lt;li&gt;Voyant Tools&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;creating-network-diagrams-from-historical-sources-8&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/creating-network-diagrams-from-historical-sources&quot;&gt;creating-network-diagrams-from-historical-sources&lt;/a&gt; (8)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Gephi&lt;/li&gt;
  &lt;li&gt;nodegoat&lt;/li&gt;
  &lt;li&gt;NodeXL&lt;/li&gt;
  &lt;li&gt;Pajek&lt;/li&gt;
  &lt;li&gt;Palladio&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;UCINET&lt;/li&gt;
  &lt;li&gt;VennMaker&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;geoparsing-text-with-edinburgh-7&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/geoparsing-text-with-edinburgh&quot;&gt;geoparsing-text-with-edinburgh&lt;/a&gt; (7)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;Firefox&lt;/li&gt;
  &lt;li&gt;GapVis&lt;/li&gt;
  &lt;li&gt;GeoNames&lt;/li&gt;
  &lt;li&gt;Google Maps&lt;/li&gt;
  &lt;li&gt;QGIS&lt;/li&gt;
  &lt;li&gt;Twitter&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;text-mining-with-extracted-features-7&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/text-mining-with-extracted-features&quot;&gt;text-mining-with-extracted-features&lt;/a&gt; (7)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;HT-Bookworm&lt;/li&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;GitHub&lt;/li&gt;
  &lt;li&gt;HathiTrust&lt;/li&gt;
  &lt;li&gt;Matlab&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;R&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;beginners-guide-to-twitter-data-6&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/beginners-guide-to-twitter-data&quot;&gt;beginners-guide-to-twitter-data&lt;/a&gt; (6)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Cytoscape&lt;/li&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;Gephi&lt;/li&gt;
  &lt;li&gt;Palladio&lt;/li&gt;
  &lt;li&gt;Tableau&lt;/li&gt;
  &lt;li&gt;Twitter&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;creating-apis-with-python-and-flask-6&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/creating-apis-with-python-and-flask&quot;&gt;creating-apis-with-python-and-flask&lt;/a&gt; (6)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;BBEdit&lt;/li&gt;
  &lt;li&gt;Firefox&lt;/li&gt;
  &lt;li&gt;JavaScript&lt;/li&gt;
  &lt;li&gt;MediaWiki&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;Twitter&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;introduction-to-ffmpeg-6&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/introduction-to-ffmpeg&quot;&gt;introduction-to-ffmpeg&lt;/a&gt; (6)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Audacity&lt;/li&gt;
  &lt;li&gt;Chrome&lt;/li&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;plot.ly&lt;/li&gt;
  &lt;li&gt;RAW&lt;/li&gt;
  &lt;li&gt;Twitter&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;sonification-6&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/sonification&quot;&gt;sonification&lt;/a&gt; (6)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;GitHub&lt;/li&gt;
  &lt;li&gt;Pandoc&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;R&lt;/li&gt;
  &lt;li&gt;Ruby&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;visualizing-with-bokeh-6&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/visualizing-with-bokeh&quot;&gt;visualizing-with-bokeh&lt;/a&gt; (6)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Bokeh&lt;/li&gt;
  &lt;li&gt;CartoDB&lt;/li&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;Leaflet&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;QGIS&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;exploring-and-analyzing-network-data-with-python-5&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/exploring-and-analyzing-network-data-with-python&quot;&gt;exploring-and-analyzing-network-data-with-python&lt;/a&gt; (5)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Gephi&lt;/li&gt;
  &lt;li&gt;NetworkX&lt;/li&gt;
  &lt;li&gt;Palladio&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;R&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;geocoding-qgis-5&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/geocoding-qgis&quot;&gt;geocoding-qgis&lt;/a&gt; (5)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;ArcGIS&lt;/li&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;Google Maps&lt;/li&gt;
  &lt;li&gt;OpenStreetMap&lt;/li&gt;
  &lt;li&gt;QGIS&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;getting-started-with-github-desktop-retired-5&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/retired/getting-started-with-github-desktop&quot;&gt;getting-started-with-github-desktop&lt;/a&gt; (retired) (5)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Dropbox&lt;/li&gt;
  &lt;li&gt;Google Drive&lt;/li&gt;
  &lt;li&gt;Jekyll&lt;/li&gt;
  &lt;li&gt;Pandoc&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;getting-started-with-mysql-using-r-5&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/getting-started-with-mysql-using-r&quot;&gt;getting-started-with-mysql-using-r&lt;/a&gt; (5)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;MySQL&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;R&lt;/li&gt;
  &lt;li&gt;RStudio&lt;/li&gt;
  &lt;li&gt;SQL Server&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;googlemaps-googleearth-5&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/googlemaps-googleearth&quot;&gt;googlemaps-googleearth&lt;/a&gt; (5)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;ArcGIS&lt;/li&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;Google Maps&lt;/li&gt;
  &lt;li&gt;QGIS&lt;/li&gt;
  &lt;li&gt;Twitter&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;introduction-to-populating-a-website-with-api-data-5&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/introduction-to-populating-a-website-with-api-data&quot;&gt;introduction-to-populating-a-website-with-api-data&lt;/a&gt; (5)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Chrome&lt;/li&gt;
  &lt;li&gt;Firefox&lt;/li&gt;
  &lt;li&gt;GeoNames&lt;/li&gt;
  &lt;li&gt;JavaScript&lt;/li&gt;
  &lt;li&gt;Skype&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;introduction-to-stylometry-with-python-5&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/introduction-to-stylometry-with-python&quot;&gt;introduction-to-stylometry-with-python&lt;/a&gt; (5)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Natural Language Toolkit · NLTK&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;R&lt;/li&gt;
  &lt;li&gt;stylo&lt;/li&gt;
  &lt;li&gt;Zotero&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;json-and-jq-5&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/json-and-jq&quot;&gt;json-and-jq&lt;/a&gt; (5)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;JavaScript&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;R&lt;/li&gt;
  &lt;li&gt;Twitter&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;qgis-layers-5&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/qgis-layers&quot;&gt;qgis-layers&lt;/a&gt; (5)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;ArcGIS&lt;/li&gt;
  &lt;li&gt;GDAL&lt;/li&gt;
  &lt;li&gt;OpenLayers&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;QGIS&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;sustainable-authorship-in-plain-text-using-pandoc-and-markdown-5&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/sustainable-authorship-in-plain-text-using-pandoc-and-markdown&quot;&gt;sustainable-authorship-in-plain-text-using-pandoc-and-markdown&lt;/a&gt; (5)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Google Docs&lt;/li&gt;
  &lt;li&gt;Jekyll&lt;/li&gt;
  &lt;li&gt;Pandoc&lt;/li&gt;
  &lt;li&gt;WordPress&lt;/li&gt;
  &lt;li&gt;Zotero&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;temporal-network-analysis-with-r-5&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/temporal-network-analysis-with-r&quot;&gt;temporal-network-analysis-with-r&lt;/a&gt; (5)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Gephi&lt;/li&gt;
  &lt;li&gt;NetworkX&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;R&lt;/li&gt;
  &lt;li&gt;RStudio&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;analyzing-documents-with-tfidf-4&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/analyzing-documents-with-tfidf&quot;&gt;analyzing-documents-with-tfidf&lt;/a&gt; (4)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Bokeh&lt;/li&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;Overview&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;extracting-illustrated-pages-4&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/extracting-illustrated-pages&quot;&gt;extracting-illustrated-pages&lt;/a&gt; (4)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;HathiTrust&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;R&lt;/li&gt;
  &lt;li&gt;Tesseract&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;extracting-keywords-4&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/extracting-keywords&quot;&gt;extracting-keywords&lt;/a&gt; (4)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;OpenRefine&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;QGIS&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;fetch-and-parse-data-with-openrefine-4&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/fetch-and-parse-data-with-openrefine&quot;&gt;fetch-and-parse-data-with-openrefine&lt;/a&gt; (4)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;Natural Language Toolkit · NLTK&lt;/li&gt;
  &lt;li&gt;OpenRefine&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;graph-databases-and-sparql-retired-4&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/retired/graph-databases-and-SPARQL&quot;&gt;graph-databases-and-SPARQL&lt;/a&gt; (retired) (4)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;GeoNames&lt;/li&gt;
  &lt;li&gt;Palladio&lt;/li&gt;
  &lt;li&gt;plot.ly&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;intro-to-twitterbots-4&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/intro-to-twitterbots&quot;&gt;intro-to-twitterbots&lt;/a&gt; (4)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;GitHub&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;Slack&lt;/li&gt;
  &lt;li&gt;Twitter&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;keywords-in-context-using-n-grams-4&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/keywords-in-context-using-n-grams&quot;&gt;keywords-in-context-using-n-grams&lt;/a&gt; (4)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Firefox&lt;/li&gt;
  &lt;li&gt;KWIC&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;Zotero&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;sentiment-analysis-4&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/sentiment-analysis&quot;&gt;sentiment-analysis&lt;/a&gt; (4)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;GitHub&lt;/li&gt;
  &lt;li&gt;Natural Language Toolkit · NLTK&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;Twitter&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;transforming-xml-with-xsl-4&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/transforming-xml-with-xsl&quot;&gt;transforming-xml-with-xsl&lt;/a&gt; (4)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Chrome&lt;/li&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;Firefox&lt;/li&gt;
  &lt;li&gt;GitHub&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;counting-frequencies-from-zotero-items-retired-3&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/retired/counting-frequencies-from-zotero-items&quot;&gt;counting-frequencies-from-zotero-items&lt;/a&gt; (retired) (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Google Books&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;Zotero&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;creating-and-viewing-html-files-with-python-3&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/creating-and-viewing-html-files-with-python&quot;&gt;creating-and-viewing-html-files-with-python&lt;/a&gt; (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Firefox&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;Zotero&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;data-mining-the-internet-archive-3&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/data-mining-the-internet-archive&quot;&gt;data-mining-the-internet-archive&lt;/a&gt; (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Google Maps&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;Wordle&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;data_wrangling_and_management_in_r-3&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/data_wrangling_and_management_in_R&quot;&gt;data_wrangling_and_management_in_R&lt;/a&gt; (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;R&lt;/li&gt;
  &lt;li&gt;RStudio&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;generating-an-ordered-data-set-from-an-ocr-text-file-3&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/generating-an-ordered-data-set-from-an-OCR-text-file&quot;&gt;generating-an-ordered-data-set-from-an-OCR-text-file&lt;/a&gt; (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;JavaScript&lt;/li&gt;
  &lt;li&gt;Perl&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;georeferencing-qgis-3&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/georeferencing-qgis&quot;&gt;georeferencing-qgis&lt;/a&gt; (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;GDAL&lt;/li&gt;
  &lt;li&gt;Google Maps&lt;/li&gt;
  &lt;li&gt;QGIS&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;geospatial-data-analysis-3&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/geospatial-data-analysis&quot;&gt;geospatial-data-analysis&lt;/a&gt; (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Google Maps&lt;/li&gt;
  &lt;li&gt;plot.ly&lt;/li&gt;
  &lt;li&gt;R&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;gravity-model-3&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/gravity-model&quot;&gt;gravity-model&lt;/a&gt; (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;R&lt;/li&gt;
  &lt;li&gt;RStudio&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;installing-omeka-3&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/installing-omeka&quot;&gt;installing-omeka&lt;/a&gt; (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;MySQL&lt;/li&gt;
  &lt;li&gt;Omeka&lt;/li&gt;
  &lt;li&gt;WordPress&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;intro-to-powershell-3&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/intro-to-powershell&quot;&gt;intro-to-powershell&lt;/a&gt; (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Mallet&lt;/li&gt;
  &lt;li&gt;Pandoc&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;introduction-and-installation-3&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/introduction-and-installation&quot;&gt;introduction-and-installation&lt;/a&gt; (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Beautiful Soup&lt;/li&gt;
  &lt;li&gt;Dropbox&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;output-data-as-html-file-3&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/output-data-as-html-file&quot;&gt;output-data-as-html-file&lt;/a&gt; (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Firefox&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;Zotero&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;output-keywords-in-context-in-html-file-3&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/output-keywords-in-context-in-html-file&quot;&gt;output-keywords-in-context-in-html-file&lt;/a&gt; (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Firefox&lt;/li&gt;
  &lt;li&gt;KWIC&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;topic-modeling-and-mallet-3&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/topic-modeling-and-mallet&quot;&gt;topic-modeling-and-mallet&lt;/a&gt; (3)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;Mallet&lt;/li&gt;
  &lt;li&gt;Voyant Tools&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;automated-downloading-with-wget-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/automated-downloading-with-wget&quot;&gt;automated-downloading-with-wget&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;Ruby&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;basic-text-processing-in-r-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/basic-text-processing-in-r&quot;&gt;basic-text-processing-in-r&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;R&lt;/li&gt;
  &lt;li&gt;RStudio&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;cleaning-data-with-openrefine-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/cleaning-data-with-openrefine&quot;&gt;cleaning-data-with-openrefine&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;OpenRefine&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;cleaning-ocrd-text-with-regular-expressions-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/cleaning-ocrd-text-with-regular-expressions&quot;&gt;cleaning-ocrd-text-with-regular-expressions&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;correspondence-analysis-in-r-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/correspondence-analysis-in-R&quot;&gt;correspondence-analysis-in-R&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;R&lt;/li&gt;
  &lt;li&gt;Zenodo&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;creating-new-items-in-zotero-retired-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/retired/creating-new-items-in-zotero&quot;&gt;creating-new-items-in-zotero&lt;/a&gt; (retired) (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;Zotero&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;dealing-with-big-data-and-network-analysis-using-neo4j-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/dealing-with-big-data-and-network-analysis-using-neo4j&quot;&gt;dealing-with-big-data-and-network-analysis-using-neo4j&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;editing-audio-with-audacity-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/editing-audio-with-audacity&quot;&gt;editing-audio-with-audacity&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Audacity&lt;/li&gt;
  &lt;li&gt;Soundflower&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;getting-started-with-markdown-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/getting-started-with-markdown&quot;&gt;getting-started-with-markdown&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Pandoc&lt;/li&gt;
  &lt;li&gt;Perl&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;intro-to-beautiful-soup-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/intro-to-beautiful-soup&quot;&gt;intro-to-beautiful-soup&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Beautiful Soup&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;intro-to-the-zotero-api-retired-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/retired/intro-to-the-zotero-api&quot;&gt;intro-to-the-zotero-api&lt;/a&gt; (retired) (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;Zotero&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;mac-installation-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/mac-installation&quot;&gt;mac-installation&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Beautiful Soup&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;naive-bayesian-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/naive-bayesian&quot;&gt;naive-bayesian&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Beautiful Soup&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;ocr-with-tesseract-and-scantailor-retired-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/retired/OCR-with-Tesseract-and-ScanTailor&quot;&gt;OCR-with-Tesseract-and-ScanTailor&lt;/a&gt; (retired) (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Tesseract&lt;/li&gt;
  &lt;li&gt;Zotero&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;preserving-your-research-data-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/preserving-your-research-data&quot;&gt;preserving-your-research-data&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;WordPress&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;r-basics-with-tabular-data-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/r-basics-with-tabular-data&quot;&gt;r-basics-with-tabular-data&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Excel&lt;/li&gt;
  &lt;li&gt;R&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;transliterating-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/transliterating&quot;&gt;transliterating&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Beautiful Soup&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;understanding-regular-expressions-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/understanding-regular-expressions&quot;&gt;understanding-regular-expressions&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
  &lt;li&gt;Ruby&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;vector-layers-qgis-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/vector-layers-qgis&quot;&gt;vector-layers-qgis&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Google Maps&lt;/li&gt;
  &lt;li&gt;QGIS&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;working-with-web-pages-2&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/working-with-web-pages&quot;&gt;working-with-web-pages&lt;/a&gt; (2)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Firefox&lt;/li&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;applied-archival-downloading-with-wget-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/applied-archival-downloading-with-wget&quot;&gt;applied-archival-downloading-with-wget&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;code-reuse-and-modularity-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/code-reuse-and-modularity&quot;&gt;code-reuse-and-modularity&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;counting-frequencies-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/counting-frequencies&quot;&gt;counting-frequencies&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;creating-an-omeka-exhibit-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/creating-an-omeka-exhibit&quot;&gt;creating-an-omeka-exhibit&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Omeka&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;creating-mobile-augmented-reality-experiences-in-unity-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/creating-mobile-augmented-reality-experiences-in-unity&quot;&gt;creating-mobile-augmented-reality-experiences-in-unity&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;C&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;downloading-multiple-records-using-query-strings-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/downloading-multiple-records-using-query-strings&quot;&gt;downloading-multiple-records-using-query-strings&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;from-html-to-list-of-words-1-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/from-html-to-list-of-words-1&quot;&gt;from-html-to-list-of-words-1&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;from-html-to-list-of-words-2-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/from-html-to-list-of-words-2&quot;&gt;from-html-to-list-of-words-2&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;installing-python-modules-pip-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/installing-python-modules-pip&quot;&gt;installing-python-modules-pip&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;intro-to-bash-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/intro-to-bash&quot;&gt;intro-to-bash&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Pandoc&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;intro-to-linked-data-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/intro-to-linked-data&quot;&gt;intro-to-linked-data&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;GeoNames&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;linux-installation-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/linux-installation&quot;&gt;linux-installation&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;manipulating-strings-in-python-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/manipulating-strings-in-python&quot;&gt;manipulating-strings-in-python&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;normalizing-data-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/normalizing-data&quot;&gt;normalizing-data&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;research-data-with-unix-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/research-data-with-unix&quot;&gt;research-data-with-unix&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Excel&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;up-and-running-with-omeka-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/up-and-running-with-omeka&quot;&gt;up-and-running-with-omeka&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Omeka&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;viewing-html-files-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/viewing-html-files&quot;&gt;viewing-html-files&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Firefox&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;windows-installation-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/windows-installation&quot;&gt;windows-installation&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;working-with-text-files-1&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/working-with-text-files&quot;&gt;working-with-text-files&lt;/a&gt; (1)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Python&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;intro-to-augmented-reality-with-unity-retired-0&quot;&gt;&lt;a href=&quot;https://programminghistorian.org/en/lessons/retired/intro-to-augmented-reality-with-unity&quot;&gt;intro-to-augmented-reality-with-unity&lt;/a&gt; (retired) (0)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;(none)&lt;/li&gt;
&lt;/ul&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer, 
	
          
          Yoann Moranville
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Which DH Tools Are Actually Used in Research?</title>
    <link href="https://weltliteratur.net/dh-tools-used-in-research/"/>
    <updated>2019-12-06T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/dh-tools-used-in-research</id>
    <content type="html">&lt;p&gt;Not that we didn’t know that, but Gephi is the most popular DH tool actually used in research work. Followed by Omeka, stylo, MALLET, Excel, D3.js, the NLTK, WordPress, Drupal, TextGrid, CollateX, GeoNames, TXM, Solr and Voyant Tools.&lt;/p&gt;

&lt;p&gt;We know this because we did some counting. So let’s explain:&lt;/p&gt;

&lt;p&gt;The longest-standing tool directory in the Digital Humanities, the Canadian portal &lt;a href=&quot;http://tapor.ca/home&quot;&gt;TAPoR&lt;/a&gt; led by Geoffrey Rockwell, has around 1.500 DH tools in its database (including historic ones). We were wondering how this richness of means and utilities in our field is manifested in actual research work. To gain some first insights, we decided to extract the names of tools from TAPoR (which is easy thanks to their API) and match them with the proceedings of the largest and broadest event series in the Digital Humanities, ADHO’s annual DH conferences. The proceedings from DH2015 to DH2019 are &lt;a href=&quot;https://github.com/ADHO/&quot;&gt;freely available&lt;/a&gt; (licensed under CC BY 4.0, thx to Fabio Ciotti for giving us early access to the DH2019 data), so that was our chosen source:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;DH2015, Sydney&lt;/li&gt;
  &lt;li&gt;DH2016, Kraków&lt;/li&gt;
  &lt;li&gt;DH2017, Montréal&lt;/li&gt;
  &lt;li&gt;DH2018, Ciudad de México&lt;/li&gt;
  &lt;li&gt;DH2019, Utrecht&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Altogether, &lt;strong&gt;238 tools&lt;/strong&gt; were mentioned at least once in these five years, and we counted &lt;strong&gt;1.498 mentions&lt;/strong&gt; of tools altogether in all the proceedings. To extract this kind of data, we wrote a little Java tool called &lt;a href=&quot;https://github.com/lehkost/ToolXtractor/&quot;&gt;&lt;strong&gt;ToolXtractor&lt;/strong&gt;&lt;/a&gt;, which you can run yourself with alternative source material if you like and also with different lists of tools. Since this is simple string matching, we had to throw out some false positives manually (not too many, though).&lt;/p&gt;

&lt;p&gt;We set up another website providing a complete list of all mentioned tools including a link to corresponding conference abstracts for verification: &lt;a href=&quot;https://lehkost.github.io/tools-dh-proceedings/index.html&quot;&gt;“Tools mentioned in the proceedings of the annual ADHO conferences (2015–2019)”&lt;/a&gt;&lt;/p&gt;

&lt;p&gt;Please also check the &lt;a href=&quot;#ranking&quot;&gt;&lt;strong&gt;ranking&lt;/strong&gt;&lt;/a&gt; of all tools at the very bottom of this blog post for further insights.&lt;/p&gt;

&lt;h2 id=&quot;40-most-used-tools&quot;&gt;40 Most Used Tools&lt;/h2&gt;

&lt;p&gt;Here are the 40 tools with most mentions (an interactive version is here: &lt;a href=&quot;https://datawrapper.dwcdn.net/L96xg/&quot;&gt;https://datawrapper.dwcdn.net/L96xg/&lt;/a&gt;):&lt;/p&gt;

&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/dh-tools-used-in-research/40-most-used-tools.png&quot; alt=&quot;40 most used tools (via Datawrapper)&quot; style=&quot;width:1000px;&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;(&lt;strong&gt;Ed.&lt;/strong&gt; We know that this line chart visually suggests a connection between tools that isn’t there. We still consider this type of display to be beneficial – especially in the &lt;a href=&quot;https://datawrapper.dwcdn.net/L96xg/&quot;&gt;interactive version&lt;/a&gt; – since you can compare occurrences per tool per year. By hovering over the lines you will see that, e.g., 2016 was the year of &lt;strong&gt;stylo&lt;/strong&gt; and &lt;strong&gt;TextGrid&lt;/strong&gt;, 2017 the year of &lt;strong&gt;Leaflet&lt;/strong&gt;, 2018 the year of &lt;strong&gt;Omeka&lt;/strong&gt; and 2019 the year of &lt;strong&gt;GeoNames&lt;/strong&gt; and &lt;strong&gt;Transkribus&lt;/strong&gt;. – Big thx to Jan Horstmann for bringing this up.)&lt;/p&gt;

&lt;h2 id=&quot;known-issues&quot;&gt;Known Issues&lt;/h2&gt;

&lt;p&gt;Before we continue with more visualisations, some caveats:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;Of course, this is not a complete list of all tools mentioned in the proceedings, only of those tools comprised in the TAPoR database (the vast majority should be covered, though).&lt;/li&gt;
  &lt;li&gt;We decided to leave programming languages like Python or JavaScript in, just like social networks (Twitter) and other things that are not DH tools in a narrower sense. Given that we published the data, it is easy to take them out to get a clearer view at actual DH tools.&lt;/li&gt;
  &lt;li&gt;We had to use a stopword list containing tools with names that are also frequent terms, like &lt;a href=&quot;https://processing.org/&quot;&gt;“Processing”&lt;/a&gt;, and we also left out one-letter programming languages, like R and C. Especially R is quite popular in our community and it would be great to compare it to its biggest contender: Python. But this would be future work. (Stopword list etc. to be found in the &lt;a href=&quot;https://github.com/lehkost/ToolXtractor/&quot;&gt;ToolXtractor repo&lt;/a&gt;.)&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Btw, everyone is welcome to do their own viz, &lt;a href=&quot;https://lehkost.github.io/tools-dh-proceedings/tools-dh-proceedings.csv&quot;&gt;the dataset is freely available (in CSV format)&lt;/a&gt;.&lt;/p&gt;

&lt;h2 id=&quot;streamgraph&quot;&gt;Streamgraph&lt;/h2&gt;

&lt;p&gt;This graph shows all tools and their mentions per year (but rather than staring at the graph below you should check out the interactive version of it featuring dynamic captions: &lt;a href=&quot;https://rpubs.com/Pozdniakov/stream_dh&quot;&gt;https://rpubs.com/Pozdniakov/stream_dh&lt;/a&gt;):&lt;/p&gt;

&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/dh-tools-used-in-research/streamgraph.png&quot; alt=&quot;Streamgraph&quot; style=&quot;width:925px;&quot; /&gt;
&lt;/figure&gt;

&lt;h2 id=&quot;line-chart-and-bar-plot-top-10-tools&quot;&gt;Line Chart and Bar Plot (Top-10 Tools)&lt;/h2&gt;

&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/dh-tools-used-in-research/linechart-top-10-tools.png&quot; alt=&quot;&quot; style=&quot;width:490px;&quot; /&gt;
  &lt;img src=&quot;/images/dh-tools-used-in-research/bar-plot-top-10-tools.png&quot; alt=&quot;&quot; style=&quot;width:490px;&quot; /&gt;
&lt;/figure&gt;

&lt;h2 id=&quot;bar-plot-and-stacked-area-all-tools&quot;&gt;Bar Plot and Stacked Area (All Tools)&lt;/h2&gt;

&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/dh-tools-used-in-research/bar-plot-all-tools.png&quot; alt=&quot;&quot; style=&quot;width:490px;&quot; /&gt;
  &lt;img src=&quot;/images/dh-tools-used-in-research/stacked-area-all-tools.png&quot; alt=&quot;&quot; style=&quot;width:490px;&quot; /&gt;
&lt;/figure&gt;

&lt;h2 id=&quot;ranking&quot;&gt;Ranking&lt;/h2&gt;

&lt;p&gt;All 238 tools mentioned with number of occurrences 2015–2019:&lt;/p&gt;

&lt;ol&gt;
  &lt;li&gt;Python (125)&lt;/li&gt;
  &lt;li&gt;Twitter (82)&lt;/li&gt;
  &lt;li&gt;Gephi (60)&lt;/li&gt;
  &lt;li&gt;JavaScript (59)&lt;/li&gt;
  &lt;li&gt;Omeka (44)&lt;/li&gt;
  &lt;li&gt;GitHub (40)&lt;/li&gt;
  &lt;li&gt;HathiTrust (37)&lt;/li&gt;
  &lt;li&gt;stylo (35)&lt;/li&gt;
  &lt;li&gt;MALLET (33)&lt;/li&gt;
  &lt;li&gt;Google Books (31)&lt;/li&gt;
  &lt;li&gt;Excel (30)&lt;/li&gt;
  &lt;li&gt;MySQL (27)&lt;/li&gt;
  &lt;li&gt;D3.js (23)&lt;/li&gt;
  &lt;li&gt;Natural Language Toolkit · NLTK (23)&lt;/li&gt;
  &lt;li&gt;WordPress (20)&lt;/li&gt;
  &lt;li&gt;Drupal (19)&lt;/li&gt;
  &lt;li&gt;TextGrid (19)&lt;/li&gt;
  &lt;li&gt;CollateX (18)&lt;/li&gt;
  &lt;li&gt;GeoNames (18)&lt;/li&gt;
  &lt;li&gt;TXM (18)&lt;/li&gt;
  &lt;li&gt;Solr (17)&lt;/li&gt;
  &lt;li&gt;Voyant Tools (17)&lt;/li&gt;
  &lt;li&gt;EEBO-TCP (15)&lt;/li&gt;
  &lt;li&gt;Palladio (15)&lt;/li&gt;
  &lt;li&gt;PostgreSQL (15)&lt;/li&gt;
  &lt;li&gt;Bootstrap (14)&lt;/li&gt;
  &lt;li&gt;Leaflet (14)&lt;/li&gt;
  &lt;li&gt;OpenRefine (14)&lt;/li&gt;
  &lt;li&gt;Zotero (14)&lt;/li&gt;
  &lt;li&gt;eXist-db (13)&lt;/li&gt;
  &lt;li&gt;Google Maps (13)&lt;/li&gt;
  &lt;li&gt;Tesseract (13)&lt;/li&gt;
  &lt;li&gt;ArcGIS (12)&lt;/li&gt;
  &lt;li&gt;Music Encoding Initiative (12)&lt;/li&gt;
  &lt;li&gt;Transkribus (12)&lt;/li&gt;
  &lt;li&gt;Carto · CartoDB (11)&lt;/li&gt;
  &lt;li&gt;Juxta (10)&lt;/li&gt;
  &lt;li&gt;KWIC (10)&lt;/li&gt;
  &lt;li&gt;Perl (10)&lt;/li&gt;
  &lt;li&gt;Neatline (9)&lt;/li&gt;
  &lt;li&gt;GAMS (8)&lt;/li&gt;
  &lt;li&gt;Hypothes.is (8)&lt;/li&gt;
  &lt;li&gt;Jekyll (8)&lt;/li&gt;
  &lt;li&gt;MediaWiki (8)&lt;/li&gt;
  &lt;li&gt;Recogito (8)&lt;/li&gt;
  &lt;li&gt;Ruby · Ruby on Rails (8)&lt;/li&gt;
  &lt;li&gt;Tableau · Tableau Public (8)&lt;/li&gt;
  &lt;li&gt;AntConc (7)&lt;/li&gt;
  &lt;li&gt;Chrome (7)&lt;/li&gt;
  &lt;li&gt;Firefox (7)&lt;/li&gt;
  &lt;li&gt;Lucene (7)&lt;/li&gt;
  &lt;li&gt;OpenStreetMap (7)&lt;/li&gt;
  &lt;li&gt;Pundit (7)&lt;/li&gt;
  &lt;li&gt;Tesserae (7)&lt;/li&gt;
  &lt;li&gt;TRACER (7)&lt;/li&gt;
  &lt;li&gt;Zenodo (7)&lt;/li&gt;
  &lt;li&gt;Annotation Studio (6)&lt;/li&gt;
  &lt;li&gt;Cytoscape (6)&lt;/li&gt;
  &lt;li&gt;GATE (6)&lt;/li&gt;
  &lt;li&gt;Lexos (6)&lt;/li&gt;
  &lt;li&gt;Neo4j (6)&lt;/li&gt;
  &lt;li&gt;Photogrammar (6)&lt;/li&gt;
  &lt;li&gt;TEI Boilerplate (6)&lt;/li&gt;
  &lt;li&gt;WebLicht (6)&lt;/li&gt;
  &lt;li&gt;WorldCat (6)&lt;/li&gt;
  &lt;li&gt;CATMA (5)&lt;/li&gt;
  &lt;li&gt;Google Scholar (5)&lt;/li&gt;
  &lt;li&gt;LIWC · Linguistic Inquiry and Word Count (5)&lt;/li&gt;
  &lt;li&gt;QGIS (5)&lt;/li&gt;
  &lt;li&gt;Blacklight (4)&lt;/li&gt;
  &lt;li&gt;Bookworm (4)&lt;/li&gt;
  &lt;li&gt;Dropbox (4)&lt;/li&gt;
  &lt;li&gt;Freebase (4)&lt;/li&gt;
  &lt;li&gt;Google Ngram Viewer (4)&lt;/li&gt;
  &lt;li&gt;JGAAP (4)&lt;/li&gt;
  &lt;li&gt;nodegoat (4)&lt;/li&gt;
  &lt;li&gt;Oral History Metadata Synchronizer (4)&lt;/li&gt;
  &lt;li&gt;Protégé (4)&lt;/li&gt;
  &lt;li&gt;Textual Communities (4)&lt;/li&gt;
  &lt;li&gt;Tumblr (4)&lt;/li&gt;
  &lt;li&gt;VARD · VARD 2 (4)&lt;/li&gt;
  &lt;li&gt;Weka (4)&lt;/li&gt;
  &lt;li&gt;Wordle (4)&lt;/li&gt;
  &lt;li&gt;WordSmith (4)&lt;/li&gt;
  &lt;li&gt;Blender (3)&lt;/li&gt;
  &lt;li&gt;Commons In A Box (3)&lt;/li&gt;
  &lt;li&gt;CQPweb (3)&lt;/li&gt;
  &lt;li&gt;ELAN (3)&lt;/li&gt;
  &lt;li&gt;Google Drive (3)&lt;/li&gt;
  &lt;li&gt;NetworkX (3)&lt;/li&gt;
  &lt;li&gt;NodeXL (3)&lt;/li&gt;
  &lt;li&gt;OxGarage (3)&lt;/li&gt;
  &lt;li&gt;PhiloLine (3)&lt;/li&gt;
  &lt;li&gt;RStudio (3)&lt;/li&gt;
  &lt;li&gt;Skype (3)&lt;/li&gt;
  &lt;li&gt;TILE (3)&lt;/li&gt;
  &lt;li&gt;TileMill (3)&lt;/li&gt;
  &lt;li&gt;Tropy (3)&lt;/li&gt;
  &lt;li&gt;Versioning Machine (3)&lt;/li&gt;
  &lt;li&gt;ABBYY FineReader (2)&lt;/li&gt;
  &lt;li&gt;ANNIS (2)&lt;/li&gt;
  &lt;li&gt;Archive-It (2)&lt;/li&gt;
  &lt;li&gt;arts-humanities.net (2)&lt;/li&gt;
  &lt;li&gt;ATLAS.ti (2)&lt;/li&gt;
  &lt;li&gt;Bokeh (2)&lt;/li&gt;
  &lt;li&gt;Dataverse (2)&lt;/li&gt;
  &lt;li&gt;DiRT Directory (2)&lt;/li&gt;
  &lt;li&gt;Diva.js (2)&lt;/li&gt;
  &lt;li&gt;DSpace (2)&lt;/li&gt;
  &lt;li&gt;EATS (2)&lt;/li&gt;
  &lt;li&gt;ediarum (2)&lt;/li&gt;
  &lt;li&gt;Fedora Commons (2)&lt;/li&gt;
  &lt;li&gt;FromThePage (2)&lt;/li&gt;
  &lt;li&gt;Gamera (2)&lt;/li&gt;
  &lt;li&gt;igraph (2)&lt;/li&gt;
  &lt;li&gt;ImageJ (2)&lt;/li&gt;
  &lt;li&gt;Inkscape (2)&lt;/li&gt;
  &lt;li&gt;Islandora (2)&lt;/li&gt;
  &lt;li&gt;Joomla (2)&lt;/li&gt;
  &lt;li&gt;Map Warper (2)&lt;/li&gt;
  &lt;li&gt;Matlab (2)&lt;/li&gt;
  &lt;li&gt;MAXQDA (2)&lt;/li&gt;
  &lt;li&gt;MorphAdorner (2)&lt;/li&gt;
  &lt;li&gt;oXygen (2)&lt;/li&gt;
  &lt;li&gt;Pandoc (2)&lt;/li&gt;
  &lt;li&gt;Paper Machines (2)&lt;/li&gt;
  &lt;li&gt;PhiloLogic (2)&lt;/li&gt;
  &lt;li&gt;Protovis (2)&lt;/li&gt;
  &lt;li&gt;RelFinder (2)&lt;/li&gt;
  &lt;li&gt;RoSE (2)&lt;/li&gt;
  &lt;li&gt;Slack (2)&lt;/li&gt;
  &lt;li&gt;Spotify (2)&lt;/li&gt;
  &lt;li&gt;SylvaDB (2)&lt;/li&gt;
  &lt;li&gt;T-PEN (2)&lt;/li&gt;
  &lt;li&gt;TAPoR (2)&lt;/li&gt;
  &lt;li&gt;TRAViz (2)&lt;/li&gt;
  &lt;li&gt;TypeWright (2)&lt;/li&gt;
  &lt;li&gt;UIMA (2)&lt;/li&gt;
  &lt;li&gt;VSim (2)&lt;/li&gt;
  &lt;li&gt;Abbot (1)&lt;/li&gt;
  &lt;li&gt;Academia.edu (1)&lt;/li&gt;
  &lt;li&gt;Adobe After Effects (1)&lt;/li&gt;
  &lt;li&gt;Adobe Flash (1)&lt;/li&gt;
  &lt;li&gt;Adobe Illustrator (1)&lt;/li&gt;
  &lt;li&gt;Adobe InDesign (1)&lt;/li&gt;
  &lt;li&gt;Advene (1)&lt;/li&gt;
  &lt;li&gt;Alpheios (1)&lt;/li&gt;
  &lt;li&gt;Alveo (1)&lt;/li&gt;
  &lt;li&gt;Annotorious (1)&lt;/li&gt;
  &lt;li&gt;Anthologize (1)&lt;/li&gt;
  &lt;li&gt;Anvil (1)&lt;/li&gt;
  &lt;li&gt;AskSam (1)&lt;/li&gt;
  &lt;li&gt;Audacity (1)&lt;/li&gt;
  &lt;li&gt;AustESE (1)&lt;/li&gt;
  &lt;li&gt;Basecamp (1)&lt;/li&gt;
  &lt;li&gt;Beautiful Soup (1)&lt;/li&gt;
  &lt;li&gt;BuddyPress (1)&lt;/li&gt;
  &lt;li&gt;Chart.js (1)&lt;/li&gt;
  &lt;li&gt;Chronos Timeline (1)&lt;/li&gt;
  &lt;li&gt;COBOL (1)&lt;/li&gt;
  &lt;li&gt;Collex (1)&lt;/li&gt;
  &lt;li&gt;ColorBrewer (1)&lt;/li&gt;
  &lt;li&gt;Confluence (1)&lt;/li&gt;
  &lt;li&gt;CONTENTdm (1)&lt;/li&gt;
  &lt;li&gt;Dedoose (1)&lt;/li&gt;
  &lt;li&gt;DH Press (1)&lt;/li&gt;
  &lt;li&gt;eLaborate (1)&lt;/li&gt;
  &lt;li&gt;EndNote (1)&lt;/li&gt;
  &lt;li&gt;Evernote (1)&lt;/li&gt;
  &lt;li&gt;EVI-LINHD (1)&lt;/li&gt;
  &lt;li&gt;Exhibit 3.0 (1)&lt;/li&gt;
  &lt;li&gt;ezlinavis (1)&lt;/li&gt;
  &lt;li&gt;FAIMS Mobile Platform (1)&lt;/li&gt;
  &lt;li&gt;Finale (1)&lt;/li&gt;
  &lt;li&gt;GapVis (1)&lt;/li&gt;
  &lt;li&gt;GeoTemCo (1)&lt;/li&gt;
  &lt;li&gt;gFacet (1)&lt;/li&gt;
  &lt;li&gt;Google Docs (1)&lt;/li&gt;
  &lt;li&gt;Graphviz (1)&lt;/li&gt;
  &lt;li&gt;H-Net (1)&lt;/li&gt;
  &lt;li&gt;HyperImage (1)&lt;/li&gt;
  &lt;li&gt;HyperPo (1)&lt;/li&gt;
  &lt;li&gt;Icon (1)&lt;/li&gt;
  &lt;li&gt;ImagePlot (1)&lt;/li&gt;
  &lt;li&gt;IRaMuTeQ (1)&lt;/li&gt;
  &lt;li&gt;IsaViz (1)&lt;/li&gt;
  &lt;li&gt;jsLDA (1)&lt;/li&gt;
  &lt;li&gt;Koha (1)&lt;/li&gt;
  &lt;li&gt;KORA (1)&lt;/li&gt;
  &lt;li&gt;Lexomics (1)&lt;/li&gt;
  &lt;li&gt;LilyPond (1)&lt;/li&gt;
  &lt;li&gt;LimeSurvey (1)&lt;/li&gt;
  &lt;li&gt;LodLive (1)&lt;/li&gt;
  &lt;li&gt;Mediathread (1)&lt;/li&gt;
  &lt;li&gt;Mendeley (1)&lt;/li&gt;
  &lt;li&gt;Mukurtu CMS (1)&lt;/li&gt;
  &lt;li&gt;MuseScore (1)&lt;/li&gt;
  &lt;li&gt;music21 (1)&lt;/li&gt;
  &lt;li&gt;NeOn Toolkit (1)&lt;/li&gt;
  &lt;li&gt;NetDraw (1)&lt;/li&gt;
  &lt;li&gt;NVivo (1)&lt;/li&gt;
  &lt;li&gt;Old Maps Online (1)&lt;/li&gt;
  &lt;li&gt;OmniPage (1)&lt;/li&gt;
  &lt;li&gt;OntoViz (1)&lt;/li&gt;
  &lt;li&gt;OpenLayers (1)&lt;/li&gt;
  &lt;li&gt;Pliny (1)&lt;/li&gt;
  &lt;li&gt;Prefuse (1)&lt;/li&gt;
  &lt;li&gt;Prism (1)&lt;/li&gt;
  &lt;li&gt;Pro Tools (1)&lt;/li&gt;
  &lt;li&gt;ProcessingJS (1)&lt;/li&gt;
  &lt;li&gt;Project Quincy (1)&lt;/li&gt;
  &lt;li&gt;Prolog (1)&lt;/li&gt;
  &lt;li&gt;PyDelta (1)&lt;/li&gt;
  &lt;li&gt;RDF Gravity (1)&lt;/li&gt;
  &lt;li&gt;SARIT (1)&lt;/li&gt;
  &lt;li&gt;Scripto (1)&lt;/li&gt;
  &lt;li&gt;SemLens (1)&lt;/li&gt;
  &lt;li&gt;Serendip (1)&lt;/li&gt;
  &lt;li&gt;Serendip-o-matic (1)&lt;/li&gt;
  &lt;li&gt;Sibelius (1)&lt;/li&gt;
  &lt;li&gt;StoryMapJS (1)&lt;/li&gt;
  &lt;li&gt;Textable (1)&lt;/li&gt;
  &lt;li&gt;Textal (1)&lt;/li&gt;
  &lt;li&gt;Textexture (1)&lt;/li&gt;
  &lt;li&gt;TextRazor (1)&lt;/li&gt;
  &lt;li&gt;tFacet (1)&lt;/li&gt;
  &lt;li&gt;Transana (1)&lt;/li&gt;
  &lt;li&gt;Trello (1)&lt;/li&gt;
  &lt;li&gt;TUSTEP (1)&lt;/li&gt;
  &lt;li&gt;VexFlow (1)&lt;/li&gt;
  &lt;li&gt;Visual Browser (1)&lt;/li&gt;
  &lt;li&gt;VisualEyes (1)&lt;/li&gt;
  &lt;li&gt;VIVO (1)&lt;/li&gt;
  &lt;li&gt;Wavesurfer (1)&lt;/li&gt;
  &lt;li&gt;Wiki Map Project (1)&lt;/li&gt;
  &lt;li&gt;WordCruncher (1)&lt;/li&gt;
  &lt;li&gt;WordHoard (1)&lt;/li&gt;
  &lt;li&gt;YAGO (1)&lt;/li&gt;
&lt;/ol&gt;
</content>
    <author>
      <name>
	
          
          Laure Barbot, 
	
          
          Frank Fischer, 
	
          
          Yoann Moranville, 
	
          
          Ivan Pozdniakov
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Using webweb and netwulf with DraCor's API to Generate Interactive Literary Networks</title>
    <link href="https://weltliteratur.net/netwulf-webweb/"/>
    <updated>2019-09-10T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/netwulf-webweb</id>
    <content type="html">&lt;p&gt;&lt;strong&gt;DraCor&lt;/strong&gt; is the Drama Corpora platform at &lt;a href=&quot;https://dracor.org/&quot;&gt;&lt;strong&gt;dracor.org&lt;/strong&gt;&lt;/a&gt; which holds a growing number of collections of plays of different languages, countries and times (German, Russian, Spanish, Swedish, Roman, Greek, with more to come). All plays are encoded in &lt;a href=&quot;https://en.wikipedia.org/wiki/Text_Encoding_Initiative&quot;&gt;XML-TEI&lt;/a&gt;, but you don’t have to meddle with the XML since the &lt;a href=&quot;https://dracor.org/documentation/api/&quot;&gt;DraCor API&lt;/a&gt; makes it really easy to access structured information, e.g., network data based on the co-occurrence of characters per scene (as an example, here’s the social network extracted from &lt;a href=&quot;https://dracor.org/shake/hamlet#network&quot;&gt;Shakespeare’s “Hamlet”&lt;/a&gt;).&lt;/p&gt;

&lt;p&gt;You can do all kinds of things with this network data, e.g., feed them into network visualisation tools. So, a couple of weeks ago, &lt;a href=&quot;https://twitter.com/DanLarremore/status/1161003051726958592&quot;&gt;Twitter&lt;/a&gt; pointed us to two interesting Python packages, &lt;strong&gt;webweb&lt;/strong&gt; and &lt;strong&gt;netwulf&lt;/strong&gt;:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;&lt;strong&gt;webweb&lt;/strong&gt; is “a tool for creating, displaying, and sharing interactive network visualizations on the web” (&lt;a href=&quot;http://dx.doi.org/10.21105/joss.01458&quot;&gt;DOI:10.21105/joss.01458&lt;/a&gt;),&lt;/li&gt;
  &lt;li&gt;&lt;strong&gt;netwulf&lt;/strong&gt; is “an interactive visualization tool for networkx Graph-objects, that allows you to produce beautifully looking network visualizations” (&lt;a href=&quot;https://netwulf.readthedocs.io/en/latest/about.html&quot;&gt;netwulf.readthedocs.io&lt;/a&gt;).&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;It’s really easy to get started and we decided to give these two a spin, thereby also demonstrating the versatility of our DraCor API which provides network data for all TEI-encoded plays of any corpus on the fly.&lt;/p&gt;

&lt;p&gt;We prepared &lt;a href=&quot;https://github.com/ldsad7/netwulf_and_webweb_for_rusdracor&quot;&gt;&lt;strong&gt;a Jupyter notebook&lt;/strong&gt;&lt;/a&gt; with the code ready to be executed.&lt;/p&gt;

&lt;p&gt;If you want to play with drama networks without doing the legwork, we uploaded some ready-to-use HTML pages generated by &lt;strong&gt;webweb&lt;/strong&gt;:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;/webweb/pushkin-boris-godunov.html&quot;&gt;Pushkin: Boris Godunov&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;/webweb/lessing-emilia-galotti.html&quot;&gt;Lessing: Emilia Galotti&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;/webweb/rus.html&quot;&gt;Russian Drama Corpus&lt;/a&gt; (you can choose between 190 plays in the left upper corner)&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;/webweb/ger.html&quot;&gt;German Drama Corpus&lt;/a&gt; (dito, between 474 plays)&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;But let’s start to toy around with the other library first:&lt;/p&gt;

&lt;h2 id=&quot;netwulf&quot;&gt;netwulf&lt;/h2&gt;

&lt;p&gt;We created a universal function (&lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;netwulf_representation&lt;/code&gt; in &lt;a href=&quot;https://github.com/ldsad7/netwulf_and_webweb_for_rusdracor&quot;&gt;our Jupyter Notebook&lt;/a&gt;) and a function to retrieve the titles of all plays of a corpus (&lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;retrieve_all_plays&lt;/code&gt;).&lt;/p&gt;

&lt;p&gt;Now let’s take Pushkin’s historical tragedy &lt;strong&gt;“Boris Godunov” (c. 1825)&lt;/strong&gt; as an example:&lt;/p&gt;

&lt;div class=&quot;language-python highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;c1&quot;&gt;# 1st network: gender groups without labels (other attributes are default)
&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;netwulf_representation&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;rus&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;s&quot;&gt;&apos;pushkin-boris-godunov&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;group&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;gender&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;show_node_labels&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;bp&quot;&gt;False&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt;
&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;
&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/netwulf-webweb/godunov1.png&quot; alt=&quot;Boris Godunov (1)&quot; style=&quot;width:500px;&quot; /&gt;
&lt;/figure&gt;

&lt;div class=&quot;language-python highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;c1&quot;&gt;# 2ns network: gender groups with labels
&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;netwulf_representation&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;rus&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;s&quot;&gt;&apos;pushkin-boris-godunov&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;group&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;gender&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt;
&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/netwulf-webweb/godunov2.png&quot; alt=&quot;Boris Godunov (2)&quot; style=&quot;width:500px;&quot; /&gt;
&lt;/figure&gt;

&lt;div class=&quot;language-python highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;c1&quot;&gt;# 3rd network: isGroup groups with labels
&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;netwulf_representation&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;rus&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;s&quot;&gt;&apos;pushkin-boris-godunov&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;group&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;isGroup&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt;
&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/netwulf-webweb/godunov3.png&quot; alt=&quot;Boris Godunov (3)&quot; style=&quot;width:500px;&quot; /&gt;
&lt;/figure&gt;

&lt;div class=&quot;language-python highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;c1&quot;&gt;# 4th network: gender groups with numOfSpeechActs size without labels
&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;netwulf_representation&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;rus&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;s&quot;&gt;&apos;pushkin-boris-godunov&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;group&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;gender&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;size&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;numOfSpeechActs&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;show_node_labels&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;bp&quot;&gt;False&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt;
&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/netwulf-webweb/godunov4.png&quot; alt=&quot;Boris Godunov (4)&quot; style=&quot;width:500px;&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;By the way, in order to fully regnerate the network (the radii of nodes are not stored by default by netwulf) we need to add a ‘size’ field to the nodes:&lt;/p&gt;

&lt;div class=&quot;language-python highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;k&quot;&gt;for&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;node&lt;/span&gt; &lt;span class=&quot;ow&quot;&gt;in&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;network&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;[&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;nodes&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;]:&lt;/span&gt;
    &lt;span class=&quot;n&quot;&gt;node&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;[&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;size&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;]&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;=&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;node&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;[&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;radius&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;]&lt;/span&gt;
&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;p&gt;Now, as second example, let’s turn to German drama and try the usual suspect, &lt;strong&gt;Lessing’s “Emilia Galotti” (1772)&lt;/strong&gt;:&lt;/p&gt;

&lt;div class=&quot;language-python highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;c1&quot;&gt;# 1st network: gender groups without labels
&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;netwulf_representation&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;ger&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;s&quot;&gt;&apos;lessing-emilia-galotti&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;group&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;gender&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;show_node_labels&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;bp&quot;&gt;False&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt;  &lt;span class=&quot;c1&quot;&gt;# other attributes are default (see above)
&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/netwulf-webweb/galotti1.png&quot; alt=&quot;Emilia Galotti (1)&quot; style=&quot;width:500px;&quot; /&gt;
&lt;/figure&gt;

&lt;div class=&quot;language-python highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;c1&quot;&gt;# 2nd network: isGroup groups with labels and degree sizes
&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;netwulf_representation&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;ger&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;s&quot;&gt;&apos;lessing-emilia-galotti&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;group&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;isGroup&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;size&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;degree&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt;
&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/netwulf-webweb/galotti2.png&quot; alt=&quot;Emilia Galotti (2)&quot; style=&quot;width:500px;&quot; /&gt;
&lt;/figure&gt;

&lt;div class=&quot;language-python highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;c1&quot;&gt;# 3rd network: gender groups with node and link labels and numOfWords size
&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;netwulf_representation&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;ger&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;s&quot;&gt;&apos;lessing-emilia-galotti&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;group&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;gender&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;size&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;numOfWords&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;show_link_labels&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;bp&quot;&gt;True&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt;
&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/netwulf-webweb/galotti3.png&quot; alt=&quot;Emilia Galotti (3)&quot; style=&quot;width:500px;&quot; /&gt;
&lt;/figure&gt;

&lt;div class=&quot;language-python highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;c1&quot;&gt;# 4th network: all characters with numOfScenes size with node and link labels
&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;netwulf_representation&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;ger&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;s&quot;&gt;&apos;lessing-emilia-galotti&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;size&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;numOfSpeechActs&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt;
&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/netwulf-webweb/galotti4.png&quot; alt=&quot;Emilia Galotti (4)&quot; style=&quot;width:500px;&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;It is also possible to apply our script to all plays of a corpus (e.g., the Russian one):&lt;/p&gt;

&lt;div class=&quot;language-python highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;n&quot;&gt;playnames&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;=&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;retrieve_all_plays&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;rus&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt;
&lt;span class=&quot;k&quot;&gt;for&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;playname&lt;/span&gt; &lt;span class=&quot;ow&quot;&gt;in&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;playnames&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;
    &lt;span class=&quot;n&quot;&gt;netwulf_representation&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;rus&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;playname&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;size&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;weightedDegree&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;group&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;s&quot;&gt;&apos;gender&apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;show_link_labels&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;bp&quot;&gt;True&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;is_test&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;=&lt;/span&gt;&lt;span class=&quot;bp&quot;&gt;True&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt;
&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;p&gt;With this piece of code, netwulf will launch the networks one by one. It will display each network for 5 seconds, followed by a 3-second fade-out before opening the next one. So if you are really bored, you can let hundreds of Russian drama networks pass by your eye. Our collection of currently 190 plays will steal 25 minutes of your time.&lt;/p&gt;

&lt;h2 id=&quot;webweb&quot;&gt;webweb&lt;/h2&gt;

&lt;p&gt;Dan Larremore’s webweb library offers a similar approach to visualise network data (see their &lt;a href=&quot;https://webwebpage.github.io/&quot;&gt;documentation&lt;/a&gt;). We wrote two scripts: the first one is for representing only one play, the second one for representing several plays, for example, a whole corpus, or translations or different editions of the same play.&lt;/p&gt;

&lt;p&gt;We created a universal function to integrate DraCor data with webweb (&lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;webweb_representation&lt;/code&gt; in our &lt;a href=&quot;https://github.com/ldsad7/netwulf_and_webweb_for_rusdracor&quot;&gt;our Jupyter Notebook&lt;/a&gt;). This script is a bit easier and uses less additional parameters. That is because the library provides us an opportunity to fine-tune the settings such as ‘color nodes by’/’scale node sizes by’ interactively in the generated HTML file. Definitely an advantage since you don’t have to restart the program.&lt;/p&gt;

&lt;p&gt;Other advantages as follows:&lt;/p&gt;
&lt;ol&gt;
  &lt;li&gt;interactive captions where you can interactively scale node sizes (although they are limited to 5 sizes),&lt;/li&gt;
  &lt;li&gt;ability to highlight a node (but it’s not possible to highlight groups of nodes, only if these nodes have a common letter sequence in their labels), yet in netwulf we can highlight any number of nodes (need to tap twice on each of them),&lt;/li&gt;
  &lt;li&gt;choose between a collection of networks (ability to compare them on one page), dynamic networks (ability to see change over time),&lt;/li&gt;
  &lt;li&gt;ability to color nodes,&lt;/li&gt;
  &lt;li&gt;SVG output.&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;In our opinion, webweb also has some disadvantages in comparison to the netwulf library:&lt;/p&gt;
&lt;ol&gt;
  &lt;li&gt;inability to zoom in and out,&lt;/li&gt;
  &lt;li&gt;no possibility to freely change parameters (e.g., colour sets, node radius, etc.) as with netwulf library,&lt;/li&gt;
  &lt;li&gt;some parameters (like the link opacity) are binary while in netwulf they are continuous, some parameters are missing (e.g., stroke width, and no ‘wiggle’ option),&lt;/li&gt;
  &lt;li&gt;inability to restore a graph (not in SVG format) in the exact same way.&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;Let’s look at the graphs produced by webweb. First &lt;strong&gt;Pushkin’s “Boris Godunov”&lt;/strong&gt;:&lt;/p&gt;

&lt;div class=&quot;language-python highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;c1&quot;&gt;# first network: gender groups with labels (size == weightedDegree)
&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;
&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/netwulf-webweb/webweb1.png&quot; alt=&quot;webweb (1)&quot; style=&quot;width:500px;&quot; /&gt;
&lt;/figure&gt;

&lt;div class=&quot;language-python highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;c1&quot;&gt;# second network: isGroup groups with labels (size == numOfSpeechActs)
&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/netwulf-webweb/webweb2.png&quot; alt=&quot;webweb (2)&quot; style=&quot;width:500px;&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;And now &lt;strong&gt;Lessing’s “Emilia Galotti”&lt;/strong&gt; again:&lt;/p&gt;

&lt;div class=&quot;language-python highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;c1&quot;&gt;# third network: gender groups with labels (size == weightedDegree)
&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/netwulf-webweb/webweb3.png&quot; alt=&quot;webweb (3)&quot; style=&quot;width:500px;&quot; /&gt;
&lt;/figure&gt;

&lt;div class=&quot;language-python highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;c1&quot;&gt;# fourth network: numOfScenes groups with labels (size == strength)
&lt;/span&gt;&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;figure style=&quot;text-align:left;&quot;&gt;
  &lt;img src=&quot;/images/netwulf-webweb/webweb4.png&quot; alt=&quot;webweb (4)&quot; style=&quot;width:500px;&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;As stated above, the HTML files generated by webweb embed all the JavaScript you need and you can save them as standalone pages:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;/webweb/pushkin-boris-godunov.html&quot;&gt;Pushkin: Boris Godunov&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;/webweb/lessing-emilia-galotti.html&quot;&gt;Lessing: Emilia Galotti&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;/webweb/rus.html&quot;&gt;Russian Drama Corpus&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;/webweb/ger.html&quot;&gt;German Drama Corpus&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;
</content>
    <author>
      <name>
	
          
          Eduard Grigoryev, 
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>"Moscow Formalism and Literary History" (De Gruyter)</title>
    <link href="https://weltliteratur.net/moscow-formalism-and-literary-history/"/>
    <updated>2019-06-01T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/moscow-formalism-and-literary-history</id>
    <content type="html">&lt;h2 id=&quot;just-out-new-translations-of-early-data-driven-formalism-yarkho-gasparov-shapir&quot;&gt;Just Out: New Translations of Early Data-Driven Formalism (Yarkho, Gasparov, Shapir)&lt;/h2&gt;

&lt;p&gt;A new volume of the “Journal of Literary Theory” was published in March 2019 (vol. 13, no. 1: &lt;a href=&quot;https://www.degruyter.com/view/j/jlt.2019.13.issue-1/issue-files/jlt.2019.13.issue-1.xml&quot;&gt;“Moscow Formalism and Literary History”&lt;/a&gt;), a special issue prepared and edited by us comprising three translations from the Russian plus preface:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;Frank Fischer, Marina Akimova, Boris Orekhov: &lt;strong&gt;Preface: Data-Driven Formalism&lt;/strong&gt; (DOI:&lt;a href=&quot;https://doi.org/10.1515/jlt-2019-0001&quot;&gt;10.1515/jlt-2019-0001&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Boris I. Yarkho: &lt;strong&gt;Speech Distribution in Five-Act Tragedies (A Question of Classicism and Romanticism)&lt;/strong&gt; (1935–1938, first published 1997) (DOI:&lt;a href=&quot;https://doi.org/10.1515/jlt-2019-0002&quot;&gt;10.1515/jlt-2019-0002&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Mikhail L. Gasparov: &lt;strong&gt;The Evolution of Russian Rhyme&lt;/strong&gt; (1984) (DOI:&lt;a href=&quot;https://doi.org/10.1515/jlt-2019-0003&quot;&gt;10.1515/jlt-2019-0003&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;Maksim I. Shapir: &lt;strong&gt;“For Thee There Is No Weight nor Measure”. The Possibilities and Limitations of “Exact” Methods in the Humanities&lt;/strong&gt; (2005) (DOI:&lt;a href=&quot;https://doi.org/10.1515/jlt-2019-0004&quot;&gt;10.1515/jlt-2019-0004&lt;/a&gt;)&lt;/li&gt;
&lt;/ul&gt;

&lt;figure style=&quot;text-align:center;&quot;&gt;
&lt;img src=&quot;/images/moscow-formalism/moscow-formalism-cover.jpg&quot; alt=&quot;Figure 1a&quot; style=&quot;height:600px;&quot; /&gt;
&lt;img src=&quot;/images/moscow-formalism/moscow-formalism-cover-toc.png&quot; alt=&quot;Figure 1b&quot; style=&quot;height:600px;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 1.&lt;/b&gt; Book cover and table of contents.&lt;/center&gt;

&lt;p&gt;The idea for this volume goes back to a chat that Asya Bonch-Osmolovskaya and I had with Fotis Jannidis at the DH2015 conference in Sydney. As one of its editors, Fotis suggested to equip a special edition of “Journal of Literary Theory” with untranslated works of Russian formalists, especially quantitative literary studies. The idea was to build a bridge from early formalism to digital literary studies of today. The outcomes of the Stanford conference on &lt;a href=&quot;https://digitalhumanities.stanford.edu/russian-formalism-digital-humanities&quot;&gt;“Russian Formalism &amp;amp; the Digital Humanities”&lt;/a&gt; (2015) suggested that it is not so easy to reconcile the two. That was, I would argue, due to the fact that some decisive texts were still untranslated.&lt;/p&gt;

&lt;h2 id=&quot;boris-yarkho&quot;&gt;Boris Yarkho&lt;/h2&gt;

&lt;p&gt;So we assembled a small team: Marina Akimova and Boris Orekhov joined as co-editors and Craig Saunders as translator. The first few pages of Yarkho’s text arrived in November 2017 with Craig admitting to “feeling a little charmed by the way he [Yarkho] writes”.&lt;/p&gt;

&lt;figure style=&quot;text-align:center;&quot;&gt;
  &lt;img src=&quot;/images/moscow-formalism/moscow-formalism-yarkho.jpg&quot; alt=&quot;Figure 2&quot; style=&quot;width:450px;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 2.&lt;/b&gt; Boris Yarkho (1889–1942). (source: &lt;a href=&quot;https://urokiistorii.ru/article/52560&quot;&gt;https://urokiistorii.ru/article/52560&lt;/a&gt;, cropped)&lt;/center&gt;

&lt;p&gt;Except for a draft plan of Yarkho’s “Methodology for a Precise Science of Literature” (transl. by L. M. O’Toole, publ. 1977 in vol. 4 of &lt;a href=&quot;https://weltliteratur.net/russian-poetics-in-translation/&quot;&gt;“Russian Poetics in Translation”&lt;/a&gt;) and Gasparov’s early account of &lt;a href=&quot;https://doi.org/10.12697/smp.2016.3.2.05&quot;&gt;“Boris Yarkho’s works on literary theory”&lt;/a&gt; (transl. by Michael Lavery and Marina Tarlinskaja, publ. 2016 in “Studia Metrica et Poetica”), there’s not much first-hand info on Yarkho available in English. “Speech Distribution in Five-Act Tragedies” is the first translation of an article by Yarkho and it can be expected that it will leave its mark over time.&lt;/p&gt;

&lt;p&gt;We know from a reliable source how the manuscript was first edited in the 1990s, because one of the co-editors of our volume was directly involved. Marina received the microfilm with copies of Yarkho’s manuscripts from Gasparov and manually transcribed them into a computer, projecting the film on a wall in her flat.&lt;/p&gt;

&lt;p&gt;Yarkho’s description of his data and experimental setup are so accurate that it was easy to implement them. DraCor, our platform for research on drama (&lt;a href=&quot;https://dracor.org/&quot;&gt;dracor.org&lt;/a&gt;), automatically extracts the speech-distribution pattern in the way Yarkho described it (monologues, dialogues, three-way dialogues and so forth) for every corpus and play connected to the platform.&lt;/p&gt;

&lt;p&gt;Here are two examples:&lt;/p&gt;

&lt;figure style=&quot;text-align:center;&quot;&gt;
  &lt;img src=&quot;/images/moscow-formalism/speech-distribution-lessing-emilia-galotti.png&quot; alt=&quot;Figure 3a&quot; style=&quot;width:500px;&quot; /&gt;
  &lt;img src=&quot;/images/moscow-formalism/speech-distribution-ozerov-dmitrij-donskoj.png&quot; alt=&quot;Figure 3b&quot; style=&quot;width:500px;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 3.&lt;/b&gt; Speech distributions for Lessing’s &quot;Emilia Galotti&quot; (source: &lt;a href=&quot;https://dracor.org/ger/lessing-emilia-galotti#speech&quot;&gt;dracor.org&lt;/a&gt;) and &lt;br /&gt;Ozerov’s &quot;Dmitry Donskoy&quot; (source: &lt;a href=&quot;https://dracor.org/rus/ozerov-dmitrij-donskoj#speech&quot;&gt;dracor.org&lt;/a&gt;)&lt;/center&gt;

&lt;h2 id=&quot;summer-in-moscow&quot;&gt;Summer in Moscow&lt;/h2&gt;

&lt;p&gt;After we had finished work on Yarkho, we tackled the remaining two articles. The main difficulty in Gasparov was to put the stress on the many examples and rhyming patterns, which were not present in the Russian original. And finally, translating Shapir proved particularly challenging. It also looks like we were the first to actually translate a Shapir article into English.&lt;/p&gt;

&lt;p&gt;Most of the work was done by the summer of 2018 when the World Cup was being held in Russia. I remember one of our last meetings one afternoon in June when Morocco had just been eliminated from the World Cup by a Ronaldo goal. I was sitting in the Metro on my way to a cafe close to &lt;a href=&quot;https://en.wikipedia.org/wiki/Russian_State_Library&quot;&gt;Leninka&lt;/a&gt; and saw many melancholic Moroccan fans with their flags trying to cheer each other up.&lt;/p&gt;

&lt;p&gt;After we received feedback from the editors, the printing proofs were largely completed in November.&lt;/p&gt;

&lt;h2 id=&quot;panel-discussion-at-zfl-berlin-june-12th-2019&quot;&gt;Panel Discussion at ZfL Berlin: June, 12th, 2019&lt;/h2&gt;

&lt;p&gt;Thanks to the Leibniz Center for Literary and Cultural Research (ZfL) in Berlin, we will have the opportunity to present the volume and talk about the larger context in a panel discussion on June, 12th, 2019. Marina and I will be joined on stage by ZfL’s Siarhei Biareyshik, and Zaal Andronikashvili will chair the event.&lt;/p&gt;

&lt;p&gt;More information:&lt;br /&gt;
&lt;a href=&quot;http://www.zfl-berlin.org/veranstaltungen-detail/items/moscow-formalism-and-literary-history.html&quot;&gt;http://www.zfl-berlin.org/veranstaltungen-detail/items/moscow-formalism-and-literary-history.html&lt;/a&gt;&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Vossian Antonomasia in "The New York Times"</title>
    <link href="https://weltliteratur.net/vossian-antonomasia-new-york-times/"/>
    <updated>2019-02-19T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/vossian-antonomasia-new-york-times</id>
    <content type="html">&lt;h2 id=&quot;new-paper-out&quot;&gt;New Paper Out!&lt;/h2&gt;

&lt;p&gt;We’ve been collecting and studying &lt;strong&gt;Vossian antonomasia&lt;/strong&gt; (VA, or, vossanto) for about a decade now. This stylistic device, named after Dutch humanist &lt;a href=&quot;https://en.wikipedia.org/wiki/Gerardus_Vossius&quot;&gt;Gerardus Vossius&lt;/a&gt;, occurs when someone is called &lt;a href=&quot;https://www.nytimes.com/2006/12/31/movies/31farb.html&quot;&gt;“the Jay Leno of Tel Aviv”&lt;/a&gt; or &lt;a href=&quot;https://www.nytimes.com/1999/09/22/arts/tahia-carioca-79-dies-a-renowned-belly-dancer.html&quot;&gt;“the Marilyn Monroe of the Arab world”&lt;/a&gt; etc. etc.&lt;/p&gt;

&lt;p&gt;We just published a new paper about this phenomenon …&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;Frank Fischer; Robert Jäschke: &lt;strong&gt;‘The Michael Jordan of greatness’—Extracting Vossian antonomasia from two decades of The New York Times, 1987–2007.&lt;/strong&gt; &lt;em&gt;Digital Scholarship in the Humanities.&lt;/em&gt; 2019. (DOI:&lt;a href=&quot;https://doi.org/10.1093/llc/fqy087&quot;&gt;10.1093/llc/fqy087&lt;/a&gt;) (Preprint: &lt;a href=&quot;https://arxiv.org/abs/1902.06428&quot;&gt;arXiv:1902.06428&lt;/a&gt;)&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;… plus a website with our code and some more interesting results for you to explore:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://vossanto.weltliteratur.net/&quot;&gt;https://vossanto.weltliteratur.net/&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;But let’s rewind a bit. Some years ago we published a &lt;a href=&quot;https://www.umblaetterer.de/datenzentrum/vossianische-antonomasien.html&quot;&gt;hand-picked list&lt;/a&gt; with our favourite VA expressions from 2009 to 2014 (mainly from German-language media) and published what we thought would be our &lt;a href=&quot;https://www.umblaetterer.de/wp-content/uploads/2014/12/vossanto_fas.png&quot;&gt;final article on the subject&lt;/a&gt; in &lt;em&gt;Frankfurter Allgemeine Sonntagszeitung&lt;/em&gt; in December 2014.&lt;/p&gt;

&lt;p&gt;There is a lot of research out there on all types of antonomasia, even on Vossian antonomasia (or ‘metaphorical antonomasia’, or ‘paragons’). But when we started looking, we didn’t come across any convincing automatic approach regarding the bulk extraction of VA to learn more about general patterns. All papers we found used hand-picked examples and took it from there. Of course, there are the type of websites like &lt;a href=&quot;https://graphics.wsj.com/michael-jordan-of/&quot;&gt;“The Michael Jordan of…”&lt;/a&gt;, which collected 1,505 &lt;em&gt;Michael Jordans of something&lt;/em&gt;. As interesting as that is, from an extraction point of view it is trivial.&lt;/p&gt;

&lt;h2 id=&quot;vossian-antonomasia-next-level&quot;&gt;Vossian Antonomasia, Next Level&lt;/h2&gt;

&lt;p&gt;So before long we were back in the game and started hacking. We wanted to find a way to extract all (well, as many as possible 😊) occurrences of Vossian antonomasia from larger newspaper corpora, not only based on usual suspects like Michael Jordan (for what it’s worth, our research has confirmed on much more reliable grounds that he’s the ruler of the Vossian world). We wanted to find all the unexpected examples, so we fetched some newspaper corpora and our first idea was to do some POS tagging and NER and then try to match the pattern “the [name] of”. We worked with two full-text corpora, the &lt;em&gt;New York Times&lt;/em&gt; 1987–2007 and German weekly &lt;em&gt;Die Zeit&lt;/em&gt; 1995–2011. We &lt;a href=&quot;https://lehkost.github.io/slides/2017-bern/&quot;&gt;presented&lt;/a&gt; this work at DHd2017 in Bern, but were not really content with the results. Especially NER let us down big time. But it was a start.&lt;/p&gt;

&lt;p&gt;A couple of months later we had the idea that took our research to a new level. Thanks to a research scholarship granted to me by the &lt;a href=&quot;https://www.sheffield.ac.uk/is&quot;&gt;Information School at the University of Sheffield&lt;/a&gt;, Robert and I were able to spend two weeks hacking together in May, 2017. And then, probably during one of our lunches at the &lt;a href=&quot;http://www.red-deer-sheffield.co.uk/&quot;&gt;Red Deer&lt;/a&gt; (Vegetarian Haggis ftw!), we came to talk about &lt;strong&gt;Wikidata&lt;/strong&gt; and then it occurred to us that it could help us out of our NER concerns. Here are the two things we brought together:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;1. We want to extract paragons, that is, well-known, infamous people. People that would definitely have a Wikipedia entry in some language, or else they wouldn’t be well-known enough to serve as source for Vossian antonomasia.&lt;/li&gt;
  &lt;li&gt;2. Wikidata inherits these famous people from all the different Wikipedia language versions and has all kinds of facts about them in a well-structured format, including nicknames and alternative spelling of names (which we would need).&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Please note that this solution was really carved out for our specific research question, because Wikidata arguably holds exactly those named entities that we would need. Yet it probably wouldn’t be the best choice for other NER purposes.&lt;/p&gt;

&lt;p&gt;The rest was pretty simple. It was now enough to use a regular expression to extract the general pattern “the … of”, and then we would match the part in-between against our newly gained persons database. And that’s what we did. The corpus at hand was the &lt;a href=&quot;https://catalog.ldc.upenn.edu/LDC2008T19&quot;&gt;&lt;em&gt;New York Times&lt;/em&gt; one&lt;/a&gt;, and whenever we initiated another extraction process (which always took some minutes), we would pay a visit to the Red Deer to have just another ginger beer.&lt;/p&gt;

&lt;h2 id=&quot;results&quot;&gt;Results&lt;/h2&gt;

&lt;p&gt;Vossian antonomasia is a rare phenomenon. Annotating a gold set of 500 randomly selected articles might not throw you even one result. So, while we cannot provide numbers for recall due to the lack of a gold standard, we can provide our precision scores, because we were able to personally assess the extraction results. Altogether we managed to extract &lt;strong&gt;2,646 VA expressions&lt;/strong&gt; from 20 years of the &lt;em&gt;New York Times&lt;/em&gt; (1987–2007).&lt;/p&gt;

&lt;p&gt;The &lt;a href=&quot;https://arxiv.org/abs/1902.06428&quot;&gt;paper&lt;/a&gt; features our most notable results, but we also provide some more things to explore on the project website:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;&lt;a href=&quot;https://vossanto.weltliteratur.net/theof/humans/statistics.html&quot;&gt;Some More Statistics&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;&lt;a href=&quot;https://vossanto.weltliteratur.net/theof/humans/vossantos.html&quot;&gt;Complete List of Extracted VA&lt;/a&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;Like mentioned above, Michael Jordan is the boss of it all, which can be witnessed in the &lt;a href=&quot;https://vossanto.weltliteratur.net/theof/humans/statistics.html#top-40-va-sources&quot;&gt;table with the top-40 source entities&lt;/a&gt; for Vossian antonomasia we extracted. But Jordan is closely followed by &lt;a href=&quot;https://en.wikipedia.org/wiki/Rodney_Dangerfield&quot;&gt;Rodney Dangerfield&lt;/a&gt;, a US-American comedian infamous for the sentence “I don’t get no respect!”. I think that being the runner-up in Vossian Olympics actually might earn him some respect after all.&lt;/p&gt;

&lt;p&gt;We also generated a gallery of the 40 top scorers by pulling their principal images from Wikidata (&lt;a href=&quot;https://www.wikidata.org/wiki/Property:P18&quot;&gt;Property:P18&lt;/a&gt;):&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/gallery-top-39-sources-vossian-antonomasia.jpg&quot; alt=&quot;Top-39 sources for Vossian antonomiasa in the NYT 1987–2007&quot; style=&quot;width:1200px;&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;(Btw, if you did the counting you’ll find only 39 images. One of the entities does not feature an image in Wikidata.)&lt;/p&gt;

&lt;h2 id=&quot;next-steps&quot;&gt;Next Steps&lt;/h2&gt;

&lt;p&gt;Our research continues. We’re generally fine-graining our method because there are things we missed. And although “the … of” is arguably the most frequent pattern, we didn’t check other patterns yet, like “a … of”, or the type where the VA source is preceded by a national adjective, as in “the Vietnamese Mozart” or “the German Donald Rumsfeld”. We’re on it. And we’re open for collaborations with other scholars. Because one other thing we wanna do in the future is to broaden our research to corpora in other languages than English with richer morphology, which will add a layer of complexity.&lt;/p&gt;

&lt;p&gt;Oh, if you find interesting Vossian antonomasia, feel free to ping me on Twitter (&lt;a href=&quot;https://twitter.com/umblaetterer&quot;&gt;@umblaetterer&lt;/a&gt;) if you like.&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>"Russian Poetics in Translation" (10 Volumes, 1975–1983)</title>
    <link href="https://weltliteratur.net/russian-poetics-in-translation/"/>
    <updated>2018-08-24T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/russian-poetics-in-translation</id>
    <content type="html">&lt;p&gt;Of course, the interest in Russian poetics is not new. What is new, however, is a growing interest within the Digital Humanities in transforming formalist and structuralist theory into algorithms and add a quantitative angle to it. An important milestone in this respect was the international symposium in Stanford in April 2015, &lt;a href=&quot;https://digitalhumanities.stanford.edu/russian-formalism-digital-humanities-abstracts&quot;&gt;“The Legacy of Russian Formalism and the Rise of the Digital Humanities”&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;In Russia, the core texts are very present in everyday university life. Beyond Russian borders, though, it might look a bit different. Although there are a number of bibliographies of translations, it is difficult to keep track of what’s there and what’s not. Since we are currently &lt;a href=&quot;https://weltliteratur.net/moscow-formalism-and-literary-history/&quot;&gt;preparing some new translations ourselves&lt;/a&gt;, we have tried to gain some kind of overview. In the course of doing so, we came across the excellent collection &lt;strong&gt;“Russian Poetics in Translation”&lt;/strong&gt;. The 10 volumes, edited by L. M. O’Toole and Ann Shukman, were published between 1975 and 1983 (here’s an &lt;a href=&quot;https://doi.org/10.1515/jlse.1982.11.2.U&quot;&gt;advertisement from 1982&lt;/a&gt;). Voilà, the front pages of the first five:&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/russian-poetics-vols-1-to-5.jpg&quot; alt=&quot;Volumes 1 to 5.&quot; style=&quot;width:625px;&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;The collection is not available in a lot of libraries, it seems. If you’re lucky, you find one or two volumes in your nearest library, but probably not the whole collection. This is why a shout-out is due to &lt;a href=&quot;https://de.wikipedia.org/wiki/Niederösterreichische_Landesbibliothek&quot;&gt;Niederösterreichische Landesbibliothek (NÖLB)&lt;/a&gt; in St. Pölten which holds almost the entire set, donated by &lt;a href=&quot;https://de.wikipedia.org/wiki/Joseph_P._Strelka&quot;&gt;Joseph P. Strelka&lt;/a&gt;, a former professor at SUNY Albany. Thus, we didn’t have to go far to get hold of most of the volumes. The rest we gathered worldwide with the help of friends, and before long we were proud owners of (copies of) all volumes.&lt;/p&gt;

&lt;p&gt;Now, in our experience, even online it is hard to throw a glance at the tables of contents per volume. Having said that, the volumes seem to be available as e-books via Amazon since 2013 (incl. the “look inside” option), yet we didn’t check these versions and are unsure about how they were converted and how readable/quotable they ultimately are (if you have any experience with them, feel free to tell us).&lt;/p&gt;

&lt;p&gt;In any case, to convey the richness of this collection, we decided to create a bibliography of all articles in all 10 volumes and publish it on this web page. We also created a &lt;a href=&quot;https://www.zotero.org/groups/2113340/russian_poetics_in_translation/items&quot;&gt;Zotero group&lt;/a&gt; in which we listed all articles. We hope this will inspire those of you who have never heard of this collection before.&lt;/p&gt;

&lt;center&gt;&lt;b&gt;* * *&lt;/b&gt;&lt;br /&gt; &lt;/center&gt;

&lt;h1 id=&quot;russian-poetics-in-translation-10-volumes-19751983&quot;&gt;Russian Poetics in Translation (10 Volumes, 1975–1983)&lt;/h1&gt;

&lt;p&gt;&lt;em&gt;Table of contents of all volumes.&lt;/em&gt;&lt;/p&gt;

&lt;h2 id=&quot;vol-1-generating-the-literary-text-1975-77pp&quot;&gt;Vol. 1. Generating the Literary Text (1975, 77pp.)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Yu. K. Shcheglov; A. K. Zholkovskii: &lt;strong&gt;Towards A “Theme – (Expression Devices) – Text” Model of Literary Structure [1971].&lt;/strong&gt; Translated by L. M. O’Toole. (pp. 3–50)&lt;/li&gt;
  &lt;li&gt;Yu. K. Shcheglov: &lt;strong&gt;Towards a Description of Detective Story Structure [1968].&lt;/strong&gt; Translated by L. M. O’Toole. (pp. 51–77)&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;vol-2-poetry-and-prose-1976-84pp&quot;&gt;Vol. 2. Poetry and Prose (1976, 84pp.)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Yu. I. Levin: &lt;strong&gt;The Logic of Metaphor: Synthesis, Semantics, Transformations [1965/1969].&lt;/strong&gt; Translated by Christopher English. (pp. 5–21)&lt;/li&gt;
  &lt;li&gt;V. N. Toporov: &lt;strong&gt;The Semiotics of Prediction in Suetonius’s “The Lives of the Twelve Caesars” [1965].&lt;/strong&gt; Translated by Halina Willetts. (pp. 22–33)&lt;/li&gt;
  &lt;li&gt;Vyach. Vs. Ivanov: &lt;strong&gt;The Structure of Khlebnikov’s Poem “Menya pronosyat na slonovykh …” [1967].&lt;/strong&gt; Translated by Ann Shukman. (pp. 34–49)&lt;/li&gt;
  &lt;li&gt;Yu. K. Shcheglov: &lt;strong&gt;Some Features of the Structure of Ovid’s “Metamorphoses” [1962].&lt;/strong&gt; Translated by Julian Graffy. (pp. 50–65)&lt;/li&gt;
  &lt;li&gt;Yu. M. Lotman: &lt;strong&gt;The Structure of Ideas in Pushkin’s Poem “Andzhelo” [1973].&lt;/strong&gt; Translated by Ann Shukman. (pp. 66–84)&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;vol-3-general-semiotics-1976-81pp&quot;&gt;Vol. 3. General Semiotics (1976, 81pp.)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Yu. M. Lotman: &lt;strong&gt;The Modelling Significance of the Concepts ‘End’ and ‘Beginning’ in Artistic Texts [1966].&lt;/strong&gt; Translated by Wendy Rosslyn. (pp. 7–11)&lt;/li&gt;
  &lt;li&gt;G. Permyakov: &lt;strong&gt;The Logico-Semiotic Level of Proverbs and Sayings (Towards a Classification of the Genre) [1967].&lt;/strong&gt; Translated by Doris Bradbury. (pp. 12–32)&lt;/li&gt;
  &lt;li&gt;P. G. Bogatyrev: &lt;strong&gt;Stage Setting, Artistic Space and Time in the Folk Theatre [1968].&lt;/strong&gt; Translated by L. M. O’Toole. (pp. 33–37)&lt;/li&gt;
  &lt;li&gt;V. N. Toporov: &lt;strong&gt;On the Cosmological Origins of Early Historical Descriptions [1973].&lt;/strong&gt; Translated by Christopher English. (pp. 38–81)&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;vol-4-formalist-theory-1977-108pp&quot;&gt;Vol. 4. Formalist Theory (1977, 108pp.)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Brief Biographical and Bibliographical Notes on Leading Formalists. Translated by Ann Shukman. (pp. 1–12)&lt;/li&gt;
  &lt;li&gt;A Contextual Glossary of Formalist Terminology. Translated by Ann Shukman and L. M. O’Toole. (pp. 13–48)&lt;/li&gt;
  &lt;li&gt;Yu. Tynyanov; R. Jakobson: &lt;strong&gt;Problems of Research in Literature and Language [1928].&lt;/strong&gt; Translated by L. M. O’Toole. (pp. 49–51)&lt;/li&gt;
  &lt;li&gt;B. Yarkho: &lt;strong&gt;Methodology for a Precise Science of Literature (Draft Plan) [1936/1969] [Excerpts].&lt;/strong&gt; Translated by L. M. O’Toole. (pp. 52–70)&lt;/li&gt;
  &lt;li&gt;S. Bernshtein: &lt;strong&gt;Aesthetic Presuppositions for a Theory of Declamation [1927].&lt;/strong&gt; Translated by Ann Shukman. (pp. 71–89)&lt;/li&gt;
  &lt;li&gt;Osip Brik: &lt;strong&gt;The So-Called Formal Method [1923].&lt;/strong&gt; Translated by Ann Shukman. (pp. 90–91)&lt;/li&gt;
  &lt;li&gt;V. Shklovsky: &lt;strong&gt;In Defence of the Sociological Method [1927].&lt;/strong&gt; Translated by Ann Shukman. (pp. 92–99)&lt;/li&gt;
  &lt;li&gt;Ann Shukman: Russian Formalism: A Bibliography of Translations and Commentaries (Works in English, French, German and Italian). (pp. 100–108)&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;vol-5-formalism-history-comparison-genre-1978-93pp&quot;&gt;Vol. 5. Formalism: History, Comparison, Genre (1978, 93pp.)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Boris Eikhenbaum: &lt;strong&gt;Some Principles of Literary History: The Study of Lermontov [1924].&lt;/strong&gt; Translated by Christopher Pike. (pp. 1–8)&lt;/li&gt;
  &lt;li&gt;Yu. N. Tynyanov: &lt;strong&gt;Tyutchev and Heine [1929].&lt;/strong&gt; Translated by Christopher Pike. (pp. 9–19)&lt;/li&gt;
  &lt;li&gt;Yu. N. Tynyanov: &lt;strong&gt;Plot and Story-Line in the Cinema [1926].&lt;/strong&gt; Translated by Ann Shukman. (pp. 20–21)&lt;/li&gt;
  &lt;li&gt;Yu. N. Tynyanov: &lt;strong&gt;Preface to “The Problem of Verse Semantics” [1923].&lt;/strong&gt; Translated by Ann Shukman. (pp. 22–23)&lt;/li&gt;
  &lt;li&gt;V. V. Vinogradov: &lt;strong&gt;The Tasks Facing Stylistics [1923]. Translated by Ann Shukman.&lt;/strong&gt; (pp. 24–29)&lt;/li&gt;
  &lt;li&gt;Olga Freidenberg: &lt;strong&gt;Three Plots or the Semantics of One: Shakespeare’s “The Taming of the Shrew” [1925/1936].&lt;/strong&gt; Translated by Ann Shukman and Halina Willetts. (pp. 30–51)&lt;/li&gt;
  &lt;li&gt;Boris Tomashevsky: &lt;strong&gt;Literary Genres [I. Dramatic Genres; II. Lyric Poetry; III. Narrative Genres] [1928].&lt;/strong&gt; Translated by L. M. O’Toole. (pp. 52–93)&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;vol-6-dramatic-structure-poetic-and-cognitive-semantics-1979-96pp&quot;&gt;Vol. 6. Dramatic Structure; Poetic and Cognitive Semantics (1979, 96pp.)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Yu. K. Scheglov: &lt;strong&gt;The Poetics of Molière’s Comedies [1977].&lt;/strong&gt; Part One: The Structural Invariants. Translated by L. M. O’Toole. Part Two: The Structure of One Comedy: &lt;em&gt;M. de Pourceaugnac&lt;/em&gt;. Translated by N. F. C. Owen. (pp. 1–83)&lt;/li&gt;
  &lt;li&gt;Yu. M. Lotman: &lt;strong&gt;Culture as Collective Intellect and the Problems of Artificial Intelligence.&lt;/strong&gt; Translated by Ann Shukman. (pp. 84–96) (&lt;strong&gt;&lt;a href=&quot;https://www.flfi.ut.ee/sites/default/files/www_ut/lotman_1979_culture_as_collective_intellect.pdf&quot;&gt;PDF&lt;/a&gt;&lt;/strong&gt;)&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;vol-7-metre-rhythm-stanza-rhyme-1980-105pp&quot;&gt;Vol. 7. Metre, Rhythm, Stanza, Rhyme (1980, 105pp.)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;M. L. Gasparov: &lt;strong&gt;Quantitative Methods in Russian Metrics: Achievements and Prospects [1974].&lt;/strong&gt; Translated by G. S. Smith. (pp. 1–19)&lt;/li&gt;
  &lt;li&gt;K. F. Taranovsky: &lt;strong&gt;The Rhythmical Structure of Russian Binary Metres [1971].&lt;/strong&gt; Translated by G. S. Smith. (pp. 20–30)&lt;/li&gt;
  &lt;li&gt;M. L. Gasparov: &lt;strong&gt;Light and Heavy Verse Lines [1977].&lt;/strong&gt; Translated by G. S. Smith. (pp. 31–44)&lt;/li&gt;
  &lt;li&gt;M. G. Tarlinskaya: &lt;strong&gt;The Problem of English Versification: A Re-Examination [1976].&lt;/strong&gt; Translated by G. S. Smith. (pp. 45–60)&lt;/li&gt;
  &lt;li&gt;M. L. Gasparov: &lt;strong&gt;Towards an Analysis of Russian Inexact Rhyme [1975].&lt;/strong&gt; Translated by G. S. Smith. (pp. 61–75)&lt;/li&gt;
  &lt;li&gt;K. D. Vishnevsky: &lt;strong&gt;The Law of Rhythmical Correspondence in Heterogeneous Stanzas [1976].&lt;/strong&gt; Translated by G. S. Smith. (pp. 76–85)&lt;/li&gt;
  &lt;li&gt;Yu. M. Lotman: &lt;strong&gt;The Natural Language/Metre Inter-Relationship in the Mechanism of Verse [1974].&lt;/strong&gt; Translated by G. S. Smith. (pp. 86–89)&lt;/li&gt;
  &lt;li&gt;References and Bibliography. (pp. 90–105)&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;vol-8-film-theory-and-general-semiotics-1981-107pp&quot;&gt;Vol. 8. Film Theory and General Semiotics (1981, 107pp.)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;V. V. Ivanov: &lt;strong&gt;Functions and Categories of Film Language [1971/1975].&lt;/strong&gt; Translated by Stephen Rudy. (pp. 1–35)&lt;/li&gt;
  &lt;li&gt;Yu. M. Lotman: &lt;strong&gt;On the Language of Animated Cartoons [1979].&lt;/strong&gt; Translated by Ruth Sobel. (pp. 36–39)&lt;/li&gt;
  &lt;li&gt;A. K. Zholkovsky: &lt;strong&gt;Generative Poetics in the Writings of Eisenstein [1969].&lt;/strong&gt; Translated by L. M. O’Toole. (pp. 40–61)&lt;/li&gt;
  &lt;li&gt;A. K. Zholkovsky: &lt;strong&gt;Pushkin’s “Poetic World” [1976].&lt;/strong&gt; Translated by L. M. O’Toole. (pp. 62–107)&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;vol-9-the-poetics-of-cinema-1982-126pp&quot;&gt;Vol. 9. The Poetics of Cinema (1982, 126pp.)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;K. Shutko: &lt;strong&gt;Preface.&lt;/strong&gt; Translated by Richard Taylor. (pp. 1–4)&lt;/li&gt;
  &lt;li&gt;B. Eikhenbaum: &lt;strong&gt;Problems of Cine-Stylistics.&lt;/strong&gt; Translated by Richard Sherwood. (pp. 5–31)&lt;/li&gt;
  &lt;li&gt;Yury Tynyanov: &lt;strong&gt;The Fundamentals of Cinema.&lt;/strong&gt; Translated by L. M. O’Toole. (pp. 32–54)&lt;/li&gt;
  &lt;li&gt;B. Kazansky: &lt;strong&gt;The Nature of Cinema.&lt;/strong&gt; Translated by Joe Andrew. (pp. 55–86)&lt;/li&gt;
  &lt;li&gt;V. Shklovsky: &lt;strong&gt;Poetry and Prose in the Cinema.&lt;/strong&gt; Translated by Richard Taylor. (pp. 87–89)&lt;/li&gt;
  &lt;li&gt;A. Piotrovsky: &lt;strong&gt;Towards a Theory of Film Genres.&lt;/strong&gt; Translated by Richard Taylor. (pp. 90–106)&lt;/li&gt;
  &lt;li&gt;E. Mikhailov; A. Moskvin: &lt;strong&gt;The Cameraman’s Part in Making a Film.&lt;/strong&gt; Translated by Ann Shukman. (pp. 107–118)&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;vol-10-bakhtin-school-papers-1983-156pp&quot;&gt;Vol. 10. Bakhtin School Papers (1983, 156pp.)&lt;/h2&gt;
&lt;ul&gt;
  &lt;li&gt;Ann Shukman: Introduction. (pp. 1–4)&lt;/li&gt;
  &lt;li&gt;V. N. Voloshinov [M. M. Bakhtin]: &lt;strong&gt;Discourse in Life and Discourse in Poetry. Questions of Sociological Poetics.&lt;/strong&gt; Translated by John Richmond. (pp. 5–30)&lt;/li&gt;
  &lt;li&gt;V. N. Voloshinov [M. M. Bakhtin]: &lt;strong&gt;The Latest Trends in Linguistic Thought in the West.&lt;/strong&gt; Translated by Noel Owen. (pp. 31–49)&lt;/li&gt;
  &lt;li&gt;P. N. Medvedev: &lt;strong&gt;The Formal (Morphological) Method or Scholarly Salieri-ism.&lt;/strong&gt; Translated by Ann Shukman. (pp. 51–65)&lt;/li&gt;
  &lt;li&gt;P. N. Medvedev: &lt;strong&gt;Sociologism Without Sociology (On the Methodological Works of P. N. Sakulin).&lt;/strong&gt; Translated by C. R. Pike. (pp. 67–74)&lt;/li&gt;
  &lt;li&gt;P. N. Medvedev [M. M. Bakhtin]: &lt;strong&gt;The Immediate Tasks Facing Literary-Historical Science.&lt;/strong&gt; Translated by C. R. Pike. (pp. 75–91)&lt;/li&gt;
  &lt;li&gt;V. N. Voloshinov [M. M. Bakhtin]: &lt;strong&gt;Literary Stylistics 1: What is Language?&lt;/strong&gt; Translated by Noel Owen. (pp. 93–113)&lt;/li&gt;
  &lt;li&gt;V. N. Voloshinov [M. M. Bakhtin]: &lt;strong&gt;Literary Stylistics 2: The Construction of the Utterance.&lt;/strong&gt; Translated by Noel Owen. (pp. 114–138)&lt;/li&gt;
  &lt;li&gt;V. N. Voloshinov [M. M. Bakhtin]: &lt;strong&gt;Literary Stylistics 3: The Word and its Social Function.&lt;/strong&gt; Translated by Joe Andrew. (pp. 139–152)&lt;/li&gt;
  &lt;li&gt;Glossary of Key Terms. (pp. 153–155)&lt;/li&gt;
  &lt;li&gt;Bibliography. (p. 156)&lt;/li&gt;
&lt;/ul&gt;
</content>
    <author>
      <name>
	
          
          Ingo Börner, 
	
          
          Angelika Hechtl, 
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Underwater Basket Weaving: Russian Writers and Their Professions According to Wikidata</title>
    <link href="https://weltliteratur.net/russian-writers-and-their-professions/"/>
    <updated>2018-08-20T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/russian-writers-and-their-professions</id>
    <content type="html">&lt;h1 id=&quot;1-introduction&quot;&gt;1. Introduction&lt;/h1&gt;

&lt;h2 id=&quot;11-tldr&quot;&gt;1.1. &lt;a href=&quot;https://en.wikipedia.org/wiki/TL;DR&quot;&gt;TL;DR&lt;/a&gt;&lt;/h2&gt;

&lt;p&gt;This is a little experiment which we conducted in a limited amount of time (one week). &lt;strong&gt;We wanted to find out how the range of jobs practised by Russian writers throughout history is represented in &lt;a href=&quot;https://www.Wikidata.org/&quot;&gt;Wikidata&lt;/a&gt;.&lt;/strong&gt; Since Wikidata is primarily filled in an automated fashion, we wanted to look for systematic errors in that process. However, we also dared to interpret the data on Russian authors at least a little, even if at this point it is neither complete nor specifically curated within Wikidata. As an appetiser, have a look at fig. 3:&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/russian-writers/figure_3.jpg&quot; alt=&quot;Figure 3&quot; style=&quot;width:1200px;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 3.&lt;/b&gt; The 16 most popular professions among Russian writers between 1700 and 2000 (in 50-year spans).&lt;/center&gt;

&lt;p&gt;No surprise there, writers generally tended to (also) be journalists, translators, literary critics. But now onto some details of our exercise…&lt;/p&gt;

&lt;h2 id=&quot;12-underwater-basket-weaving&quot;&gt;1.2. “Underwater Basket Weaving”&lt;/h2&gt;

&lt;p&gt;Originally, &lt;a href=&quot;https://en.wikipedia.org/wiki/Underwater_basket_weaving&quot;&gt;according to Wikipedia&lt;/a&gt;, “underwater basket weaving” is an idiom “referring in a negative way to supposedly useless or absurd college or university courses”, or it can serve as an “intentionally humorous generic answer to questions about an academic degree”. As we were working on our project, we encountered a writer whose occupation was marked as “basket weaving”. In search for a good Russian translation we’ve encountered the idiom mentioned above. It seemed so funny and to the point that we named our project after it. During Soviet times, many famous writers were not officially recognised as such. Hence, they had to do some pseudo or meaningless work, something along the lines of “underwater basket weaving”.&lt;/p&gt;

&lt;h2 id=&quot;13-russian-writers-professions-and-languages&quot;&gt;1.3. Russian Writers’ Professions and Languages&lt;/h2&gt;

&lt;p&gt;The original goal of our project was to check the variety of jobs held by Russian writers. This is of interest since for the most part of Russian history it was hard to make a living off of just being a writer. Next to that question, and since we were experimenting with Wikidata, we also thought it would be interesting to see what languages except Russian the writers were able to speak. This could be interesting as an indicator of (foreign) cultural influences on Russian literary history.&lt;/p&gt;

&lt;h2 id=&quot;14-using-wikidata-to-answer-scientific-questions&quot;&gt;1.4. Using Wikidata to Answer Scientific Questions&lt;/h2&gt;

&lt;p&gt;Using Wikidata as data source is fun, but problematic. In order to draw conclusions from analysing this data, you have to understand how data enters Wikidata and how it is maintained. Although human users can still alter single facts just like on any other project of the Wikimedia family, the interesting part of the Wikidata experience is that most of the processes are automated. So instead of correcting single facts, it is far more interesting to look for systematic errors.&lt;/p&gt;

&lt;p&gt;While there are heavily curated subsets within Wikidata (like the one on &lt;a href=&quot;https://blog.ehri-project.eu/2018/02/12/using-Wikidata/&quot;&gt;Holocaust-era ghettos&lt;/a&gt;), our chosen subset of Russian writers is what it is at the moment and hasn’t seen any concerted action regarding its curation (at least not yet). However, we did dare to interpret the data on Russian authors at least a little, even if they are at this point neither complete nor specifically curated in Wikidata.&lt;/p&gt;

&lt;h2 id=&quot;15-ways-to-automatically-improve-wikidata-using-bots&quot;&gt;1.5. Ways to Automatically Improve Wikidata Using Bots&lt;/h2&gt;

&lt;p&gt;After collecting systematic errors, we also looked into ways to automatise the process using bots. Since we didn’t have a lot of time, we didn’t succeed to actually apply a bot, but we experimented a little and plan to explore this further, maybe in shape of another course work in the coming semester.&lt;/p&gt;

&lt;h1 id=&quot;2-description-of-data-source&quot;&gt;2. Description of Data Source&lt;/h1&gt;

&lt;h2 id=&quot;21-wikidata&quot;&gt;2.1. Wikidata&lt;/h2&gt;

&lt;p&gt;Wikidata is a free and open knowledge base that can be read and edited by both humans and machines. Wikidata acts as central storage point for structured data of all projects of the Wikimedia family (but is not limited to that). Each entity in Wikidata has a label (a name, for example), an identifier and several statements (such as “occupation”, “day of birth”, etc. for people).&lt;/p&gt;

&lt;h2 id=&quot;22-russian-wikipedia-for-fact-checking&quot;&gt;2.2. Russian Wikipedia for Fact-Checking&lt;/h2&gt;

&lt;p&gt;When we spotted something suspicious in Wikidata, we cross-checked this information with the Russian Wikipedia. There were a lot of misattributions and systematic mistranslations, so we checked the original articles in Russian to find the source for this kind of mistakes.&lt;/p&gt;

&lt;h2 id=&quot;23-pywikibot-library&quot;&gt;2.3. Pywikibot Library&lt;/h2&gt;

&lt;p&gt;As Wikidata provides the instruments for automated editing, we set out to fix some Wikidata mistakes using a bot. There is a special Python library for that, &lt;a href=&quot;https://www.mediawiki.org/wiki/Manual:Pywikibot&quot;&gt;Pywikibot&lt;/a&gt;, which can be used to create a bot. It can be launched in different ways among which we chose to host it on the Wikimedia labs environment using &lt;a href=&quot;https://www.mediawiki.org/wiki/PAWS&quot;&gt;PAWS: A Web Shell&lt;/a&gt;.&lt;/p&gt;

&lt;h1 id=&quot;3-workflow&quot;&gt;3. Workflow&lt;/h1&gt;

&lt;h2 id=&quot;31-querying-wikidata&quot;&gt;3.1. Querying Wikidata&lt;/h2&gt;

&lt;p&gt;We used Wikidata to identify Russian writers and their respective occupations, languages spoken, gender and year of death. To be able to look into temporal distribution, we chose the year of death as marker because we thought this was closer to their active time as writers than their birth date. The obvious limitation that comes with this approach is that writers still alive or ones lacking a year of death were not included into the analysis. But this way we were also closer to the canon of Russian literature, making us a bit more independent of Wikipedia’s “contemporary bias”.&lt;/p&gt;

&lt;h2 id=&quot;32-dataset-creation&quot;&gt;3.2. Dataset Creation&lt;/h2&gt;

&lt;p&gt;Now, how did we find our set of Russian writers? We included people labeled as writers in Wikidata. They had to be either Russian by their nationality, have Russian as a native language or have to lived in the Russian Empire, the USSR or the Russian Federation. We excluded all people who did not feature Russian among their languages.&lt;/p&gt;

&lt;h2 id=&quot;33-systematic-problems-in-wikidata&quot;&gt;3.3. Systematic Problems in Wikidata&lt;/h2&gt;

&lt;p&gt;As we reviewed our raw data, we noticed a couple of errors and tried to classify them. We also made a table of all the mistakes we’ve encountered in order to then check whether they could be fixed by a bot (a link to our GitHub is below).&lt;/p&gt;

&lt;h2 id=&quot;34-attempt-to-create-a-wikibot&quot;&gt;3.4. Attempt to Create a Wikibot&lt;/h2&gt;

&lt;p&gt;As Wikidata is an open resource with an enormous amount of data, using bots is actually dangerous. So, there are some mandatory stages you have to go through when &lt;a href=&quot;https://www.wikidata.org/wiki/Wikidata:Bots&quot;&gt;creating a bot&lt;/a&gt;. We set up a regular Wikimedia account. To clarify that it is a bot you just have to mention the word “bot” at the end of username. To be approved to work on Wikidata your bot has to make several operations which will be assessed. If everything is good, the bot gets approved. When writing your bot script you can test it on the &lt;a href=&quot;https://www.wikidata.org/wiki/Wikidata:Sandbox&quot;&gt;Wikidata:Sandbox&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;Ultimately, we did not get to the stage of approval this time around because of struggling with several problems in the preparation process. Details are described below.&lt;/p&gt;

&lt;h2 id=&quot;35-manual-corrections&quot;&gt;3.5. Manual Corrections&lt;/h2&gt;

&lt;p&gt;Due to the limited amount of time, we did not get to the point were we could correct systematic errors. Since our dataset is very limited, we manually corrected all mistakes spotted, and then repeated our queries.&lt;/p&gt;

&lt;h2 id=&quot;36-json-unification-in-python&quot;&gt;3.6. JSON Unification in Python&lt;/h2&gt;

&lt;p&gt;Our queries initially had every writer occur as many times as he had professions plus spoken languages. To make the JSON more convenient, we wrote a Python script to transform raw data into a dictionary with unique names under which was stored a multi-layered dictionary of occupations, known languages, gender and year of death.&lt;/p&gt;

&lt;h2 id=&quot;37-visualisation-in-python&quot;&gt;3.7. Visualisation in Python&lt;/h2&gt;

&lt;p&gt;The created dictionary was used to visualise information with by help of the following Python libraries: pandas, seaborn and matplotlib. At first, the data was reorganised into tables in order to create pandas dataframes. Thereafter, the seaborn graphs were compiled based on the aforementioned dataframes. Visual aspects were set up such as labels of the axes, captions, colour palettes, fonts, etc. Finally, all the charts were saved and displayed via matplotlib library.&lt;/p&gt;

&lt;h1 id=&quot;4-wikidata-errors&quot;&gt;4. Wikidata Errors&lt;/h1&gt;

&lt;h2 id=&quot;41-mistranslations&quot;&gt;4.1. Mistranslations&lt;/h2&gt;

&lt;p&gt;Most of the problems we faced in our research were systematic, for example, mistranslations. It is quite obvious that &lt;a href=&quot;https://en.wikipedia.org/wiki/Yanka_Kupala&quot;&gt;Yanka Kupala&lt;/a&gt; was not an “imagemaker”. The error occurred due to automatic translation. In the Russian WP he was marked as “публицист” (transcribed “publicist” = “opinion journalist”), but the automatic translated the publicist back to &lt;a href=&quot;https://www.wikidata.org/wiki/Q4178004&quot;&gt;“имиджмейкер”&lt;/a&gt; (“imagemaker”).&lt;/p&gt;

&lt;h2 id=&quot;42-misclassification-due-to-mistranslation&quot;&gt;4.2. Misclassification Due to Mistranslation&lt;/h2&gt;

&lt;p&gt;In many cases there was a generic name instead of the actual job title (mostly due to mistranslation). For example, according to Wikidata, Bulgakov &lt;em&gt;is&lt;/em&gt; &lt;a href=&quot;https://www.wikidata.org/w/index.php?title=Q835&amp;amp;oldid=692018680&quot;&gt;speculative fiction&lt;/a&gt; and Gennady Malakhov &lt;em&gt;is&lt;/em&gt; &lt;a href=&quot;https://www.wikidata.org/w/index.php?title=Q1997859&amp;amp;oldid=681291374&quot;&gt;urine therapy&lt;/a&gt;. Such mistakes can be found easily, though. If as an occupation there is an entity which is not an ‘instance of profession’ but rather ‘instance of genre’ or something similar, it is an error. Yet in some cases there is no easy way to derive the original mistranslated job title as sometimes the mistranslation is too vague, such as ‘Orthodox Christianity’, making it harder to find out what the original job title was.&lt;/p&gt;

&lt;h2 id=&quot;43-redundancies&quot;&gt;4.3. Redundancies&lt;/h2&gt;

&lt;p&gt;The last systematic problem we came across was the replication or synonymous labelling of entities such as “police officer” (Q361593 and Q384593) or ethnologist (Q1371378) versus ethnographer (Q12347522). Yet sometimes it is not quite clear whether articles should be merged. It is not obvious whether “official” (“должностное лицо”, Q599151), “official” (“чиновник”, Q382887) and “civil servant” (“государственный служащий”, Q212238) convey the same meaning.&lt;/p&gt;

&lt;h2 id=&quot;44-wikibot-for-error-correction&quot;&gt;4.4. Wikibot for Error Correction&lt;/h2&gt;

&lt;p&gt;Bots can be useful in solving some of the problems described above, especially in cases of misclassification. This involves the automated changing of statements, which is more complicated than just changing labels like for mistranslations.&lt;/p&gt;

&lt;p&gt;However, the main problem we encountered was connected to bot testing within the sandbox. The structure of the sandbox is very different from the actual one, the two do not contain any identical entities, meaning there are no pages of Russian writers in the sandbox equal to the ones in the original Wikidata. In order to test our bot we had to create like 10 to 20 pages, imitating the actual Wikidata pages with all the properties we need, and only then try to work with our bot on this material. In the short amount of time that we had to prepare our project, we did not succeed to do so, but this is where we would have to continue next time around.&lt;/p&gt;

&lt;h1 id=&quot;5-results-and-tentative-interpretations&quot;&gt;5. Results and Tentative Interpretations&lt;/h1&gt;

&lt;p&gt;We identified 8160 Russian writers in Wikidata meeting our eligibility criteria. Not all the fields were filled in for all of them, so those who lacked some pieces of information were excluded from the corresponding analysis. For example, we couldn’t put writers without year of death in our timeline. The amount of data on those who already have attributed a year of death can be seen in figure 1. As expected, the period best covered is the 20th century. The amount of data for the 19th century was significantly lower, and the first half of the 21st century appears to be under-represented, too, for obvious reasons.&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/russian-writers/figure_1.png&quot; alt=&quot;Figure 1&quot; style=&quot;width:1200px;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 1.&lt;/b&gt; Amount of Russian writers in Wikidata.&lt;/center&gt;

&lt;h2 id=&quot;51-occupations&quot;&gt;5.1. Occupations&lt;/h2&gt;

&lt;h3 id=&quot;511-general&quot;&gt;5.1.1. General&lt;/h3&gt;

&lt;p&gt;Figure 2 shows the distribution of different occupations of Russian writers. Unsurprisingly, the most popular professions are those where the ability to write well is a prerequisite: journalism, translation, literary criticism.&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/russian-writers/figure_2.jpg&quot; alt=&quot;Figure 2&quot; style=&quot;width:1200px;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 2.&lt;/b&gt; Percentage of different occupations among Russian authors.&lt;/center&gt;

&lt;h3 id=&quot;512-by-time&quot;&gt;5.1.2. By Time&lt;/h3&gt;

&lt;p&gt;Figure 3 only shows the professions that had more than 5 people assigned to them in each period. We witness the decline of linguistics, philosophy, criticism, translation, diplomacy, history, education over the centuries, possibly also due to a diversifying landscape of professions, and the rise of journalism.&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/russian-writers/figure_3.jpg&quot; alt=&quot;Figure 3&quot; style=&quot;width:1200px;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 3.&lt;/b&gt; The 16 most popular professions among Russian writers.&lt;/center&gt;

&lt;h3 id=&quot;513-by-gender&quot;&gt;5.1.3. By Gender&lt;/h3&gt;

&lt;p&gt;Figure 4 shows that women writers had fewer professional options in past centuries, mostly limited to “work from home” options. They worked as interpreters, linguists, philologists, artists or were salonnières. An exception to this rule are the actresses among women writers.&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/russian-writers/figure_4.jpg&quot; alt=&quot;Figure 4&quot; style=&quot;width:1200px;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 4.&lt;/b&gt; Popularity of professions among male and female writers.&lt;br /&gt; &lt;/center&gt;

&lt;h3 id=&quot;514-just-a-writer-by-time--gender&quot;&gt;5.1.4. Just a Writer (by Time &amp;amp; Gender)&lt;/h3&gt;

&lt;p&gt;Figure 5 shows the percentage of writers only marked as &lt;em&gt;writers&lt;/em&gt; on Wikidata without additional occupations. The situation remains quite stable considering men, whereas the percentage of women falls from 100% to approximately 70% by the beginning of 19th century and then gradually declines to 45–50%. We tend to explain this by the fact that women gradually entered the labour market. Towards the 21st century, there is a 50% chance that being a writer means that a person also has an additional profession, according to Wikidata.&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/russian-writers/figure_5.png&quot; alt=&quot;Figure 5&quot; style=&quot;width:1200px;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 5.&lt;/b&gt; Percentage of writers whose only profession in Wikidata is writer.&lt;br /&gt; &lt;/center&gt;

&lt;h2 id=&quot;52-languages-spoken&quot;&gt;5.2. Languages Spoken&lt;/h2&gt;
&lt;h3 id=&quot;521-general&quot;&gt;5.2.1. General&lt;/h3&gt;

&lt;p&gt;Figure 6 shows the distribution of additional languages Russian writers spoke or wrote in. There are two distinct types of languages: native regional languages spoken in territories of the former Russian Empire and the USSR (Ukrainian, Bashkir, Yiddish, etc.) and secondary European languages (French, German, English, etc.).&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/russian-writers/figure_6.jpg&quot; alt=&quot;Figure 6&quot; style=&quot;width:1200px;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 6.&lt;/b&gt; Popularity of additional languages mastered by Russian writers.&lt;br /&gt; &lt;/center&gt;

&lt;h3 id=&quot;522-by-gender&quot;&gt;5.2.2. By Gender&lt;/h3&gt;

&lt;p&gt;Figure 7 uncovers some interesting tendencies. Women writers in our data set mastered more European languages, such as English, French, German, and Italian, than male writers did. Women also knew better than men some smaller, less well-represented languages, like Armyan, Buryat, Hungarian, Saami.&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/russian-writers/figure_7.jpg&quot; alt=&quot;Figure 7&quot; style=&quot;width:1200px;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 7.&lt;/b&gt; Languages spoken by Russian writers, divided by gender.&lt;br /&gt; &lt;/center&gt;

&lt;h3 id=&quot;523-by-time&quot;&gt;5.2.3. By Time&lt;/h3&gt;

&lt;p&gt;The distribution of languages spoken by Russian writers by time (figure 8) can reflect cultural influences of given periods. Between 1800–1850, the most popular language was German, which might be due to the German-orientated culture established by Peter the Great. During the second half of the 19th century (1850–1900), the most popular language was French, which corresponds to the big influence that French culture had on Russian culture at the time. At the beginning of the 20th century (1900–1950), the most popular languages were Ukrainian, French and Yiddish. During the second half of the 20th century (1950–2000), English joined the parade. In the most contemporary period, the most popular additional languages are French, Ukrainian and English. It is worth noting that the gap between popular and unpopular languages decreased over time, possibly an effect of globalisation, leading to a mix of contemporary cultures and languages exerting an influence on Russian literary life.&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/russian-writers/figure_8.jpg&quot; alt=&quot;Figure 8&quot; style=&quot;width:1200px;&quot; /&gt;
&lt;/figure&gt;
&lt;center&gt;&lt;b&gt;Fig. 8.&lt;/b&gt; Languages spoken by Russian writers, divided by time.&lt;br /&gt; &lt;/center&gt;

&lt;h1 id=&quot;6-appendix&quot;&gt;6. Appendix&lt;/h1&gt;

&lt;p&gt;All files, Python code, images, Wikidata search queries, etc. can be found &lt;a href=&quot;https://github.com/brouhahaha/dh_final_project&quot;&gt;on our GitHub&lt;/a&gt;.&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Galina Ryazanskaya, 
	
          
          Alena Shchevyeva, 
	
          
          Maria Suvorova
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Network Analysis of the 9th Season of "Curb Your Enthusiasm"</title>
    <link href="https://weltliteratur.net/curb-your-enthusiasm-season-9-network-analysis/"/>
    <updated>2018-01-01T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/curb-your-enthusiasm-season-9-network-analysis</id>
    <content type="html">&lt;p&gt;2017 brought us the ninth season of &lt;a href=&quot;https://en.wikipedia.org/wiki/Curb_Your_Enthusiasm&quot;&gt;“Curb Your Enthusiasm”&lt;/a&gt; (six whole years after season eight). I wanted to do a quick network analysis of the new season to check what we can learn from such a formalisation. To those of you who have watched it, this list of all ten episodes might serve as a quick recap, as “Curb” episode titles are usually good summaries, too:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;9x01: Foisted!&lt;/li&gt;
  &lt;li&gt;9x02: The Pickle Gambit&lt;/li&gt;
  &lt;li&gt;9x03: A Disturbance in the Kitchen&lt;/li&gt;
  &lt;li&gt;9x04: Running with the Bulls&lt;/li&gt;
  &lt;li&gt;9x05: Thank You for Your Service&lt;/li&gt;
  &lt;li&gt;9x06: The Accidental Text on Purpose&lt;/li&gt;
  &lt;li&gt;9x07: Namaste&lt;/li&gt;
  &lt;li&gt;9x08: Never Wait for Seconds!&lt;/li&gt;
  &lt;li&gt;9x09: The Shucker&lt;/li&gt;
  &lt;li&gt;9x10: Fatwa!&lt;/li&gt;
&lt;/ul&gt;

&lt;h1 id=&quot;extraction-method-and-visualisation&quot;&gt;Extraction Method and Visualisation&lt;/h1&gt;

&lt;p&gt;To link two characters to each other I used the same approach as in our research on social networks in dramatic texts: “Two characters interact with one another if they perform a speech act within the same segment of a drama (usually a ‘scene’).” (quoted from &lt;a href=&quot;https://dlina.github.io/presentations/2016-krakow/#/2/3&quot;&gt;these slides&lt;/a&gt;; just replace ‘drama’ by ‘TV series’)&lt;/p&gt;

&lt;p&gt;I segmented all 10 episodes by hand (414 segments across all ten episodes) and put them in the very simple format that our own tool &lt;a href=&quot;https://ezlinavis.dracor.org/&quot;&gt;ezlinavis&lt;/a&gt; parses to calculate network relations. And here’s the result after Gephi’ing it up a bit (licensed under &lt;a href=&quot;https://creativecommons.org/licenses/by/4.0/&quot;&gt;CC BY 4.0&lt;/a&gt;):&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/curb-season-9-network-graph-1024px.png&quot; alt=&quot;Curb Your Enthusiasm, Season 9, Network Graph&quot; style=&quot;width:1024px;&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;No surprise there, “Curb” by all means is a Larry-centric show. There is hardly any scene without him, and each character of the show is only defined by his or her relation to Larry. We don’t know what Richard Lewis is doing or what Jeff and Susie are talking about when Larry’s not around. A notable exception in the above graph can be spotted in the left upper corner, Morsi’s visits to characters of older seasons of the show in episode 9x08.&lt;/p&gt;

&lt;p&gt;While the Larry-centricity should be no news to watchers of the show, it is another thing to see this visualised for clarity.&lt;/p&gt;

&lt;h1 id=&quot;some-network-related-rankings&quot;&gt;Some Network-Related Rankings&lt;/h1&gt;

&lt;p&gt;Let’s look at some network-analytical metrics to get some more insights. The whole graph has 152 nodes and 427 edges. Most of the nodes are individual characters (with proper or generic names), some of them are groups of people (like the Male Hotel Guests trying to open a pickle jar in episode 9x02 or the Bus People in episode 9x07).&lt;/p&gt;

&lt;p&gt;The network density is 0.037, a very low value, which has to do with the fact that the show is not interested in establishing relations between all its characters. Like mentioned above, all characters only come to life in their relation to Larry. The average path length is 2.197, not really a very meaningful value either for this network graph. It just means that, hypothetically, every character can reach any other character of the network via Larry, in no more than two steps, with a few exceptions.&lt;/p&gt;

&lt;p&gt;More interesting are some node-related metrics, let’s have a look at the 10 characters with the highest degree:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;Character&lt;/th&gt;
      &lt;th&gt;Degree&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;Larry&lt;/td&gt;
      &lt;td&gt;132&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Jeff&lt;/td&gt;
      &lt;td&gt;36&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Susie&lt;/td&gt;
      &lt;td&gt;29&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Leon&lt;/td&gt;
      &lt;td&gt;27&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Funkhouser&lt;/td&gt;
      &lt;td&gt;20&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Lin&lt;/td&gt;
      &lt;td&gt;19&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Sign Language Interpreter&lt;/td&gt;
      &lt;td&gt;19&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Victor&lt;/td&gt;
      &lt;td&gt;17&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Ted&lt;/td&gt;
      &lt;td&gt;16&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Cheryl&lt;/td&gt;
      &lt;td&gt;15&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;So out of 151 possible connections to other characters in the network, Larry directly relates to 132 of them. A bit surprising is the high-ranking position of the Sign Language Interpreter who only appears in the last episode 9x10 (spot her in the bottom right corner of the network visualisation). Her appearance was a bit difficult to formalise. Her interpretation for the deaf is recognised by many of the main characters, which is why I established a link in each of these cases. Her degree value is high up also because she is present at two different events with two different audiences, the rehearsal for the musical and Sammi’s wedding.&lt;/p&gt;

&lt;p&gt;But there’s a way to rank characters by their actual presence throughout the whole season. Let’s check their weighted degree (or, ‘strength’). The top ten looks a tad different here:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;Character&lt;/th&gt;
      &lt;th&gt;Weighted Degree&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;Larry&lt;/td&gt;
      &lt;td&gt;573&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Jeff&lt;/td&gt;
      &lt;td&gt;234&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Susie&lt;/td&gt;
      &lt;td&gt;153&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Leon&lt;/td&gt;
      &lt;td&gt;115&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Funkhouser&lt;/td&gt;
      &lt;td&gt;73&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Ted&lt;/td&gt;
      &lt;td&gt;69&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Cheryl&lt;/td&gt;
      &lt;td&gt;68&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Lin&lt;/td&gt;
      &lt;td&gt;67&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Bridget&lt;/td&gt;
      &lt;td&gt;66&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;Richard&lt;/td&gt;
      &lt;td&gt;53&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;The Sign Language Interpreter as well as Sammi’s husband-to-be Victor fall out of this ranking, while Bridget (Larry’s girlfriend during some of the episodes) and his old pal Richard Lewis enter the picture. Ted and Cheryl, who are now an item, also rank higher, showing that they actually did occur on different occasions during the season, also if they didn’t play a central part.&lt;/p&gt;

&lt;h1 id=&quot;mrs-templeton-and-other-intricacies&quot;&gt;Mrs. Templeton and Other Intricacies&lt;/h1&gt;

&lt;p&gt;Segmenting the episodes and identifying and placing all characters can be challenging. In general, I tried to use the cast lists &lt;a href=&quot;www.imdb.com/title/tt0264235/episodes?season=9&quot;&gt;on IMDb&lt;/a&gt; to find the right names for all characters, especially if they aren’t given actual names (but are listed under generic terms like “Hotel Day Manager” or “Receptionist”). If known characters appear under different names (Larry as “Buck Dancer” and Leon as “Chappie Johnson” in episode 9x02), I ignored that. In some few cases I had to introduce group names like “Male Hotel Guests”, because the characters didn’t act as distinguishable individuals but as part of a group.&lt;/p&gt;

&lt;p&gt;And then there’s Mrs. Templeton, wife of Larry’s therapist Dr. Templeton in episode 9x04. We never see her talking at all, but I interpreted her wordless snubbing of Larry as speech act, which is why she appears in the graph. All in all, I tried to be as objective as possible. Feel free to have a look at the segmentation files.&lt;/p&gt;

&lt;h1 id=&quot;data-and-software-used&quot;&gt;Data and Software Used&lt;/h1&gt;

&lt;p&gt;Speaking of which, I opened a repo on GitHub for this small project, see &lt;a href=&quot;https://github.com/lehkost/curb-your-enthusiasm&quot;&gt;here&lt;/a&gt;, where you can find the formalisation for each episode. After formalising, I threw the files into &lt;a href=&quot;https://ezlinavis.dracor.org/&quot;&gt;ezlinavis&lt;/a&gt; and saved the resulting network data in an CSV file, which I then opened in &lt;a href=&quot;https://gephi.org/&quot;&gt;Gephi&lt;/a&gt; to calculate and visualise everything you’ve seen above.&lt;/p&gt;

&lt;p&gt;As for “Curb”, word has it that we won’t have to wait another six years for the next season, production for season ten is said to start &lt;a href=&quot;https://nyti.ms/2jWj3Mx&quot;&gt;next spring&lt;/a&gt;.&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>A Giant 1890 Flowchart of Foreign Influences on German Literature (Spanning 12 Centuries)</title>
    <link href="https://weltliteratur.net/A-Giant-1890-Flowchart-of-Foreign-Influences-on-German-Literature/"/>
    <updated>2016-11-03T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/A-Giant-1890-Flowchart-of-Foreign-Influences-on-German-Literature</id>
    <content type="html">&lt;p&gt;Exploring the influence that authors have (or might have) exerted on other authors is among the standard procedures pursued by scholars of (Comparative) Literary Studies. Provided the right type of data, these kinds of relations can be easily visualised in a network graph with arrows pointing in myriad directions: X influenced Y and, possibly, vice versa. What we have here, though, is a collection of unidirectional pointers which – in the form of streams of movements, events, works, and authors of world literature – exert their influence on one particular, ever-broader river: German literature, its manifold protagonists and epochs.&lt;/p&gt;

&lt;p&gt;It is Cäsar Flaischlen’s &lt;strong&gt;“Graphische Litteratur-Tafel”&lt;/strong&gt; (Graphic Literature Table) we’re talking about, a 58×86.5-cm poster published in 1890, and a flamboyantly beautiful one:&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/graphische_litteratur-tafel_1040px.jpg&quot; alt=&quot;Graphische Litteratur-Tafel (1890)&quot; style=&quot;width:1040px;&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;We came across this table just a couple of weeks ago, and quickly purchased a copy via ZVAB.com, scanned it, OCRed the 8-column preface, encoded both the text and the graphic table in TEI and put up a website to present the data, at &lt;strong&gt;&lt;a href=&quot;http://litteratur-tafel.weltliteratur.net/&quot;&gt;litteratur-tafel.weltliteratur.net&lt;/a&gt;&lt;/strong&gt; (still in alpha!). Simultaneously, we had another copy scanned by the Digitisation Centre of Göttingen’s University and State Library; you can download their hi-res version &lt;a href=&quot;http://gdz.sub.uni-goettingen.de/pdfcache/PPN860488233/PPN860488233___LOG_0004.pdf&quot;&gt;here (PDF format, 4.5 MB)&lt;/a&gt;.&lt;/p&gt;

&lt;h1 id=&quot;the-author&quot;&gt;The Author&lt;/h1&gt;

&lt;p&gt;How did this chart come about? Cäsar Flaischlen (1864–1920) was not much of a practising literary scholar, nor an academic. The dissertation he wrote (on the Enlightenment playwright Otto von Gemmingen, with a nod to Diderot) was not well-received by his colleagues (one of which said it was &lt;a href=&quot;http://www.digizeitschriften.de/dms/img/?PID=PPN345204123_0035%7Clog84&amp;amp;physid=phys600#navi&quot;&gt;“ziemlich nachlässig und schleuderhaft”&lt;/a&gt;, i.e., “quite sloppy and negligent”). Flaischlen’s dissertation was published in 1890, the same year in which his “Graphic Table of Literature” saw the light of day. Then he left academia to continue writing (dialect) poetry, novels and plays, while working as an editor for arts and literary magazines to pay the bills. (There’s &lt;a href=&quot;https://books.google.com/books?id=mLI6ZB6vrFEC&amp;amp;pg=PA1996&quot;&gt;an article&lt;/a&gt; on him in “Deutsches Literatur-Lexikon. Das 20. Jahrhundert”, vol. 9, 2006.)&lt;/p&gt;

&lt;h1 id=&quot;context-i-positivism&quot;&gt;Context I: Positivism&lt;/h1&gt;

&lt;p&gt;Flaischlen does not mention any of the sources from which he compiled his graphic table. The canon of authors and other entities he mentions is not in the least controversial and was part of the literary-historical consensus taught at universities at the time. Flaischlen’s impressive visualisation calls upon the power (and shortcomings, too) of positivist thinking (there are remarks on the outdatedness of Flaischlen’s approach in &lt;a href=&quot;https://books.google.de/books?id=XT7nBQAAQBAJ&amp;amp;pg=PA36&quot;&gt;this 2012 introduction to Comparative Literature&lt;/a&gt;).&lt;/p&gt;

&lt;p&gt;Positivism is largely seen in a negative light today, but one of the aspects behind it has been vividly discussed recently in the attempt to reinvent the Humanities along the lines of the natural sciences. “The natural sciences ride triumphantly in the chariot of victory to which we are all chained”, &lt;a href=&quot;https://books.google.ru/books?id=QJuJAgAAQBAJ&amp;amp;pg=PA47&quot;&gt;wrote Wilhelm Scherer&lt;/a&gt;, a central representative of positivism in the 1870s and 1880s. Trying to establish Literary Studies as an “exact science” was by no means unusual, as we can see in another quote by Shakespeare scholar Wilhelm Wetz who, in the 1890s, wanted to prove that Literary Studies, too, &lt;a href=&quot;https://archive.org/stream/shakespearevoms01wetzgoog#page/n12/mode/2up&quot;&gt;“can rise to the ranks of the exact sciences” (p. VII)&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;At the same time, Flaischlen’s flowchart stands in opposition to a popular positivist undercurrent of his time, biographism. Instead of looking painstakingly for possibly undiscovered details in an author’s life according to the biographistic paradigm, he dabbles in a kind of macroanalysis, an abstract model of the big picture, the organic connection.&lt;/p&gt;

&lt;p&gt;Another aspect of Flaischlen’s chart is how he deals with the question of crafting a history of literature. The first half of the 19th century had a more philosophical approach. Literary history was understood as self-realisation of a mind, of an idea. This changed in the second half of the century. Erich Schmidt (who had studied under Scherer) described the new paradigm like this (in his inaugural address at the University of Vienna in 1880):&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;“Litteraturgeschichte soll ein Stück Entwicklungsgeschichte des geistigen Lebens eines Volkes mit vergleichenden Ausblicken auf die anderen Nationallitteraturen sein. Sie erkennt das Sein aus dem Werden und untersucht wie die neuere Naturwissenschaft Vererbung und Anpassung und wieder Vererbung und so fort in fester Kette.” (&lt;a href=&quot;https://archive.org/stream/charakteristiken01schmuoft#page/490/mode/2up&quot;&gt;p. 491&lt;/a&gt;)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;This approach seeks to understand a national literature in its interconnectedness to the literature of other countries. In the same article, Schmidt continues:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;“Wie steht man zum Ausland? Der Begriff der Nationallitteratur duldet gleichwohl keinen engherzigen Schutzzoll; im geistigen Leben sind wir freihändlerisch. Aber ist Selbständigkeit oder Unselbständigkeit, größere Receptivität oder Productivität, wahre oder falsche Aneignung sichtbar, und wie hat die deutsche Litteratur sich allmählich zu universalistischer Antheilnahme emporgearbeitet?” (&lt;a href=&quot;https://archive.org/stream/charakteristiken01schmuoft#page/492/mode/2up&quot;&gt;p. 493&lt;/a&gt;)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;By and large, that is the point of Flaischlen’s flowchart.&lt;/p&gt;

&lt;p&gt;At the time, one of the leading ideas about the history of literature was Scherer’s “Wellentheorie” (wave theory). This theory assumes that literature blooms and decays periodically, thus adopting the metaphorical form of a wave (or waves), peaking roughly every 600 years. Supposedly, the last heyday of German literature happened at around 1800. So according to the theory, contemporary literature (i.e., around 1890) would find itself in a period of decay, something that will not have found Flaischlen’s approval. He was an adherent to the naturalist movement of his time and a poet and novelist himself. His river model of German literature can thus be seen as a polemical remark about Scherer’s theory. Although the river narrows after 1200 (maintaining accordance with Scherer’s view), it becomes ever broader even after 1800, foreign influences stream into it, and the watery body of Germanophone literature absorbs, accumulates, amalgamates everything.&lt;/p&gt;

&lt;h1 id=&quot;context-ii-visualising-time&quot;&gt;Context II: Visualising Time&lt;/h1&gt;

&lt;p&gt;Flaischlen was not the first to come up with the idea of representing a timeline as a river. This visualisation metaphor goes back to Austrian historiographer &lt;a href=&quot;http://www.fnz.geschichte.uni-muenchen.de/forschung/autoren-tabellenwerke/strass/index.html&quot;&gt;Friedrich Strass&lt;/a&gt; (1766–1845) who in 1804 published his highly influential “Strom der Zeiten” (“Stream of Time”). &lt;a href=&quot;http://daten.digitale-sammlungen.de/~db/0009/bsb00094935/images/index.html?id=00094935&amp;amp;fip=193.174.98.30&amp;amp;seite=2&quot;&gt;A hi-res scan of this chart&lt;/a&gt; was made by the Munich Digitisation Center (MDZ).&lt;/p&gt;

&lt;p&gt;Rosenberg and Grafton discuss “Der Strom der Zeiten” in their fabulous compendium “Cartographies of Time” (2010, pp. 143–149). And they quote William Bell, the English translator of Strass’s impressive chart, who (in 1810) emphasises the qualities of the river metaphor:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;“However natural it may be to assist the perceptive faculty, in its assumption of abstract time, by the idea of a &lt;em&gt;line&lt;/em&gt;, (…) it is astonishing that (…) the image of a &lt;em&gt;Stream&lt;/em&gt; should not have presented itself to any one (…). The expressions of &lt;em&gt;gliding&lt;/em&gt;, and &lt;em&gt;rolling on&lt;/em&gt;; or of the &lt;em&gt;rapid current&lt;/em&gt;, applied to time, are equally familiar to us with those of &lt;em&gt;long&lt;/em&gt; and &lt;em&gt;short&lt;/em&gt;. Neither does it require any great discernment to trace (…) in the &lt;em&gt;rise&lt;/em&gt; and &lt;em&gt;fall&lt;/em&gt; of empire, an allusion to the source of a river, and to the increasing rapidity of its currents, in proportion with the declivity of their channels towards the engulfing ocean. Nay, this metaphor (…) gives greater liveliness to the ideas, and impresses events more forcibly upon the mind, than the stiff regularity of the straight line. Its diversified power, likewise, of separating the various currents into subordinate branches, or of uniting them into one vast ocean of power (…) tends to render the idea by its beauty more attractive, by its simplicity more perspicuous, and by its resemblance more consistent.” (pp. 143, 147; &lt;a href=&quot;https://books.google.com/books?id=qhVXAAAAcAAJ&amp;amp;pg=PA8&quot;&gt;original quote here&lt;/a&gt;)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;As a positivist approach, Flaischlen’s flowchart is non-controversial. The only thing that could appear controversial is the impact (breadth) of each of the influxes. He himself stresses in the last paragraph of his preface “that the breadth of the main river is not mathematically calculated”. His chart is not based on “data”; it is not yet part of those 19th-century movements that started to use numerical statistics to transform “knowledge” into “data knowledge” (“Datenwissen”, cf. Andreas Bernard: &lt;a href=&quot;https://www.merkur-zeitschrift.de/2016/02/01/das-totale-archiv/&quot;&gt;Das totale Archiv&lt;/a&gt;, Merkur, 2016). But his “Graphische Litteratur-Tafel” is an inspiring predecessor of our attempts today to use &lt;a href=&quot;https://www.versobooks.com/books/261-graphs-maps-trees&quot;&gt;“graphs, maps, trees”&lt;/a&gt; to visualise and explore literary data.&lt;/p&gt;

&lt;h1 id=&quot;bibliography&quot;&gt;Bibliography&lt;/h1&gt;

&lt;p&gt;Bernard, Andreas: &lt;strong&gt;Das totale Archiv. Zur Funktion des Nicht-Wissens in der digitalen Kultur.&lt;/strong&gt; Merkur, vol. 801 (February, 2016), pp. 5–17. (&lt;a href=&quot;https://www.merkur-zeitschrift.de/2016/02/01/das-totale-archiv/&quot;&gt;merkur-zeitschrift.de&lt;/a&gt;)&lt;/p&gt;

&lt;p&gt;Flaischlen, Cäsar: &lt;strong&gt;Graphische Litteratur-Tafel. Die deutsche Litteratur und der Einfluß fremder Litteraturen auf ihren Verlauf von Beginn einer schriftlichen Überlieferung an bis heute in graphischer Darstellung.&lt;/strong&gt; Stuttgart, Göschen, 1890. (2. Tausend: Stuttgart, Göschen, 1890; 3. Tausend: Berlin, Behr, 1890)&lt;/p&gt;

&lt;p&gt;Minor, Jakob: &lt;strong&gt;[Review of Flaischlen’s Dissertation on Otto von Gemmingen.]&lt;/strong&gt; Anzeiger für deutsches Alterthum und deutsche Litteratur, XVII, 2 (April 1891), pp. 147–149. (&lt;a href=&quot;http://www.digizeitschriften.de/dms/img/?PID=PPN345204123_0035%7Clog84&amp;amp;physid=phys600#navi&quot;&gt;digizeitschriften.de&lt;/a&gt;)&lt;/p&gt;

&lt;p&gt;Nebrig, Alexander: &lt;strong&gt;Vergleichen als Wissenschaft: Zur Fachgeschichte.&lt;/strong&gt; Evi Zemanek, Alexander Nebrig (eds.): Komparatistik. Berlin, Akademie Verlag, 2012, pp. 35–36. (&lt;a href=&quot;https://books.google.com/books?id=XT7nBQAAQBAJ&amp;amp;pg=PA35&quot;&gt;books.google.com&lt;/a&gt;)&lt;/p&gt;

&lt;p&gt;Rosenberg, Daniel; Grafton, Anthony: &lt;strong&gt;Cartographies of Time.&lt;/strong&gt; Princeton Architectural Press, New York, 2010, pp. 143–149.&lt;/p&gt;

&lt;p&gt;Schmidt, Erich: &lt;strong&gt;Wege und Ziele der deutschen Literaturgeschichte. Eine Antrittsvorlesung.&lt;/strong&gt; Charakteristiken. Vol. I. Berlin, Weidmann, 1886, pp. 480–498. (&lt;a href=&quot;https://archive.org/stream/charakteristiken01schmuoft#page/480/mode/2up&quot;&gt;archive.org&lt;/a&gt;)&lt;/p&gt;

&lt;p&gt;Weschenfelder, Anke: &lt;strong&gt;Flaischlen, Cäsar (Otto Hugo).&lt;/strong&gt; Deutsches Literatur-Lexikon. Das 20. Jahrhundert. Vol. 9. Zurich/Munich, K.G. Saur, 2006, cols. 33–37. (&lt;a href=&quot;https://books.google.com/books?id=mLI6ZB6vrFEC&amp;amp;pg=PA1996&quot;&gt;books.google.com&lt;/a&gt;)&lt;/p&gt;

&lt;p&gt;Wetz, Wilhelm: &lt;strong&gt;Shakespeare vom Standpunkte der vergleichenden Litteraturgeschichte.&lt;/strong&gt; Vol. I: &lt;strong&gt;Die Menschen in Shakespeare Dramen.&lt;/strong&gt; Hamburg, Haendke &amp;amp; Lehmkuhl, 1897. (&lt;a href=&quot;https://archive.org/details/shakespearevoms01wetzgoog&quot;&gt;archive.org&lt;/a&gt;)&lt;/p&gt;

&lt;h1 id=&quot;thanks-a-lot-&quot;&gt;Thanks a lot …&lt;/h1&gt;
&lt;p&gt;… to Kurt Ubelhoer for copyediting this post!!!1!&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Ingo Börner, 
	
          
          Angelika Hechtl, 
	
          
          Frank Fischer, 
	
          
          Peer Trilcke
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Mistranslated Dates in Literature</title>
    <link href="https://weltliteratur.net/Mistranslated-Dates-in-Literature/"/>
    <updated>2016-09-29T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/Mistranslated-Dates-in-Literature</id>
    <content type="html">&lt;p&gt;(This posting is about a phenomenon that I can only partly explain. It seems to be more than just a couple of random incidents. Maybe somebody else has some hints?) (&lt;strong&gt;Update&lt;/strong&gt; Oct 3, 2016: Jürgen Hermes closely examined the “Count of Monte Cristo” example and explains some of the “mistranslations”, scroll to the bottom of the page for more info.)&lt;/p&gt;

&lt;p&gt;When translating numbers and dates, strange things happen at times. Some of them are still explicable, one of which is the following. Take “Othello”, Act I, beginning of Scene 3 (&lt;a href=&quot;https://en.wikisource.org/wiki/The_Tragedy_of_Othello,_The_Moor_of_Venice#ACT_1._SCENE_III._A_council-chamber.&quot;&gt;Wikisource&lt;/a&gt;), just look at the numbers:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;DUKE OF VENICE&lt;br /&gt;
There is no composition in these news&lt;br /&gt;
That gives them credit.&lt;br /&gt;
&lt;br /&gt;
FIRST SENATOR&lt;br /&gt;
Indeed, they are disproportion’d;&lt;br /&gt;
My letters say a &lt;strong&gt;hundred and seven galleys&lt;/strong&gt;.&lt;br /&gt;
&lt;br /&gt;
DUKE OF VENICE&lt;br /&gt;
And mine, a hundred and forty.&lt;br /&gt;
&lt;br /&gt;
SECOND SENATOR&lt;br /&gt;
And mine, two hundred:&lt;br /&gt;
(…)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Here’s how Wolf Graf Baudissin translated the passage into German (&lt;a href=&quot;http://gutenberg.spiegel.de/buch/2185/1&quot;&gt;Gutenberg-DE&lt;/a&gt;):&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;HERZOG&lt;br /&gt;
In diesen Briefen fehlt Zusammenhang,&lt;br /&gt;
Der sie glaubwürdig machte.&lt;br /&gt;
&lt;br /&gt;
ERSTER SENATOR&lt;br /&gt;
Jawohl, sie weichen voneinander ab;&lt;br /&gt;
Mein Schreiben nennt mir &lt;strong&gt;hundertsechs Galeeren&lt;/strong&gt;.&lt;br /&gt;
&lt;br /&gt;
HERZOG&lt;br /&gt;
Und meines hundertvierzig.&lt;br /&gt;
&lt;br /&gt;
ZWEITER SENATOR&lt;br /&gt;
Meins zweihundert.&lt;br /&gt;
(…)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Ok, none of the characters seems to know the exact number of ships anyway, so why not translate &lt;strong&gt;106&lt;/strong&gt; instead of &lt;strong&gt;107&lt;/strong&gt;? Voss also translated 106 (&lt;a href=&quot;https://books.google.com/books?id=7AQ5AQAAMAAJ&amp;amp;pg=PA24&quot;&gt;Google Books&lt;/a&gt;), while Wieland sticked to the 107 (&lt;a href=&quot;https://de.wikisource.org/wiki/Seite:Wieland_Shakespear_Theatralische_Werke_VII.djvu/198&quot;&gt;Wikisource&lt;/a&gt;).&lt;/p&gt;

&lt;p&gt;As insinuated before, that one’s easy to solve. Wieland chose to translate “Othello” in prose, the other two wanted to stick to Shakespeare’s blank verse. So for metrical reasons, they had to get rid of the second syllable of “sieben”, and how do you do that? Make it a “sechs”. Things like that happen a lot in metrical translations. (And I have to thank &lt;a href=&quot;http://asv.informatik.uni-leipzig.de/staff/Thomas_Efer&quot;&gt;Thomas Efer&lt;/a&gt; for this and many other examples.)&lt;/p&gt;

&lt;p&gt;So far, so good. Let’s now move from wrong numbers to wrong dates.&lt;/p&gt;

&lt;h2 id=&quot;strange-case-of-dr-jekyll-and-mr-hyde&quot;&gt;Strange Case of Dr Jekyll and Mr Hyde&lt;/h2&gt;

&lt;p&gt;At the beginning of chapter nine of R. L. Stevenson’s infamous novella, Dr Lanyon receives a letter from Jekyll. Lanyon says he received it on the 9th of January and that he had dined with Jekyll the day before, so obviously, the letter must have been written on the 8th or 9th of January. Yet it says in all editions of Stevenson’s text: “10th December, 18—” (&lt;a href=&quot;https://en.wikisource.org/wiki/Strange_Case_of_Dr_Jekyll_and_Mr_Hyde/Dr._Lanyon%27s_Narrative&quot;&gt;Wikisource&lt;/a&gt;). This inconsistency has received some coverage (&lt;a href=&quot;https://books.google.com/books?id=sk7zCQAAQBAJ&amp;amp;pg=PA64&quot;&gt;like here&lt;/a&gt;) and, for one, has been corrected in the Spanish translation which is also &lt;a href=&quot;https://es.wikisource.org/wiki/El_extraño_caso_del_Dr._Jekyll_y_Mr._Hyde:_Capítulo_IX&quot;&gt;on Wikisource&lt;/a&gt;, the letter there dates from “9 de enero de 18…”. The German translation &lt;a href=&quot;http://gutenberg.spiegel.de/buch/der-seltsame-fall-des-doktor-jekyll-und-des-herrn-hyde-8629/9&quot;&gt;on Gutenberg-DE&lt;/a&gt; chose to omit the dating to mitigate the temporal inconsistency.&lt;/p&gt;

&lt;p&gt;So these changes are still utterly explainable. But there are other cases.&lt;/p&gt;

&lt;h2 id=&quot;the-count-of-monte-cristo&quot;&gt;The Count of Monte Cristo&lt;/h2&gt;

&lt;p&gt;I’m aware that the doctrines of literary translations changed over the centuries, and this is really a research field of its own. But the oddities I want to talk about go beyond that. Let’s have a look at the opening sentence of “Le Comte de Monte-Christo” by Alexandre Dumas:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;Le 24 février 1815, la vigie de Notre-Dame de la Garde signala le trois-mâts le Pharaon, venant de Smyrne, Trieste et Naples. (&lt;a href=&quot;https://fr.wikisource.org/wiki/Le_Comte_de_Monte-Cristo/Chapitre_1&quot;&gt;Wikisource&lt;/a&gt;)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;The Spanish translation goes along with this, no problem:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;El 24 de febrero de 1815, el vigía de Nuestra Señora de la Guarda dio la señal de que se hallaba a la vista el bergantín El Faraón procedente de Esmirna, Trieste y Nápoles. (&lt;a href=&quot;https://es.wikisource.org/wiki/El_conde_de_Montecristo:_1-01&quot;&gt;Wikisource&lt;/a&gt;)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Yet the German translation by Max Pannwitz, hosted &lt;a href=&quot;http://gutenberg.spiegel.de/buch/6319/2&quot;&gt;on Gutenberg-DE&lt;/a&gt;, starts with this sentence:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;Am 25. Februar 1815 fuhr der Dreimaster Pharao langsam und wie zögernd in den Hafen von Marseille.&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;This is in line with an earlier translation by August Zoller (see &lt;a href=&quot;https://books.google.com/books?id=AIS_skQYStEC&amp;amp;pg=PA7&quot;&gt;Google Books&lt;/a&gt;).&lt;/p&gt;

&lt;p&gt;And the Polish translation &lt;a href=&quot;https://pl.wikisource.org/wiki/Strona:PL_Aleksander_Dumas_-_Hrabia_Monte_Christo_01.djvu/007&quot;&gt;on Wikisource&lt;/a&gt; has this:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;W słonecznym i bardzo ciepłym dniu 27 lutego 1815 r. dano z wieży kościoła Notre Dame de la Garde, w Marsylji, hasło, zapowiadające powrót z podróży do Smyrny, Tryjestu i Neapolu — trójżaglowca “Faraon”.&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Now what? 24th, 25th, 27th of February? I don’t see any reason to change the date of the arrival of the ship.&lt;/p&gt;

&lt;p&gt;This is also not the fault of buggy OCR like in the case of &lt;a href=&quot;http://drhagen.com/blog/the-missing-11th-of-the-month/&quot;&gt;the (non-existant) underreprentation of the 11th day of the month in the Google Ngram corpus&lt;/a&gt;.&lt;/p&gt;

&lt;h2 id=&quot;anna-karenina&quot;&gt;Anna Karenina&lt;/h2&gt;

&lt;p&gt;Let’s check on Tolstoy, “Anna Karenina”, Part 3, Chapter 23, which in its original shape looks like this (&lt;a href=&quot;https://ru.wikisource.org/wiki/Анна_Каренина_(Толстой)/Часть_III/Глава_XXIII&quot;&gt;Wikisource&lt;/a&gt;):&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;В понедельник было обычное заседание комиссии 2-го июня.&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Yet the German translation hosted &lt;a href=&quot;http://gutenberg.spiegel.de/buch/4043/93&quot;&gt;on Gutenberg-DE&lt;/a&gt; has this:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;Montags war die gewöhnliche Sitzung der Kommission vom zweiten Juli.&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;2nd of June becomes 2nd of July. Which could be an honest and simple mistake, because of the similarity of the month names.&lt;/p&gt;

&lt;p&gt;Let’s stay with the same book and jump to Part 4, Chapter 6, first sentence:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;Алексей Александрович одержал блестящую победу в заседании комиссии семнадцатого августа, (…). (&lt;a href=&quot;https://ru.wikisource.org/wiki/Анна_Каренина_(Толстой)/Часть_IV/Глава_VI&quot;&gt;Wikisource&lt;/a&gt;)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;The Spanish translation &lt;a href=&quot;https://es.wikisource.org/wiki/Ana_Karenina_IV:_Capítulo_VI&quot;&gt;on Wikisource&lt;/a&gt; has this:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;Karenin obtuvo una brillante victoria en la sesión celebrada por la Comisión el 1 de agosto, (…)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;17th of August becomes 1st of August. Explanation could be simple again: Maybe the translator just forgot a number here?&lt;/p&gt;

&lt;p&gt;This is also not a Julian/Gregorian calendar issue. The Julian calendar was 12 days behind in the 19th century, which doesn’t explain the difference between the two dates.&lt;/p&gt;

&lt;h2 id=&quot;a-descent-into-the-maelström&quot;&gt;A Descent into the Maelström&lt;/h2&gt;

&lt;p&gt;Next example, Poe’s “Descent into the Maelström”, where it says:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;It was on the tenth day of July, 18—, a day which the people of this part of the world will never forget—for it was one in which blew the most terrible hurricane that ever came out of the heavens. (&lt;a href=&quot;https://en.wikisource.org/wiki/Tales_(Poe)/A_Descent_into_the_Maelstr%C3%B6m&quot;&gt;Wikisource&lt;/a&gt;)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;A Spanish translation goes like this:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;Dentro de pocos días se cumplirán tres años desde que sucedió lo que voy a relataros. Era el 10 de agosto de 18—, día que la gente de este lado del mundo jamás olvidará, porque se desató el huracán más formidable que jamás envió el cielo. (&lt;a href=&quot;https://es.wikisource.org/wiki/Un_descenso_por_el_Maelstr%C3%B6m&quot;&gt;Wikisource&lt;/a&gt;)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;July? August!&lt;/p&gt;

&lt;h2 id=&quot;sherlock-holmes&quot;&gt;Sherlock Holmes&lt;/h2&gt;

&lt;p&gt;Let’s do some Sherlock Holmes. The Hound of the Baskervilles!&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;Dr. Mortimer drew a folded newspaper out of his pocket.&lt;br /&gt;
“Now, Mr. Holmes, we will give you something a little more recent. This is the Devon County Chronicle of May 14th of this year. (…)” (&lt;a href=&quot;https://www.gutenberg.org/ebooks/3070&quot;&gt;gutenberg.org&lt;/a&gt;)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;The same passage in Spanish translation:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;El doctor Mortimer se sacó del bolsillo un periódico doblado.&lt;br /&gt;
–Ahora, señor Holmes, voy a leerle una noticia un poco más reciente, publicada en el Devon County Chronicle del 14 de junio de este año. (…) (&lt;a href=&quot;http://www.biblioteca.org.ar/libros/130087.pdf&quot;&gt;biblioteca.org.ar&lt;/a&gt;)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;May? June! – Another shift of one month when translating from English to Spanish, maybe a pattern? Well, maybe not. Let’s have some Stendhal:&lt;/p&gt;

&lt;h2 id=&quot;the-charterhouse-of-parma&quot;&gt;The Charterhouse of Parma&lt;/h2&gt;

&lt;blockquote&gt;
  &lt;p&gt;Le 7 mars 1815, les dames étaient de retour, depuis l’avant-veille, d’un charmant petit voyage de Milan ; elles se promenaient dans la belle allée de platanes, récemment prolongée sur l’extrême bord du lac. (&lt;a href=&quot;https://fr.wikisource.org/wiki/La_Chartreuse_de_Parme_(%C3%A9dition_Martineau,_1927)/Chapitre_II&quot;&gt;Wikisource&lt;/a&gt;)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Which, in a Spanish translation, looks like this:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;El 7 de mayo de 1815 hacía dos días que las señoras habían vuelto de un precioso viajecito a Milán; estaban paseándose por la hermosa avenida de plátanos que había sido prolongada hacía poco tiempo hasta el borde mismo del lago. (&lt;a href=&quot;http://www.edu.mec.gub.uy/biblioteca_digital/libros/s/Stendhal%20-%20La%20cartuja%20de%20Parma.pdf&quot;&gt;edu.mec.gub.uy&lt;/a&gt;)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;March? May!&lt;/p&gt;

&lt;h2 id=&quot;jules-verne&quot;&gt;Jules Verne&lt;/h2&gt;

&lt;blockquote&gt;
  &lt;p&gt;Le 29 mai de cette année-là, un berger surveillait son troupeau à la lisière d’un plateau verdoyant, au pied du Retyezat, qui domine une vallée fertile, boisée d’arbres à tiges droites, enrichie de belles cultures. (&lt;a href=&quot;https://fr.wikisource.org/wiki/Le_Ch%C3%A2teau_des_Carpathes/1&quot;&gt;Wikisource&lt;/a&gt;)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;The same passage in Spanish:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;El 19 de mayo de aquel año, un pastor apacentaba su rebaño a la orilla de un verde prado, al pie del Retyezat, que domina un valle fértil, cubierto de árboles de rama-je recto y enriquecido con bellas plantaciones. (&lt;a href=&quot;http://www.biblioteca.org.ar/libros/656189.pdf&quot;&gt;biblioteca.org.ar&lt;/a&gt;)&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;29th? 19th!&lt;/p&gt;

&lt;h2 id=&quot;explanations&quot;&gt;Explanation(s)?&lt;/h2&gt;

&lt;p&gt;For all we know, this last example could be another honest mistake (mistyping a number, happens). Maybe all these examples can be traced back to simple mistakes that go unnoticed by the common reader (unless you’re part of a research project whose purpose it is to extract dates from literary corpora :-)… Anyway, we got more findings like this. You’ll find some examples in &lt;a href=&quot;/Introducing-TIWOLI/&quot;&gt;TIWOLI, our app for Android and iOS&lt;/a&gt;. We didn’t dare to correct wrong dates, we wanted to preserve them as they were. After all, that’s how every reader of these elder translations encountered them.&lt;/p&gt;

&lt;p&gt;Many of the translations are from the 19th or early 20th century when transations didn’t yet meet our idea of a proper, authentic translation. For example, the first German translation of Huxley’s “Brave New World” changed the location &lt;a href=&quot;https://de.wikipedia.org/wiki/Sch%C3%B6ne_neue_Welt#Deutsche_.C3.9Cbersetzung&quot;&gt;from “London” to “Berlin”&lt;/a&gt; (which was authorised by the author, by the way; but this wouldn’t pull through nowadays, I guess).&lt;/p&gt;

&lt;p&gt;Some sort of weird adaptation could be the reason for some of the above-cited examples. But I really don’t know. The question is: Are all these examples plain mistakes? Or are there reasons for the way those dates were mistranslated?&lt;/p&gt;

&lt;p&gt;(For hints, drop me a line &lt;a href=&quot;https://twitter.com/umblaetterer&quot;&gt;on Twitter&lt;/a&gt;, via mail, or use the comment field below.)&lt;/p&gt;

&lt;h2 id=&quot;update&quot;&gt;Update&lt;/h2&gt;

&lt;p&gt;On October 3rd, Jürgen Hermes published a great follow-up to this post (in German): &lt;strong&gt;&lt;a href=&quot;https://texperimentales.hypotheses.org/1813&quot;&gt;“Zum Tag der Daten-Einheit”&lt;/a&gt;&lt;/strong&gt;. In which he interprets the “mistranslations” relating to “The Count of Monte Cristo” as “belated copyediting” by German and Polish translators. Especially the Polish translator applied some effort to repair the flawed timeline of the original novel. The same effort might be behind some of the other mistranslations, too. The example from “Jekyll and Hyde” was also pointing into that direction.&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Introducing TIWOLI ("Today in World Literature"), a Calendar App of Fictional Events</title>
    <link href="https://weltliteratur.net/Introducing-TIWOLI/"/>
    <updated>2016-09-07T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/Introducing-TIWOLI</id>
    <content type="html">&lt;p&gt;This could become a whole new series: &lt;strong&gt;Fancy by-products of Digital Humanities research projects…&lt;/strong&gt;&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Short version:&lt;/strong&gt; Download TIWOLI for Android &lt;strong&gt;&lt;a href=&quot;https://play.google.com/store/apps/details?id=de.wannauchimmer.apps.tiwoli&quot;&gt;here&lt;/a&gt;&lt;/strong&gt;. Download TIWOLI for iOS &lt;strong&gt;&lt;a href=&quot;https://appsto.re/de/I69sfb.i&quot;&gt;here&lt;/a&gt;&lt;/strong&gt;.&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Update (Oct 29, 2016):&lt;/strong&gt; The iOS version is finally available in the App Store. The article was updated accordingly.&lt;/p&gt;

&lt;p&gt;And now let’s jump right into it and name seven famous days in world literature:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;The day of James Joyce’s “Ulysses”? (Of course, &lt;a href=&quot;https://en.wikipedia.org/wiki/Bloomsday&quot;&gt;Bloomsday&lt;/a&gt;, &lt;a href=&quot;https://ga.wikipedia.org/wiki/L%C3%A1_Bloom&quot;&gt;Lá Bloom&lt;/a&gt;, June 16!)&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;That is an easy start, but the world beyond this über-day of world literature is trickier. Most people will know about the fictional events we’re going to itemise, but won’t typically have the exact dates in mind. Let’s check:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;On which day did Robinson Crusoe set foot on his island? (September 30) On which day did he eventually return to England, after 35 years of absence? (June 11)&lt;/li&gt;
  &lt;li&gt;When is the deadline for Phileas Fogg’s travel around the world in eighty days? (December 21)&lt;/li&gt;
  &lt;li&gt;What’s Tristram Shandy’s birthday? (November 5) – Funny enough, there still seem to be &lt;a href=&quot;http://edwardiansisterhood.blogspot.de/2009/11/hill-times.html&quot;&gt;people drinking to Tristram Shandy’s health&lt;/a&gt;.&lt;/li&gt;
  &lt;li&gt;When did Goethe’s “Young Werther” commit suicide? (December 21)&lt;/li&gt;
  &lt;li&gt;On what day did Charlie visit the Chocolate Factory? (February 1) – Especially interesting to note here that &lt;a href=&quot;https://en.wikipedia.org/wiki/Willy_Wonka_%26_the_Chocolate_Factory&quot;&gt;in the 1971 movie&lt;/a&gt;, the golden ticket refers to “the first day of October” for whatever reason.&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;So this is the subject of our blogpost, &lt;strong&gt;(in)famous days in world-literary fiction&lt;/strong&gt;.&lt;/p&gt;

&lt;p&gt;How did this come about? At the annual ADHO Digital Humanities conference 2015 (&lt;a href=&quot;http://dh2015.org/&quot;&gt;DH2015&lt;/a&gt;) in Sydney, we delivered a talk called “When does (German) literature take place?” (&lt;a href=&quot;http://dbs.ifi.uni-heidelberg.de/fileadmin/Team/jannik/publications/dh2015-sydney-fischer-stroetgen-german-literature-slides.pdf&quot;&gt;slides&lt;/a&gt;) In order to answer that unusual question, we extracted exact dates and months from a corpus of 2,700+ works of German-language fiction from 1510 to the 1940s. Our results suggested that &lt;a href=&quot;https://de.wikisource.org/wiki/Im_wundersch%C3%B6nen_Monat_Mai&quot;&gt;“the marvellous month of May”&lt;/a&gt; (Heinrich Heine) is preeminent in our corpus of German fiction. There was already enough evidence to claim that May is the favourite month of poets (at least in the northern hemisphere), and Heine is just one example. But May’s preeminence in prose was a real finding and there are a bunch of explanations, but this (along with many other insights yielded by this whole date-extraction-from-fiction business) shall happen in the corresponding research paper, not here.&lt;/p&gt;

&lt;p&gt;As a by-product (and no more than that), we coded a native Android app to showcase the more exciting and self-explanatory of our results. We called our app &lt;strong&gt;TIWOLI&lt;/strong&gt; (short for “Today in World Literature”) and made it &lt;strong&gt;&lt;a href=&quot;https://play.google.com/store/apps/details?id=de.wannauchimmer.apps.tiwoli&quot;&gt;available via the Google Play Store&lt;/a&gt;&lt;/strong&gt;. This was followed by &lt;strong&gt;&lt;a href=&quot;https://appsto.re/de/I69sfb.i&quot;&gt;the iOS version&lt;/a&gt;&lt;/strong&gt;, done by our dear friend &lt;a href=&quot;http://thboegel.de/project/&quot;&gt;Thomas Bögel&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;Now what TIWOLI does is give you the opportunity to zigzag through an entire world-literary year, a calendar filled with fictional happenings. Here some screenshots with some of the examples from above:&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-06-16-joyce-ulysses.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/4300&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-02-01-dahl-charlie.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://books.google.com/books?id=v7WeJp2CU-QC&amp;amp;pg=PT55&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-11-05-sterne-tristram.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/1079&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-12-20-verne-eighty.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/103&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;We’re currently curating three corpora for the app, in &lt;strong&gt;English&lt;/strong&gt;, &lt;strong&gt;Spanish&lt;/strong&gt; and &lt;strong&gt;German&lt;/strong&gt;. In some cases, you can even switch between translations of one and the same passage (not in all cases, though, the contents of our corpora are very different per language). This is a 1st-of-January example from Balzac’s novel “Eugénie Grandet”:&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-01-01-balzac-eugenie.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/1715&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-es-01-01-balzac-eugenia.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://es.wikisource.org/wiki/Eugenia_Grandet&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-01-01-balzac-eugenie.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://gutenberg.spiegel.de/buch/4850/3&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;The interface was also translated into the three languages (it will fall back to English if your system language cannot be matched). Here are the three versions of our welcome screen, plus the English version of the settings page:&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-welcome.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-bienvenido.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-willkommen.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-settings.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;If you like you can set the alarm so it informs you about fictional events that “happened” on the current day.&lt;/p&gt;

&lt;p&gt;It was not all that easy to fill up a whole year with suitable quotes for every day. To replenish the English corpus with nice enough quotes we are very grateful to Mark Algee-Hewitt and the Stanford Literary Lab who helped us out with their corpus of English-language fiction. Let’s give you some more:&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-01-15-bronte-jane-eyre.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/1260&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-02-01-twain-wilson.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/102&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-02-22-fitzgerald-beautiful.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/9830&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-03-04-doyle-scarlet.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/244&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;For each corpus, the app contains quotes from translated world-literary works, too, because after all, weltliteratur is the raison d’être of this blog:&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-03-16-novalis-ofterdingen.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/31873&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-03-25-hugo-hunchback.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/2610&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-04-07-gogol-nose.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/36238&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-05-20-turgenev-fathers.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/30723&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;The dates from the Russian 19th-century novels point to the Julian calendar, of course (which was 12 days behind the Gregorian back then). Some more translated literature:&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-06-12-hamsun-mysteries.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://books.google.com/books?id=xB6nAgAAQBAJ&amp;amp;pg=PT4&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-07-05-voltaire-micromegas.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://en.wikisource.org/wiki/Micromegas/Chapter_3&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-07-27-zweig-chess.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://books.google.com/books?id=Uv4UAgAAQBAJ&amp;amp;pg=PT31&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-09-23-hoffmann-pot.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://gutenberg.net.au/ebooks06/0605801h.html&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;For every quote we provide a link to the full-text source (gutenberg.org, gutenberg.net.au, archive.org, books.google.com, …). Epitexts and links/metadata are generated automatically. The portraits/pictures were not chosen by us, we’re relying on the ‘principal image’ provided by Wikidata, a great way to easily enrich an app. Let’s give you the Defoe examples from above, accompanied by Poe (for the sake of rhyming names):&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-06-22-poe-roget.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/2147&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-07-10-poe-maelstrom.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://en.wikisource.org/wiki/Tales_(Poe)/A_Descent_into_the_Maelstr%C3%B6m&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-09-30-defoe-robinson.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/521&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-12-19-defoe-robinson.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/521&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;Some more German literature in translation, including the aforementioned destiny of storm-and-stressy Werther:&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-09-27-fontane-effi.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://archive.org/stream/EffiBriest/fontane_theodor_1819_1898_effi_briest#page/n91/mode/2up&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-10-03-fontane-effi.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://archive.org/stream/EffiBriest/fontane_theodor_1819_1898_effi_briest#page/n31/mode/2up&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-10-30-hoffmann-sandman.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/32046&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-12-21-goethe-werther.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/2527&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;The first two quotes are from Fontane’s “Effi Briest”, in many respects the righteous German counterpart to “Madame Bovary”. The second quote alludes to Effi’s and Instetten’s wedding which coincides with the national day of Germany (not a good omen, given the tragic outcome of the novel 😇).&lt;/p&gt;

&lt;p&gt;Four final examples from the English corpus:&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-07-31-shelley-frankenstein.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/84&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-11-02-eliot-bede.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/507&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-11-09-wilde-dorian.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/174&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-12-10-stevenson-hyde.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/43&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;The last quote, from “Jekyll &amp;amp; Hyde”, is interesting insofar as the given date can be regarded a mistake by the author that went into several print editions (as &lt;a href=&quot;https://books.google.com/books?id=sk7zCQAAQBAJ&amp;amp;pg=PA64&quot;&gt;discussed here&lt;/a&gt;, for example).&lt;/p&gt;

&lt;p&gt;We also harvested some newer books as you can see in the Auster example below (the picture is not the best, but somebody might upload a better one to Wikimedia Commons someday). There’s also one more Brontë quote and two nice quotes from the time-travel novel by Bellamy (where times and dates are the actual subject):&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-03-20-bronte-heights.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/768&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-05-19-auster-glass.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://books.google.com/books?id=VRfe-QG_ls4C&amp;amp;pg=PA10&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-05-30-bellamy-2000.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/624&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-en-09-10-bellamy-2000.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://www.gutenberg.org/ebooks/624&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;Now let’s sneak a peek at the Spanish corpus. You can easily switch to the other corpora at any point from within the app.&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-es-04-22-valera-pepita.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://es.wikisource.org/wiki/Pepita_Jim%C3%A9nez:_08&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-es-06-15-caballero-gaviota.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://es.wikisource.org/wiki/La_gaviota_(Caballero):_17&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-es-07-12-baroja-silvestre.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://oreneta.com/kalebeul/2006/06/14/silvester-paradox-meets-mr-macbeth/&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-es-07-20-cervantes-quijote.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://cvc.cervantes.es/literatura/clasicos/quijote/edicion/parte2/cap36/default.htm&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;We can also follow the chronology of Borges’ cuento “La muerte y la brújula” (&lt;a href=&quot;https://en.wikipedia.org/wiki/Death_and_the_Compass&quot;&gt;“Death and the Compass”&lt;/a&gt;):&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-es-12-03-borges-brujula.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://www.literatura.us/borges/lamuerte.html&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-es-01-03-borges-brujula.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://www.literatura.us/borges/lamuerte.html&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-es-02-03-borges-brujula.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://www.literatura.us/borges/lamuerte.html&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-es-03-01-borges-brujula.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://www.literatura.us/borges/lamuerte.html&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;Some concluding Spanish examples:&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-es-02-28-dumas-montecristo.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://es.wikisource.org/wiki/El_conde_de_Montecristo:_1-15&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-es-12-21-goethe-werther.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://ciudadseva.com/texto/werther/&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-es-10-11-marquez-soledad.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://mgarci.aas.duke.edu/cibertextos/EDICIONES-BILINGUES/GARCIA-MARQUEZ-G/CIEN-ANOS-SOLEDAD/CAP13.HTM&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-es-12-31-aira-fantasmas.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;https://books.google.com/books?id=LJPxQvuMd-oC&amp;amp;pg=PA7&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;On to the German-language corpus:&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-01-20-buechner-lenz.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://www.zeno.org/Literatur/M/B%C3%BCchner,+Georg/Erz%C3%A4hlung/Lenz&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-04-04-schnitzler-gustl.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://gutenberg.spiegel.de/buch/5342/1&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-05-12-lasswitz-planeten.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://gutenberg.spiegel.de/buch/8473/17&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-07-27-zweig-schachnovelle.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://gutenberg.spiegel.de/buch/7318/2&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;The first quotation, from Georg Büchner’s novella fragment &lt;a href=&quot;https://en.wikipedia.org/wiki/Lenz_(fragment)&quot;&gt;“Lenz”&lt;/a&gt;, is probably one of the most famous first sentences. On the second screenshot, Schnitzler’s &lt;a href=&quot;http://www.encyclopedia.com/article-1G2-3408300617/none-but-brave-leutnant.html&quot;&gt;“Leutnant Gustl”&lt;/a&gt; is pondering his destiny on 4th of April, resulting in &lt;strong&gt;not&lt;/strong&gt; killing himself. On the third one, the government of the Martian States takes control over our planet, because we earthlings are not capable of maintaining peace (the author of the passage, &lt;a href=&quot;https://en.wikipedia.org/wiki/Kurd_Lasswitz&quot;&gt;Kurd Lasswitz&lt;/a&gt;, can be regarded a pioneer of German sci-fi lit). The fourth quote is especially interesting, since the “27th of July” is also discussed in a “meta” way, as a combination of a printed number and a word, which serves as intellectual nourishment for the imprisoned intellectual (if you didn’t install the app yet, check the English translation of this passage &lt;a href=&quot;https://books.google.com/books?id=Uv4UAgAAQBAJ&amp;amp;pg=PT31&quot;&gt;on Google Books&lt;/a&gt;).&lt;/p&gt;

&lt;p&gt;We already featured Theodor Fontane with some paragraphs translated to English. But what is it with the Prussian author and the 3rd of October? It is an eventful day in many of his novels, and we’re not sure if anyone has looked into that yet:&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-10-03-fontane-birnbaum.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://gutenberg.spiegel.de/buch/4437/20&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-10-03-fontane-effi.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://gutenberg.spiegel.de/buch/4446/5&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-10-03-fontane-stechlin.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://gutenberg.spiegel.de/buch/4434/2&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;Fitfully, we included some quotations from contemporary fiction. The given sources contain page numbers as they point to print editions of the works:&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-07-19-zeh-unterleuten.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://d-nb.info/1078450226&quot;&gt;source&lt;/a&gt;, p. 287)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-07-23-goetz-holtrop.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://d-nb.info/1017662762&quot;&gt;source&lt;/a&gt;, p. 275)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-08-15-werner-abgang.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://d-nb.info/840734891&quot;&gt;source&lt;/a&gt;, p. 116)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-12-12-enzensberger-robert.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://d-nb.info/954601262&quot;&gt;source&lt;/a&gt;, p. 31)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;Some more:&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-10-07-schmidt-atheisten.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://d-nb.info/942534050&quot;&gt;source&lt;/a&gt;, p. 11)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-10-12-stein-rabbi.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://d-nb.info/1045299928&quot;&gt;source&lt;/a&gt;, pp. 72)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-11-09-richter-gran-via.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://d-nb.info/1008856525&quot;&gt;source&lt;/a&gt;, p. 144)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;The first quote from Arno Schmidt’s 1972 novel &lt;a href=&quot;https://en.wikipedia.org/wiki/The_School_for_Atheists&quot;&gt;“The School for Atheists”&lt;/a&gt; points to a future date, the 7th of October, 2014. When we traversed this point in time two years ago, at least one major newspaper published an article about this long-imminent event (German weekly “Die Zeit”: &lt;a href=&quot;http://www.zeit.de/2014/43/arno-schmidt-tellingstedt-die-schule-der-atheisten&quot;&gt;“The story starts now.”&lt;/a&gt;).&lt;/p&gt;

&lt;p&gt;How about &lt;a href=&quot;https://en.wikipedia.org/wiki/Jean_Paul&quot;&gt;Jean Paul&lt;/a&gt;, the greatest of all May lovers:&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-05-01-jean-paul-hesperus.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://gutenberg.spiegel.de/buch/3193/6&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-05-03-jean-paul-quintus.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://gutenberg.spiegel.de/buch/3209/31&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-05-16-jean-paul-quintus.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://gutenberg.spiegel.de/buch/3209/34&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;
&lt;div style=&quot;clear:left;&quot;&gt;&lt;/div&gt;

&lt;p&gt;And last not least, some Christmas quotes:&lt;/p&gt;

&lt;div class=&quot;tiwoligallery&quot;&gt;
 &lt;ul&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-12-24-droste-huelshoff-judenbuche.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://gutenberg.spiegel.de/buch/2837/7&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-12-24-hesse-unterm-rad.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://d-nb.info/1025308913&quot;&gt;source&lt;/a&gt;, p. 204)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-12-24-hoffmann-nussknacker-und-mausekoenig.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://gutenberg.spiegel.de/buch/3083/1&quot;&gt;source&lt;/a&gt;)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
  &lt;li&gt;&lt;figure&gt;&lt;img src=&quot;https://weltliteratur.net/images/tiwoli/tiwoli-de-12-24-mann-buddenbrooks.png&quot; alt=&quot;Screenshot from TIWOLI app.&quot; /&gt;&lt;figcaption&gt;(&lt;a href=&quot;http://d-nb.info/970666128&quot;&gt;source&lt;/a&gt;, p. 582)&lt;/figcaption&gt;&lt;/figure&gt;&lt;/li&gt;
 &lt;/ul&gt;
&lt;/div&gt;

&lt;p&gt;So much for our app, TIWOLI. Also if it wasn’t part of the original research, it helped to sharpen our notion of dates and month names in fictional texts. Literature as a genre has an inclination to waive exact dates and prefers temporal settings like “at the beginning of November”, &lt;a href=&quot;https://en.wikipedia.org/wiki/It_was_a_dark_and_stormy_night&quot;&gt;“It was a dark and stormy night”&lt;/a&gt;, etc. But sometimes exact dates do occur, of course. And if they do, they have a tendency to mean something. (Unless, of course, they occur in epistolary, adventure or historical novels which are usually full of exact dates. We tried to exclude them from the app and just used some as fill-ins. Goethe’s “Werther”, all Jules Verne novels and Benito Pérez Galdós proved especially useful. 😊)&lt;/p&gt;

&lt;p&gt;Before we conclude this admittedly BuzzFeed-like über-post with far too many pics we have two cliffhangers for you:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;Our friends at the University of Cologne (&lt;a href=&quot;http://www.spinfo.phil-fak.uni-koeln.de/&quot;&gt;Spinfo&lt;/a&gt;/&lt;a href=&quot;http://www.cceh.uni-koeln.de/&quot;&gt;CCeH&lt;/a&gt;) are currently building a web interface with our data. &lt;a href=&quot;https://twitter.com/spinfocl/status/794147295084314624&quot;&gt;Prototype looks very promising.&lt;/a&gt;&lt;/li&gt;
  &lt;li&gt;There’s an &lt;strong&gt;easter egg&lt;/strong&gt; somewhere in the app. Whoever finds it first is our hero/ine!&lt;/li&gt;
&lt;/ul&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer, 
	
          
          Jannik Strötgen
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Contributors to World Literature – Identifying Writers in Wikipedia, Part II</title>
    <link href="https://weltliteratur.net/Contributors-to-World-Literature-Identifying-Writers-in-Wikipedia-Part-II/"/>
    <updated>2016-08-19T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/Contributors-to-World-Literature-Identifying-Writers-in-Wikipedia-Part-II</id>
    <content type="html">&lt;p&gt;In
&lt;a href=&quot;/Contributors-to-World-Literature-Identifying-Writers-in-Wikipedia-Part-I/&quot;&gt;the first post&lt;/a&gt;
of this 2-part series we introduced six features of Wikipedia pages on
writers that might help us figure out if those pages are, in fact,
about (literary) writers. Today we want to discuss which of these
features are actually suited to reliably identify writers in
Wikipedia. Our goal is to build a set of contributors to world
literature that we can further investigate (and keep updated, as
Wikipedia evolves). Instead of providing a conclusive answer here,
this post will concentrate on observations that could eventually help
to implement an approach for the automated identification of literary
writers in Wikipedia.&lt;/p&gt;

&lt;h2 id=&quot;first-sentence&quot;&gt;First Sentence&lt;/h2&gt;

&lt;p&gt;&lt;img src=&quot;/images/wp_john_irving_first_sentence.png&quot; alt=&quot;The first sentence on the Wikipedia Page of John Irving&quot; /&gt;&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;This feature is available throughout all language versions.&lt;/li&gt;
  &lt;li&gt;The automated analysis of sentences requires some sophisticated
natural language processing which should also be available for
different languages. The quality of such approaches highly
depends on the chosen language.&lt;/li&gt;
  &lt;li&gt;The results are quite ambiguous, since even within one language
version there is no standardised way to describe writers.&lt;/li&gt;
  &lt;li&gt;DBpedia has implemented a
&lt;a href=&quot;https://github.com/dbpedia/fact-extractor&quot;&gt;framework for fact extraction&lt;/a&gt;
which could potentially extract this information.
&lt;!-- DBpedia provides datasets that classify persons based on data
extracted from the running text of the article. **point to this
dataset and briefly explain it** --&gt;&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;table-of-contents&quot;&gt;Table of Contents&lt;/h2&gt;

&lt;p&gt;&lt;img src=&quot;/images/wp_john_irving_toc.png&quot; alt=&quot;The table of contents on the Wikipedia page of John Irving&quot; /&gt;&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;At first sight, a “Bibliography” section seems to be the prevalent
way to list original works of an author in the English Wikipedia.
But not all writers have one. Also, the name of the section is
not consistent (e.g., in the article on
&lt;a href=&quot;https://en.wikipedia.org/wiki/John_Milton#Works&quot;&gt;John Milton&lt;/a&gt; it
is called “Works”). “Bibliography” can also refer to a list of
secondary literature (see the article on
&lt;a href=&quot;https://en.wikipedia.org/wiki/Gabriel_García_Márquez#Bibliography&quot;&gt;Gabriel García Márquez&lt;/a&gt;).&lt;/li&gt;
  &lt;li&gt;Albeit its ambiguity, finding the “Bibliography” headline in the
wiki source code is easy. Finding equivalent headings in other
language versions would require some effort, though.&lt;/li&gt;
  &lt;li&gt;It is unclear if and how this feature is used in other language
versions.&lt;/li&gt;
  &lt;li&gt;The section structure of articles is currently not provided by
DBpedia.&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;occupation-property&quot;&gt;Occupation Property&lt;/h2&gt;

&lt;p&gt;&lt;img src=&quot;/images/wp_john_irving_infobox.png&quot; alt=&quot;The infobox on the Wikipedia page of John Irving&quot; /&gt;&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;This property often contains several (ambiguous) values without
further distinctions (“writer”, “novelist”, “poet”, etc.).&lt;/li&gt;
  &lt;li&gt;Values of this property are highly language-dependent, it is rather
difficult to make use of it in a cross-language environment.&lt;/li&gt;
  &lt;li&gt;Not all language editions permit infoboxes for writers (among
others, the German edition).&lt;/li&gt;
  &lt;li&gt;If available, the infobox properties are extracted by DBpedia. They
can be found in the &lt;a href=&quot;http://wiki.dbpedia.org/services-resources/documentation/datasets#infoboxproperties&quot;&gt;infobox properties dataset&lt;/a&gt;
and in the
&lt;a href=&quot;http://wiki.dbpedia.org/services-resources/documentation/datasets#mappingbasedliterals&quot;&gt;mapping-based properties dataset&lt;/a&gt;. The
&lt;a href=&quot;http://wiki.dbpedia.org/Downloads2015-10&quot;&gt;latest dataset&lt;/a&gt; contains
this information in the infobox properties dataset:&lt;/li&gt;
&lt;/ul&gt;

&lt;div class=&quot;language-xml highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;nt&quot;&gt;&amp;lt;http:&lt;/span&gt;&lt;span class=&quot;err&quot;&gt;//dbpedia.org/resource/John_Irving&lt;/span&gt;&lt;span class=&quot;nt&quot;&gt;&amp;gt;&lt;/span&gt; &lt;span class=&quot;nt&quot;&gt;&amp;lt;http:&lt;/span&gt;&lt;span class=&quot;err&quot;&gt;//dbpedia.org/property/occupation&lt;/span&gt;&lt;span class=&quot;nt&quot;&gt;&amp;gt;&lt;/span&gt; &quot;Novelist&quot;@en .
&lt;span class=&quot;nt&quot;&gt;&amp;lt;http:&lt;/span&gt;&lt;span class=&quot;err&quot;&gt;//dbpedia.org/resource/John_Irving&lt;/span&gt;&lt;span class=&quot;nt&quot;&gt;&amp;gt;&lt;/span&gt; &lt;span class=&quot;nt&quot;&gt;&amp;lt;http:&lt;/span&gt;&lt;span class=&quot;err&quot;&gt;//dbpedia.org/property/occupation&lt;/span&gt;&lt;span class=&quot;nt&quot;&gt;&amp;gt;&lt;/span&gt; &quot;Screenwriter&quot;@en .
&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;h2 id=&quot;list-of-works&quot;&gt;List of Works&lt;/h2&gt;

&lt;p&gt;&lt;img src=&quot;/images/wp_john_irving_template_john_irving.png&quot; alt=&quot;The &amp;quot;John Irving&amp;quot; template on the Wikipedia page of John Irving&quot; /&gt;&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;In our example page on
&lt;a href=&quot;https://en.wikipedia.org/wiki/John_Irving&quot;&gt;John Irving&lt;/a&gt;, this
feature was implemented using a specific
&lt;a href=&quot;https://en.wikipedia.org/wiki/Template:John_Irving&quot;&gt;John Irving template&lt;/a&gt;.
Other popular/canonised writers also have that kind of template, yet
in general, this feature is not very widespread.&lt;/li&gt;
  &lt;li&gt;There are categories for such templates, like
&lt;a href=&quot;https://en.wikipedia.org/wiki/Category:Novelist_navigational_boxes&quot;&gt;Category:Novelist navigational boxes&lt;/a&gt;,
making it easy to check which writers have such a template.&lt;/li&gt;
  &lt;li&gt;Apparently, there are only two other language editions with such a
template,
&lt;a href=&quot;https://ro.wikipedia.org/wiki/Categorie:Formate_romancieri&quot;&gt;Romanian&lt;/a&gt;
and &lt;a href=&quot;https://fa.wikipedia.org/wiki/%D8%B1%D8%AF%D9%87:%D8%AC%D8%B9%D8%A8%D9%87%E2%80%8C%D9%87%D8%A7%DB%8C_%D9%86%D8%A7%D9%88%D8%A8%D8%B1%DB%8C_%D8%B1%D9%85%D8%A7%D9%86%E2%80%8C%D9%86%D9%88%DB%8C%D8%B3&quot;&gt;Farsi&lt;/a&gt;.&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;categories&quot;&gt;Categories&lt;/h2&gt;

&lt;p&gt;&lt;img src=&quot;/images/wp_john_irving_categories.png&quot; alt=&quot;The categories of the Wikipedia page of John Irving&quot; /&gt;&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;The category graph is quite inconsistent. In particular, it is not a
tree. (cf. O. Medelyan, D. Milne, C. Legg, I. H. Witten (2009),
&lt;a href=&quot;http://dx.doi.org/10.1016/j.ijhcs.2009.05.004&quot;&gt;Mining meaning from Wikipedia&lt;/a&gt;)&lt;/li&gt;
  &lt;li&gt;As we have seen, there are many categories that could be used to
identify a person as a writer, but they do not cover all
writers. Categories higher up the hierarchy (e.g.,
&lt;a href=&quot;https://en.wikipedia.org/wiki/Category:Writers_by_century&quot;&gt;Writers by century&lt;/a&gt;)
would cover more writers but due to some dubious subcategories
(e.g.,
&lt;a href=&quot;https://en.wikipedia.org/wiki/Category:Baconian_theory_of_Shakespeare_authorship&quot;&gt;Baconian theory of Shakespeare authorship&lt;/a&gt;)
also cover pages that are definitely not about writers (e.g.,
&lt;a href=&quot;https://en.wikipedia.org/wiki/Honorificabilitudinitatibus&quot;&gt;Honorificabilitudinitatibus&lt;/a&gt;).&lt;/li&gt;
  &lt;li&gt;For instance, starting the traversal at the
&lt;a href=&quot;https://en.wikipedia.org/wiki/Category:Writers_by_century&quot;&gt;Writers by Century&lt;/a&gt;
category in the English version, the set would contain persons like
&lt;a href=&quot;&quot;&gt;Winston Churchill&lt;/a&gt; and &lt;a href=&quot;&quot;&gt;Leonardo da Vinci&lt;/a&gt;, while omitting
writers like &lt;a href=&quot;&quot;&gt;Gertrude Stein&lt;/a&gt; and &lt;a href=&quot;&quot;&gt;Heinrich Heine&lt;/a&gt;.&lt;/li&gt;
  &lt;li&gt;Each language version has its own category structure. Thus, the
effort of identifying corresponding categories and removing false
positives would need to be repeated for each language.&lt;/li&gt;
  &lt;li&gt;Articles are usually assigned to several categories which would enable
identification of writers which also had other occupations.&lt;/li&gt;
  &lt;li&gt;DBpedia provides the &lt;a href=&quot;http://wiki.dbpedia.org/services-resources/documentation/datasets#articlecategories&quot;&gt;article_categories&lt;/a&gt; dataset
which contains the assignment of categories to articles and the
&lt;a href=&quot;http://wiki.dbpedia.org/services-resources/documentation/datasets#skoscategories&quot;&gt;skos_categories&lt;/a&gt; dataset which contains the category
graph.&lt;/li&gt;
&lt;/ul&gt;

&lt;h2 id=&quot;writer-template&quot;&gt;Writer Template&lt;/h2&gt;

&lt;p&gt;&lt;img src=&quot;/images/wp_john_irving_template_writer.png&quot; alt=&quot;The source code of the &amp;quot;writer&amp;quot; template of the Wikipedia page of John Irving&quot; /&gt;&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;Not all language editions have or use such template. E.g., the
German Wikipedia does not permit the writer template and, thus, its page
on
&lt;a href=&quot;https://de.wikipedia.org/wiki/Johann_Wolfgang_von_Goethe&quot;&gt;Johann Wolfgang von Goethe&lt;/a&gt;
lacks an infobox.&lt;/li&gt;
  &lt;li&gt;Each page can be assigned to exactly one template, which means that
some persons do not have a writer template, although we’d clearly
like to have them in our set. This includes
&lt;a href=&quot;https://en.wikipedia.org/wiki/Franz_Kafka&quot;&gt;Franz Kafka&lt;/a&gt; (neutral
&lt;a href=&quot;https://en.wikipedia.org/wiki/Template:Infobox_person&quot;&gt;person template&lt;/a&gt;)
and &lt;a href=&quot;https://en.wikipedia.org/wiki/Umberto_Eco&quot;&gt;Umberto Eco&lt;/a&gt; (tied
to a
&lt;a href=&quot;https://en.wikipedia.org/wiki/Template:Infobox_philosopher&quot;&gt;philosopher template&lt;/a&gt;).&lt;/li&gt;
  &lt;li&gt;The association of Wikipedia pages to templates has been extracted
by DBpedia in the &lt;a href=&quot;http://wiki.dbpedia.org/services-resources/documentation/datasets#instancetypes&quot;&gt;instance types dataset&lt;/a&gt;.
The value is provided in the &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;type&lt;/code&gt; property.  For John Irving, it
contains the following information:&lt;/li&gt;
&lt;/ul&gt;

&lt;div class=&quot;language-xml highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;nt&quot;&gt;&amp;lt;http:&lt;/span&gt;&lt;span class=&quot;err&quot;&gt;//dbpedia.org/resource/John_Irving&lt;/span&gt;&lt;span class=&quot;nt&quot;&gt;&amp;gt;&lt;/span&gt; &lt;span class=&quot;nt&quot;&gt;&amp;lt;http:&lt;/span&gt;&lt;span class=&quot;err&quot;&gt;//www.w3.org/1999/02/22-rdf-syntax-ns#type&lt;/span&gt;&lt;span class=&quot;nt&quot;&gt;&amp;gt;&lt;/span&gt; &lt;span class=&quot;nt&quot;&gt;&amp;lt;http:&lt;/span&gt;&lt;span class=&quot;err&quot;&gt;//dbpedia.org/ontology/Writer&lt;/span&gt;&lt;span class=&quot;nt&quot;&gt;&amp;gt;&lt;/span&gt; .
&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;h1 id=&quot;conclusion&quot;&gt;Conclusion&lt;/h1&gt;

&lt;p&gt;Summarising the above findings in a table, we get the following result:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;feature&lt;/th&gt;
      &lt;th&gt;language editions&lt;/th&gt;
      &lt;th&gt;usage within language editions&lt;/th&gt;
      &lt;th&gt;DBpedia dataset&lt;/th&gt;
      &lt;th&gt;effort&lt;/th&gt;
      &lt;th&gt;ambiguity&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;first sentence&lt;/td&gt;
      &lt;td&gt;all&lt;/td&gt;
      &lt;td&gt;all&lt;/td&gt;
      &lt;td&gt;-&lt;/td&gt;
      &lt;td&gt;high&lt;/td&gt;
      &lt;td&gt;high&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;table of contents&lt;/td&gt;
      &lt;td&gt;all&lt;/td&gt;
      &lt;td&gt;some&lt;/td&gt;
      &lt;td&gt;-&lt;/td&gt;
      &lt;td&gt;low&lt;/td&gt;
      &lt;td&gt;medium&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;occupation property&lt;/td&gt;
      &lt;td&gt;some&lt;/td&gt;
      &lt;td&gt;many&lt;/td&gt;
      &lt;td&gt;&lt;a href=&quot;http://downloads.dbpedia.org/2015-10/core-i18n/en/infobox_properties_en.ttl.bz2&quot;&gt;infobox properties&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;medium&lt;/td&gt;
      &lt;td&gt;high&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;list of works&lt;/td&gt;
      &lt;td&gt;few&lt;/td&gt;
      &lt;td&gt;few&lt;/td&gt;
      &lt;td&gt;-&lt;/td&gt;
      &lt;td&gt;low&lt;/td&gt;
      &lt;td&gt;low&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;categories&lt;/td&gt;
      &lt;td&gt;all&lt;/td&gt;
      &lt;td&gt;many&lt;/td&gt;
      &lt;td&gt;&lt;a href=&quot;http://downloads.dbpedia.org/2015-10/core-i18n/en/article_categories_en.ttl.bz2&quot;&gt;article categories&lt;/a&gt; and &lt;a href=&quot;http://downloads.dbpedia.org/2015-10/core-i18n/en/skos_categories_en.ttl.bz2&quot;&gt;skos categories&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;high&lt;/td&gt;
      &lt;td&gt;high&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;writer template&lt;/td&gt;
      &lt;td&gt;some&lt;/td&gt;
      &lt;td&gt;many&lt;/td&gt;
      &lt;td&gt;&lt;a href=&quot;http://downloads.dbpedia.org/2015-10/core-i18n/en/instance_types_en.ttl.bz2&quot;&gt;instance types&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;low&lt;/td&gt;
      &lt;td&gt;low&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;Given the large variety of properties the different features have, it
is quite difficult to devise an approach to identify writers on
Wikipedia that works across different language versions. This 2-part
blog post was a preliminary introduction to get a sense of the problem.
We proposed and applied a solution that basically works across
different language versions in a still unpublished paper (on which we
will keep you posted) and will introduce a pragmatic and manageable
set that can be used for a variety of purposes in one of our next
blog posts.&lt;/p&gt;

&lt;p&gt;&lt;img src=&quot;http://vg07.met.vgwort.de/na/9c7df397f1404f1e9d7b43220540abf8&quot; width=&quot;1&quot; height=&quot;1&quot; alt=&quot;&quot; /&gt;&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Robert Jäschke, 
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Contributors to World Literature – Identifying Writers in Wikipedia, Part I</title>
    <link href="https://weltliteratur.net/Contributors-to-World-Literature-Identifying-Writers-in-Wikipedia-Part-I/"/>
    <updated>2016-08-18T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/Contributors-to-World-Literature-Identifying-Writers-in-Wikipedia-Part-I</id>
    <content type="html">&lt;p&gt;The first challenge we face when trying to analyse the representation
of world literature on Wikipedia is the automatic identification of
literary writers and works across all 280+ language versions. Our last
posts have shown how to extract &lt;strong&gt;works&lt;/strong&gt; that are possible candidates
for a Wikipedia-inherent world-literary canon
(&lt;a href=&quot;/Wikidata-Meets-World-Literature/&quot;&gt;Wikidata-based&lt;/a&gt; and
&lt;a href=&quot;/DBpedia-and-World-Literature/&quot;&gt;DBpedia-based&lt;/a&gt;). In this 2-part post
we focus on the extraction of &lt;strong&gt;writers&lt;/strong&gt;, i.e., possible contributors
to world literature.&lt;/p&gt;

&lt;p&gt;For a start, let us have a detailed look at the
&lt;a href=&quot;https://en.wikipedia.org/wiki/John_Irving&quot;&gt;Wikipedia page of John Irving&lt;/a&gt;:&lt;/p&gt;

&lt;p&gt;&lt;img src=&quot;/images/wp_john_irving_all.png&quot; alt=&quot;Wikipedia Page of John Irving&quot; /&gt;&lt;/p&gt;

&lt;p&gt;This web page has several features indicating that John Irving is a
writer we want to have in our set. Let’s introduce and discuss six of
them:&lt;/p&gt;

&lt;ol&gt;
  &lt;li&gt;
    &lt;p&gt;The &lt;strong&gt;first sentence&lt;/strong&gt; says that John Irving is a novelist and
screenwriter:&lt;/p&gt;

    &lt;p&gt;&lt;img src=&quot;/images/wp_john_irving_first_sentence.png&quot; alt=&quot;The first sentence on the Wikipedia Page of John Irving&quot; /&gt;&lt;/p&gt;
  &lt;/li&gt;
  &lt;li&gt;
    &lt;p&gt;The &lt;strong&gt;table of contents&lt;/strong&gt; lists a bibliography:&lt;/p&gt;

    &lt;p&gt;&lt;img src=&quot;/images/wp_john_irving_toc.png&quot; alt=&quot;The table of contents on the Wikipedia page of John Irving&quot; /&gt;&lt;/p&gt;
  &lt;/li&gt;
  &lt;li&gt;
    &lt;p&gt;The infobox on the right has an &lt;strong&gt;occupation property&lt;/strong&gt; with
values “Novelist” and “Screenwriter”:&lt;/p&gt;

    &lt;p&gt;&lt;img src=&quot;/images/wp_john_irving_infobox.png&quot; alt=&quot;The infobox on the Wikipedia page of John Irving&quot; /&gt;&lt;/p&gt;
  &lt;/li&gt;
  &lt;li&gt;
    &lt;p&gt;Scrolling down, the
&lt;a href=&quot;https://en.wikipedia.org/wiki/Template:John_Irving&quot;&gt;John Irving template&lt;/a&gt;
assembles a &lt;strong&gt;list of works&lt;/strong&gt; by the author:&lt;/p&gt;

    &lt;p&gt;&lt;img src=&quot;/images/wp_john_irving_template_john_irving.png&quot; alt=&quot;The &amp;quot;John Irving&amp;quot; template on the Wikipedia page of John Irving&quot; /&gt;&lt;/p&gt;
  &lt;/li&gt;
  &lt;li&gt;
    &lt;p&gt;Scrolling down further, we find a list of &lt;strong&gt;categories&lt;/strong&gt;, among
them
&lt;a href=&quot;https://en.wikipedia.org/wiki/Category:20th-century_American_novelists&quot;&gt;20th-century American novelists&lt;/a&gt;,
&lt;a href=&quot;https://en.wikipedia.org/wiki/Category:21st-century_American_novelists&quot;&gt;21st-century American novelists&lt;/a&gt;,
&lt;a href=&quot;https://en.wikipedia.org/wiki/Category:American_feminist_writers&quot;&gt;American feminist writers&lt;/a&gt;,
and
&lt;a href=&quot;https://en.wikipedia.org/wiki/Category:American_male_screenwriters&quot;&gt;American male screenwriters&lt;/a&gt;:&lt;/p&gt;

    &lt;p&gt;&lt;img src=&quot;/images/wp_john_irving_categories.png&quot; alt=&quot;The categories of the Wikipedia page of John Irving&quot; /&gt;&lt;/p&gt;
  &lt;/li&gt;
  &lt;li&gt;
    &lt;p&gt;In addition, by looking at
&lt;a href=&quot;https://en.wikipedia.org/w/index.php?title=John_Irving&amp;amp;action=edit&amp;amp;editintro=Template:BLP_editintro&quot;&gt;the wiki source code of the page&lt;/a&gt; we realise that the
infobox uses the
&lt;strong&gt;&lt;a href=&quot;https://en.wikipedia.org/wiki/Template:Infobox_writer&quot;&gt;writer template&lt;/a&gt;&lt;/strong&gt;:&lt;/p&gt;

    &lt;p&gt;&lt;img src=&quot;/images/wp_john_irving_template_writer.png&quot; alt=&quot;The source code of the &amp;quot;writer&amp;quot; template of the Wikipedia page of John Irving&quot; /&gt;&lt;/p&gt;
  &lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;To leverage one of these features to automatically identify John
Irving (and others, of course) as a writer, we have to answer these
questions:&lt;/p&gt;

&lt;ol&gt;
  &lt;li&gt;Which of the information is universally available (i.e., in
different language versions)?&lt;/li&gt;
  &lt;li&gt;Which of the information is available through DBpedia or Wikidata?&lt;/li&gt;
  &lt;li&gt;How widespread is the information used?&lt;/li&gt;
  &lt;li&gt;How easy is it to actually use the information to identify writers?&lt;/li&gt;
&lt;/ol&gt;

&lt;p&gt;We will discuss these questions tomorrow in the 2nd part of this blog post.&lt;/p&gt;

&lt;p&gt;&lt;img src=&quot;http://vg07.met.vgwort.de/na/dc38a9fd7d4c4b3d85d1fb78941ebb1b&quot; width=&quot;1&quot; height=&quot;1&quot; alt=&quot;&quot; /&gt;&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Robert Jäschke, 
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>DBpedia and World Literature</title>
    <link href="https://weltliteratur.net/DBpedia-and-World-Literature/"/>
    <updated>2016-05-30T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/DBpedia-and-World-Literature</id>
    <content type="html">&lt;p&gt;After our blog post
&lt;a href=&quot;/Wikidata-Meets-World-Literature/&quot;&gt;“Wikidata Meets World Literature”&lt;/a&gt;
some might have mumbled into their tea cups, “ok yeah, but what about
DBpedia?” Accordingly, let us add some technological diversity to our
initial experiment. &lt;a href=&quot;http://www.dbpedia.org/&quot;&gt;DBpedia&lt;/a&gt; is a great
project and follows its very own approach, it is also a few years
older than Wikidata. More on the differences between the two projects
can be found
&lt;a href=&quot;https://www.quora.com/What-is-the-difference-between-Wikidata-and-DBpedia&quot;&gt;on Quora&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;So, DBpedia and World Literature. Let’s kick off by translating/simplifying the original SPARQL query to match the DBpedia ontology:&lt;/p&gt;

&lt;div class=&quot;language-sql highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;k&quot;&gt;SELECT&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;desc&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;authorlabel&lt;/span&gt;
&lt;span class=&quot;k&quot;&gt;WHERE&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;{&lt;/span&gt;
  &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;rdf&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;type&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;dbo&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;Book&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;
  &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;dbo&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;author&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;author&lt;/span&gt;
  &lt;span class=&quot;n&quot;&gt;OPTIONAL&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;{&lt;/span&gt;
    &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;rdfs&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;label&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;desc&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;FILTER&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;lang&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;desc&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;=&lt;/span&gt; &lt;span class=&quot;nv&quot;&gt;&quot;en&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;).&lt;/span&gt;
  &lt;span class=&quot;p&quot;&gt;}&lt;/span&gt;
  &lt;span class=&quot;n&quot;&gt;OPTIONAL&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;{&lt;/span&gt;
    &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;author&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;rdfs&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;label&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;authorlabel&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;FILTER&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;lang&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;authorlabel&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;=&lt;/span&gt; &lt;span class=&quot;nv&quot;&gt;&quot;en&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;).&lt;/span&gt;
  &lt;span class=&quot;p&quot;&gt;}&lt;/span&gt;
&lt;span class=&quot;p&quot;&gt;}&lt;/span&gt;
&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;p&gt;This query &lt;strong&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?query=SELECT+%3Fs+%3Fdesc+%3Fauthorlabel%0D%0AWHERE+{++%3Fs+rdf%3Atype+dbo%3ABook+.%0D%0A++%3Fs+dbo%3Aauthor+%3Fauthor%0D%0A++OPTIONAL+{++++%3Fs+rdfs%3Alabel+%3Fdesc+FILTER+%28lang%28%3Fdesc%29+%3D+%22en%22%29.%0D%0A++}%0D%0A++OPTIONAL+{++++%3Fauthor+rdfs%3Alabel+%3Fauthorlabel+FILTER+%28lang%28%3Fauthorlabel%29+%3D+%22en%22%29.%0D%0A++}%0D%0A}%0D%0A&quot;&gt;returns&lt;/a&gt;&lt;/strong&gt; a list of resources that have
been assigned the class
&lt;a href=&quot;http://mappings.dbpedia.org/server/ontology/classes/Book&quot;&gt;Book&lt;/a&gt; in
the
&lt;a href=&quot;http://mappings.dbpedia.org/server/ontology/classes/&quot;&gt;DBpedia ontology&lt;/a&gt;,
along with their authors. (Yes, just &lt;strong&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?query=SELECT+%3Fs+%3Fdesc+%3Fauthorlabel%0D%0AWHERE+{++%3Fs+rdf%3Atype+dbo%3ABook+.%0D%0A++%3Fs+dbo%3Aauthor+%3Fauthor%0D%0A++OPTIONAL+{++++%3Fs+rdfs%3Alabel+%3Fdesc+FILTER+%28lang%28%3Fdesc%29+%3D+%22en%22%29.%0D%0A++}%0D%0A++OPTIONAL+{++++%3Fauthor+rdfs%3Alabel+%3Fauthorlabel+FILTER+%28lang%28%3Fauthorlabel%29+%3D+%22en%22%29.%0D%0A++}%0D%0A}%0D%0A&quot;&gt;click&lt;/a&gt;&lt;/strong&gt;, this
will auto-execute the query and show you the results.)&lt;/p&gt;

&lt;p&gt;In our original post we used the number of Wikipedia language versions
per book to rank them, with “One Thousand and One Nights” taking away
all the glory (for the time being, that is). We can try something
similar in DBpedia by counting the number of labels per book. Each
Wikipedia language edition from which a resource was extracted by
DBpedia has its label stored using the
&lt;a href=&quot;https://www.w3.org/TR/2004/REC-rdf-schema-20040210/#ch_label&quot;&gt;rdfs:label&lt;/a&gt;
property. This is our query:&lt;/p&gt;

&lt;div class=&quot;language-sql highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;k&quot;&gt;SELECT&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;desc&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;authorlabel&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;COUNT&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;DISTINCT&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;label&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;as&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;labelcount&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt;
&lt;span class=&quot;k&quot;&gt;WHERE&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;{&lt;/span&gt;
  &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;rdf&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;type&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;dbo&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;Book&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;
  &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;rdfs&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;label&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;label&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;
  &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;dbo&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;author&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;author&lt;/span&gt;
  &lt;span class=&quot;n&quot;&gt;OPTIONAL&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;{&lt;/span&gt;
    &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;rdfs&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;label&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;desc&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;FILTER&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;lang&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;desc&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;=&lt;/span&gt; &lt;span class=&quot;nv&quot;&gt;&quot;en&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;).&lt;/span&gt;
  &lt;span class=&quot;p&quot;&gt;}&lt;/span&gt;
  &lt;span class=&quot;n&quot;&gt;OPTIONAL&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;{&lt;/span&gt;
    &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;author&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;rdfs&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;label&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;authorlabel&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;FILTER&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;lang&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;authorlabel&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;=&lt;/span&gt; &lt;span class=&quot;nv&quot;&gt;&quot;en&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;).&lt;/span&gt;
  &lt;span class=&quot;p&quot;&gt;}&lt;/span&gt;
&lt;span class=&quot;p&quot;&gt;}&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;GROUP&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;BY&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;desc&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;authorlabel&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;ORDER&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;BY&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;DESC&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;labelcount&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt;
&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;p&gt;Unfortunately, &lt;a href=&quot;http://dbpedia.org/snorql/?query=SELECT+%3Fs+%3Fdesc+%3Fauthorlabel+%28COUNT%28DISTINCT+%3Flabel%29+as+%3Flabelcount%29%0D%0AWHERE+{++%3Fs+rdf%3Atype+dbo%3ABook+.%0D%0A++%3Fs+rdfs%3Alabel+%3Flabel+.%0D%0A++%3Fs+dbo%3Aauthor+%3Fauthor%0D%0A++OPTIONAL+{++++%3Fs+rdfs%3Alabel+%3Fdesc+FILTER+%28lang%28%3Fdesc%29+%3D+%22en%22%29.%0D%0A++}%0D%0A++OPTIONAL+{++++%3Fauthor+rdfs%3Alabel+%3Fauthorlabel+FILTER+%28lang%28%3Fauthorlabel%29+%3D+%22en%22%29.%0D%0A++}%0D%0A}+GROUP+BY+%3Fs+%3Fdesc+%3Fauthorlabel+ORDER+BY+DESC%28%3Flabelcount%29+LIMIT+20&quot;&gt;the output&lt;/a&gt; is a bit disappointing:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;s&lt;/th&gt;
      &lt;th&gt;desc&lt;/th&gt;
      &lt;th&gt;authorlabel&lt;/th&gt;
      &lt;th&gt;labelcount&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/The_Adventures_of_Tom_Sawyer&quot;&gt;:The_Adventures_of_Tom_Sawyer&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“The Adventures of Tom Sawyer”@en&lt;/td&gt;
      &lt;td&gt;“Mark Twain”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Strange_Case_of_Dr_Jekyll_and_Mr_Hyde&quot;&gt;:Strange_Case_of_Dr_Jekyll_and_Mr_Hyde&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Strange Case of Dr Jekyll and Mr Hyde”@en&lt;/td&gt;
      &lt;td&gt;“Robert Louis Stevenson”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Fifty_Shades_of_Grey&quot;&gt;:Fifty_Shades_of_Grey&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Fifty Shades of Grey”@en&lt;/td&gt;
      &lt;td&gt;“E. L. James”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Gray&apos;s_Anatomy&quot;&gt;:Gray’s_Anatomy&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Gray’s Anatomy”@en&lt;/td&gt;
      &lt;td&gt;“Henry Gray”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Sense_and_Sensibility&quot;&gt;:Sense_and_Sensibility&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Sense and Sensibility”@en&lt;/td&gt;
      &lt;td&gt;“Jane Austen”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/A_Brief_History_of_Time&quot;&gt;:A_Brief_History_of_Time&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“A Brief History of Time”@en&lt;/td&gt;
      &lt;td&gt;“Stephen Hawking”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Mein_Kampf&quot;&gt;:Mein_Kampf&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Mein Kampf”@en&lt;/td&gt;
      &lt;td&gt;“Adolf Hitler”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Crime_and_Punishment&quot;&gt;:Crime_and_Punishment&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Crime and Punishment”@en&lt;/td&gt;
      &lt;td&gt;“Fyodor Dostoyevsky”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Catching_Fire&quot;&gt;:Catching_Fire&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Catching Fire”@en&lt;/td&gt;
      &lt;td&gt;“Suzanne Collins”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/David_Copperfield&quot;&gt;:David_Copperfield&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“David Copperfield”@en&lt;/td&gt;
      &lt;td&gt;“Charles Dickens”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Mansfield_Park&quot;&gt;:Mansfield_Park&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Mansfield Park”@en&lt;/td&gt;
      &lt;td&gt;“Jane Austen”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Les_Liaisons_dangereuses&quot;&gt;:Les_Liaisons_dangereuses&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Les Liaisons dangereuses”@en&lt;/td&gt;
      &lt;td&gt;“Pierre Choderlos de Laclos”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Murder_on_the_Orient_Express&quot;&gt;:Murder_on_the_Orient_Express&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Murder on the Orient Express”@en&lt;/td&gt;
      &lt;td&gt;“Agatha Christie”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/The_Sign_of_the_Four&quot;&gt;:The_Sign_of_the_Four&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“The Sign of the Four”@en&lt;/td&gt;
      &lt;td&gt;“Arthur Conan Doyle”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/The_Republic_(Plato)&quot;&gt;:The_Republic_(Plato)&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“The Republic (Plato)”@en&lt;/td&gt;
      &lt;td&gt;“Plato”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/The_Communist_Manifesto&quot;&gt;:The_Communist_Manifesto&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“The Communist Manifesto”@en&lt;/td&gt;
      &lt;td&gt;“Friedrich Engels”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/From_the_Earth_to_the_Moon&quot;&gt;:From_the_Earth_to_the_Moon&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“From the Earth to the Moon”@en&lt;/td&gt;
      &lt;td&gt;“Jules Verne”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Buddenbrooks&quot;&gt;:Buddenbrooks&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Buddenbrooks”@en&lt;/td&gt;
      &lt;td&gt;“Thomas Mann”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/The_Brothers_Karamazov&quot;&gt;:The_Brothers_Karamazov&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“The Brothers Karamazov”@en&lt;/td&gt;
      &lt;td&gt;“Fyodor Dostoyevsky”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/The_Trial&quot;&gt;:The_Trial&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“The Trial”@en&lt;/td&gt;
      &lt;td&gt;“Franz Kafka”@en&lt;/td&gt;
      &lt;td&gt;12&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;Seems that books have been extracted from only 12 language editions
(ar, de, en, es, fr, it, ja, nl, pl, pt, ru, zh). Or, only these 12
languages feature a
&lt;a href=&quot;https://en.wikipedia.org/wiki/Help:Template&quot;&gt;page template for books&lt;/a&gt;.
Or … well, let’s stop speculating and have a look: DBpedia features a
page with
&lt;a href=&quot;http://wiki.dbpedia.org/services-resources/datasets/cross-language-overlap-statistics&quot;&gt;statistics about the extracted data&lt;/a&gt;
and we can see in the “Cross-Language Instance Overlap” table that
there are, for instance, 268 books appearing in 16 language editions.
However, not all datasets from all language versions are available at
the
&lt;a href=&quot;http://wiki.dbpedia.org/OnlineAccess#1.1%20Public%20SPARQL%20Endpoint&quot;&gt;public SPARQL endpoint&lt;/a&gt;.
&lt;a href=&quot;http://downloads.dbpedia.org/2015-04/core/&quot;&gt;This list&lt;/a&gt; shows that
currently a “labels” dataset has been loaded for exactly the 12
language versions we mentioned above. For the same reason, other
properties like
&lt;a href=&quot;http://dbpedia.org/snorql/?property=http%3A//dbpedia.org/ontology/abstract&quot;&gt;dbo:abstract&lt;/a&gt;
or &lt;a href=&quot;http://www.w3.org/2002/07/owl#sameAs&quot;&gt;owl:sameAs&lt;/a&gt; that could be
linked to the number of language editions show the same behaviour. (For
an overview of potential properties take a look at the DBpedia entry
on
&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/The_Adventures_of_Tom_Sawyer&quot;&gt;“The Adventures of Tom Sawyer”&lt;/a&gt;.)&lt;/p&gt;

&lt;p&gt;When looking for alternatives we found the &lt;a href=&quot;http://people.aifb.kit.edu/ath/&quot;&gt;PageRank dataset&lt;/a&gt; by Andreas Thalhammer. Fortunately, it is deployed on the official DBpedia SPARQL endpoint and so, instead of counting the number of language editions, we can easily use the &lt;a href=&quot;https://en.wikipedia.org/wiki/PageRank&quot;&gt;PageRank&lt;/a&gt; of each page within the English Wikipedia as a measure of importance:&lt;/p&gt;

&lt;div class=&quot;language-sql highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;k&quot;&gt;PREFIX&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;vrank&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;&amp;lt;&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;http&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;//&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;purl&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;org&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;/&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;voc&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;/&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;vrank&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;#&amp;gt;&lt;/span&gt;

&lt;span class=&quot;k&quot;&gt;SELECT&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;SAMPLE&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;desc&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;AS&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;label&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;GROUP_CONCAT&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;authorlabel&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;,&lt;/span&gt; &lt;span class=&quot;s1&quot;&gt;&apos;, &apos;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;AS&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;author&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;MAX&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;v&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;AS&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;rank&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt;
&lt;span class=&quot;k&quot;&gt;FROM&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;&amp;lt;&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;http&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;//&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;dbpedia&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;org&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;&amp;gt;&lt;/span&gt;
&lt;span class=&quot;k&quot;&gt;FROM&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;&amp;lt;&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;http&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;//&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;people&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;aifb&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;kit&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;edu&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;/&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;ath&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;/#&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;DBpedia_PageRank&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;&amp;gt;&lt;/span&gt;
&lt;span class=&quot;k&quot;&gt;WHERE&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;{&lt;/span&gt;
  &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;rdf&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;type&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;dbo&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;Book&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;
  &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;vrank&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;hasRank&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;/&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;vrank&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;rankValue&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;v&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;
  &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;dbo&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;author&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;author&lt;/span&gt;
  &lt;span class=&quot;n&quot;&gt;OPTIONAL&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;{&lt;/span&gt;
    &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;rdfs&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;label&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;desc&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;FILTER&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;lang&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;desc&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;=&lt;/span&gt; &lt;span class=&quot;nv&quot;&gt;&quot;en&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;).&lt;/span&gt;
  &lt;span class=&quot;p&quot;&gt;}&lt;/span&gt;
  &lt;span class=&quot;n&quot;&gt;OPTIONAL&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;{&lt;/span&gt;
    &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;author&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;rdfs&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;label&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;authorlabel&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;FILTER&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;lang&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;authorlabel&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;=&lt;/span&gt; &lt;span class=&quot;nv&quot;&gt;&quot;en&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;).&lt;/span&gt;
  &lt;span class=&quot;p&quot;&gt;}&lt;/span&gt;
&lt;span class=&quot;p&quot;&gt;}&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;GROUP&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;BY&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;ORDER&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;BY&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;DESC&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;rank&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt;
&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;p&gt;We modified the query so it groups authors of multi-author books into one row. Despite some Wikipedia-related distortions, the result is much more meaningful:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;s&lt;/th&gt;
      &lt;th&gt;label&lt;/th&gt;
      &lt;th&gt;author&lt;/th&gt;
      &lt;th&gt;rank&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/The_World_Factbook&quot;&gt;:The_World_Factbook&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“The World Factbook”@en&lt;/td&gt;
      &lt;td&gt;“Central Intelligence Agency”&lt;/td&gt;
      &lt;td&gt;146.277&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Systema_Naturae&quot;&gt;:Systema_Naturae&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Systema Naturae”@en&lt;/td&gt;
      &lt;td&gt;“Carl Linnaeus”&lt;/td&gt;
      &lt;td&gt;68.6522&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Natural_History_(Pliny)&quot;&gt;:Natural_History_(Pliny)&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Natural History (Pliny)”@en&lt;/td&gt;
      &lt;td&gt;“Pliny the Elder”&lt;/td&gt;
      &lt;td&gt;64.9683&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/On_the_Origin_of_Species&quot;&gt;:On_the_Origin_of_Species&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“On the Origin of Species”@en&lt;/td&gt;
      &lt;td&gt;“Charles Darwin”&lt;/td&gt;
      &lt;td&gt;56.7624&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/The_Rolling_Stone_Album_Guide&quot;&gt;:The_Rolling_Stone_Album_Guide&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“The Rolling Stone Album Guide”@en&lt;/td&gt;
      &lt;td&gt;“Anthony DeCurtis, Dave Marsh”&lt;/td&gt;
      &lt;td&gt;47.9486&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Don_Quixote&quot;&gt;:Don_Quixote&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Don Quixote”@en&lt;/td&gt;
      &lt;td&gt;“Miguel de Cervantes”&lt;/td&gt;
      &lt;td&gt;45.3332&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/All_Music_Guide_to_Jazz&quot;&gt;:All_Music_Guide_to_Jazz&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“All Music Guide to Jazz”@en&lt;/td&gt;
      &lt;td&gt;“Vladimir Bogdanov (editor), Stephen Thomas Erlewine”&lt;/td&gt;
      &lt;td&gt;44.641&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Alice&apos;s_Adventures_in_Wonderland&quot;&gt;:Alice’s_Adventures_in_Wonderland&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Alice’s Adventures in Wonderland”@en&lt;/td&gt;
      &lt;td&gt;“Lewis Carroll”&lt;/td&gt;
      &lt;td&gt;42.4669&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/All_Music_Guide_to_the_Blues&quot;&gt;:All_Music_Guide_to_the_Blues&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“All Music Guide to the Blues”@en&lt;/td&gt;
      &lt;td&gt;“Vladimir Bogdanov, Stephen Thomas Erlewine”&lt;/td&gt;
      &lt;td&gt;40.9857&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Records_of_the_Grand_Historian&quot;&gt;:Records_of_the_Grand_Historian&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Records of the Grand Historian”@en&lt;/td&gt;
      &lt;td&gt;“Sima Qian”&lt;/td&gt;
      &lt;td&gt;40.262&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Nineteen_Eighty-Four&quot;&gt;:Nineteen_Eighty-Four&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Nineteen Eighty-Four”@en&lt;/td&gt;
      &lt;td&gt;“George Orwell”&lt;/td&gt;
      &lt;td&gt;39.9243&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/The_Republic_(Plato)&quot;&gt;:The_Republic_(Plato)&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“The Republic (Plato)”@en&lt;/td&gt;
      &lt;td&gt;“Plato”&lt;/td&gt;
      &lt;td&gt;38.9124&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/The_Wealth_of_Nations&quot;&gt;:The_Wealth_of_Nations&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“The Wealth of Nations”@en&lt;/td&gt;
      &lt;td&gt;“Adam Smith”&lt;/td&gt;
      &lt;td&gt;37.4529&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Euclid&apos;s_Elements&quot;&gt;:Euclid’s_Elements&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Euclid’s Elements”@en&lt;/td&gt;
      &lt;td&gt;“Euclid”&lt;/td&gt;
      &lt;td&gt;36.0581&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Paradise_Lost&quot;&gt;:Paradise_Lost&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Paradise Lost”@en&lt;/td&gt;
      &lt;td&gt;“John Milton”&lt;/td&gt;
      &lt;td&gt;32.8596&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Moby-Dick&quot;&gt;:Moby-Dick&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Moby-Dick”@en&lt;/td&gt;
      &lt;td&gt;“Herman Melville”&lt;/td&gt;
      &lt;td&gt;32.632&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Encyclopædia_Iranica&quot;&gt;:Encyclopædia_Iranica&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Encyclopædia Iranica”@en&lt;/td&gt;
      &lt;td&gt;“Ehsan Yarshater”&lt;/td&gt;
      &lt;td&gt;30.9694&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Dracula&quot;&gt;:Dracula&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Dracula”@en&lt;/td&gt;
      &lt;td&gt;“Bram Stoker”&lt;/td&gt;
      &lt;td&gt;29.6592&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/Histories_(Herodotus)&quot;&gt;:Histories_(Herodotus)&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“Histories (Herodotus)”@en&lt;/td&gt;
      &lt;td&gt;“Herodotus”&lt;/td&gt;
      &lt;td&gt;29.4831&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://dbpedia.org/snorql/?describe=http%3A//dbpedia.org/resource/The_Hobbit&quot;&gt;:The_Hobbit&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;“The Hobbit”@en&lt;/td&gt;
      &lt;td&gt;“J. R. R. Tolkien”&lt;/td&gt;
      &lt;td&gt;29.4576&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;These are the top 20 books, ranked by their PageRank within the English Wikipedia. One thing still blocking the view a bit are the popular encyclopedic works, of course. The many, many in-links earned by &lt;a href=&quot;https://en.wikipedia.org/w/index.php?title=Special:WhatLinksHere/The_World_Factbook&amp;amp;limit=500&quot;&gt;the “World Factbook”&lt;/a&gt; or &lt;a href=&quot;https://en.wikipedia.org/w/index.php?title=Special:WhatLinksHere/The_Rolling_Stone_Album_Guide&amp;amp;limit=500&quot;&gt;the “Rolling Stone Album Guide”&lt;/a&gt; will not speak for their unrivalled literary quality, supposedly. Sorting out the literary works from this set of (all kinds of) books would be the obvious next step.&lt;/p&gt;

&lt;p&gt;&lt;img src=&quot;http://vg07.met.vgwort.de/na/d738ac9ec4fb4d3c93330958bddcd03c&quot; width=&quot;1&quot; height=&quot;1&quot; alt=&quot;&quot; /&gt;&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Robert Jäschke, 
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Empirical Data on Over-Length Books</title>
    <link href="https://weltliteratur.net/Empirical-Data-on-Over-Length-Books/"/>
    <updated>2016-05-25T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/Empirical-Data-on-Over-Length-Books</id>
    <content type="html">&lt;p&gt;(Little note upfront: The technical part of this posting involves Docker, RDF→XSLT→JSON, Elasticsearch and Kibana. More on this at the bottom of this page or without further ado &lt;a href=&quot;https://github.com/lehkost/DNBTitel-Elasticsearch&quot;&gt;on our GitHub repo&lt;/a&gt;. Our solution is the result of nothing more than a 4-hour hackathon, so don’t expect anything polished.)&lt;/p&gt;

&lt;p&gt;As a preliminary for a &lt;strong&gt;study of over-length books&lt;/strong&gt; (novels, primarily) we wanted to gather more empirical data. There is still this &lt;a href=&quot;https://en.wikipedia.org/wiki/List_of_longest_novels&quot;&gt;“List of longest novels”&lt;/a&gt; in the English Wikipedia, but this list is problematic, because it’s obviously canon-driven and uses completely different measures to define the extent of books. We also gathered some evidence on our own, a &lt;a href=&quot;http://www.umblaetterer.de/2014/10/21/tausendseiter/&quot;&gt;list with novels of more than a thousand pages&lt;/a&gt;, on another blog (the explanatory text there is in German, watch out). Our list is sorted chronologically, since this is one aspect we have to take into account when planning to pen something like a “History of the Over-Length Novel”.&lt;/p&gt;

&lt;p&gt;Using the number of pages as measurement is, of course, part of the problem. It would be better to work with number of letters or words per work, but this is not (yet?) part of bibliographic metadata. So, number of pages. How do we access them, large-scale?&lt;/p&gt;

&lt;h2 id=&quot;german-national-library-enters-the-stage&quot;&gt;German National Library Enters the Stage&lt;/h2&gt;

&lt;p&gt;Let’s start with this panoramic view of the German Library in Leipzig, predecessor and now part of the German National Library (DNB) – please also note the book towers on the left side:&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/dnb-german-library-leipzig-2008.jpg&quot; alt=&quot;Panoramic view of the German national Library, source: Wikimedia Commons.&quot; /&gt;
  &lt;figcaption&gt;Source: &lt;a href=&quot;https://commons.wikimedia.org/wiki/File:Deutsche_Buecherei_(German_Library)_2008-Feb.jpg&quot;&gt;Wikimedia Commons&lt;/a&gt; (CC BY-SA 3.0)&lt;/figcaption&gt;
&lt;/figure&gt;

&lt;p&gt;Really nice, yes, but it’s another kind of picture of the actual German National Library that we’re going to show you, one that is much more to the point. When we came across the &lt;a href=&quot;http://datendienst.dnb.de/cgi-bin/mabit.pl?userID=opendata&amp;amp;pass=opendata&amp;amp;cmd=login&quot;&gt;data-service page of the German National Library&lt;/a&gt;, we were almost enthusiastic. They offer different kinds of sets in different formats (all licenced under CC0 1.0!), from which we chose the “DNBTitel.rdf.gz” one comprising the records for all books/items archived at the DNB library (the dump was generated on March 10, 2016, and is 1,5 GB in size, which makes for an uncompressed 21,3 GB). There’s no SPARQL endpoint (yet?) by which users could query the catalogue data directly (but you can register for access via &lt;a href=&quot;http://www.dnb.de/oai&quot;&gt;OAI&lt;/a&gt; and &lt;a href=&quot;http://www.dnb.de/sru&quot;&gt;SRU&lt;/a&gt;). So, as said before, we decided to download the RDF file and started to build our own query thing.&lt;/p&gt;

&lt;p&gt;Now, the other picture of the German National Library we wanted to show you is this:&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/dnb-matrix-view-subjects-objects.png&quot; alt=&quot;HDT matrix front view: subjects meet objects.&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;This is the matrix view of the RDF file we downloaded, a 3D scatter plot of triples generated with &lt;a href=&quot;http://www.rdfhdt.org/&quot;&gt;HDT-it!&lt;/a&gt;. Each predicate there has a different colour. In this front view, subjects meet objects, but we can see a pink predicate gleaming through, we’re looking at ISBD element &lt;a href=&quot;http://iflastandards.info/ns/isbd/elements/P1053&quot;&gt;P1053&lt;/a&gt; (“has extent”), the thing we’ll be looking at in the following:&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/dnb-matrix-view-pink-predicate.png&quot; alt=&quot;HDT matrix front view: pink predicate gleaming through.&quot; /&gt;
&lt;/figure&gt;

&lt;h2 id=&quot;some-results&quot;&gt;Some Results&lt;/h2&gt;

&lt;p&gt;Btw, we started this little project as part of a spontaneous 4-hour hackathon last July, at the R&amp;amp;D department of the Göttingen State and University Library. Now, just about a year later, we decided to wrap it up a bit and show you how we made it work.&lt;/p&gt;

&lt;p&gt;First and foremost, you have to be aware of the history of the German National Library to know what you can expect from any query result. The gist of it is: They were a bit late compared to other national libraries in Europe and began collecting “all German and German-language publications from 1913, foreign publications about Germany, translations of German works, and the works of German-speaking emigrants published abroad between 1933 and 1945” (quoting the official &lt;a href=&quot;http://www.dnb.de/EN/Wir/wir_node.html&quot;&gt;“About us” page&lt;/a&gt;).&lt;/p&gt;

&lt;p&gt;Before we bore you with how we did it, let’s go for some results. As a proof of concept, let’s see which authors are the ones with the most books in the catalogue (&lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;dcterms:creator&lt;/code&gt;), let’s generate a top 25:&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/dnb-25-authors-with-most-books.png&quot; alt=&quot;Bar chart: 25 authors with most books in the German National Library.&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;In the bar chart above, the authors are identified by their GND records. Very well then, let’s resolve them:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;DNB Identifier&lt;/th&gt;
      &lt;th&gt;Number of Records&lt;/th&gt;
      &lt;th&gt;Author&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118540238&quot;&gt;118540238&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;5792&lt;/td&gt;
      &lt;td&gt;Goethe, Johann Wolfgang von&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118617443&quot;&gt;118617443&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;3881&lt;/td&gt;
      &lt;td&gt;Steiner, Rudolf&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118577166&quot;&gt;118577166&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;3402&lt;/td&gt;
      &lt;td&gt;Mann, Thomas&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/11855042X&quot;&gt;11855042X&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;3336&lt;/td&gt;
      &lt;td&gt;Hesse, Hermann&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/11856515X&quot;&gt;11856515X&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;3227&lt;/td&gt;
      &lt;td&gt;Konsalik, Heinz G.&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118637479&quot;&gt;118637479&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2995&lt;/td&gt;
      &lt;td&gt;Zweig, Stefan&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118542257&quot;&gt;118542257&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2840&lt;/td&gt;
      &lt;td&gt;Grimm, Jacob&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118578537&quot;&gt;118578537&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2732&lt;/td&gt;
      &lt;td&gt;Marx, Karl&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118542265&quot;&gt;118542265&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2721&lt;/td&gt;
      &lt;td&gt;Grimm, Wilhelm&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118530380&quot;&gt;118530380&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2504&lt;/td&gt;
      &lt;td&gt;Engels, Friedrich&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/12002179X&quot;&gt;12002179X&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2478&lt;/td&gt;
      &lt;td&gt;Schaal, Eric&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118607626&quot;&gt;118607626&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2405&lt;/td&gt;
      &lt;td&gt;Schiller, Friedrich&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118618725&quot;&gt;118618725&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2133&lt;/td&gt;
      &lt;td&gt;Storm, Theodor&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118514768&quot;&gt;118514768&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2132&lt;/td&gt;
      &lt;td&gt;Brecht, Bertolt&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118559230&quot;&gt;118559230&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2100&lt;/td&gt;
      &lt;td&gt;Kafka, Franz&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118559206&quot;&gt;118559206&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2086&lt;/td&gt;
      &lt;td&gt;Kästner, Erich&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118818651&quot;&gt;118818651&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2080&lt;/td&gt;
      &lt;td&gt;May, Karl&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118613723&quot;&gt;118613723&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2070&lt;/td&gt;
      &lt;td&gt;Shakespeare, William&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/11856109X&quot;&gt;11856109X&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1828&lt;/td&gt;
      &lt;td&gt;Keller, Gottfried&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118512676&quot;&gt;118512676&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1818&lt;/td&gt;
      &lt;td&gt;Böll, Heinrich&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118601024&quot;&gt;118601024&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1794&lt;/td&gt;
      &lt;td&gt;Rilke, Rainer Maria&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118587943&quot;&gt;118587943&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1781&lt;/td&gt;
      &lt;td&gt;Nietzsche, Friedrich&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118534262&quot;&gt;118534262&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1755&lt;/td&gt;
      &lt;td&gt;Fontane, Theodor&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118533436&quot;&gt;118533436&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1677&lt;/td&gt;
      &lt;td&gt;Fischer, Marie Louise&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/gnd/118520628&quot;&gt;118520628&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1674&lt;/td&gt;
      &lt;td&gt;Christie, Agatha&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;Looks plausible, in a way. And interesting enough for an interpretation (which, for the time being, we won’t deliver). One thing becomes clearer now, though, we should really talk about “items”, rather than “books”. Photographer Eric Schaal, for example, didn’t get into this top 25 by writing more than two thousand books. To be honest, we made a top 25 just to get at least some women into this board of men, ranking 24th and 25th. And stating the obvious, Goethe also didn’t write almost six thousand books since 1913, what really puts weight on the authors is the substantial number of re-editions, of course.&lt;/p&gt;

&lt;h2 id=&quot;number-of-booksitems-in-the-catalogue&quot;&gt;Number of Books/Items in the Catalogue&lt;/h2&gt;

&lt;p&gt;We’ve got 11.373.862 items altogether (some didn’t make it into the Elasticsearch index since we didn’t really address error handling or validation; the regexps in our XSLT weren’t perfect either, things we can improve next time we’re not high-speed hackathoning). Anyway, in &lt;strong&gt;5.874.504 cases&lt;/strong&gt; we successfully parsed the &lt;a href=&quot;http://iflastandards.info/ns/isbd/elements/P1053&quot;&gt;&lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;isbd:P1053&lt;/code&gt;&lt;/a&gt; element (= ”has extent”) into a usable number of pages, summing up to a total number of 969.846.170 pages. The max number of pages in this set is 2.711.111, which is obviously the result of a metadata apocalypse, somebody must have slipped on the &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;1&lt;/code&gt; key (&lt;a href=&quot;https://twitter.com/umblaetterer/status/735772888687464452&quot;&gt;&lt;strong&gt;this&lt;/strong&gt; is the book that’s said to have more than two million pages&lt;/a&gt;) (ed. 05/27/16: n° of pages &lt;a href=&quot;http://d-nb.info/965667081&quot;&gt;has been corrected today&lt;/a&gt;, they notified us &lt;a href=&quot;https://twitter.com/DNB_Aktuelles/status/736086931646205952&quot;&gt;via Twitter&lt;/a&gt;, nice!).&lt;/p&gt;

&lt;p&gt;You can glean from our &lt;a href=&quot;https://github.com/lehkost/DNBTitel-Elasticsearch/blob/master/dnb2es/rdf2json.xsl&quot;&gt;XSLT file&lt;/a&gt; that we’re only using extent information if we found a number succeeded by “ S.” (= ”pages”) in the &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;isbd:P1053&lt;/code&gt; string. So we’re not using at least 5.499.358 items out of the 11.373.862. Either they had no information on the book extent/number of pages or we didn’t parse it because we just used a very basic pattern. But with this simple method we still managed to cover 51,65% of all the books/items stored in the German National Library. We can sure improve our data extraction, but for the time being we’re good with what we have. After all, we’re still speaking about almost six million book records.&lt;/p&gt;

&lt;p&gt;Now onto some more meaningful stuff. Let’s take five major publishers and compare them just by looking at the extent of their books. We held a little powwow and, full of intentional bias, chose Aufbau, Eichborn, Hanser, Rowohlt, Suhrkamp. Let’s have a look at the number of items per publisher in the catalogue:&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/dnb-5-publishers-number-of-items.png&quot; alt=&quot;Bar chart: 5 publishers, number of items.&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;This is an interesting perspective, also if this comparison makes not much sense. For all we know, there could be thousands of re-editions involved. Well, okay.&lt;/p&gt;

&lt;h2 id=&quot;average-number-of-pages-per-book-per-publisher&quot;&gt;Average Number of Pages per Book per Publisher&lt;/h2&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/dnb-5-publishers-average-number-of-pages.png&quot; alt=&quot;Bar chart: 5 publishers, average number of pages.&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;So obviously, the average Suhrkamp book beats the average Rowohlt book beats the average Eichborn book when considering the number of pages. This comparison is intriguing, but it shouldn’t be taken too literally, we’re dealing with a certain amount of incorrect metadata as we’ll see in the next set of lists.&lt;/p&gt;

&lt;h2 id=&quot;longest-books-per-publisher&quot;&gt;Longest Books per Publisher&lt;/h2&gt;

&lt;p&gt;This brings us a bit closer to our goal. But these lists also show the problems of erroneous data and the need for additional metadata. If we want to look into over-length novels, we’ll obviously need another indicator, one of which is not provided by the current DNB dataset. And once again, the lists clarify that it’s more appropriate to speak of “item” than of “book”, but now let’s start with the rankings:&lt;/p&gt;

&lt;h3 id=&quot;aufbau&quot;&gt;Aufbau&lt;/h3&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;DNB Identifier&lt;/th&gt;
      &lt;th&gt;Number of Pages&lt;/th&gt;
      &lt;th&gt;Author: Title (Year)&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/988488205&quot;&gt;988488205&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1359 S.&lt;/td&gt;
      &lt;td&gt;Vikram Chandra: Der Pate von Bombay (2009)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/1001932447&quot;&gt;1001932447&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1291 S.&lt;/td&gt;
      &lt;td&gt;Lew Tolstoi: Krieg und Frieden (2010)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/576696420&quot;&gt;576696420&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1286 S.&lt;/td&gt;
      &lt;td&gt;Alexej Tolstoi: Der Leidensweg (2. Aufl., 1955)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/576696439&quot;&gt;576696439&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1286 S.&lt;/td&gt;
      &lt;td&gt;Alexej Tolstoi: Der Leidensweg (3. Aufl., 1959)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/1011565994&quot;&gt;1011565994&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1243 S.&lt;/td&gt;
      &lt;td&gt;Hans Fallada: Wolf unter Wölfen (2011)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/98848823X&quot;&gt;98848823X&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1227 S.&lt;/td&gt;
      &lt;td&gt;Lew Tolstoi: Anna Karenina (2008)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/945188846&quot;&gt;945188846&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1211 S.&lt;/td&gt;
      &lt;td&gt;Friedrich Gorenstein: Der Platz (1995)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/988488272&quot;&gt;988488272&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1200 S.&lt;/td&gt;
      &lt;td&gt;Fjodor Dostojewski: Die Brüder Karamasow (2008)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/949346470&quot;&gt;949346470&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1183 S.&lt;/td&gt;
      &lt;td&gt;Lew Tolstoi: Anna Karenina (1996)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/451896025&quot;&gt;451896025&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1174 S.&lt;/td&gt;
      &lt;td&gt;G. W. F. Hegel: Ästhetik (1955)&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;h3 id=&quot;eichborn&quot;&gt;Eichborn&lt;/h3&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;DNB Identifier&lt;/th&gt;
      &lt;th&gt;Number of Pages&lt;/th&gt;
      &lt;th&gt;Author: Title (Year)&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/946561486&quot;&gt;946561486&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;&lt;s&gt;1814 S.&lt;/s&gt;&lt;/td&gt;
      &lt;td&gt;(wrong number of pages)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/967526825&quot;&gt;967526825&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1222 S.&lt;/td&gt;
      &lt;td&gt;Leo Tolstoi: Krieg und Frieden (2003)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/950298603&quot;&gt;950298603&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1081 S.&lt;/td&gt;
      &lt;td&gt;Rolf Vollmann: Die wunderbaren Falschmünzer (1997)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/988571005&quot;&gt;988571005&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;991 S.&lt;/td&gt;
      &lt;td&gt;Daniel Schwartz: Schnee in Samarkand (2008)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/979687187&quot;&gt;979687187&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;954 S.&lt;/td&gt;
      &lt;td&gt;Paul Verhaeghen: Omega minor (2006)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/974540919&quot;&gt;974540919&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;855 S.&lt;/td&gt;
      &lt;td&gt;David M. Crowe: Oskar Schindler (2005)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/979691044&quot;&gt;979691044&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;852 S.&lt;/td&gt;
      &lt;td&gt;Laurence Sterne: Leben und Ansichten von Tristram Shandy, Gentleman (2006)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/840181604&quot;&gt;840181604&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;841 S.&lt;/td&gt;
      &lt;td&gt;Fred Denger: Der grosse Boss (1984)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/860929477&quot;&gt;860929477&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;841 S.&lt;/td&gt;
      &lt;td&gt;Fred Denger: Der grosse Boss (6. Aufl., 1985)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/870140876&quot;&gt;870140876&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;841 S.&lt;/td&gt;
      &lt;td&gt;Fred Denger: Der grosse Boss (5. Aufl., 1985)&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;h3 id=&quot;hanser&quot;&gt;Hanser&lt;/h3&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;DNB Identifier&lt;/th&gt;
      &lt;th&gt;Number of Pages&lt;/th&gt;
      &lt;th&gt;Author: Title (Year)&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/1022146394&quot;&gt;1022146394&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;&lt;s&gt;4587 S.&lt;/s&gt;&lt;/td&gt;
      &lt;td&gt;(wrong number of pages)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/99886398X&quot;&gt;99886398X&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1810 S.&lt;/td&gt;
      &lt;td&gt;Walter Doberenz; Thomas Gewinnus: Visual C# 2010 (2010)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/98401098X&quot;&gt;98401098X&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1806 S.&lt;/td&gt;
      &lt;td&gt;Walter Doberenz; Thomas Gewinnus: Borland Delphi 7 (2007)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/998863955&quot;&gt;998863955&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1802 S.&lt;/td&gt;
      &lt;td&gt;Walter Doberenz; Thomas Gewinnus: Visual Basic 2010 (2010)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/760043035&quot;&gt;760043035&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1672 S.&lt;/td&gt;
      &lt;td&gt;Joseph von Eichendorff: Werke (4. Aufl., 1971)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/451062442&quot;&gt;451062442&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1606 S.&lt;/td&gt;
      &lt;td&gt;Joseph von Eichendorff: Werke (2. Aufl., 1959)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/451062817&quot;&gt;451062817&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1590 S.&lt;/td&gt;
      &lt;td&gt;Joseph von Eichendorff: Werke (1. Aufl., 1955)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/453424740&quot;&gt;453424740&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1511 S.&lt;/td&gt;
      &lt;td&gt;Eduard Mörike: Sämtliche Werke (3. Aufl., 1964)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/780204875&quot;&gt;780204875&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1511 S.&lt;/td&gt;
      &lt;td&gt;Eduard Mörike: Sämtliche Werke (5. Aufl., 1976)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/970961294&quot;&gt;970961294&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1469 S.&lt;/td&gt;
      &lt;td&gt;Uwe Bünning; Jörg Krause: Windows XP Professional (3. Aufl., 2004)&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;h3 id=&quot;rowohlt&quot;&gt;Rowohlt&lt;/h3&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;DNB Identifier&lt;/th&gt;
      &lt;th&gt;Number of Pages&lt;/th&gt;
      &lt;th&gt;Author: Title (Year)&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/944325807&quot;&gt;944325807&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2253 S.&lt;/td&gt;
      &lt;td&gt;Klaus Harpprecht: Thomas Mann (1995)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/945394659&quot;&gt;945394659&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2253 S.&lt;/td&gt;
      &lt;td&gt;Klaus Harpprecht: Thomas Mann (16.–30. Tsd., 1995)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/967713358&quot;&gt;967713358&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2026 S.&lt;/td&gt;
      &lt;td&gt;Karl Corino: Robert Musil (2003)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/101781905X&quot;&gt;101781905X&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1723 S.&lt;/td&gt;
      &lt;td&gt;Péter Nádas: Parallelgeschichten (2012)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/1028105657&quot;&gt;1028105657&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1723 S.&lt;/td&gt;
      &lt;td&gt;Péter Nádas: Parallelgeschichten (Taschenbuch, 2013)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/1008548022&quot;&gt;1008548022&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1719 S.&lt;/td&gt;
      &lt;td&gt;Rolf Hochhuth: Essayistische Prosa und Gedichte (2011)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/575594950&quot;&gt;575594950&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1671 S.&lt;/td&gt;
      &lt;td&gt;Robert Musil: Der Mann ohne Eigenschaften (1952)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/961281588&quot;&gt;961281588&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1642 S.&lt;/td&gt;
      &lt;td&gt;Rolf Hochhuth: Alle Erzählungen, Gedichte und Romane (2001)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/457661054&quot;&gt;457661054&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1632 S.&lt;/td&gt;
      &lt;td&gt;Robert Musil: Der Mann ohne Eigenschaften (1970)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/575594969&quot;&gt;575594969&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1632 S.&lt;/td&gt;
      &lt;td&gt;Robert Musil: Der Mann ohne Eigenschaften (23.–29. Tsd., 1960)&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;h3 id=&quot;suhrkamp&quot;&gt;Suhrkamp&lt;/h3&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;DNB Identifier&lt;/th&gt;
      &lt;th&gt;Number of Pages&lt;/th&gt;
      &lt;th&gt;Author: Title (Year)&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/946102384&quot;&gt;946102384&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;&lt;s&gt;3980 S.&lt;/s&gt;&lt;/td&gt;
      &lt;td&gt;(wrong number of pages)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/945262094&quot;&gt;945262094&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;&lt;s&gt;2909 S.&lt;/s&gt;&lt;/td&gt;
      &lt;td&gt;(wrong number of pages)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/991420225&quot;&gt;991420225&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2569 S.&lt;/td&gt;
      &lt;td&gt;Amos Oz: Die Romane (2009)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/988814668&quot;&gt;988814668&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;2085 S.&lt;/td&gt;
      &lt;td&gt;E. M. Cioran: Werke (2008)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/986493635&quot;&gt;986493635&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1909 S.&lt;/td&gt;
      &lt;td&gt;Marguerite Duras: Die Romane (2008)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/986531766&quot;&gt;986531766&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1840 S.&lt;/td&gt;
      &lt;td&gt;Thomas Bernhard: Die Romane (2008)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/991398939&quot;&gt;991398939&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1838 S.&lt;/td&gt;
      &lt;td&gt;Hermann Hesse: Die Erzählungen und Märchen (2009)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/988840758&quot;&gt;988840758&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1782 S.&lt;/td&gt;
      &lt;td&gt;Max Frisch: Romane, Erzählungen, Tagebücher (2008)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/998413925&quot;&gt;998413925&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1782 S.&lt;/td&gt;
      &lt;td&gt;Bertolt Brecht: Prosa (2013)&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://d-nb.info/1008349852&quot;&gt;1008349852&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;1735 S.&lt;/td&gt;
      &lt;td&gt;Alejo Carpentier: Die Romane (2011)&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;We won’t comment these lists now, although they make for some interesting discussions. You can obviously do much more with what we built here and we’ll certainly get back to this later. So let’s close this blog post with a short note on how we built this bridge from the freely available DNB catalogue data to the results shown above.&lt;/p&gt;

&lt;h2 id=&quot;german-national-library-goes-elasticsearch&quot;&gt;German National Library Goes Elasticsearch&lt;/h2&gt;

&lt;p&gt;The somewhat weird original idea we had when initiating the hackathon was this: We wanted to know how much the German National Library weighs on books, and we wanted to find out by the number of pages of all the books it stores which we then would have multiplied by the average weight of a book page. Well, you can do the maths yourself now, you can find the total number of pages we counted above and then extrapolate, don’t forget to divide this by two, it will be a good enough approximation.&lt;/p&gt;

&lt;p&gt;What saves us time now is that we already described our mechanism &lt;a href=&quot;https://github.com/lehkost/DNBTitel-Elasticsearch&quot;&gt;on GitHub&lt;/a&gt; where we also provide all the info you need to rebuild our machine. We basically reorganised the whole thing as a Docker project, which will create a container running Elasticsearch/Kibana. The repo also features shell scripts for downloading the current version of the German National Library title catalogue. Some selected data fields from every book in that catalogue are then transformed into JSON and pushed to the Elasticsearch instance. After that you will be able to query the DNB catalogue data with Elasticsearch to create nice outputs with Kibana. As should be clear from the text above, the data fields we’re focusing on are mainly the number of pages per book and some book metadata (author, title, year, publisher, etc.) for identification.&lt;/p&gt;

&lt;p&gt;That’s it for now. And lest we forget, special thanks to Max Brodhun and Carsten Thiel for writing the XSLT and helping with the shell scripting, “as quick as boiled asparagus”, so to speak, it was only because of them that we could go on with what we &lt;em&gt;actually&lt;/em&gt; wanted to do. &lt;a href=&quot;https://twitter.com/umblaetterer/status/679317574740533248&quot;&gt;The next Uludağ is on us!&lt;/a&gt; ;)&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer, 
	
          
          Ubbo Veentjer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Fefe Research Institute</title>
    <link href="https://weltliteratur.net/Fefe-Research-Institute/"/>
    <updated>2016-05-18T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/Fefe-Research-Institute</id>
    <content type="html">&lt;p&gt;Ok, the headline is a joke, of course. But let’s start at the beginning. There are three things German computer scientists/IT guys must not miss: their &lt;a href=&quot;https://www.heise.de/&quot;&gt;Heise Online&lt;/a&gt; newsticker, their &lt;a href=&quot;https://xkcd.com/&quot;&gt;xkcd&lt;/a&gt;, and their daily Fefe.&lt;/p&gt;

&lt;p&gt;Fefe is the nom de guerre of &lt;a href=&quot;https://en.wikipedia.org/wiki/Felix_von_Leitner&quot;&gt;Felix von Leitner&lt;/a&gt;, a C programmer (check &lt;a href=&quot;https://en.wikipedia.org/wiki/Dietlibc&quot;&gt;dietlibc&lt;/a&gt;), IT security adept and blogger, running his own speaker’s corner &lt;strong&gt;&lt;a href=&quot;https://blog.fefe.de/&quot;&gt;“Fefes Blog”&lt;/a&gt;&lt;/strong&gt; since 2005. This blog is notorious for its “cut the crap” non-layout and his author for his brash comments on incidents in the broad field of IT security and their political implications, especially regarding the industry of mass surveillance. His reach is enormous and will surely beat that of many established newspapers (according to &lt;a href=&quot;https://blog.fefe.de/?ts=b3d1b6bf&quot;&gt;this 2011 post&lt;/a&gt;, anyway; newer numbers would be helpful, though).&lt;/p&gt;

&lt;p&gt;We don’t know if pressing F5 at least three times a day in the browser tab reserved for “Fefes Blog” makes us part of the fan base. Either way, we couldn’t abstain from sneaking a peek behind the curtain and wanted to analyse some traits of Fefe’s characteristic style and tone. They are, in fact, so characteristic that he was even &lt;a href=&quot;http://www.titanic-magazin.de/news/was-passiert-eigentlich-gerade-auf-fefes-blog-6220/&quot;&gt;mimicked by the ultimate German satire magazine “Titanic”&lt;/a&gt;, which sure counts for something.&lt;/p&gt;

&lt;p&gt;Before we get to it: There have been other attempts to analyse Fefe’s language, just take the inspiring 2012 blog post &lt;a href=&quot;http://www.security-informatics.de/blog/?p=956&quot;&gt;“Darüber lacht Fefe”&lt;/a&gt; by Joachim Scharloth. What we’re going to do is, in fact, something very basic, we’re primarily looking for word (and URL) frequencies and n-grams. For that purpose, we wrote a little Python script which does the following three things:&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;download the whole blog (easy on bandwidth, because, one: “Fefes Blog” really features the lightest HTML code imaginable, and two: &lt;code class=&quot;language-plaintext highlighter-rouge&quot;&gt;time.sleep(random.randint(1, 10))&lt;/code&gt; while scraping)&lt;/li&gt;
  &lt;li&gt;parse HTML files: extract date, post identifier and the actual post and put them into a single 3-column TSV file (this is really some of the cleanest data you can work with, what a feast!)&lt;/li&gt;
  &lt;li&gt;start analysis&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;After step two, the TSV file contained all blog entries from the beginning (March, 2005) to Mid-May, 2016, which is the cut-off time for our dataset.&lt;/p&gt;

&lt;h2 id=&quot;main-sources&quot;&gt;Main Sources&lt;/h2&gt;

&lt;p&gt;These are the most frequent top-level domains linked to in “Fefes Blog”:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;Rank&lt;/th&gt;
      &lt;th&gt;TLD&lt;/th&gt;
      &lt;th&gt;Frequency&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;spiegel.de&lt;/td&gt;
      &lt;td&gt;4042&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;2&lt;/td&gt;
      &lt;td&gt;fefe.de&lt;/td&gt;
      &lt;td&gt;3316&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;3&lt;/td&gt;
      &lt;td&gt;heise.de&lt;/td&gt;
      &lt;td&gt;2673&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;4&lt;/td&gt;
      &lt;td&gt;wikipedia.org&lt;/td&gt;
      &lt;td&gt;1319&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;5&lt;/td&gt;
      &lt;td&gt;bbc.co.uk&lt;/td&gt;
      &lt;td&gt;1102&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;6&lt;/td&gt;
      &lt;td&gt;tagesschau.de&lt;/td&gt;
      &lt;td&gt;1096&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;7&lt;/td&gt;
      &lt;td&gt;guardian.co.uk&lt;/td&gt;
      &lt;td&gt;1043&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;8&lt;/td&gt;
      &lt;td&gt;youtube.com&lt;/td&gt;
      &lt;td&gt;1010&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;9&lt;/td&gt;
      &lt;td&gt;nytimes.com&lt;/td&gt;
      &lt;td&gt;692&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;10&lt;/td&gt;
      &lt;td&gt;twitter.com&lt;/td&gt;
      &lt;td&gt;624&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;11&lt;/td&gt;
      &lt;td&gt;tagesspiegel.de&lt;/td&gt;
      &lt;td&gt;531&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;12&lt;/td&gt;
      &lt;td&gt;yahoo.com&lt;/td&gt;
      &lt;td&gt;489&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;13&lt;/td&gt;
      &lt;td&gt;rian.ru&lt;/td&gt;
      &lt;td&gt;485&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;14&lt;/td&gt;
      &lt;td&gt;zeit.de&lt;/td&gt;
      &lt;td&gt;475&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;15&lt;/td&gt;
      &lt;td&gt;faz.net&lt;/td&gt;
      &lt;td&gt;468&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;16&lt;/td&gt;
      &lt;td&gt;reuters.com&lt;/td&gt;
      &lt;td&gt;443&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;17&lt;/td&gt;
      &lt;td&gt;taz.de&lt;/td&gt;
      &lt;td&gt;408&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;18&lt;/td&gt;
      &lt;td&gt;sueddeutsche.de&lt;/td&gt;
      &lt;td&gt;400&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;19&lt;/td&gt;
      &lt;td&gt;cnn.com&lt;/td&gt;
      &lt;td&gt;376&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;20&lt;/td&gt;
      &lt;td&gt;washingtonpost.com&lt;/td&gt;
      &lt;td&gt;364&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;So this list shows Fefe’s main sources. Now, it’s part of his irony to urge his readers – in the subheader of the blog – to send him “fancy conspiracy links”. This, at least for some, leaves room for irritation and Fefe kind of clarifies this in the &lt;a href=&quot;https://blog.fefe.de/faq.html&quot;&gt;FAQ&lt;/a&gt;: “Why do you write ‘conspiracy links’ if your content is all normal news?” Reply: “Yes.”&lt;/p&gt;

&lt;h2 id=&quot;countries-mentioned&quot;&gt;Countries Mentioned&lt;/h2&gt;

&lt;p&gt;This list is a bit half-baked, since we’re only counting exact hits and only added synonyms for two countries (GB, USA). Wales, Scotland and England were, too, mapped to Great Britain. All a bit hasty, but anyway:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;Rank&lt;/th&gt;
      &lt;th&gt;Country&lt;/th&gt;
      &lt;th&gt;Frequency&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;USA&lt;/td&gt;
      &lt;td&gt;1918&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;2&lt;/td&gt;
      &lt;td&gt;Deutschland&lt;/td&gt;
      &lt;td&gt;1506&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;3&lt;/td&gt;
      &lt;td&gt;Israel&lt;/td&gt;
      &lt;td&gt;997&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;4&lt;/td&gt;
      &lt;td&gt;Iran&lt;/td&gt;
      &lt;td&gt;927&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;5&lt;/td&gt;
      &lt;td&gt;Großbritannien&lt;/td&gt;
      &lt;td&gt;708&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;6&lt;/td&gt;
      &lt;td&gt;China&lt;/td&gt;
      &lt;td&gt;657&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;7&lt;/td&gt;
      &lt;td&gt;Irak&lt;/td&gt;
      &lt;td&gt;474&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;8&lt;/td&gt;
      &lt;td&gt;Afghanistan&lt;/td&gt;
      &lt;td&gt;354&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;9&lt;/td&gt;
      &lt;td&gt;Griechenland&lt;/td&gt;
      &lt;td&gt;298&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;10&lt;/td&gt;
      &lt;td&gt;Ukraine&lt;/td&gt;
      &lt;td&gt;294&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;11&lt;/td&gt;
      &lt;td&gt;Frankreich&lt;/td&gt;
      &lt;td&gt;287&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;12&lt;/td&gt;
      &lt;td&gt;Türkei&lt;/td&gt;
      &lt;td&gt;279&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;13&lt;/td&gt;
      &lt;td&gt;Japan&lt;/td&gt;
      &lt;td&gt;253&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;14&lt;/td&gt;
      &lt;td&gt;Syrien&lt;/td&gt;
      &lt;td&gt;231&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;15&lt;/td&gt;
      &lt;td&gt;Schweiz&lt;/td&gt;
      &lt;td&gt;222&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;16&lt;/td&gt;
      &lt;td&gt;Österreich&lt;/td&gt;
      &lt;td&gt;209&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;17&lt;/td&gt;
      &lt;td&gt;Polen&lt;/td&gt;
      &lt;td&gt;177&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;18&lt;/td&gt;
      &lt;td&gt;Pakistan&lt;/td&gt;
      &lt;td&gt;172&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;19&lt;/td&gt;
      &lt;td&gt;Italien&lt;/td&gt;
      &lt;td&gt;167&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;20&lt;/td&gt;
      &lt;td&gt;Schweden&lt;/td&gt;
      &lt;td&gt;158&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;We matched our dataset against the list of countries provided &lt;a href=&quot;http://www.datendieter.de/item/Verzeichnis_aller_Staaten_der_Welt&quot;&gt;by datendieter.de&lt;/a&gt;. (Btw, it’s very probable that Daten-Dieter is a close friend of &lt;a href=&quot;https://www.youtube.com/watch?v=2Nlm2XSBJmQ&quot;&gt;MS-DOS-Manfred, BIOS-Bernhard, Hardware-Hanspeter and Lötkolben-Ludwig&lt;/a&gt;, but that’s a whole different story.)&lt;/p&gt;

&lt;h2 id=&quot;countries-mentioned-over-time&quot;&gt;Countries Mentioned Over Time&lt;/h2&gt;

&lt;p&gt;We can also add time as a component, so we let Gnuplot draw a line chart showing Fefe’s interest in the 12 most-mentioned countries over time:&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/fefe_time_countries.svg&quot; alt=&quot;Mentions of countries over time in Fefes Blog.&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;This chart we can actually try to read out loud a bit: The Snowden year, 2013, marks the beginning of an increasing coverage of the USA and Great Britain. In 2014, Russia and Ukraine are peaking, for obvious reasons. One more interesting thing is the slow but steady decline of interest in Israel and Iran since 2005. Greece is starting to be covered in late 2009 with the beginning of the Greek government-debt crisis. So as can be expected from the list of his main sources seen above, “Fefes Blog” more or less mirrors mainstream media coverage.&lt;/p&gt;

&lt;h2 id=&quot;acronyms&quot;&gt;Acronyms&lt;/h2&gt;

&lt;p&gt;Well, acronyms, or words written in capitel letters:&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;Rank&lt;/th&gt;
      &lt;th&gt;Acronym/Word&lt;/th&gt;
      &lt;th&gt;Frequency&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;USA&lt;/td&gt;
      &lt;td&gt;2332&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;2&lt;/td&gt;
      &lt;td&gt;CDU&lt;/td&gt;
      &lt;td&gt;1213&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;3&lt;/td&gt;
      &lt;td&gt;US&lt;/td&gt;
      &lt;td&gt;927&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;4&lt;/td&gt;
      &lt;td&gt;EU&lt;/td&gt;
      &lt;td&gt;926&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;5&lt;/td&gt;
      &lt;td&gt;NSA&lt;/td&gt;
      &lt;td&gt;800&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;6&lt;/td&gt;
      &lt;td&gt;SPD&lt;/td&gt;
      &lt;td&gt;723&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;7&lt;/td&gt;
      &lt;td&gt;WTF&lt;/td&gt;
      &lt;td&gt;556&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;8&lt;/td&gt;
      &lt;td&gt;BND&lt;/td&gt;
      &lt;td&gt;541&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;9&lt;/td&gt;
      &lt;td&gt;FDP&lt;/td&gt;
      &lt;td&gt;486&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;10&lt;/td&gt;
      &lt;td&gt;BKA&lt;/td&gt;
      &lt;td&gt;484&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;11&lt;/td&gt;
      &lt;td&gt;CIA&lt;/td&gt;
      &lt;td&gt;467&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;12&lt;/td&gt;
      &lt;td&gt;CCC&lt;/td&gt;
      &lt;td&gt;393&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;13&lt;/td&gt;
      &lt;td&gt;DAS&lt;/td&gt;
      &lt;td&gt;366&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;14&lt;/td&gt;
      &lt;td&gt;OK&lt;/td&gt;
      &lt;td&gt;341&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;15&lt;/td&gt;
      &lt;td&gt;FBI&lt;/td&gt;
      &lt;td&gt;313&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;16&lt;/td&gt;
      &lt;td&gt;DIE&lt;/td&gt;
      &lt;td&gt;277&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;17&lt;/td&gt;
      &lt;td&gt;BESTEN&lt;/td&gt;
      &lt;td&gt;253&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;18&lt;/td&gt;
      &lt;td&gt;IV&lt;/td&gt;
      &lt;td&gt;250&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;19&lt;/td&gt;
      &lt;td&gt;CSU&lt;/td&gt;
      &lt;td&gt;232&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;20&lt;/td&gt;
      &lt;td&gt;NIE&lt;/td&gt;
      &lt;td&gt;217&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;h2 id=&quot;3-grams&quot;&gt;3-Grams&lt;/h2&gt;

&lt;p&gt;Let’s now look at some n-grams (we used &lt;a href=&quot;http://www.laurenceanthony.net/software/antconc/&quot;&gt;AntConc&lt;/a&gt; for this). Attention, the lists were curated by us and only show the frequencies for self-contained phrases, which in our opinion are contributing to the typical Fefe sound.&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;Rank&lt;/th&gt;
      &lt;th&gt;Frequency&lt;/th&gt;
      &lt;th&gt;Phrase&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;863&lt;/td&gt;
      &lt;td&gt;in den usa&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;3&lt;/td&gt;
      &lt;td&gt;387&lt;/td&gt;
      &lt;td&gt;lacher des tages&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;28&lt;/td&gt;
      &lt;td&gt;209&lt;/td&gt;
      &lt;td&gt;das ehemalige nachrichtenmagazin&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;49&lt;/td&gt;
      &lt;td&gt;173&lt;/td&gt;
      &lt;td&gt;bug des tages&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;56&lt;/td&gt;
      &lt;td&gt;166&lt;/td&gt;
      &lt;td&gt;die amis haben&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;Runner-up among the 3-grams is “das ist ja”, and then we have “das ist ein” and “was für eine” ranking 4th and 5th. Sure, all these high-frequent 3-grams add to Fefe’s stylometric fingerprint, but since we’re not (yet) doing stylometry here, we decided to curate the n-gram lists a bit to not overthrow you with endless lists of boring syntagmas.&lt;/p&gt;

&lt;h2 id=&quot;4-grams&quot;&gt;4-Grams&lt;/h2&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;Rank&lt;/th&gt;
      &lt;th&gt;Frequency&lt;/th&gt;
      &lt;th&gt;Phrase&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;3&lt;/td&gt;
      &lt;td&gt;214&lt;/td&gt;
      &lt;td&gt;kommt ihr nie drauf&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;10&lt;/td&gt;
      &lt;td&gt;115&lt;/td&gt;
      &lt;td&gt;stellt sich raus dass&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;13&lt;/td&gt;
      &lt;td&gt;109&lt;/td&gt;
      &lt;td&gt;oh und wo wir&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;38&lt;/td&gt;
      &lt;td&gt;75&lt;/td&gt;
      &lt;td&gt;kennt ihr den schon&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;39&lt;/td&gt;
      &lt;td&gt;75&lt;/td&gt;
      &lt;td&gt;was für eine farce&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;49&lt;/td&gt;
      &lt;td&gt;68&lt;/td&gt;
      &lt;td&gt;wie arsch auf eimer&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;61&lt;/td&gt;
      &lt;td&gt;61&lt;/td&gt;
      &lt;td&gt;einmal mit profis arbeiten&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;67&lt;/td&gt;
      &lt;td&gt;60&lt;/td&gt;
      &lt;td&gt;wir werden alle störben&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;69&lt;/td&gt;
      &lt;td&gt;59&lt;/td&gt;
      &lt;td&gt;wer hätte das gedacht&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;h2 id=&quot;5-grams&quot;&gt;5-Grams&lt;/h2&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;Rank&lt;/th&gt;
      &lt;th&gt;Frequency&lt;/th&gt;
      &lt;th&gt;Phrase&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;21&lt;/td&gt;
      &lt;td&gt;62&lt;/td&gt;
      &lt;td&gt;habt ihr das auch gehört&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;42&lt;/td&gt;
      &lt;td&gt;45&lt;/td&gt;
      &lt;td&gt;bei uns ist kernkraft sicher&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;50&lt;/td&gt;
      &lt;td&gt;40&lt;/td&gt;
      &lt;td&gt;bei uns ist atomkraft sicher&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;58&lt;/td&gt;
      &lt;td&gt;37&lt;/td&gt;
      &lt;td&gt;was kann da schon passieren&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;81&lt;/td&gt;
      &lt;td&gt;30&lt;/td&gt;
      &lt;td&gt;was kann da schon schiefgehen&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;h2 id=&quot;6-grams&quot;&gt;6-Grams&lt;/h2&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;Rank&lt;/th&gt;
      &lt;th&gt;Frequency&lt;/th&gt;
      &lt;th&gt;Phrase&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;2&lt;/td&gt;
      &lt;td&gt;89&lt;/td&gt;
      &lt;td&gt;oh und wo wir gerade bei&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;7&lt;/td&gt;
      &lt;td&gt;65&lt;/td&gt;
      &lt;td&gt;die polizei dein freund und helfer&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;10&lt;/td&gt;
      &lt;td&gt;60&lt;/td&gt;
      &lt;td&gt;na dann ist ja alles gut&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;13&lt;/td&gt;
      &lt;td&gt;52&lt;/td&gt;
      &lt;td&gt;da weiß man was man hat&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;17&lt;/td&gt;
      &lt;td&gt;41&lt;/td&gt;
      &lt;td&gt;update mir mailt gerade jemand dass&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;25&lt;/td&gt;
      &lt;td&gt;35&lt;/td&gt;
      &lt;td&gt;das geht ja mal gar nicht&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;h2 id=&quot;7-grams&quot;&gt;7-Grams&lt;/h2&gt;

&lt;p&gt;The 7-grams are simply a class of their own and prove to be &lt;strong&gt;real Fefe earworms&lt;/strong&gt;, don’t they?&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;Rank&lt;/th&gt;
      &lt;th&gt;Frequency&lt;/th&gt;
      &lt;th&gt;Phrase&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;85&lt;/td&gt;
      &lt;td&gt;also damit konnte ja wohl niemand rechnen&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;2&lt;/td&gt;
      &lt;td&gt;76&lt;/td&gt;
      &lt;td&gt;die besten der besten der besten sir&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;3&lt;/td&gt;
      &lt;td&gt;63&lt;/td&gt;
      &lt;td&gt;das wird euch jetzt sicher genau so&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;4&lt;/td&gt;
      &lt;td&gt;52&lt;/td&gt;
      &lt;td&gt;kann man sich gar nicht ausdenken sowas&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;5&lt;/td&gt;
      &lt;td&gt;48&lt;/td&gt;
      &lt;td&gt;aus der beliebten kategorie bei uns ist&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;6&lt;/td&gt;
      &lt;td&gt;38&lt;/td&gt;
      &lt;td&gt;man sich mal auf der zunge zergehen&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;12&lt;/td&gt;
      &lt;td&gt;33&lt;/td&gt;
      &lt;td&gt;aus der beliebten reihe bei uns ist&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;13&lt;/td&gt;
      &lt;td&gt;33&lt;/td&gt;
      &lt;td&gt;aus der beliebten serie bei uns ist&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;15&lt;/td&gt;
      &lt;td&gt;31&lt;/td&gt;
      &lt;td&gt;wo kämen wir da auch hin wenn&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;24&lt;/td&gt;
      &lt;td&gt;25&lt;/td&gt;
      &lt;td&gt;beste demokratie die man für geld kaufen&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;30&lt;/td&gt;
      &lt;td&gt;24&lt;/td&gt;
      &lt;td&gt;da fühlt man sich doch gleich viel&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;33&lt;/td&gt;
      &lt;td&gt;21&lt;/td&gt;
      &lt;td&gt;gar nicht so viel fressen wie man&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;h2 id=&quot;fun-part-i-given-names-of-german-it-guys&quot;&gt;Fun Part I: Given Names of German IT Guys&lt;/h2&gt;

&lt;p&gt;When Fefe posts a link, hint or story that someone sent him, he gives due credit, the pattern being “(Danke, [name].)” So we also looked into 2-grams starting with “danke”, and while the result was somehow predictable, it is still quite funny, isn’t it? It not only sheds a light on who Fefe’s fiercest audience is, but also shows the opulence of names given to German boys in the 1970ies and 1980ies.&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;Rank&lt;/th&gt;
      &lt;th&gt;Frequency&lt;/th&gt;
      &lt;th&gt;2-Gram/Name&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;1&lt;/td&gt;
      &lt;td&gt;237&lt;/td&gt;
      &lt;td&gt;danke mathias&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;2&lt;/td&gt;
      &lt;td&gt;220&lt;/td&gt;
      &lt;td&gt;danke frank&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;3&lt;/td&gt;
      &lt;td&gt;183&lt;/td&gt;
      &lt;td&gt;danke christian&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;4&lt;/td&gt;
      &lt;td&gt;176&lt;/td&gt;
      &lt;td&gt;danke thomas&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;5&lt;/td&gt;
      &lt;td&gt;164&lt;/td&gt;
      &lt;td&gt;danke stefan&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;6&lt;/td&gt;
      &lt;td&gt;135&lt;/td&gt;
      &lt;td&gt;danke andreas&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;7&lt;/td&gt;
      &lt;td&gt;135&lt;/td&gt;
      &lt;td&gt;danke michael&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;8&lt;/td&gt;
      &lt;td&gt;106&lt;/td&gt;
      &lt;td&gt;danke martin&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;9&lt;/td&gt;
      &lt;td&gt;106&lt;/td&gt;
      &lt;td&gt;danke peter&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;10&lt;/td&gt;
      &lt;td&gt;100&lt;/td&gt;
      &lt;td&gt;danke daniel&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;11&lt;/td&gt;
      &lt;td&gt;100&lt;/td&gt;
      &lt;td&gt;danke klaus&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;12&lt;/td&gt;
      &lt;td&gt;95&lt;/td&gt;
      &lt;td&gt;danke matthias&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;13&lt;/td&gt;
      &lt;td&gt;94&lt;/td&gt;
      &lt;td&gt;danke florian&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;14&lt;/td&gt;
      &lt;td&gt;94&lt;/td&gt;
      &lt;td&gt;danke jan&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;15&lt;/td&gt;
      &lt;td&gt;91&lt;/td&gt;
      &lt;td&gt;danke rop&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;16&lt;/td&gt;
      &lt;td&gt;84&lt;/td&gt;
      &lt;td&gt;danke jens&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;17&lt;/td&gt;
      &lt;td&gt;73&lt;/td&gt;
      &lt;td&gt;danke timo&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;18&lt;/td&gt;
      &lt;td&gt;73&lt;/td&gt;
      &lt;td&gt;danke tobias&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;19&lt;/td&gt;
      &lt;td&gt;65&lt;/td&gt;
      &lt;td&gt;danke sebastian&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;20&lt;/td&gt;
      &lt;td&gt;64&lt;/td&gt;
      &lt;td&gt;danke markus&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;21&lt;/td&gt;
      &lt;td&gt;59&lt;/td&gt;
      &lt;td&gt;danke alexander&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;22&lt;/td&gt;
      &lt;td&gt;55&lt;/td&gt;
      &lt;td&gt;danke johannes&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;23&lt;/td&gt;
      &lt;td&gt;55&lt;/td&gt;
      &lt;td&gt;danke jörg&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;24&lt;/td&gt;
      &lt;td&gt;54&lt;/td&gt;
      &lt;td&gt;danke kris&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;25&lt;/td&gt;
      &lt;td&gt;49&lt;/td&gt;
      &lt;td&gt;danke ralf&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;26&lt;/td&gt;
      &lt;td&gt;45&lt;/td&gt;
      &lt;td&gt;danke lutz&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;27&lt;/td&gt;
      &lt;td&gt;45&lt;/td&gt;
      &lt;td&gt;danke sven&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;28&lt;/td&gt;
      &lt;td&gt;43&lt;/td&gt;
      &lt;td&gt;danke julian&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;29&lt;/td&gt;
      &lt;td&gt;42&lt;/td&gt;
      &lt;td&gt;danke christoph&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;30&lt;/td&gt;
      &lt;td&gt;39&lt;/td&gt;
      &lt;td&gt;danke philipp&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;31&lt;/td&gt;
      &lt;td&gt;39&lt;/td&gt;
      &lt;td&gt;danke stephan&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;32&lt;/td&gt;
      &lt;td&gt;37&lt;/td&gt;
      &lt;td&gt;danke gerry&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;33&lt;/td&gt;
      &lt;td&gt;37&lt;/td&gt;
      &lt;td&gt;danke hans&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;34&lt;/td&gt;
      &lt;td&gt;36&lt;/td&gt;
      &lt;td&gt;danke simon&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;35&lt;/td&gt;
      &lt;td&gt;35&lt;/td&gt;
      &lt;td&gt;danke bernd&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;Somebody suggested to put the names in a word cloud, but we hate word clouds. :)&lt;/p&gt;

&lt;p&gt;If you happen to know some of Fefe’s friends and colleagues, you can surely guess who is behind some of the credits pinned to a name (disclosure: we, the authors of this article, have been thanked at least twice, too, if we recall correctly).&lt;/p&gt;

&lt;h2 id=&quot;fun-part-ii-using-markov-chains-to-generate-fefe-texts&quot;&gt;Fun Part II: Using Markov Chains to Generate Fefe Texts&lt;/h2&gt;

&lt;p&gt;And now, this: Let’s generate some pseudo random Fefe text by using Markov chains. For this experiment we quickly forked dellis23’s &lt;a href=&quot;https://gist.github.com/dellis23/6174914&quot;&gt;markov.py script&lt;/a&gt; and adjusted it, the chain size we worked with is 3 and gives us results like this:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;Ziegen fickt”, erklärte drastisch US-Präsident Harry S. Truman einem nachdenklichen Kongressabgeordneten. “Wenn er aus und das Ende des Tages: Die Belegschaft sagt, dass die Behörden so 2009 herum angefangen, missliebige Mitbürgern nach allen Regeln der Kunst auseinandernehmen. Ich weiß, welches T-Shirt ich ab jetzt haben, wo sie den Hartz IV kürzen will, stürmen noch schnell und lautlos miterledigen. Keine weiteren Fragen. Die Mühlen der Full-Disclosure-Fraktion bei Sicherheitslücken: Der Jeep-Hack neulich wurde noch der Presse. Erstens: Nachtsicht- und GPS-Geräte kaufen -&amp;gt; 3 Jahre Haft vorgeschlagen. (Danke, Marcel) Die Russen sind schlauer als der losredete, war das in der Türkei.&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;This is far from being anything worthwhile and the mere result of toying around with this juicy corpus. But alright, let’s generate another one:&lt;/p&gt;

&lt;blockquote&gt;
  &lt;p&gt;Mit Terrorismus hat das mit Cyberangriffen deutlich anders aus, Stichwort Lawful Interception. Wieder was gelernt, diesmal über amerikanische Studenten: Ich habe ehrlich gesagt nicht so aus, dass der vor Gericht erstreiten muss, kann ich nur mit Adblocker nicht zu uns, denn in westlichen Demokratien wie der öffentliche GNU-CVS-Server, furchtbar überlastet ist, und daher muss auch sein Lebensunterhalt direkt davon abhängt, dass er jetzt alles seine Ordnung” um. Denn die Labels daran zugrunde gehen werden. Francesco-Parisi-Universität Styrum — endlich mal was tun diesmal! Sonst lassen die Regierung mit den Nazis damals funktionieren. Zweitens: Wir merken es oft nicht gelöscht würden.&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;The code is quite simple and can sure be optimised in many ways, but for today the Fefe Research Institute pulls the plug.&lt;/p&gt;

&lt;p&gt;&lt;img src=&quot;http://vg07.met.vgwort.de/na/db1f895be9d44f8c8e5d2f9f2dfe8c24&quot; width=&quot;1&quot; height=&quot;1&quot; alt=&quot;&quot; /&gt;&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Robert Jäschke, 
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Distant Reading with Foucault?</title>
    <link href="https://weltliteratur.net/Distant-Reading-with-Foucault/"/>
    <updated>2016-05-06T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/Distant-Reading-with-Foucault</id>
    <content type="html">&lt;p&gt;Half a year ago, November 14, 2015, we participated in a workshop at the Institute of Science and Art in Vienna, entitled &lt;a href=&quot;http://www.iwk.ac.at/events/distant-reading-und-diskursanalyse&quot;&gt;“Distant Reading and Discourse Analysis”&lt;/a&gt;. Our talk, which went under the headline “Distant Reading with Foucault?”, promised to give thoughts on “the practice of distant reading” and ponder potential “operationalisations of Foucauldian discourse analysis”.&lt;/p&gt;

&lt;p&gt;A revised version of the talk was published last week as a featured article in the brilliant &lt;strong&gt;foucaultblog&lt;/strong&gt;, and here comes the caveat, it’s in German: &lt;strong&gt;&lt;a href=&quot;http://doi.org/10.16995/lefou.15&quot;&gt;“Fernlesen mit Foucault?”&lt;/a&gt;&lt;/strong&gt; So, the raison d’être of this blog post is to give you a short summary of the article.&lt;/p&gt;

&lt;p&gt;What we tried to do in the first place was to define what Distant Reading is, or rather: &lt;em&gt;what it was&lt;/em&gt; in the past 15 years since Moretti coined the term in his essay &lt;em&gt;Conjectures on World Literature&lt;/em&gt;. We considered five possible answers. Distant Reading, was it …&lt;/p&gt;

&lt;ul&gt;
  &lt;li&gt;… “a joke”?&lt;/li&gt;
  &lt;li&gt;… a polemical term, a buzzword?&lt;/li&gt;
  &lt;li&gt;… a computer-based method for the analysis of literature?&lt;/li&gt;
  &lt;li&gt;… Moretti’s attempt towards a ‘canon-critical’ large-scale literary historiography?&lt;/li&gt;
  &lt;li&gt;… a failed or at least terminated project?&lt;/li&gt;
&lt;/ul&gt;

&lt;p&gt;All of these possible answers had their grains of truth in them (and are discussed thoroughly &lt;a href=&quot;http://doi.org/10.16995/lefou.15&quot;&gt;in the original German version of our article&lt;/a&gt;). However, what Distant Reading certainly did not provide in the past one and a half decade, was a reliable methodology. Scholars using this term usually thought it enough to reference Moretti. At the same time, one of Moretti’s main epistemic moves seems to have been that of following analogies, as has been suggested by scholars such as &lt;a href=&quot;https://newleftreview.org/II/34/christopher-prendergast-evolution-and-literary-history&quot;&gt;Christopher Prendergast (2005)&lt;/a&gt; or &lt;a href=&quot;http://www.literaturkritik.de/public/rezension.php?rez_id=12719&quot;&gt;Katja Mellmann (2009)&lt;/a&gt;, something Moretti is not unaware of given that he himself, for one, discussed the problematic analogy between World-Systems Theory and a “World-Literary System” (in &lt;em&gt;More Conjectures&lt;/em&gt;).&lt;/p&gt;

&lt;h2 id=&quot;old-and-new-distant-reading&quot;&gt;‘Old’ and ‘New’ Distant Reading&lt;/h2&gt;

&lt;p&gt;When rereading Moretti for this talk we were kind of surprised to find that Distant Reading in its original form has nothing at all to do with the practices of the Digital Humanities. It is hard to find Moretti talking about technological implications, he never mentions standards, protocols, scripts, databases and all the nerve-racking little problems you encounter when trying to squeeze literary data for meaningful findings. There was no freely available corpus or code or documentation that made his theses reproducible.&lt;/p&gt;

&lt;p&gt;In Vienna, while strolling up and down Berg-Gasse and &lt;a href=&quot;https://twitter.com/peertrilcke/status/666702587438243840&quot;&gt;climbing the notorious Strudlhofstiege a.k.a. Stiedlhufstroge&lt;/a&gt;, we were discussing what our quasi-obituary for Old Distant Reading meant for our own research. Because at the very same time we were preparing a data-driven poster for the annual Digital Humanities conference of the German-speaking countries which took place in March, 2016. Our poster set out to be a “Distant-Reading Showcase: 200 Years of German-Language Drama at a Glance” (&lt;a href=&quot;https://doi.org/10.6084/m9.figshare.3101203.v2&quot;&gt;you can find it on figshare&lt;/a&gt;).&lt;/p&gt;

&lt;p&gt;We wrote about the making of the poster &lt;a href=&quot;https://dlina.github.io/Distant-Reading-Showcase-Poster-DHd2016-Leipzig/&quot;&gt;in our DLINA project blog&lt;/a&gt; ([&lt;strong&gt;D&lt;/strong&gt;igital] &lt;strong&gt;Li&lt;/strong&gt;terary &lt;strong&gt;N&lt;/strong&gt;etwork &lt;strong&gt;A&lt;/strong&gt;nalysis), where we also give examples of what you can “distantly read” when looking at this obscure bulk of network graphs. What is important here is that our own project confronted us with the question: What should a renewed Distant Reading be like? While we shouldn’t cling to the term, we can state that Distant Reading (or Macroanalysis, or whatever you call it) should be reproducible, which includes all the above-mentioned aspects: freely available corpus, code, documentation, data, something that can cost months of additional work, but something we regard as essential as the eventual presentation of the results.&lt;/p&gt;

&lt;h2 id=&quot;and-now-foucault&quot;&gt;And Now, Foucault&lt;/h2&gt;

&lt;p&gt;The clarification on what Distant Reading meant or means ate up the better part of our talk, but it was not too late to check in with Foucault. So, could there be a methodologically clean ‘Distant Reading with Foucault’?&lt;/p&gt;

&lt;p&gt;It is obvious that traditional, semantically rich concepts balk – almost programmatically – at their operationalisation in contexts of quantitative, formalised research. This concerns numerous hermeneutic ideas regarding the ‘deeper’ meaning of a text or text element. However, there is no reason why the Digital Humanities shouldn’t utter ideas for possible operationalisations. For the most part, this will result in a genuinely different perception of a subject, or, to paraphrase Moretti: At the end, these terms will probably have their place within completely different theories.&lt;/p&gt;

&lt;p&gt;This becomes clear when looking at “the elementary unit of discourse”, the ‘statement’ (‘énoncé’). Foucault’s definition is perimetric, he repeatedly stresses that a ‘statement’ is something quite different from a string of characters (that could be located in a corpus based on definable rules). Instead, considering his extensive explanations in &lt;em&gt;The Archaeology of Knowledge&lt;/em&gt;, a ‘statement’ is, to a large extent, context-relative, its determination, it seems, is less a positivist than a hermeneutical act. ‘Hermeneutical’ insofar as the discourse analyst has to “at least superficially understand the meaning of statements” to be able to classify them as such (quoting &lt;a href=&quot;https://www.researchgate.net/publication/263320629&quot;&gt;Philipp Sarasin 2014, p. 66&lt;/a&gt;; our translation).&lt;/p&gt;

&lt;p&gt;Therefore, in light of the apparently very low operational potential of Foucault’s discourse analysis, we should ask ourselves whether it wouldn’t perhaps make more sense to go another way. Instead of thinking about how Foucault’s theory design could be operationalised, we could, for starters – following Moretti’s slogan “Forget programs and visions” (&lt;a href=&quot;https://litlab.stanford.edu/LiteraryLabPamphlet6.pdf&quot;&gt;PDF&lt;/a&gt;) – view and discuss the results provided by well-established techniques of text analysis. In fact, decades ago we saw the emergence of a field of discourse analysis relying on corpus-linguistic/lexicometric approaches, just think of the pioneering work on an &lt;em&gt;Analyse automatique du discours&lt;/em&gt; (1969) by &lt;a href=&quot;https://en.wikipedia.org/wiki/Michel_Pêcheux&quot;&gt;Michel Pêcheux&lt;/a&gt;.&lt;/p&gt;

&lt;p&gt;Also Moretti (in a &lt;em&gt;Pamphlet&lt;/em&gt; from his post-Distant-Reading phase, so to speak) simply applied corpus-linguistic methods to describe something like the transformation of economic discourse based on an analysis of the annual reports of the World Bank (&lt;a href=&quot;https://litlab.stanford.edu/LiteraryLabPamphlet9.pdf&quot;&gt;PDF&lt;/a&gt;). Sure, this is not an examination of ‘statements’ in a Foucauldian sense. Instead, Moretti is looking for most frequent words, collocations or the frequency of certain parts of speech and grammatical constructions, in other words, he and his co-author Dominique Pestre are looking for linguistic, not discourse-analytical ‘units’. But maybe, based on this kind of data, a new (and necessarily different) discourse analysis could be established, a revamped implementation of Foucault’s project, which could – unlike most corpus-linguistic approaches – continue what appears to us the key signature of Foucault’s “work on the discourses”: research as a practice of a critical science. ▣&lt;/p&gt;

&lt;h2 id=&quot;concluding-remarks&quot;&gt;Concluding Remarks&lt;/h2&gt;

&lt;p&gt;Given the workshop character of the event, our talk was little more than a first tentative approach to the question raised in the title. Sure enough, the workshop did spur some nice discussions that will certainly be continued. All talks of the workshop will soon be part of a separate &lt;a href=&quot;https://foucaldien.net/2/volume/2/issue/1/&quot;&gt;‘issue’&lt;/a&gt; on the &lt;strong&gt;foucaultblog&lt;/strong&gt;, edited by Simon Ganahl and Maurice Erb.&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer, 
	
          
          Peer Trilcke
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>Wikidata Meets World Literature</title>
    <link href="https://weltliteratur.net/Wikidata-Meets-World-Literature/"/>
    <updated>2016-04-25T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/Wikidata-Meets-World-Literature</id>
    <content type="html">&lt;p&gt;You can currently access and contribute to &lt;a href=&quot;https://meta.wikimedia.org/wiki/List_of_Wikipedias&quot;&gt;281 active Wikipedia language versions&lt;/a&gt;. What follows is a top-25 list of ‘books’ ranked by the number of articles dedicated to these ‘books’ in different Wikipedia language editions (results first, explanation thereafter):&lt;/p&gt;

&lt;table&gt;
  &lt;thead&gt;
    &lt;tr&gt;
      &lt;th&gt;Wikidata item&lt;/th&gt;
      &lt;th&gt;Title&lt;/th&gt;
      &lt;th&gt;Authorlabel&lt;/th&gt;
      &lt;th&gt;Linkcount&lt;/th&gt;
    &lt;/tr&gt;
  &lt;/thead&gt;
  &lt;tbody&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q8258&quot;&gt;Q8258&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;One Thousand and One Nights&lt;/td&gt;
      &lt;td&gt;anonymous&lt;/td&gt;
      &lt;td&gt;116&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q15228&quot;&gt;Q15228&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;The Lord of the Rings&lt;/td&gt;
      &lt;td&gt;J. R. R. Tolkien&lt;/td&gt;
      &lt;td&gt;115&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q480&quot;&gt;Q480&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Don Quixote&lt;/td&gt;
      &lt;td&gt;Miguel de Cervantes&lt;/td&gt;
      &lt;td&gt;113&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q25338&quot;&gt;Q25338&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Le Petit Prince&lt;/td&gt;
      &lt;td&gt;Antoine de Saint-Exupéry&lt;/td&gt;
      &lt;td&gt;110&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q40591&quot;&gt;Q40591&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;The Communist Manifesto&lt;/td&gt;
      &lt;td&gt;Karl Marx/Friedrich Engels&lt;/td&gt;
      &lt;td&gt;108&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q459842&quot;&gt;Q459842&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Book of Mormon&lt;/td&gt;
      &lt;td&gt;Joseph Smith&lt;/td&gt;
      &lt;td&gt;106&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q208460&quot;&gt;Q208460&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Nineteen Eighty-Four&lt;/td&gt;
      &lt;td&gt;George Orwell&lt;/td&gt;
      &lt;td&gt;102&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q8251&quot;&gt;Q8251&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;The Art of War&lt;/td&gt;
      &lt;td&gt;Sun Tzu&lt;/td&gt;
      &lt;td&gt;95&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q92640&quot;&gt;Q92640&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Alice’s Adventures in Wonderland&lt;/td&gt;
      &lt;td&gt;Lewis Carroll&lt;/td&gt;
      &lt;td&gt;93&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q8279&quot;&gt;Q8279&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Shahnameh&lt;/td&gt;
      &lt;td&gt;Ferdowsi&lt;/td&gt;
      &lt;td&gt;93&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q74287&quot;&gt;Q74287&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;The Hobbit&lt;/td&gt;
      &lt;td&gt;J. R. R. Tolkien&lt;/td&gt;
      &lt;td&gt;91&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q43361&quot;&gt;Q43361&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Harry Potter and the Philosopher’s Stone&lt;/td&gt;
      &lt;td&gt;J. K. Rowling&lt;/td&gt;
      &lt;td&gt;86&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q1396889&quot;&gt;Q1396889&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Animal Farm&lt;/td&gt;
      &lt;td&gt;George Orwell&lt;/td&gt;
      &lt;td&gt;85&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q8265&quot;&gt;Q8265&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Dream of the Red Chamber&lt;/td&gt;
      &lt;td&gt;Cao Xueqin&lt;/td&gt;
      &lt;td&gt;82&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q41675&quot;&gt;Q41675&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Guinness World Records&lt;/td&gt;
      &lt;td&gt;Craig Glenday&lt;/td&gt;
      &lt;td&gt;80&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q48244&quot;&gt;Q48244&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Mein Kampf&lt;/td&gt;
      &lt;td&gt;Adolf Hitler&lt;/td&gt;
      &lt;td&gt;80&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q8269&quot;&gt;Q8269&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;The Tale of Genji&lt;/td&gt;
      &lt;td&gt;Murasaki Shikibu&lt;/td&gt;
      &lt;td&gt;79&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q165318&quot;&gt;Q165318&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Crime and Punishment&lt;/td&gt;
      &lt;td&gt;Fyodor Dostoyevsky&lt;/td&gt;
      &lt;td&gt;77&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q47209&quot;&gt;Q47209&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Harry Potter and the Chamber of Secrets&lt;/td&gt;
      &lt;td&gt;J. K. Rowling&lt;/td&gt;
      &lt;td&gt;77&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q47598&quot;&gt;Q47598&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Harry Potter and the Prisoner of Azkaban&lt;/td&gt;
      &lt;td&gt;J. K. Rowling&lt;/td&gt;
      &lt;td&gt;75&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q46887&quot;&gt;Q46887&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Harry Potter and the Half-Blood Prince&lt;/td&gt;
      &lt;td&gt;J. K. Rowling&lt;/td&gt;
      &lt;td&gt;75&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q80817&quot;&gt;Q80817&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Harry Potter and the Order of the Phoenix&lt;/td&gt;
      &lt;td&gt;J. K. Rowling&lt;/td&gt;
      &lt;td&gt;75&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q123397&quot;&gt;Q123397&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;The Republic&lt;/td&gt;
      &lt;td&gt;Plato&lt;/td&gt;
      &lt;td&gt;74&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q170583&quot;&gt;Q170583&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Pride and Prejudice&lt;/td&gt;
      &lt;td&gt;Jane Austen&lt;/td&gt;
      &lt;td&gt;74&lt;/td&gt;
    &lt;/tr&gt;
    &lt;tr&gt;
      &lt;td&gt;&lt;a href=&quot;http://www.wikidata.org/entity/Q147787&quot;&gt;Q147787&lt;/a&gt;&lt;/td&gt;
      &lt;td&gt;Anna Karenina&lt;/td&gt;
      &lt;td&gt;Leo Tolstoy&lt;/td&gt;
      &lt;td&gt;73&lt;/td&gt;
    &lt;/tr&gt;
  &lt;/tbody&gt;
&lt;/table&gt;

&lt;p&gt;Sports commentator proclaims: Legendary Arabian story collection first, the Lord of the Rings a close runner-up, Spanish hack Cervantes completing the podium! But let’s curb our enthusiasm for this quirky race for a moment. You cannot be cautious enough when interpreting this kind of lists/rankings. There are a hundred things you’d have to involve when doing this.&lt;/p&gt;

&lt;p&gt;Our ranking reveals nothing on the nature or quality of, let’s say, these 110 language versions of “Le Petit Prince”. We could deal with a one-sentence article in 50 Wikipedias for all we know. Understanding the genesis of Wikipedia and measuring its contents is a research field of its own. And Wikidata, although building on Wikipedia content, of course, is yet another thing whose characteristics you’d have to take into account before jumping to any rash conclusions.&lt;/p&gt;

&lt;p&gt;Let’s start slowly. What this ranking suggests to show is a list of 25 ‘books’ (we’ll come back to this term later) that are of whatsoever importance across many countries/cultures/language communities. It is not at all a proper basis to redefine the term ‘world literature’ for the digital age. But it does include the idea of ‘world’ as part of ‘world literature’, meaning ‘books’ that transcend national and language borders, also if their literary value may vary considerably. In any case, before starting to interpret anything, you’ll always have to communicate how exactly you generated this kind of results. So …&lt;/p&gt;

&lt;p&gt;… as an example of how the Wikidata Query Service can facilitate this type of quantitative analysis, we wrote a SPARQL query that orders ‘instances of’ ‘books’ with an author by ‘site link’. This means that if a work of literature has the statement ‘instance of book’ in Wikidata, and it has a Wikipedia page, it can be ranked by the number of Wikipedias that have a page for it.&lt;/p&gt;

&lt;p&gt;This is easily reproducible for you at your own machines, just head over to &lt;a href=&quot;http://query.wikidata.org&quot;&gt;http://query.wikidata.org&lt;/a&gt; and copy &amp;amp; paste this in the upper box, then hit “Execute” (and while the query is being processed, throw a glance at Alan Liu’s &lt;a href=&quot;http://dhdebates.gc.cuny.edu/debates/text/20&quot;&gt;criticism of the concentration “on pushing the ‘execute’ button” in the Digital Humanities&lt;/a&gt;):&lt;/p&gt;

&lt;div class=&quot;language-sql highlighter-rouge&quot;&gt;&lt;div class=&quot;highlight&quot;&gt;&lt;pre class=&quot;highlight&quot;&gt;&lt;code&gt;&lt;span class=&quot;k&quot;&gt;prefix&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;schema&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;&amp;lt;&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;http&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;//&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;schema&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;org&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;/&amp;gt;&lt;/span&gt;
&lt;span class=&quot;k&quot;&gt;prefix&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;wd&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;&amp;lt;&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;http&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;//&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;www&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;wikidata&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;org&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;/&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;entity&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;/&amp;gt;&lt;/span&gt;
&lt;span class=&quot;k&quot;&gt;prefix&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;wdt&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;&amp;lt;&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;http&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;//&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;www&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;wikidata&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;org&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;/&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;prop&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;/&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;direct&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;/&amp;gt;&lt;/span&gt;

&lt;span class=&quot;k&quot;&gt;SELECT&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;desc&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;authorlabel&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;COUNT&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;DISTINCT&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;sitelink&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;as&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;linkcount&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt;
&lt;span class=&quot;k&quot;&gt;WHERE&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;{&lt;/span&gt;
  &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;wdt&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;P31&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;wd&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;Q571&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;
  &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;sitelink&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;schema&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;about&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;.&lt;/span&gt;
  &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;wdt&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;P50&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;author&lt;/span&gt;
  &lt;span class=&quot;n&quot;&gt;OPTIONAL&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;{&lt;/span&gt;
    &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;rdfs&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;label&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;desc&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;filter&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;lang&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;desc&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;=&lt;/span&gt; &lt;span class=&quot;nv&quot;&gt;&quot;en&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;).&lt;/span&gt;
  &lt;span class=&quot;p&quot;&gt;}&lt;/span&gt;
  &lt;span class=&quot;n&quot;&gt;OPTIONAL&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;{&lt;/span&gt;
    &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;author&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;rdfs&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;:&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;label&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;authorlabel&lt;/span&gt; &lt;span class=&quot;n&quot;&gt;filter&lt;/span&gt; &lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;lang&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;authorlabel&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;=&lt;/span&gt; &lt;span class=&quot;nv&quot;&gt;&quot;en&quot;&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;).&lt;/span&gt;
  &lt;span class=&quot;p&quot;&gt;}&lt;/span&gt;
&lt;span class=&quot;p&quot;&gt;}&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;GROUP&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;BY&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;s&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;k&quot;&gt;desc&lt;/span&gt; &lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;authorlabel&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;ORDER&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;BY&lt;/span&gt; &lt;span class=&quot;k&quot;&gt;DESC&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;(&lt;/span&gt;&lt;span class=&quot;o&quot;&gt;?&lt;/span&gt;&lt;span class=&quot;n&quot;&gt;linkcount&lt;/span&gt;&lt;span class=&quot;p&quot;&gt;)&lt;/span&gt;
&lt;/code&gt;&lt;/pre&gt;&lt;/div&gt;&lt;/div&gt;

&lt;p&gt;Did it work for you?&lt;/p&gt;

&lt;p&gt;By the way, changing the object criteria from ‘book’ to ‘literary work’ (wd:Q7725634) returns a different result, that includes books of the Bible (that are not classified as ‘instance of book’). Running it without the author, also yields a slightly different result, as there are a few works without an author that are very widely translated (like, for example, today’s winner, “One Thousand and One Nights”).&lt;/p&gt;

&lt;p&gt;The query that yielded the above ranking was run today (April 25, 2016) at 08:25 CEST. If you save the results of your query (&lt;a href=&quot;https://github.com/weltliteratur/blog/blob/gh-pages/data/2016-04-25_08&apos;25_CEST_wikidata_query_service_result.csv&quot;&gt;like we did, in CSV format&lt;/a&gt;, which can be done directly in Wikidata’s Query Service) and compare it to the same query conducted at a later date, you would have a useful rudiment for the analysis of how certain works gain or lose importance in the Wikipedian universe. Something we might come back to later. Until then, happy SPARQL’ing!&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Christopher Johnson, 
	
          
          Frank Fischer
	
      </name>
    </author>
  </entry>
  
  <entry>
    <title>This is weltliteratur.net</title>
    <link href="https://weltliteratur.net/This-is-weltliteratur-net/"/>
    <updated>2016-04-23T00:00:00+00:00</updated>
    <id>https://weltliteratur.net/This-is-weltliteratur-net</id>
    <content type="html">&lt;blockquote&gt;
  &lt;p&gt;“Welcome! Welcome! Welcome!”&lt;br /&gt;
– &lt;em&gt;John Oliver&lt;/em&gt;&lt;/p&gt;
&lt;/blockquote&gt;

&lt;p&gt;Now this is, believe it or not, the sky over Göttingen #OnThisDay, April 23, 2016, after just some minimal photoshopping (turning light grey into summery light blue):&lt;/p&gt;

&lt;figure&gt;
  &lt;img src=&quot;/images/sky_over_goettingen.jpg&quot; alt=&quot;Sky over Göttingen.&quot; style=&quot;width:1040px;&quot; /&gt;
&lt;/figure&gt;

&lt;p&gt;Göttingen (and Hanover) were the places where we first started to think about a blog called &lt;strong&gt;weltliteratur.net&lt;/strong&gt;, a neat and plushy think tank where we could publish some of our musings on Digital Humanities-related things that might or might not be or become part of a bigger research project.&lt;/p&gt;

&lt;p&gt;Now that some of us are scattered to the four winds (to Potsdam, Sheffield, Moscow, etc.) we’re planning to use this site as a place to continue discussing and exploring ideas emanating from our daily DH’ing. Let’s see how it goes.&lt;/p&gt;

&lt;p&gt;Today is World Book Day, thanks to Shakespeare and Cervantes, who both died exactly 400 years ago, well, Shakespeare some days later, thanks to the Julian calendar. Either way, a suitable day to start off with something called &lt;strong&gt;weltliteratur.net&lt;/strong&gt;, which also alludes to Goethe’s notion that “the epoch of world literature is at hand”, as uttered by him in 1827.&lt;/p&gt;

&lt;p&gt;Our first non-meta article will be delivered on Monday, something about Wikidata meeting World Literature, &lt;a href=&quot;/Wikidata-Meets-World-Literature/&quot;&gt;a small blogpost on how you can try to grasp what this fuzzy corpus called world literature is&lt;/a&gt;, according to Wikipedia/Wikidata.&lt;/p&gt;

&lt;p&gt;Oh, one other USP of &lt;strong&gt;weltliteratur.net&lt;/strong&gt; is the very unofficial &lt;strong&gt;&lt;a href=&quot;/dh-mixtape-2016/&quot;&gt;DH Soundtrack&lt;/a&gt;&lt;/strong&gt;, a growing list of current songs we listen to before, after and, of course: &lt;em&gt;while&lt;/em&gt; conducting our research. So if some of our entries seem to make no sense to you, try switching on our official &lt;em&gt;&lt;a href=&quot;https://de.wikipedia.org/wiki/Musikbett&quot;&gt;musikbett&lt;/a&gt;&lt;/em&gt; to ponder a supposed deeper meaning.&lt;/p&gt;

&lt;p&gt;This, by and large, is it.&lt;/p&gt;

&lt;p&gt;On behalf of everybody else,&lt;br /&gt;
Frank – Mathias – Robert&lt;/p&gt;
</content>
    <author>
      <name>
	
          
          Frank Fischer, 
	
          
          Mathias Göbel, 
	
          
          Robert Jäschke
	
      </name>
    </author>
  </entry>
  

</feed>
