<?xml version="1.0" encoding="utf-8"?><feed xmlns="http://www.w3.org/2005/Atom" xml:lang="en"><generator uri="https://jekyllrb.com/" version="4.4.1">Jekyll</generator><link href="https://tillgrallert.github.io/feed.xml" rel="self" type="application/atom+xml"/><link href="https://tillgrallert.github.io/" rel="alternate" type="text/html" hreflang="en"/><updated>2026-01-07T16:51:06+00:00</updated><id>https://tillgrallert.github.io/feed.xml</id><title type="html">Till Grallert</title><subtitle>A simple, whitespace theme for academics. Based on [*folio](https://github.com/bogoli/-folio) design. </subtitle><entry><title type="html">Integrating library data into authority file</title><link href="https://tillgrallert.github.io/blog/2022/02/16/integrating-library-data-into-authority-file/" rel="alternate" type="text/html" title="Integrating library data into authority file"/><published>2022-02-16T22:00:00+00:00</published><updated>2022-02-16T22:00:00+00:00</updated><id>https://tillgrallert.github.io/blog/2022/02/16/integrating-library-data-into-authority-file</id><content type="html" xml:base="https://tillgrallert.github.io/blog/2022/02/16/integrating-library-data-into-authority-file/"><![CDATA[]]></content><author><name></name></author><summary type="html"><![CDATA[The challenges of MARC XML and inconsistent transcription practices - A recent Twitter post from my former colleague Anne Klammt made me aware of a recent relaunch of “Zeitschriftendatenbank” (ZDB), the portal for periodical holdings in German (and Austrian) libraries. Just played around with the German Union Catalogue of Serails (ZDB) \@DNB_Aktuelles Nice graphs, maps and data. #journals #linkeddata #authorityfiles pic.twitter.com/f0qVtAWA5d&mdash; Dr. Anne Klammt (\@archaeoklammt) November 30, 2021 Part of the German National Library (Deutsche Nationalbibliothek, DNB), the website looks great and provides a lot of data-driven functionality, such as maps and timelines of holdings. The display language of the website itself, though not the bibliographic data, can be toggled between German and English. This is a welcome nod to international users and will certainly increase the visibility of this important portal. However, it must be noted that unfortunately the dataset of bibliographic data is not as accessible as the interface. Languages written in scripts other than Latin are provided in a variety of inconsistent transcriptions into Latin script for mostly historical technical reasons. This is not the fault of ZDB per se but it will prevent communities from the Global South from finding and accessing their own cultural heritage, which for various reasons are held by institutions in the Global North. This is especially relevant for Arabic material, as I will elaborate in the section on transliterations below. On the upside, however, ZDB supports historical political entities, such as the Ottoman Empire, for facetted browsing, which I have not yet seen in other library catalogues. Crucially, ZDB also provides...]]></summary></entry><entry><title type="html">Mapping Arabic periodical titles between 1799 and 1929</title><link href="https://tillgrallert.github.io/blog/2021/04/28/mapping-arabic-periodical-titles-between-1799-and-1929/" rel="alternate" type="text/html" title="Mapping Arabic periodical titles between 1799 and 1929"/><published>2021-04-28T21:00:00+00:00</published><updated>2021-04-28T21:00:00+00:00</updated><id>https://tillgrallert.github.io/blog/2021/04/28/mapping-arabic-periodical-titles-between-1799-and-1929</id><content type="html" xml:base="https://tillgrallert.github.io/blog/2021/04/28/mapping-arabic-periodical-titles-between-1799-and-1929/"><![CDATA[]]></content><author><name></name></author><summary type="html"><![CDATA[A tutorial for mapping with R - To celebrate this year’s Day of DH (#DayOfDH20201), I want to share this draft tutorial for mapping multilingual bibliographic data sets with R. NOTES This is a draft and I appreciate any comments, questions etc. using hypothes.is. This website is generated using the default GitHub pages Jekyll workflow, which does not support a citeproc plugin. The few references are, therefore, currently not formatted. The resulting maps have been published in their own repository. Introduction Large bibliographic dataset such as the one compiled by Project Jarāʾid (Mestyan, Grallert, and et al. 2020), which comprises information on more than 3300 periodical titles, are far too large and unwieldy for manual analysis. One wants to use descriptive statistics and visualisations for a first exploration of the data set, which can then guide further scrutiny and create new research questions. Basic information for each periodical in the Project Jarāʾid data set (with very few exceptions) include the date of the first publication as well as a publication location and additional languages beyond Arabic for bi- or even multilingual periodicals. Such a data set lends itself to mapping in order to see the distribution of the number of periodical titles per location for a certain period and region. In order to visualise change over time, one will want to animate the map in some way, just as the ones below: These GIFs combine multiple maps into a single animation and show the geographic distribution of new periodical titles in rolling periods of different length (year...]]></summary></entry><entry><title type="html">Annual report: OpenArabicPE in 2019</title><link href="https://tillgrallert.github.io/blog/2020/01/07/annual-report-openarabicpe-in-2019/" rel="alternate" type="text/html" title="Annual report: OpenArabicPE in 2019"/><published>2020-01-07T22:00:00+00:00</published><updated>2020-01-07T22:00:00+00:00</updated><id>https://tillgrallert.github.io/blog/2020/01/07/annual-report-openarabicpe-in-2019</id><content type="html" xml:base="https://tillgrallert.github.io/blog/2020/01/07/annual-report-openarabicpe-in-2019/"><![CDATA[]]></content><author><name></name></author><summary type="html"><![CDATA[The work on digital scholarly editions (DSE) of late Ottoman Arabic periodicals continued within the framework of OpenArabicPE (see annual report 2017). I added a number of periodicals to the existing editions of Muḥammad Kurd ʿAlī’s al-Muqtabas (Damascus and Cairo, 1906–18) and ʿAbd al-Qādir al-Iskandarānī’s al-Ḥaqāʾiq (Damascus, 1910–12) following the principles and workflows established over the last years. These are: Anastās Mārī al-Karmalī’s monthly journal Lughat al-ʿArab (Baghdad, 1911–14), Anṭūn al-Jumayyil’s monthly journal al-Zuhūr (Cairo, 1910–1913) and Abd Allāh Nadīm al-Idrīsī’s weekly journal al-Ustādh (Cairo, 1892–1893). The ability to quickly add and release a number of periodicals with full text and digital facsimiles was helped by the anonymous transcribers at al-Maktaba al-Shāmila, who reproduced the page breaks as found in the printed originals that allow us to quickly link the text to the facsimile. This is very different from both al-Muqtabas and al-Ḥaqāʾiq for which we had to add each of the 8000+ page breaks manually. Finally, I worked on a facsimile edition with transcriptions of article titles and bylines of Jirjī Niqūlā Bāz’s al-Ḥasnāʾ (Beirut, 1909–11). In the last report, I mentioned the importance of authorship attribution for the vast majority of anonymous articles if one wants to analyse the (social) network of authors and texts that form the ideosphere of the Arabic press in the Eastern Mediterranean. The, often implicit and accepted, hypothesis is that a periodical’s editors authored all articles for which they did not provide a meaningful byline. There are two issues with this hypothesis: first,...]]></summary></entry><entry><title type="html">Annual report: OpenArabicPE in 2018</title><link href="https://tillgrallert.github.io/blog/2019/01/14/annual-report-openarabicpe-in-2018/" rel="alternate" type="text/html" title="Annual report: OpenArabicPE in 2018"/><published>2019-01-14T22:00:00+00:00</published><updated>2019-01-14T22:00:00+00:00</updated><id>https://tillgrallert.github.io/blog/2019/01/14/annual-report-openarabicpe-in-2018</id><content type="html" xml:base="https://tillgrallert.github.io/blog/2019/01/14/annual-report-openarabicpe-in-2018/"><![CDATA[]]></content><author><name></name></author><summary type="html"><![CDATA[The work on the digital editions of Muḥammad Kurd ʿAlī’s journal al-Muqtabas (Damascus and Cairo, 1906–18) and ʿAbd al-Qādir al-Iskandarānī’s journal al-Ḥaqāʾiq (Damascus, 1910–12) continued within the framework of OpenArabicPE. Contributions from our interns Manzi Tanna-Händel, Xaver Kretzschmar, Klara Mayer, Tobias Sick and Hans Magne Jaatun allowed us to release a further four volumes. Readers can find a project description in the last annual report and I will focus here on the first foray into the analysis of our corpus. One of the driving research questions behind OpenArabicPE focusses on reconstructing the ideosphere of the late Ottoman and early Arabic press through establishing networks of authors and texts published and referenced in the periodicals. With regards to al-Muqtabas and al-Ḥaqāʾiq, we ask: who published what in late Ottoman Damascus? As well as, what was read in late Ottoman Damascus? Any sort of meaningful computational analysis of the global connections between authors, texts, and periodicals as a venue for publication and review requires access to reliable standardised bibliographic metadata as a bare minimum. Unfortunately, such data is practically non-existant on the article level beyond our OpenArabicPE corpus. The situation is only marginally better on the issue level. Available metadata is not commonly provided in a standard-compliant and machine-actionable format. But even then, the vast majority of articles would remain outside our analytical scopes. Many publishers did not provide (meaningful) bylines and the majority articles in journals and newspapers from Beirut, Cairo or Damascus did not credit their authors (c.f. Table). One...]]></summary></entry><entry><title type="html">v0.5 of (*majallat*) *al-Muqtabas*: volume 3</title><link href="https://tillgrallert.github.io/blog/2018/12/16/v05-of-majallat-al-muqtabas-volume-3/" rel="alternate" type="text/html" title="v0.5 of (*majallat*) *al-Muqtabas*: volume 3"/><published>2018-12-16T22:00:00+00:00</published><updated>2018-12-16T22:00:00+00:00</updated><id>https://tillgrallert.github.io/blog/2018/12/16/v05-of-majallat-al-muqtabas-volume-3</id><content type="html" xml:base="https://tillgrallert.github.io/blog/2018/12/16/v05-of-majallat-al-muqtabas-volume-3/"><![CDATA[]]></content><author><name></name></author><summary type="html"><![CDATA[We just released v0.5 of the journal al-Muqtabas. This release includes structural mark-up and page breaks based on the digital facsimiles for all issues of volume 3. As always, this release comprises all previously released files. This time, however, we have also included various improvements to TEI files and the bibliographic metadata, namely links from named entities, such as persons and places, to authority files (VIAF and GeoNames). vol. 1: complete vol. 3: complete vol. 4: complete vol. 5: complete vol. 6: complete I wish to express my gratitude to the following contributors, who helped with the mark-up of page breaks: Manzi Tanna-Händel: volume 1, issues 10 to 12. volume 3, issues 8 to 12. Layla Youssef: volume 4, issues 7 to 12. Dimitar Dragnev: volume 5, issues 2 to 12.]]></summary></entry><entry><title type="html">v0.4 of (*majallat*) *al-Muqtabas*: volume 1</title><link href="https://tillgrallert.github.io/blog/2018/09/27/v04-of-majallat-al-muqtabas-volume-1/" rel="alternate" type="text/html" title="v0.4 of (*majallat*) *al-Muqtabas*: volume 1"/><published>2018-09-27T21:00:00+00:00</published><updated>2018-09-27T21:00:00+00:00</updated><id>https://tillgrallert.github.io/blog/2018/09/27/v04-of-majallat-al-muqtabas-volume-1</id><content type="html" xml:base="https://tillgrallert.github.io/blog/2018/09/27/v04-of-majallat-al-muqtabas-volume-1/"><![CDATA[]]></content><author><name></name></author><summary type="html"><![CDATA[We just released v0.4 of the journal al-Muqtabas. This release includes structural mark-up and page breaks based on the digital facsimiles of the first edition (sic!)1 for all issues of volume 1. In addition, this release comprises all previously released files. vol. 1: complete vol. 4: complete vol. 5: complete vol. 6: complete I wish to express my gratitude to the following contributors, who helped with the mark-up of page breaks: Manzi Tanna-Händel: volume 1, issues 10 to 12. Layla Youssef: volume 4, issues 7 to 12. Dimitar Dragnev: volume 5, issues 2 to 12. The TEI boilerplate based webview will default to facsimiles of the second edition. Page breaks occasionally differ between the two editions. &#8617;]]></summary></entry><entry><title type="html">v0.3 of (*majallat*) *al-Muqtabas*: volume 4</title><link href="https://tillgrallert.github.io/blog/2018/05/08/v03-of-majallat-al-muqtabas-volume-4/" rel="alternate" type="text/html" title="v0.3 of (*majallat*) *al-Muqtabas*: volume 4"/><published>2018-05-08T21:00:00+00:00</published><updated>2018-05-08T21:00:00+00:00</updated><id>https://tillgrallert.github.io/blog/2018/05/08/v03-of-majallat-al-muqtabas-volume-4</id><content type="html" xml:base="https://tillgrallert.github.io/blog/2018/05/08/v03-of-majallat-al-muqtabas-volume-4/"><![CDATA[]]></content><author><name></name></author><summary type="html"><![CDATA[After a long gap, that saw me return to Beirut after a prolonged parental leave, we just released v0.3 of the journal al-Muqtabas. This release includes structural mark-up and page breaks based on the digital facsimiles for all issues of volume 4. In addition, this release comprises all previously released files. vol. 4: complete vol. 5: complete vol. 6: complete I wish to express my gratitude to the following contributors, who helped with the mark-up of page breaks: Dimitar Dragnev: volume 5, issues 2 to 12. Layla Youssef: volume 4, issues 7 to 12.]]></summary></entry><entry><title type="html">v0.1 of *al-Ḥaqāʾiq*: volume 1</title><link href="https://tillgrallert.github.io/blog/2018/04/26/v01-of-al-aqiq-volume-1/" rel="alternate" type="text/html" title="v0.1 of *al-Ḥaqāʾiq*: volume 1"/><published>2018-04-26T21:00:00+00:00</published><updated>2018-04-26T21:00:00+00:00</updated><id>https://tillgrallert.github.io/blog/2018/04/26/v01-of-al-aqiq-volume-1</id><content type="html" xml:base="https://tillgrallert.github.io/blog/2018/04/26/v01-of-al-aqiq-volume-1/"><![CDATA[]]></content><author><name></name></author><summary type="html"><![CDATA[We just released v0.1 of the journal al-Ḥaqāʾiq published by ʿAbd al-Qādir al-Iskandarānī. As per our release schedule this release includes the following: TEI files of all twelve issues of the first volume of al-Ḥaqāʾiq with tructural mark-up of mastheads, sections, articles, and with page breaks linked to the facsimiles; MODS and BibTeX files for all issues, sections, and articles; A local installation of the TEI Boilerplate for Arabic editions; I wish to express my gratitude to the following contributors, who helped with the mark-up of page breaks: Talha Güzel: volume 1, issues 8 and 10. Xaver Kretzschmar: volume 1, issue 11. Thanks to the efforts of CERN’s Zenodo platform we even got a DOI!]]></summary></entry><entry><title type="html">Introduction to Plain Text Workflows and Sustainable Publishing</title><link href="https://tillgrallert.github.io/blog/2017/02/22/plain-text/" rel="alternate" type="text/html" title="Introduction to Plain Text Workflows and Sustainable Publishing"/><published>2017-02-22T16:13:32+00:00</published><updated>2017-02-22T16:13:32+00:00</updated><id>https://tillgrallert.github.io/blog/2017/02/22/plain-text</id><content type="html" xml:base="https://tillgrallert.github.io/blog/2017/02/22/plain-text/"><![CDATA[<p><em>DRAFT</em></p> <p>Disclaimer 1: many of the ideas have been inspired by the work of Alex Gil (@<a href="https://twitter.com/elotroalex">elotroalex</a>) and Denis Tenen (@<a href="https://twitter.com//dennistenen">dennistenen</a>)<sup id="fnref:1"><a href="#fn:1" class="footnote" rel="footnote" role="doc-noteref">1</a></sup> and discussions with the two of them and others at DHI Beirut, THATCamp Beirut, and DHSI.</p> <p>Disclaimer 2: I am teaching short workshops on some of the ideas outlined in this post at <a href="https://dhibeirut.wordpress.com/">Digital Humanities Institute - Beirut</a> on 10 March 2017 and at <a href="https://wp.nyu.edu/dhad/">DH Abu Dhabi</a> on 10 April 2017. Basic slides are available <a href="https://tillgrallert.github.io/slides/2017-dhi-beirut/">here</a> and <a href="https://tillgrallert.github.io/slides/2017-dhad/">here</a>.</p> <h1 id="intro">Intro</h1> <p>In the world of (academic) publishing, large aggregators and indexers have turned into and acquired publishing presses and generate obscene profits by charging the public (every tax-payer worldwide) multiple times over. First by charging the predominantly publicly-funded academic for publishing the results of her publicly-funded research and by enforcing a culture of <em>pro bono</em> labour among academic reviewers and editors; second by selling this content to equally predominantly publicly-funded libraries, which then increasingly demand access fees from members of the public, who want to access their collections; and third by offloading the cost of long-term preservation to, again, publicly-funded institutions. This system not only created a hierarchy of academics and institutions in the relatively well-off “West”—two classes divided by their ability to pay for being published and accessing publications (their own and others). It also increasingly prevents anybody outside western academia from accessing cutting-edge research and participating in intellectual discourse.</p> <p>One of the opportunities afforded by the <em>digital humanities</em> and the stated goal of this endeavour is to remove the middlemen—be they technical or entrepreneurial—between authors, readers, and the library-cum-archive. There are two main obstacles to this aim:</p> <ol> <li>copyright laws and</li> <li>the alienation of us, authors and academics, from the means of production.</li> </ol> <p>We will not be able to change copyright legislation and the vested business interests in sustaining and expanding the regime of profit-generating copy and distribution rights in the foreseeable future, but we can all provide our knowledge under a <a href="https://creativecommons.org/">creative commons licence</a>. In order to do so, we need to re-claim the means of production. We argue that by doing so the main argument for restrictive copyright—namely, the provision of allegedly expensive services such as quality control and metadata curation—collapses. By the time of writing, printing costs and the global distribution of heavy and voluminous books are already negligible as the main avenue of scholarly publication, the journal, has already moved to digital online publication.</p> <h1 id="principles">Principles</h1> <p>The main principles in our effort to (re)claim the means of production are: accessibility, simplicity, sustainability, and credibility. They shall pertain both to the intellectual endeavour and to the tools employed.</p> <ol> <li><em>Accessibility</em>. Accessible means first of all free and open to use and re-purpose for all that have the technical ability to do. Therefore we will have to forfeit all proprietary software and formats. From the imperative of openness and accessibility derive the second and third principles:</li> <li><em>Simplicity</em>: apart from the necessity to be able read and write in a specific human language, there should be as few tools, hardware and software requirements as possible. Ideally everything should work on ten-year old hardware and nothing but the software packaged with the operating system—or even better: the core technologies should work without a computer.</li> <li><em>Sustainability</em>: only simple systems adhering to widely accepted standards can be sustained with minimal / reasonable effort.</li> <li><em>Credibility</em>: Credibility is at the core of scholarly production. In addition to transparency as to the sources, methodology, and tools used, authorship needs to be ascertained and acknowledged as the main tool of scholarly quality control.</li> </ol> <h1 id="tools--formats">Tools / formats</h1> <h2 id="1-writing">1. Writing</h2> <p>Looking at the most broadly-employed software in academic contexts and beyond, Microsoft’s Word, a piece of bloated and expensive proprietary software, what are the functions we need to replace?</p> <h3 id="1-composition--authoring">1. composition / authoring</h3> <p>To avoid overtly complex software and proprietary formats, form and content must be separated. Structural / semantic and representational information as well as metadata must be embedded in the text itself in order to be inseparable. We suggest using plain text with rudimentary markup following the conventions of MarkDown and a short block of metadata written in Yaml as the format of choice.</p> <ul> <li><em>Format</em>: plain text. At their core, all files are simple strings of letters. In the case of plain text files (<code class="language-plaintext highlighter-rouge">TXT</code>), this string of letters happens to be human readable. Plain text has been with us since the early days of computing and <code class="language-plaintext highlighter-rouge">TXT</code> files could be viewed and edited with 1980s hard- and software. We can therefore assume that this basic format will remain accessible for the years to come.</li> <li><em>Syntax</em>: <a href="http://daringfireball.net/projects/markdown/syntax">MarkDown</a> and its derivatives provide a simple but formal way of providing structural (headings, lists, notes, block quotes etc.) and formating (italics, bold) information that is both human and machine readable. Think of it as the old way of underlining a line of text on a type-writer to indicate a heading.<sup id="fnref:4"><a href="#fn:4" class="footnote" rel="footnote" role="doc-noteref">2</a></sup></li> <li><em>Metadata</em>: <a href="http://yaml.org/">Yaml (Yaml ain’t markup language)</a> provides a simple way of including structured metadata (data about data) at the beginning of a text file recording information on authors, title, date etc..</li> <li><em>Tools</em>: All major operating systems come with basic text editors sufficient for reading and editing plain text files (“Notepad” on Windows, “TextEdit” on Mac{» comment on Linux«}). However, some additional features, such as syntax highlighting and basic customisation of the writing interface (you have to look at it for substantial periods of time), will significantly enhance the writing experience. <ul> <li><a href="https://www.sublimetext.com/">Sublime Text</a> (Windows, Mac, Linux)</li> <li><a href="https://notepad-plus-plus.org/">NotePad++</a> (Windows only)</li> </ul> </li> </ul> <h3 id="2-ascertaining-authorship-version-control-and-archiving">2. ascertaining authorship, version control, and archiving</h3> <p>Writing is a process and subject to change. We need to be able to try out different structures and formulations and, more often than one would like to, we discover that yesterday’s deletions would have been worth keeping. Not to mention an external editor or collaborators that quickly make any approach involving ever-longer file names futile (we all have our folders full of <code class="language-plaintext highlighter-rouge">text.docx</code>, <code class="language-plaintext highlighter-rouge">text-new-version.docx</code>, <code class="language-plaintext highlighter-rouge">text-new-version-2004-01-01.docx</code>, <code class="language-plaintext highlighter-rouge">text-new-version-2004-01-01-comments-by-tg-2.docx</code> etc.).</p> <ul> <li><a href="https://git-scm.com/"><em>git</em></a>: git is an open <em>version control system</em> (vcs) that works with any file type. It traces every change with clear information on author and a timestamp.</li> <li><a href="https://www.github.com"><em>GitHub</em></a>: GitHub is probably the most popular <em>distributed version control system</em> (dvcs) and code-sharing platform based on git.<sup id="fnref:2"><a href="#fn:2" class="footnote" rel="footnote" role="doc-noteref">3</a></sup> While GitHub is a commercial company, it offers free accounts and unlimited public repositories to everyone and free private repositories to academics. Once a change has been committed to a GitHub repository, authors have a public proof of their authorship with a unique identifier that can be referenced. In addition, the repository forms a redundant online back-up of your work. <ul> <li><em>ssh</em> (secure shell): in order to communicate with a server without a local client software, some knowledge of ssh is necessary. To avoid this, one can use any of the numerous GitHub clients.<sup id="fnref:3"><a href="#fn:3" class="footnote" rel="footnote" role="doc-noteref">4</a></sup></li> </ul> </li> <li><a href="https://zenodo.org/"><em>Zenodo</em></a>: Zenodo is an open science platform developed and operated by CERN. Some of its many features are the provision of DOIs (<a href="https://www.doi.org/"><em>digital object identifiers</em></a>) and long-term storage thus providing stable links and protection against <a href="https://en.wikipedia.org/wiki/Link_rot">link-rot</a>. It also hooks into GitHub.</li> </ul> <h3 id="3-collaboration--external-review">3. collaboration / external review</h3> <ul> <li>git / GitHub: git allows for branching and forking of text. All changes / suggestions can be reviewed and the author can decide whether to accept or reject them.</li> </ul> <p>{»We need additional collaborative tools«}</p> <ul> <li><em><a href="http://prose.io/">prose.io</a></em>: author text on GitHub online.</li> </ul> <h2 id="2-publishing">2. Publishing</h2> <h3 id="1-generate-an-accessible-representation-of-your-text">1. Generate an accessible representation of your text</h3> <p>While it would be absolutely sufficient to publish / distribute the plain-text files by means of a USB key, a graphical user interface (GUI) that translates the structural information into a formatted and aesthetically pleasing layout is often {==advisable==}{»better wording?«}—be they a printed paper copy or a website:</p> <ul> <li><a href="https://pandoc.org/"><em>Pandoc</em></a>: Pandoc is a tool for converting plain text and markdown into multiple (hence <em>pan</em>) target formats (thus <em>doc</em>), such as, but not limited to <code class="language-plaintext highlighter-rouge">HTML</code>, <code class="language-plaintext highlighter-rouge">DOCX</code>, and <code class="language-plaintext highlighter-rouge">PDF</code>. Pandoc, which is under active development by John MacFarlane, a professor of philosophy at UC Berkeley, also supports the formatting of references using <a href="http://www.bibtex.org/">BibTeX</a> (another plain text format for storing structured bibliographic information) and <a href="http://citationstyles.org/">CSL (citation style language)</a>.</li> <li><em><code class="language-plaintext highlighter-rouge">HTML</code> and <code class="language-plaintext highlighter-rouge">CSS</code></em>: <code class="language-plaintext highlighter-rouge">HTML</code> and <code class="language-plaintext highlighter-rouge">CSS</code> are well-established standards to separate form and content maintained by the <a href="https://www.w3.org/">World Wide Web Consortium (W3C)</a>. While the <em>hyper text markup language</em> (<code class="language-plaintext highlighter-rouge">HTML</code>) carries the content of our text as well as all the structural and formating information and the metadata, <em>cascading stylesheets</em> (<code class="language-plaintext highlighter-rouge">CSS</code>) provide the actual layout. This combination allows to provide different layouts for different contexts and devices using the same content.</li> <li><em>Metadata</em>: <code class="language-plaintext highlighter-rouge">HTML5</code> supports semantic tags and machine-readable metadata following various standards, such as those provided by <a href="http://dublincore.org/">Dublin Core (DC)</a> or <a href="https://schema.org">schema.org</a>. If one includes Dublin Core metadata in the head of <code class="language-plaintext highlighter-rouge">HTML</code> files, aggregators, search engines, and reference managers can find and extract the structured information on author, title, publication date, keywords, etc.</li> </ul> <h3 id="2-think-about-and-provide-a-licence">2. Think about and provide a licence</h3> <p>{»mention copyright«}</p> <p>A licence is formal agreement that specifies the rights and duties of both the <em>licensor</em> (e.g. us as authors) and the <em>licencee</em> (e.g. us as readers). Its most important purpose within our discussion is to assure the readers of our texts of their rights to read, copy, and cite them. An open licence might for instance allow reproduction of the text but might prohibit charging for accessing the reproduction.</p> <p>Formulating one’s own licence text is a challenge and one might not be familiar enough with the necessary “legalese” to write a text readers can rely on. In consequence we suggest looking at established licences and having made a case for open access {»open science, open knowledge etc.«}, we suggest to start with <a href="https://creativecommons.org/">creative commons licences</a>.<sup id="fnref:5"><a href="#fn:5" class="footnote" rel="footnote" role="doc-noteref">5</a></sup></p> <h3 id="3-publish--distribute-the-content">3. Publish / distribute the content</h3> <p>Websites of more than a single page of text tend to be technically complex and require an infrastructure of data storage, internet connections, web addresses, some <em>content management system</em> (cms), and databases that rarely come for free and without the need of maintenance. In addition, contemporary dynamic websites are almost impossible to archive or download in their entirety. To reduce complexity and the number of technologies, we suggest using static, self-contained websites without a content management system and no database.</p> <ul> <li><em>jekyll</em>: <a href="https://jekyllrb.com/">Jekyll</a> is currently one of the most popular open tools to generate “dynamic” static, self-contained websites from plain text files containing MarkDown and Yaml, using nothing but <code class="language-plaintext highlighter-rouge">HTML</code> and <code class="language-plaintext highlighter-rouge">CSS</code> and nested files and folders. These can either be hosted online or distributed on a USB key.</li> <li><em>GitHub pages</em>: There are, of course, plenty of open, free or commercial providers to host your website. But having already signed up to GitHub and keeping all our texts in a GitHub repository, it is worthwhile to look at <a href="https://pages.github.com/">GitHub pages</a>, which provide free hosting of version-controlled content and supports jekyll out of the box. This means one can directly publish Markdown-formatted plain text files through GitHub pages and jekyll.</li> </ul> <h1 id="challenges">Challenges:</h1> <p>How to deal with sensitive data / material, that should not be publicly accessible, such as ethnographic field notes?</p> <ul> <li><a href="http://openpgp.org/">OpenPGP</a>: PGP stands for “pretty good privacy” and is widely used for encrypting emails. OpenPGP is a proposed standard in <a href="https://tools.ietf.org/html/rfc4880">RFC 4480</a>. It works with plain text and thus with all the tools covered in this {==course==}{»workshop«}.</li> </ul> <h1 id="resources--literature">Resources / literature:</h1> <ul> <li><a href="http://programminghistorian.org/">The Programming Historian</a>: open-access, peer-reviewed suite of tutorials that help humanists learn a wide range of digital tools, techniques, and workflows to facilitate their research.</li> </ul> <p>{»add relevant literature«}</p> <h2 id="use-of-git-and-github-in-the-humanities">Use of git and GitHub in the humanities</h2> <ul> <li>makerlab at UVic: <a href="https://github.com/uvicmakerlab/dhsi2015/blob/master/git.md">git/github in 20 steps</a></li> <li>Chad Black’s <a href="https://parezcoydigo.wordpress.com/2013/08/26/getting-started-with-github-and-prose-io/">getting started with github and prose.io</a> on using a combination of GitHub, jekyll, and prose.io for a collaborative seminar <a href="http://dh.chadblack.net/info/all_posts/">blog</a>.</li> <li><a href="http://www.hastac.org/blogs/harrisonm/2013/10/12/github-academia-and-collaborative-writing">GitHub, Academia, and Collaborative Writing</a></li> <li><a href="http://www.hybridpedagogy.com/journal/push-pull-fork-github-for-academics/">Push, Pull, Fork: GitHub for Academics</a></li> <li><a href="http://blogs.lse.ac.uk/impactofsocialsciences/2013/06/04/github-for-academics/">GitHub for Academics: the open-source way to host, create and curate knowledge</a></li> <li><a href="http://savageminds.org/2013/02/15/living-in-a-plain-text-world-tools-we-use/">Living in a Plain Text World (Tools We Use)</a></li> <li><a href="http://technical.ly/brooklyn/2014/09/12/authorea-software-for-academic-paper-writing-is-inspired-by-git/">This software for academic paper writing is inspired by Git</a></li> <li><a href="http://www.lifehack.org/articles/technology/how-to-use-git-to-version-your-writing.html">How To: Use Git to Version Your Writing</a></li> <li><a href="https://www.martineve.com/2013/08/18/using-git-in-my-writing-workflow/">Using git in my writing workflow</a></li> </ul> <div class="footnotes" role="doc-endnotes"> <ol> <li id="fn:1"> <p>Tenen, Dennis and Grant Wythoff. “<a href="http://programminghistorian.org/lessons/sustainable-authorship-in-plain-text-using-pandoc-and-markdown">Sustainable Authorship in Plain Text Using Pandoc and Markdown</a>.” <a href="#fnref:1" class="reversefootnote" role="doc-backlink">&#8617;</a></p> </li> <li id="fn:4"> <p>MarkDown really is only a convention and <a href="http://daringfireball.net/projects/markdown/syntax">John Gruber’s (and Adam Swartz’ [yes, <em>the</em> Adam Swartz]) canonical description of the Markdown syntax</a> is at least partially ambiguous and and lacks some core functionality for academic writing, such as support for tables and footnotes. In consequence, a plethora of formats (MultiMarkdown, GitHub flavored markdown, etc.) and software implementations have proliferated inspired by and based on Markdown that make the actual rendering of Markdown in <code class="language-plaintext highlighter-rouge">HTML</code> rather unpredictable beyond the core functionality. In recent years, a group of people involving <a href="http://johnmacfarlane.net/">John MacFarlane</a>, professor of philosophy at UC Berkeley and author of Pandoc, proposed and developed are more rigid standard which they call <a href="http://commonmark.org/">CommonMark</a>. <a href="#fnref:4" class="reversefootnote" role="doc-backlink">&#8617;</a></p> </li> <li id="fn:2"> <p>Other options based on git are <a href="https://bitbucket.org/">BitBucket</a> and <a href="https://about.gitlab.com/">GitLab</a>. <a href="#fnref:2" class="reversefootnote" role="doc-backlink">&#8617;</a></p> </li> <li id="fn:3"> <p>GitHub provides its own clients for <a href="https://central.github.com/mac/latest">Mac</a> and <a href="https://github-windows.s3.amazonaws.com/GitHubSetup.exe">Windows</a> <a href="#fnref:3" class="reversefootnote" role="doc-backlink">&#8617;</a></p> </li> <li id="fn:5"> <p>There is a great list of open source licences, including links to their full texts, at <a href="https://opensource.org/licenses/category">opensource.org</a> <a href="#fnref:5" class="reversefootnote" role="doc-backlink">&#8617;</a></p> </li> </ol> </div>]]></content><author><name>Till Grallert</name></author><category term="blog"/><category term="project_dh"/><category term="teaching"/><category term="draft"/><category term="teaching"/><category term="tools"/><category term="plain text"/><category term="markdown"/><category term="pandoc"/><category term="git"/><category term="github"/><category term="jekyll"/><summary type="html"><![CDATA[DRAFT Disclaimer 1: many of the ideas have been inspired by the work of Alex Gil (@elotroalex) and Denis Tenen (@dennistenen)1 and discussions with the two of them and others at DHI Beirut, THATCamp Beirut, and DHSI. Disclaimer 2: I am teaching short workshops on some of the ideas outlined in this post at Digital Humanities Institute - Beirut on 10 March 2017 and at DH Abu Dhabi on 10 April 2017. Basic slides are available here and here. Intro In the world of (academic) publishing, large aggregators and indexers have turned into and acquired publishing presses and generate obscene profits by charging the public (every tax-payer worldwide) multiple times over. First by charging the predominantly publicly-funded academic for publishing the results of her publicly-funded research and by enforcing a culture of pro bono labour among academic reviewers and editors; second by selling this content to equally predominantly publicly-funded libraries, which then increasingly demand access fees from members of the public, who want to access their collections; and third by offloading the cost of long-term preservation to, again, publicly-funded institutions. This system not only created a hierarchy of academics and institutions in the relatively well-off “West”—two classes divided by their ability to pay for being published and accessing publications (their own and others). It also increasingly prevents anybody outside western academia from accessing cutting-edge research and participating in intellectual discourse. Tenen, Dennis and Grant Wythoff. “Sustainable Authorship in Plain Text Using Pandoc and Markdown.” &#8617;]]></summary></entry><entry><title type="html">Presentation at ‘Dangerous Classes’ conference in Oxford</title><link href="https://tillgrallert.github.io/blog/2017/02/18/dangerous-classes/" rel="alternate" type="text/html" title="Presentation at ‘Dangerous Classes’ conference in Oxford"/><published>2017-02-18T17:06:47+00:00</published><updated>2017-02-18T17:06:47+00:00</updated><id>https://tillgrallert.github.io/blog/2017/02/18/dangerous-classes</id><content type="html" xml:base="https://tillgrallert.github.io/blog/2017/02/18/dangerous-classes/"><![CDATA[<p>On 26 January this year I had a chance to present a paper on food riots titled “Women in the streets! Urban food riots in late Ottoman <em>Bilād al-Shām</em>” at the conference <a href="https://www.sant.ox.ac.uk/events/conference-%E2%80%9Cdangerous-classes%E2%80%9D-middle-east-and-north-africa">“The ‘Dangerous Classes’ in the Middle East and North Africa”</a> organised by Stephanie Cronin at St. Antony’s college, University of Oxford.</p> <p>As always, I have made the slides available on <a href="https://tillgrallert.github.io/slides/2017-Oxford">GitHub</a>.</p>]]></content><author><name>Till Grallert</name></author><category term="presentation"/><category term="project_food-riots"/><category term="presentation"/><category term="conferences"/><category term="food riots"/><summary type="html"><![CDATA[On 26 January this year I had a chance to present a paper on food riots titled “Women in the streets! Urban food riots in late Ottoman Bilād al-Shām” at the conference “The ‘Dangerous Classes’ in the Middle East and North Africa” organised by Stephanie Cronin at St. Antony’s college, University of Oxford.]]></summary></entry></feed>