{"id":14668,"date":"2017-12-27T11:05:17","date_gmt":"2017-12-27T10:05:17","guid":{"rendered":"https:\/\/digitisation.eu\/?p=14668"},"modified":"2023-06-19T09:44:32","modified_gmt":"2023-06-19T07:44:32","slug":"call-participation-hacking-news-workshop","status":"publish","type":"post","link":"https:\/\/digitisation.eu\/?p=14668","title":{"rendered":"Call for participation: Hacking the news workshop"},"content":{"rendered":"<p><b>Monday 5 March 2018 (12:30) &#8211; Tuesday 6 March 2018 (17:00)<\/b><\/p>\n<p><b>Hacking the News: from digitised newspapers to the archived-web:\u00a0<\/b><b>an introductory workshop to text and data-mining <\/b><b>#DHNhacknews<\/b><\/p>\n<p><span style=\"font-weight: 400;\">Libraries have been digitising historical newspapers since the early 2000\u2019s. However, to what extent are these digitised newspaper archives being used in digital humanities research? Web-archiving began in 1996 with the <\/span><a href=\"https:\/\/archive.org\/\"><span style=\"font-weight: 400;\">Internet Archive initiative<\/span><\/a><span style=\"font-weight: 400;\"> and its well-known digital archive \u2018<\/span><a href=\"https:\/\/archive.org\/web\/\"><span style=\"font-weight: 400;\">The Wayback Machine<\/span><\/a><span style=\"font-weight: 400;\">\u2019. Since then a multitude of web-archiving initiatives have been established to continue these efforts. However, the true potential of digital newspaper corpora and web-archives is as yet under-exploited. <\/span><i><span style=\"font-weight: 400;\">Hacking the news: from digitised newspapers to the archived-web: an introductory workshop<\/span><\/i> <i><span style=\"font-weight: 400;\">to text and data-mining<\/span><\/i><span style=\"font-weight: 400;\"> is intended to help redress this balance.<\/span><\/p>\n<p><i><span style=\"font-weight: 400;\">Hacking the news<\/span><\/i><span style=\"font-weight: 400;\"> is a 1.5 day workshop, prior to<\/span><a href=\"http:\/\/dig-hum-nord.eu\/dhn-2018\"> <span style=\"font-weight: 400;\">DHN 2018<\/span><\/a><span style=\"font-weight: 400;\">. Primarily intended for, but not limited to, early career researchers, the aim of this workshop is to provide an introduction to a range of topics to consider when undertaking digital analysis of newspaper corpora and analysing web-archives for research. \u00a0A <\/span><a href=\"https:\/\/docs.google.com\/document\/d\/10TAVfwLtEPCofbd_kjjh0nzXTB3q3XdiMkjXkU7RxUg\/edit?usp=sharing\"><span style=\"font-weight: 400;\">draft programme<\/span><\/a><span style=\"font-weight: 400;\"> for the workshop is available.<\/span><\/p>\n<p><b>Day one<\/b><span style=\"font-weight: 400;\"> of the workshop will focus on setting the context. Topics such as: How are digital newspaper corpora created? What is Optical Character Recognition? How does that differ from Optical Layout Recognition? How does news on the archived web differ from digitised newspapers? What data formats are used for the archived web and how do we analyse Web Archive datasets<\/span><span style=\"font-weight: 400;\">? <\/span><\/p>\n<p><b>Day two<\/b><span style=\"font-weight: 400;\"> of <\/span><span style=\"font-weight: 400;\">workshop will provide opportunity for participants to get their (digital) hands dirty, by working with digital newspaper and web-archives. Corpora will be provided in a number of languages, as far as possible, based on the needs of the workshop participants. As well as<\/span> <span style=\"font-weight: 400;\">corpus preparation, where issues such as data cleaning will be explored, there will be opportunity to test a range of<\/span> <span style=\"font-weight: 400;\">text and data mining tools for analysing digital corpora<\/span><i><span style=\"font-weight: 400;\">. <\/span><\/i><\/p>\n<p><b>Call for Participation: <\/b><span style=\"font-weight: 400;\">To participate in the workshop (ca. 25-30 participants), please <\/span><a href=\"https:\/\/docs.google.com\/forms\/d\/e\/1FAIpQLSfILiaUoXTf7xQq8e5MAomYGK95nKllB1VJKOUmaTiCNeoVBg\/viewform\"><span style=\"font-weight: 400;\">complete this form<\/span><\/a><span style=\"font-weight: 400;\"> with details of your research interests, preferred languages for the digital corpora, level of technical experience, motivation for participating in the workshop plus a short biographical note by <\/span><del><b>Monday 15 January 2018<\/b><\/del><b>\u00a0Wednesday 31 January 2018<\/b><span style=\"font-weight: 400;\">. <\/span><\/p>\n<p><b>Venue: <\/b><i><span style=\"font-weight: 400;\">Hacking the News<\/span><\/i><span style=\"font-weight: 400;\"> is hosted by the <\/span><a href=\"https:\/\/www.kansalliskirjasto.fi\/en\/directions-to-library-network-services\"><span style=\"font-weight: 400;\">National Library of Finland\u2019s Network Services Division<\/span><\/a><span style=\"font-weight: 400;\">, Kaikukatu 4, Helsinki.<\/span><\/p>\n<p><b>Organisers: <\/b><span style=\"font-weight: 400;\">This workshop is co-organised by the Ghent Centre for Digital Humanities (GhentCDH), the International Internet Preservation Consortium (IIPC) and the National Library of Finland, in collaboration with the Helsinki Centre for Digital Humanities (HELDIG), Digital Humanities Lab (DIGHUMLAB), Denmark, the DH Lab of \u00c9cole polytechnique f\u00e9d\u00e9rale de Lausanne (EPFL), the Luxembourg Centre for Contemporary and Digital History (C2DH), Alan Turing Institute, Platform DH, University of Antwerp and the Leuven Centre for Digital Humanities. The digitised newspaper collections and web-archives will be provided by a number of National and University Libraries. The workshop is supported by DARIAH (Digital Research Infrastructure for the Arts and Humanities), CLARIN (European Research Infrastructure for Language Resources and Technology) and the IMPACT Centre of Competence. <\/span><\/p>\n<p><b>Contact information: <\/b><a href=\"mailto:dhnhacknews@gmail.com\"><span style=\"font-weight: 400;\">dhnhacknews@gmail.com<\/span><\/a><\/p>\n<ul>\n<li><span style=\"font-weight: 400;\">Sally Chambers, Ghent Centre for Digital Humanities<\/span><\/li>\n<li><span style=\"font-weight: 400;\">Olga Holownia, International Internet Preservation Consortium<\/span><\/li>\n<li><span style=\"font-weight: 400;\">Lassi Lager, National Library of Finland<\/span><\/li>\n<\/ul>\n","protected":false},"excerpt":{"rendered":"<p>Monday 5 March 2018 (12:30) &#8211; Tuesday 6 March 2018 (17:00) Hacking the News: from digitised newspapers to the archived-web:\u00a0an introductory workshop to text and data-mining #DHNhacknews Libraries have been digitising historical newspapers since the early 2000\u2019s. However, to what extent are these digitised newspaper archives being used in digital humanities research? Web-archiving began in &hellip; <a href=\"https:\/\/digitisation.eu\/?p=14668\" class=\"more-link\">Continue reading<span class=\"screen-reader-text\"> &#8220;Call for participation: Hacking the news workshop&#8221;<\/span><\/a><\/p>\n","protected":false},"author":476,"featured_media":14671,"comment_status":"closed","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[448,156,153],"tags":[],"class_list":["post-14668","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-dhn2018","category-events","category-news"],"acf":[],"_links":{"self":[{"href":"https:\/\/digitisation.eu\/index.php?rest_route=\/wp\/v2\/posts\/14668","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/digitisation.eu\/index.php?rest_route=\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/digitisation.eu\/index.php?rest_route=\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/digitisation.eu\/index.php?rest_route=\/wp\/v2\/users\/476"}],"replies":[{"embeddable":true,"href":"https:\/\/digitisation.eu\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=14668"}],"version-history":[{"count":5,"href":"https:\/\/digitisation.eu\/index.php?rest_route=\/wp\/v2\/posts\/14668\/revisions"}],"predecessor-version":[{"id":18020,"href":"https:\/\/digitisation.eu\/index.php?rest_route=\/wp\/v2\/posts\/14668\/revisions\/18020"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/digitisation.eu\/index.php?rest_route=\/wp\/v2\/media\/14671"}],"wp:attachment":[{"href":"https:\/\/digitisation.eu\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=14668"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/digitisation.eu\/index.php?rest_route=%2Fwp%2Fv2%2Fcategories&post=14668"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/digitisation.eu\/index.php?rest_route=%2Fwp%2Fv2%2Ftags&post=14668"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}