Category: Discussions
IMPACT Final Conference – The Functional Extension Parser: A Document Understanding Platform with Günter Mühlberger
Günter Mühlberger chaired the 3rd block: Tools for Improved Text Recognition. However, as time ran out in the session he chose to postpone the delivery of his talk on the FEP until the second day Research Parallel Sessions.
IMPACT Final Conference – Evaluation of lexicon supported OCR and information retrieval
Jesse De Does from the INL gave a brief but rich presentation on the evaluation of lexicon supported OCR and the project’s recent improvements. To evaluate lexica in OCR, the FineReader SDK 10 is used. In short, the software measures OCR with a default included dictionary, and, for each word or fuzzy set, it gives … Continue reading "IMPACT Final Conference – Evaluation of lexicon supported OCR and information retrieval"
IMPACT Final Conference-Overview of language work in IMPACT
Katrien Depuydt provided a brief overview of the IMPACT project’s work packages devoted to creating language tools and lexicon to aid in both information retrieval and OCR processing. How might one measure successful improvement to the access of text? She cleverly posits that the key will be in asking ourselves: “Can we handle the “world”? … Continue reading "IMPACT Final Conference-Overview of language work in IMPACT"
IMPACT Final Conference – Post-Correction on IMPACT with Ulrich Reffle
Developing a unique user tool with a team led by Prof. Schulz, Ulrich Reffle along with Annette Gotscharek, Christoph Ringlstetter and Thorsten Vobl. at the University of Munich have potentially revolutionised the speed at which researchers can analyse texts.
IMPACT Final Conference-Crowdsourcing in the Digitalkoot Project
DIGI-to digitise Talkoot-people gathering to work together voluntarily (without payment). Majilis Bremer-Laamanen (National Library of Finland) shares their unique experiment in crowdsourcing OCR correction through gaming with their Digitalkoot Project launched February 2011.
IMPACT Final Conference – IBM Adaptive OCR Engine and CONCERT Cooperative Correction
Asaf Tzadok (IBM Haifa Research Lab) showed us IBM’s CONCERT tool which facilitates collaborative OCR correction. CONCERT (Cooperative Engine for the Correction of Extracted Text) works in three steps: character session, word session and page-level session.
IMPACT Final Conference: 1st Keynote: The Strategic Digital Overview
Richard Boulderstone, Director of eStrategy and Programs at the British Library, kicked off the IMPACT Conference this morning with a suitably impactful statement of scope: the British Library, he estimates, has nearly 5 billion physical pages in a 150 million object collection.
BSB Demo Day: Impressionen und überarbeitete Artikel
One week after our very successful Demo Day at the Bavarian State Library, we had a look at all articles and reworded them were necessary, to weed out factual, grammatical and spelling errors. Turns out blogging live just isn’t that easy. Also, the presentation slides were added to all talks. As promised, we also caught … Continue reading "BSB Demo Day: Impressionen und überarbeitete Artikel"
Kollaborative Korrektur
Doris Škarić from the Bavarian State Library reported about collaborative correction of OCR results by volunteers. She presented the IMPACT tool CONCERT (the COllaborative eNgine for the CorREction of Texts) and reported about the findings of a pilot test of the tool at the Bavarian State Library.
Dokumentstrukturerkennung
Günter Mühlberger from the University and Regional Library of Tyrol in Innsbruck presented the Functional Extension Parser (FEP), a tool for the OCR-based structural analysis of printed texts.
