LREC 2018 Proceedings

Summary of the paper

Title	ILCM - A Virtual Research Infrastructure for Large-Scale Qualitative Data
Authors	Andreas Niekler, Arnim Bleier, Christian Kahmann, Lisa Posch, Gregor Wiedemann, Kenan Erdogan, Gerhard Heyer and Markus Strohmaier
Abstract	The iLCM project pursues the development of an integrated research environment for the analysis of structured and unstructured data in a “Software as a Service” architecture (SaaS). The research environment addresses requirements for the quantitative evaluation of large amounts of qualitative data with text mining methods as well as requirements for the reproducibility of data-driven research designs in the social sciences. For this, the iLCM research environment comprises two central components. First, the Leipzig Corpus Miner (LCM), a decentralized SaaS application for the analysis of large amounts of news texts developed in a previous Digital Humanities project. Second, the text mining tools implemented in the LCM are extended by an “Open Research Computing” (ORC) environment for executable script documents, so-called “notebooks”. This novel integration allows to combine generic, high-performance methods to process large amounts of unstructured text data and with individual program scripts to address specific research requirements in computational social science and digital humanities. ilcm.informatik.uni-leipzig.de
Topics	Text Mining, Corpus (Creation, Annotation, Etc.), Lr Infrastructures And Architectures
Full paper	ILCM - A Virtual Research Infrastructure for Large-Scale Qualitative Data
Bibtex	@InProceedings{NIEKLER18.734, author = {Andreas Niekler and Arnim Bleier and Christian Kahmann and Lisa Posch and Gregor Wiedemann and Kenan Erdogan and Gerhard Heyer and Markus Strohmaier}, title = "{ILCM - A Virtual Research Infrastructure for Large-Scale Qualitative Data}", booktitle = {Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC 2018)}, year = {2018}, month = {May 7-12, 2018}, address = {Miyazaki, Japan}, editor = {Nicoletta Calzolari (Conference chair) and Khalid Choukri and Christopher Cieri and Thierry Declerck and Sara Goggi and Koiti Hasida and Hitoshi Isahara and Bente Maegaard and Joseph Mariani and Hélène Mazo and Asuncion Moreno and Jan Odijk and Stelios Piperidis and Takenobu Tokunaga}, publisher = {European Language Resources Association (ELRA)}, isbn = {979-10-95546-00-9}, language = {english} }