For bachelor students we offer German lectures on database systems in addition with paper- or project-oriented seminars. Within a one-year bachelor project students finalize their studies in cooperation with external partners. For master students we offer courses on information integration, data profiling, search engines and information retrieval enhanced by specialized seminars, master projects and advised master theses.
The Web Science group focuses on various topics related to the Web, such as Information Retrieval, Natural Language Processing, Data Mining, Knowledge Discovery, Social Network Analysis, Entity Linking, and Recommender Systems. The group is particularly interested in Text Mining to deal with the vast amount of unstructured and semi-structured information available on the Web.
Most of our research is conducted in the context of larger research projects, in collaboration across students, across groups, and across universities. We strive to make available most of our data sets and source code.
In the past, we have built various large and small data integration systems. They are no longer maintained and many if not most of them are not longer actively running. Please contact Felix Naumann to learn more.
GovWILD: An integrated set of government data to explore nepotism in politics and economy.
Aladin: A system to perform almost automatic integration of datasets.
MAC / Hummer: A system to integrate heterogeneous datasets, including schema matching, deduplication and data fusion steps.