Text and Data Mining - Extracting Knowledge from Unstructured Data

Text and data mining comprises the development and application of methods which are designed to extract knowledge that is relevant to the social sciences from unstructured texts or data streams.

Our research on Text and Data Mining

Detection of statistical regularities in data and text and alignment of these regularities with variables of interest such as political leaning or gender
Combine digital behavioral data and survey data to create new types of user models
Semantic enrichment and analysis of collaboratively generated documents (e.g. wikipedia articles or scientific publications) and the social dynamics of the creation process (e.g. conflicts, productivity)
Statistical modelling of sequential human behavior (e.g., the decisions made when navigating on the web or individual movement in urban surroundings)
Detection, disambiguation and linking of entities which are of interest for the social sciences in academic publications (especially references to research data)
Extraction of key information from texts and (semi-)automatic indexing

Title	Start	End	Funder
Kompetenzzentrum Datenqualität in den Sozialwissenschaften (KODAQS)	2023-11-15	2026-11-14	Bund
NFDI for Data Science and Artificial Intelligence (NFDI4DS)	2021-10-01	2026-09-30	DFG
NFDI for Business, Economic and Related Data (BERD@NFDI)	2021-10-01	2026-09-30	DFG
Dehumanization Online: Measurement and Consequences (Professorinnenprogramm) (DeHum)	2021-01-01	2026-09-30	SAW (Leibniz)

Find out more about our consulting and services:

Analyzing Digital Behavioral Data
Methods, tools, frameworks and infrastructures for analyzing digital behavioral data.
GESIS Guides to Digital Behavioral Data
Expertise and hands-on advice on the acquisition and analysis of digital behavioral data and the computational methods needed.

Contact persons

Wagner, Prof. Dr. Claudia

Computational Social Science
Head of department

+49 (0221) 47694-224
claudia•wagner[at]gesis•org
vCard

Zapilko, Dr. Benjamin

Knowledge Technologies for the Social Sciences
Information Extraction and Linking

+49 (0221) 47694-515
benjamin•zapilko[at]gesis•org
vCard