Quantification and Analysis of Scientific Language Variation Across Research Fields

12/04/2018
by   Pei Zhou, et al.
0

Quantifying differences in terminologies from various academic domains has been a longstanding problem yet to be solved. We propose a computational approach for analyzing linguistic variation among scientific research fields by capturing the semantic change of terms based on a neural language model. The model is trained on a large collection of literature in five computer science research fields, for which we obtain field-specific vector representations for key terms, and global vector representations for other words. Several quantitative approaches are introduced to identify the terms whose semantics have drastically changed, or remain unchanged across different research fields. We also propose a metric to quantify the overall linguistic variation of research fields. After quantitative evaluation on human annotated data and qualitative comparison with other methods, we show that our model can improve cross-disciplinary data collaboration by identifying terms that potentially induce confusion during interdisciplinary studies.

READ FULL TEXT

page 1

page 4

research
10/17/2017

Role of Interdisciplinarity in Computer Sciences: Quantification, Impact and Life Trajectory

The tremendous advances in computer science in the last few decades have...
research
02/07/2023

The Effect of Metadata on Scientific Literature Tagging: A Cross-Field Cross-Model Study

Due to the exponential growth of scientific publications on the Web, the...
research
06/27/2017

Using text analysis to quantify the similarity and evolution of scientific disciplines

We use an information-theoretic measure of linguistic similarity to inve...
research
04/29/2020

Analysing Lexical Semantic Change with Contextualised Word Representations

This paper presents the first unsupervised approach to lexical semantic ...
research
04/22/2021

Combining dissimilarity measure for the study of evolution in scientific fields

The evolution of scientific fields has been attracting much attention in...
research
05/03/2021

Metaphor Research in the 21st Century: A Bibliographic Analysis

Metaphor is widely used in human communication. The cohort of scholars s...
research
06/02/2021

The data paper as a socio-linguistic epistemic object: A content analysis on the rhetorical moves used in data paper abstracts

The data paper is an emerging academic genre that focuses on the descrip...

Please sign up or login with your details

Forgot password? Click here to reset