SearcharxivSearch

arXiv subjects

Aliakbar Akbaritabar

Publications and source records attributed to Aliakbar Akbaritabar.

6 recordsLinked to original sources

Differentiating Emigration from Return Migration of Scholars Using Name-Based Nationality Detection Models

Most web and digital trace data do not include information about an individual's nationality due to privacy concerns. The lack of data on nationality can create challenges for migration research. It can lead to a left-censoring issue since we are uncertain about the migrant's country of origin. Once we observe an emigration event, if we know the nationality, we can differentiate it from return migration. We propose methods to detect the nationality with the least available data, i.e., full names. We use the detected nationality in comparison with the country of academic origin, which is a common approach in studying the migration of researchers. We gathered 2.6 million unique name-nationality pairs from Wikipedia and categorized them into families of nationalities with three granularity levels to use as our training data. Using a character-based machine learning model, we achieved a weighted F1 score of 84% for the broadest and 67% for the most granular, country-level categorization. In our empirical study, we used the trained and tested model to assign nationality to 8+ million scholars' full names in Scopus data. Our results show that using the country of first publication as a proxy for nationality underestimates the size of return flows, especially for countries with a more diverse academic workforce, such as the USA, Australia, and Canada. We found that around 48% of emigration from the USA was return migration once we used the country of name origin, in contrast to 33% based on academic origin. In the most recent period, 79% of scholars whose affiliation has consistently changed from the USA to China, and are considered emigrants, have Chinese names in contrast to 41% with a Chinese academic origin. Our proposed methods for addressing left-censoring issues are beneficial for other research that uses digital trace data to study migration.

cs.DL

Methodological monotheism across fields of science in contemporary quantitative research

The importance of research teams' diversity for the progress of science is highlighted extensively. Despite the seemingly hegemonic role of hypothesis testing in modern quantitative research, little attention has been devoted to the diversity of quantitative methods, epitomized by the linear model framework of analysis. Using bibliometric data from the Web of Science, we conduct a large-scale and cross-disciplinary assessment of the prevalence of linear-model-based research from 1990 to 2022. In absolute terms, linear models are widely used across all fields of science. In relative terms, three patterns suggest linear models are hegemonic among Social Sciences. First, there is a high and growing prevalence of linear-model-based research. Second, global patterns of linear-model-based research prevalence align with global inequalities in knowledge production. Third, there was a citation premium to linear-model-based research until 2012 for articles' number of citations and for the entire period in terms of having at least one citation. Previous research suggests that the confluence of these patterns may be detrimental to the Social Sciences as it potentially marginalizes theories incompatible with the linear models' framework, lowers the diversity of narratives about social phenomena, and prevents innovative and path-breaking research, limiting the breadth of research.

cs.DL

A study of referencing changes in preprint-publication pairs across multiple fields

Manuscripts have a complex development process with multiple influencing factors. Reconstructing this process is difficult without large-scale, comparable data on different versions of manuscripts. Preprints are increasingly available and may provide access to the earliest manuscript versions. Here, we matched 6,024 preprint-publication pairs across multiple fields and examined changes in their reference lists between the manuscript versions as one aspect of manuscripts' development. We also qualitatively analysed the context of references to investigate the potential reasons for changes. We found that 90 percent of references were unchanged between versions and 8 percent were newly added. We found that manuscripts in the natural and medical sciences undergo more extensive reframing of the literature while changes in engineering mostly focused on methodological details. Our qualitative analysis suggests that peer review increases the methodological soundness of scientific claims, improves the communication of findings, and ensures appropriate credit for previous research.

cs.DL

Quantitative View of the Structure of Institutional Scientific Collaborations Using the Examples of Halle, Jena and Leipzig

Examining effectiveness of institutional scientific coalitions can inform future policies. This is a study on the structure of scientific collaborations in three cities in central Germany. Since 1995, the three universities of this region have formed and maintained a coalition which led to the establishment of an interdisciplinary center in 2012, i.e., German Center for Integrative Biodiversity Research (iDiv). We investigate whether the impact of the former coalition is evident in the region's structure of scientific collaborations and the scientific output of the new center. Using publications data from 1996-2018, we build co-authorship networks and identify the most cohesive communities in terms of collaboration, and compare them with communities identified based on publications presented as the scientific outcome of the coalition and new center on their website. Our results show that despite the highly cohesive structure of collaborations presented on the coalition website, there is still much potential to be realized. The newly established center has bridged the member institutions but not to a particularly strong level. We see that geographical proximity, collaboration policies, funding, and organizational structure alone do not ensure prosperous scientific collaboration structures. When new center's scientific output is compared with its regional context, observed trends become less conspicuous. Nevertheless, the level of success the coalition achieved could inform policy makers regarding other regions' scientific development plans.

cs.DL

Berlin: A Quantitative View of the Structure of Institutional Scientific Collaborations

This paper examines the structure of scientific collaborations in a large European metropolitan area. It aims to identify strategic coalitions among organizations in Berlin as a specific case with high institutional and sectoral diversity. By adopting a global, regional and organization based approach we provide a quantitative, exploratory and macro view of this diversity. We use publications data with at least one organization located in Berlin from 1996-2017. We further investigate four members of the Berlin University Alliance (BUA) through their self-represented research profiles comparing it with empirical results of OECD disciplines. Using a bipartite network modeling framework, we are able to move beyond the uncontested trend towards team science and increasing internationalization. Our results show that BUA members shape the structure of scientific collaborations in the region. However, they are not collaborating cohesively in all disciplines. Larger divides exist in some disciplines e.g., Agricultural Sciences and Humanities. Only Medical and Health Sciences have cohesive intraregional collaborations which signals the success of regional cooperation established in 2003. We explain possible underlying factors shaping the observed trends and sectoral and intra-regional groupings. A major methodological contribution of this paper is evaluating coverage and accuracy of different organization name disambiguation techniques.

cs.DL

Merits and Limits: Applying open data to monitor open access publications in bibliometric databases

Identifying and monitoring Open Access (OA) publications might seem a trivial task while practical efforts prove otherwise. Contradictory information arise often depending on metadata employed. We strive to assign OA status to publications in Web of Science (WOS) and Scopus while complementing it with different sources of OA information to resolve contradicting cases. We linked publications from WOS and Scopus via DOIs and ISSNs to Unpaywall, Crossref, DOAJ and ROAD. Only about 50% of articles and reviews from WOS and Scopus could be matched via a DOI to Unpaywall. Matching with Crossref brought 56 distinct licences, which define in many cases the legally binding access status of publications. But only 44% of publications hold only a single licence on Crossref, while more than 50% have no licence information submitted to Crossref. Contrasting OA information from Crossref licences with Unpaywall we found contradictory cases overall amounting to more than 25%, which might be partially explained by (ex-)including green OA. A further manual check found about 17% of OA publications that are not accessible and 15% non-OA publications that are accessible through publishers' websites. These preliminary results suggest that identification of OA state of publications denotes a difficult and currently unfulfilled task.

cs.DL