SearcharxivSearch

arXiv subjects

Rodrigo Costas

Publications and source records attributed to Rodrigo Costas.

At least 37 records · Page 2Linked to original sources

Studying the characteristics of scientific communities using individual-level bibliometrics: the case of Big Data research

Unlike most bibliometric studies focusing on publications, taking Big Data research as a case study, we introduce a novel bibliometric approach to unfold the status of a given scientific community from an individual level perspective. We study the academic age, production, and research focus of the community of authors active in Big Data research. Artificial Intelligence (AI) is selected as a reference area for comparative purposes. Results show that the academic realm of "Big Data" is a growing topic with an expanding community of authors, particularly of new authors every year. Compared to AI, Big Data attracts authors with a longer academic age, who can be regarded to have accumulated some publishing experience before entering the community. Despite the highly skewed distribution of productivity amongst researchers in both communities, Big Data authors have higher values of both research focus and production than those of AI. Considering the community size, overall academic age, and persistence of publishing on the topic, our results support the idea of Big Data as a research topic with attractiveness for researchers. We argue that the community-focused indicators proposed in this study could be generalized to investigate the development and dynamics of other research fields and topics.

cs.DL

Exploring the relevance of ORCID as a source of study of data sharing activities at the individual-level: a methodological discussion

ORCID is a scientific infrastructure created to solve the problem of author name ambiguity. Over the years ORCID has also become a useful source for studying academic activities reported by researchers. Our objective in this research was to use ORCID to analyze one of these research activities: the publication of datasets. We illustrate how the identification of datasets that shared in researchers' ORCID profiles enables the study of the characteristics of the researchers who have produced them. To explore the relevance of ORCID to study data sharing practices we obtained all ORCID profiles reporting at least one dataset in their "works" list, together with information related to the individual researchers producing the datasets. The retrieved data was organized and analyzed in a SQL database hosted at CWTS. Our results indicate that DataCite is by far the most important data source for providing information about datasets recorded in ORCID. There is also a substantial overlap between DataCite records with other repositories (Figshare, Dryad, and Zenodo). The analysis of the distribution of researchers producing datasets shows that the top six countries with more data producers, also have a relatively higher percentage of people who have produced datasets out of total researchers with datasets than researchers in the total ORCID. By disciplines, researchers that belong to the areas of Natural Sciences and Medicine and Life Sciences are those with the largest amount of reported datasets. Finally, we observed that researchers who have started their PhD around 2015 published their first dataset earlier that those researchers that started their PhD before. The work concludes with some reflections of the possibilities of ORCID as a relevant source for research on data sharing practices.

cs.DL

Scholars mobility and its impact on the knowledge producers' workforce of European regions

Knowledge production increasingly relies on mobility. However, its role as a mechanism for knowledge recombination and dissemination remains largely unknown. Based on 1,244,080 Web of Science publications from 1,435,729 authors that we used to construct a panel dataset, we study the impact of inter-regional publishing and scientists' mobility in fostering the workforce composition of European countries during 2008-2017. Specifically, we collect information on scientists who have published in one region and then published elsewhere, and explore some determinants of regional and international mobility. Preliminary findings suggest that while talent pools of researchers are increasingly international, their movements seem to be steered by geographical structures. Future research will investigate the impact of mobility on the regional structure of scientific fields by accounting for the appearance and disappearance of research topics over time.

cs.DL

The role of scientific output in public debates in times of crisis: A case study of the reopening of schools during the COVID-19 pandemic

Situations in which no scientific consensus has been reached due to either insufficient, inconclusive or contradicting findings place strain on governments and public organizations which are forced to take action under circumstances of uncertainty. In this chapter, we focus on the case of COVID-19, its effects on children and the public debate around the reopening of schools. The aim is to better understand the relationship between policy interventions in the face of an uncertain and rapidly changing knowledge landscape and the subsequent use of scientific information in public debates related to the policy interventions. Our approach is to combine scientific information from journal articles and preprints with their appearance in the popular media, including social media. First, we provide a picture of the different scientific areas and approaches, by which the effects of COVID-19 on children are being studied. Second, we identify news media and social media attention around the COVID-19 scientific output related to children and schools. We focus on policies and media responses in three countries: Spain, South Africa and the Netherlands. These countries have followed very different policy actions with regard to the reopening of schools and represent very different policy approaches to the same problem. We analyse the activity in (social) media around the debate between COVID-19, children and school closures by focusing on the use of references to scientific information in the debate. Finally, we analyse the dominant topics that emerge in the news outlets and the online debates. We draw attention to illustrative cases of miscommunication related to scientific output and conclude the chapter by discussing how information from scientific publication, the media and policy actions shape the public discussion in the context of a global health pandemic.

cs.DL

An extensive analysis of the presence of altmetric data for Web of Science publications across subject fields and research topics

Sufficient data presence is one of the key preconditions for applying metrics in practice. Based on both Altmetric.com data and Mendeley data collected up to 2019, this paper presents a state-of-the-art analysis of the presence of 12 kinds of altmetric events for nearly 12.3 million Web of Science publications published between 2012 and 2018. Results show that even though an upward trend of data presence can be observed over time, except for Mendeley readers and Twitter mentions, the overall presence of most altmetric data is still low. The majority of altmetric events go to publications in the fields of Biomedical and Health Sciences, Social Sciences and Humanities, and Life and Earth Sciences. As to research topics, the level of attention received by research topics varies across altmetric data, and specific altmetric data show different preferences for research topics, on the basis of which a framework for identifying hot research topics is proposed and applied to detect research topics with higher levels of attention garnered on certain altmetric data source. Twitter mentions and policy document citations were selected as two examples to identify hot research topics of interest of Twitter users and policy-makers, respectively, shedding light on the potential of altmetric data in monitoring research trends of specific social attention.

cs.DL

Tracking the Twitter attention around the research efforts on the COVID-19 pandemic

The outbreak of the COVID-19 pandemic has been accompanied by a bulk of scientific research and related Twitter discussions. To unravel the public concerns about the COVID-19 crisis reflected in the science-based Twitter conversations, this study tracked the Twitter attention around the COVID-19 research efforts during the first three months of 2020. On the basis of nearly 1.4 million Twitter mentions of 6,162 COVID-19-related scientific publications, we investigated the temporal tweeting dynamic and the Twitter users involved in the online discussions around COVID-19-related research. The results show that the quantity of Twitter mentions of COVID-19-related publications was on rising. Scholarly-oriented Twitter users played an influential role in disseminating research outputs on COVID-19, with their tweets being frequently retweeted. Over time, a change in the focus of the Twitter discussions can be observed, from the initial attention to virological and clinical research to more practical topics, such as the potential treatments, the countermeasures by the governments, the healthcare measures, and the influences on the economy and society, in more recent times.

cs.DL

Open Access uptake by universities worldwide

The implementation of policies promoting the adoption of an Open Science culture must be accompanied by indicators that allow monitoring the penetration of such policies and their potential effects on research publishing and sharing practices. This study presents indicators of Open Access (OA) penetration at the institutional level for universities worldwide. By combining data from Web of Science, Unpaywall and the Leiden Ranking disambiguation of institutions, we track OA coverage of universities' output for 963 institutions. This paper presents the methodological challenges, conceptual discrepancies and limitations and discusses further steps needed to move forward the discussion on fostering Open Access and Open Science practices and policies.

cs.DL

How do academic topics shift across altmetric sources? A case study of the research area of Big Data

Taking the research area of Big Data as a case study, we propose an approach for exploring how academic topics shift through the interactions among audiences across different altmetric sources. Data used is obtained from Web of Science (WoS) and Altmetric.com, with a focus on Blog, News, Policy, Wikipedia, and Twitter. Author keywords from publications and terms from online events are extracted as the main topics of the publications and the online discussion of their audiences at Altmetric. Different measures are applied to determine the (dis)similarities between the topics put forward by the publication authors and those by the online audiences. Results show that overall there are substantial differences between the two sets of topics around Big Data scientific research. The main exception is Twitter, where high-frequency hashtags in tweets have a stronger concordance with the author keywords in publications. Among the online communities, Blogs and News show a strong similarity in the terms commonly used, while Policy documents and Wikipedia articles exhibit the strongest dissimilarity in considering and interpreting Big Data related research. Specifically, the audiences not only focus on more easy-to-understand academic topics related to social or general issues, but also extend them to a broader range of topics in their online discussions. This study lays the foundations for further investigations about the role of online audiences in the transformation of academic topics across altmetric sources, and the degree of concern and reception of scholarly contents by online communities.

cs.DL

Unveiling the research landscape of Sustainable Development Goals and their inclusion in Higher Education Institutions and Research Centers: major trends in 2000-2017

Sustainable Development Goals are the blueprint to achieve a better and more sustainable future for society. Its legacy is linked with the Millennium Development Goals, set up in 2000. A bibliometric analysis was conducted to 1) measure "core" research output from 2000-2017, with the aim to map the global research of sustainability goals, 2) describe thematic specialization based on keywords co-occurrence analysis and strongest citation burst, 3) present a methodology to classify scientific output (based on an ad-hoc glossary) and assess SDGs interconnections. Sustainability goals publications (core+expand based on direct citations) were identified in-house CWTS Web of Science by using search terms in titles, abstracts, and keywords. 25,299 bibliographic records were analyzed, from which 21,653 (85.59%) are from HEIs and research centres (RC). The purpose of this paper is to analyze the role of these organizations in sustainability research. The findings reveal the increasing participation of these organizations in this research (660 institutions in 2000-2005 to 1744 institutions involved in 2012-2017). In terms of specialization, some institutions present a higher production and specialization on the topic (e.g., London School of Hygiene & Tropical Medicine and World Health Organization); however, others present less production but higher specialization (e.g., Stockholm Environment Institute). Regarding the topics, health (especially in developing countries), women and socio-economic aspects are the most prominent ones. Moreover, it is observed the interlinked nature of SDGs between some SDGs in research output (e.g., SDG11 and SDG3). This study provides important orientation for HEIs and RCs in terms of Research, Development and Innovation (R&D+i) to respond to major societal challenges and could be useful for the policymakers in order to promote the research agenda on this topic.

cs.DL

The stability of Twitter metrics: A study on unavailable Twitter mentions of scientific publications

This paper investigates the stability of Twitter counts of scientific publications over time. For this, we conducted an analysis of the availability statuses of over 2.6 million Twitter mentions received by the 1,154 most tweeted scientific publications recorded by Altmetric.com up to October 2017. Results show that of the Twitter mentions for these highly tweeted publications, about 14.3% have become unavailable by April 2019. Deletion of tweets by users is the main reason for unavailability, followed by suspension and protection of Twitter user accounts. This study proposes two measures for describing the Twitter dissemination structures of publications: Degree of Originality (i.e., the proportion of original tweets received by a paper) and Degree of Concentration (i.e., the degree to which retweets concentrate on a single original tweet). Twitter metrics of publications with relatively low Degree of Originality and relatively high Degree of Concentration are observed to be at greater risk of becoming unstable due to the potential disappearance of their Twitter mentions. In light of these results, we emphasize the importance of paying attention to the potential risk of unstable Twitter counts, and the significance of identifying the different Twitter dissemination structures when studying the Twitter metrics of scientific publications.

cs.DL

Indicators of Open Access for universities

This paper presents a first attempt to analyse Open Access integration at the institutional level. For this, we combine information from Unpaywall and the Leiden Ranking to offer basic OA indicators for universities. We calculate the overall number of Open Access publications for 930 universities worldwide. OA indicators are also disaggregated by green, gold and hybrid Open Access. We then explore differences between and within countries and offer a general ranking of universities based on the proportion of their output which is openly accessible.

cs.DL

Making sense of global collaboration dynamics: Developing a methodological framework to study (dis)similarities between country disciplinary profiles and choice of collaboration partners

This paper presents a novel methodological framework by which the effects of globalization on international collaboration can be studied and understood. Using the cosine similarity of the disciplinary and partner profiles of countries by collaboration types it is possible to analyse the effects of globalization and the costs and benefits of an increasing global networked research system.

cs.DL

The many faces of mobility: Using bibliometric data to measure the movement of scientists

This paper presents a methodological framework for developing scientific mobility indicators based on bibliometric data. We identify nearly 16 million individual authors from publications covered in the Web of Science for the 2008-2015 period. Based on the information provided across individuals' publication records, we propose a general classification for analyzing scientific mobility using institutional affiliation changes. We distinguish between migrants--authors who have ruptures with their country of origin--and travelers--authors who gain additional affiliations while maintaining affiliation with their country of origin. We find that 3.7 percent of researchers who have published at least one paper over the period are mobile. Travelers represent 72.7 percent of all mobile scholars, but migrants have higher scientific impact. We apply this classification at the country level, expanding the classification to incorporate the directionality of scientists' mobility (i.e., incoming and outgoing). We provide a brief analysis to highlight the utility of the proposed taxonomy to study scholarly mobility and discuss the implications for science policy.

cs.DL

Unbundling Open Access dimensions: a conceptual discussion to reduce terminology inconsistencies

The current ways in which documents are made freely accessible in the Web no longer adhere to the models established Budapest/Bethesda/Berlin (BBB) definitions of Open Access (OA). Since those definitions were established, OA-related terminology has expanded, trying to keep up with all the variants of OA publishing that are out there. However, the inconsistent and arbitrary terminology that is being used to refer to these variants are complicating communication about OA-related issues. This study intends to initiate a discussion on this issue, by proposing a conceptual model of OA. Our model features six different dimensions (prestige, user rights, stability, immediacy, peer-review, and cost). Each dimension allows for a range of different options. We believe that by combining the options in these six dimensions, we can arrive at all the current variants of OA, while avoiding ambiguous and/or arbitrary terminology. This model can be an useful tool for funders and policy makers who need to decide exactly which aspects of OA are necessary for each specific scenario.

cs.DL

Evidence of Open Access of scientific publications in Google Scholar: a large-scale analysis

This article uses Google Scholar (GS) as a source of data to analyse Open Access (OA) levels across all countries and fields of research. All articles and reviews with a DOI and published in 2009 or 2014 and covered by the three main citation indexes in the Web of Science (2,269,022 documents) were selected for study. The links to freely available versions of these documents displayed in GS were collected. To differentiate between more reliable (sustainable and legal) forms of access and less reliable ones, the data extracted from GS was combined with information available in DOAJ, CrossRef, OpenDOAR, and ROAR. This allowed us to distinguish the percentage of documents in our sample that are made OA by the publisher (23.1%, including Gold, Hybrid, Delayed, and Bronze OA) from those available as Green OA (17.6%), and those available from other sources (40.6%, mainly due to ResearchGate). The data shows an overall free availability of 54.6%, with important differences at the country and subject category levels. The data extracted from GS yielded very similar results to those found by other studies that analysed similar samples of documents, but employed different methods to find evidence of OA, thus suggesting a relative consistency among methods.

cs.DL

Social media metrics for new research evaluation

This chapter approaches, both from a theoretical and practical perspective, the most important principles and conceptual frameworks that can be considered in the application of social media metrics for scientific evaluation. We propose conceptually valid uses for social media metrics in research evaluation. The chapter discusses frameworks and uses of these metrics as well as principles and recommendations for the consideration and application of current (and potentially new) metrics in research evaluation.

cs.DL

Scientific mobility indicators in practice: International mobility profiles at the country level

This paper presents and describes the methodological opportunities offered by bibliometric data to produce indicators of scientific mobility. Large bibliographic datasets of disambiguated authors and their affiliations allow for the possibility of tracking the affiliation changes of scientists. Using the Web of Science as data source, we analyze the distribution of types of mobile scientists for a selection of countries. We explore the possibility of creating profiles of international mobility at the country level, and discuss potential interpretations and caveats. Five countries (Canada, The Netherlands, South Africa, Spain, and the United States) are used as examples. These profiles enable us to characterize these countries in terms of their strongest links with other countries. This type of analysis reveals circulation among and between countries with strong policy implications.

cs.DL

Developing indicators on Open Access by combining evidence from diverse data sources

In the last couple of years, the role of Open Access (OA) publishing has become central in science management and research policy. In the UK and the Netherlands, national OA mandates require the scientific community to seriously consider publishing research outputs in OA forms. At the same time, other elements of Open Science are becoming also part of the debate, thus including not only publishing research outputs but also other related aspects of the chain of scientific knowledge production such as open peer review and open data. From a research management point of view, it is important to keep track of the progress made in the OA publishing debate. Until now, this has been quite problematic, given the fact that OA as a topic is hard to grasp by bibliometric methods, as most databases supporting bibliometric data lack exhaustive and accurate open access labelling of scientific publications. In this study, we present a methodology that systematically creates OA labels for large sets of publications processed in the Web of Science database. The methodology is based on the combination of diverse data sources that provide evidence of publications being OA

cs.DL