The function of the organism hinges on the performance of its information-processing networks, which convey information via molecular recognition. Many paths within these networks utilize molecular codebooks, such as the genetic code, to translate information written in one class of molecules into another molecular "language" . The present paper examines the emergence and evolution of molecular codes in terms of rate-distortion theory and reviews recent results of this approach. We discuss how the biological problem of maximizing the fitness of an organism by optimizing its molecular coding machinery is equivalent to the communication engineering problem of designing an optimal information channel. The fitness of a molecular code takes into account the interplay between the quality of the channel and the cost of resources which the organism needs to invest in its construction and maintenance. We analyze the dynamics of a population of organisms that compete according to the fitness of their codes. The model suggests a generic mechanism for the emergence of molecular codes as a phase transition in an information channel. This mechanism is put into biological context and demonstrated
We argue that the communication structures in the Chinese social sciences have not yet been sufficiently reformed. Citation patterns among Chinese domestic journals in three subject areas -- political science and marxism, library and information science, and economics -- are compared with their counterparts internationally. Like their colleagues in the natural and life sciences, Chinese scholars in the social sciences provide fewer references to journal publications than their international counterparts; like their international colleagues, social scientists provide fewer references than natural sciences. The resulting citation networks, therefore, are sparse. Nevertheless, the citation structures clearly suggest that the Chinese social sciences are far less specialized in terms of disciplinary delineations than their international counterparts. Marxism studies are more established than political science in China. In terms of the impact of the Chinese political system on academic fields, disciplines closely related to the political system are less specialized than those weakly related. In the discussion section, we explore reasons that may cause the current stagnation and provide p
The Journal Citation Reports of the Science Citation Index 2004 were used to delineate a core set of nanotechnology journals and a nanotechnology-relevant set. In comparison with 2003, the core set has grown and the relevant set has decreased. This suggests a higher degree of codification in the field of nanotechnology: the field has become more focused in terms of citation practices. Using the citing patterns among journals at the aggregate level, a core group of ten nanotechnology journals in the vector space can be delineated on the criterion of betweenness centrality. National contributions to this core group of journals are evaluated for the years 2003, 2004, and 2005. Additionally, the specific class of nanotechnology patents in the database of the U.S. Patent and Trade Office (USPTO) is analyzed to determine if non-patent literature references can be used as a source for the delineation of the knowledge base in terms of scientific journals. The references are primarily to general science journals and letters, and therefore not specific enough for the purpose of delineating a journal set.
In recent decades, the relevance of polarimetry in planetary sciences and astronomy has increased rapidly. Polarization is a fundamental property of light and can be modified by any scattering event. As such, polarization yields additional information that cannot be obtained by only assessing light's scalar properties. For instance, the polarization state of starlight scattered by planetary surfaces can provide useful insights on the composition, size, morphology, and porosity of regolith particles and might even indicate the presence of life. Beside being useful for characterization, polarimetry can also greatly enhance the detection of exoplanets. Here, polarization can be harnessed to enhance the contrast between the bright light of a star, which can be considered to be fully unpolarized, and the very dim but polarized light reflected by an exoplanet. In this paper, we discuss and review the current developments and advances in optical polarimetry and polarimetric instrumentation in Switzerland within the framework of the National Centre of Competence in Research PlanetS. We focus on their implications for the vast range of science cases that polarimetry can address within the r
This study compares publication pattern dynamics in the social sciences and humanities in five European countries. Three are Central and Eastern European countries that share a similar cultural and political heritage (the Czech Republic, Slovakia, and Poland). The other two are Flanders (Belgium) and Norway, representing Western Europe and the Nordics, respectively. We analysed 449,409 publications from 2013-2016 and found that, despite persisting differences between the two groups of countries across all disciplines, publication patterns in the Central and Eastern European countries are becoming more similar to those in their Western and Nordic counterparts. Articles from the Central and Eastern European countries are increasingly published in journals indexed in Web of Science and also in journals with the highest citation impacts. There are, however, clear differences between social science and humanities disciplines, which need to be considered in research evaluation and science policy.
Small bodies, the unaccreted leftovers of planetary formation, are often mistaken for the leftovers of planetary science in the sense that they are everything else after the planets and their satellites (or sometimes just their regular satellites) are accounted for. This mistaken view elides the great diversity of compositions, histories, and present-day conditions and processes found in the small bodies, and the interdisciplinary nature of their study. Understanding small bodies is critical to planetary science as a field, and we urge planetary scientists and our decision makers to continue to support science-based mission selections and to recognize that while small bodies have been grouped together for convenience, the diversity of these objects in terms of composition, mass, differentiation, evolution, activity, dynamical state, physical structure, thermal environment, thermal history, and formation vastly exceeds the observed variability in the major planets and their satellites. Treating them as a monolithic group with interchangeable members does a grave injustice to the range of fundamental questions they address. We advocate for a deep and ongoing program of missions, tele
The journal structure in the China Scientific and Technical Papers and Citations Database (CSTPCD) is analysed from three perspectives: the database level, the specialty level and the institutional level (i.e., university journals versus journals issued by the Chinese Academy of Sciences). The results are compared with those for (Chinese) journals included in the Science Citation Index. The frequency of journal-journal citation relations in the CSTPCD is an order of magnitude lower than in the SCI. Chinese journals, especially high-quality journals, prefer to cite international journals rather than domestic ones. However, Chinese journals do not get an equivalent reception from their international counterparts. The international visibility of Chinese journals is low, but varies among fields of science. Journals of the Chinese Academy of Sciences (CAS) have a better reception in the international scientific community than university journals.
In the field of scientometrics, the subject classification system of academic journals holds great importance. Accurate identification and classification of "multidisciplinary" journals are crucial in revealing the scientific structure and evaluating journals. Based on data from the Web of Science database from 2016 to 2020, we calculated the disciplinary diversity of journals using the paper-level subject classification system, then conducted a systematic analysis of JCR multidisciplinary journals. Studies showed that most multidisciplinary journals have high disciplinary diversity, while non-multidisciplinary journals tend to have relatively lower diversity. Some multidisciplinary journals with low disciplinary diversities may misclassify disciplines. In addition, there are inconsistencies in the diversity of journal disciplines at different granularities. Our study also visually analyzed the four types of diversity distribution tendencies of multidisciplinary journals. Moreover, ten potential multidisciplinary journals were found in non-multidisciplinary categories.
Do different fields of knowledge require different research strategies? A numerical model exploring different virtual knowledge landscapes, revealed two diverging optimal search strategies. Trend following is maximized when the popularity of new discoveries determine the number of individuals researching it. This strategy works best when many researchers explore few large areas of knowledge. In contrast, individuals or small groups of researchers are better in discovering small bits of information in dispersed knowledge landscapes. Bibliometric data of scientific publications showed a continuous bipolar distribution of these strategies, ranging from natural sciences, with highly cited publications in journals containing a large number of articles, to the social sciences, with rarely cited publications in many journals containing a small number of articles. The natural sciences seem to adapt their research strategies to landscapes with large concentrated knowledge clusters, whereas social sciences seem to have adapted to search in landscapes with many small isolated knowledge clusters. Similar bipolar distributions were obtained when comparing levels of insularity estimated by indic
Current research in High Energy Cosmic Ray Physics touches on fundamental questions regarding the origin of cosmic rays, their composition, the acceleration mechanisms, and their production. Unambiguous measurements of the energy spectra and of the composition of cosmic rays at the "knee" region could provide some of the answers to the above questions. So far only ground based observations, which rely on sophisticated models describing high energy interactions in the earth's atmosphere, have been possible due to the extremely low particle rates at these energies. A calorimetry based space experiment that could provide not only flux measurements but also energy spectra and particle identification, would certainly overcome some of the uncertainties of ground based experiments. Given the expected particle fluxes, a very large acceptance is needed to collect a sufficient quantity of data, in a time compatible with the duration of a space mission. This in turn, contrasts with the lightness and compactness requirements for space based experiments. We present a novel idea in calorimetry which addresses these issues whilst limiting the mass and volume of the detector. In this paper we repo
Symbolic regression (SR) has emerged as a powerful method for uncovering interpretable mathematical relationships from data, offering a novel route to both scientific discovery and efficient empirical modelling. This article introduces the Special Issue on Symbolic Regression for the Physical Sciences, motivated by the Royal Society discussion meeting held in April 2025. The contributions collected here span applications from automated equation discovery and emergent-phenomena modelling to the construction of compact emulators for computationally expensive simulations. The introductory review outlines the conceptual foundations of SR, contrasts it with conventional regression approaches, and surveys its main use cases in the physical sciences, including the derivation of effective theories, empirical functional forms and surrogate models. We summarise methodological considerations such as search-space design, operator selection, complexity control, feature selection, and integration with modern AI approaches. We also highlight ongoing challenges, including scalability, robustness to noise, overfitting and computational complexity. Finally we emphasise emerging directions, particula
We study gender representation on the editorial boards of 435 journals in the mathematical sciences. Women are known to comprise approximately 15% of tenure-stream faculty positions in doctoral-granting mathematical sciences departments in the United States. Compared to this pool, the likely source of journal editorships, we find that 8.9% of the 13067 editorships in our study are held by women. We describe group variations within the editorships by identifying specific journals, subfields, publishers, and countries that significantly exceed or fall short of this average. To enable our study, we develop a semi-automated method for inferring gender that has an estimated accuracy of 97.5%. Our findings provide the first measure of gender distribution on editorial boards in the mathematical sciences, offer insights that suggest future studies in the mathematical sciences, and introduce new methods that enable large-scale studies of gender distribution in other fields.
The internationalization characteristics of two Malaysian journals, Bulletin of the Malaysian Mathematical Sciences Society (indexed by ISI) and the Malaysian Journal of Computer Science (indexed by Inspec and Scopus) is observed. All issues for the years 2000 to 2007 were looked at to obtain the following information, (i) total articles published between 2000 and 2007; (ii) the distribution of foreign and Malaysian authors publishing in the journals; (iii) the distribution of articles by country and (iv) the geographical distribution of authors citing articles published in the journals. Citation to articles is derived from information given by Google scholar. The results indicate that both journals exhibit average internationalization characteristics as they are current in their publications but with between 19% -30% international composition of reviewers or editorials, publish between 36%-79% of foreign articles and receive between 60%-70% of citations from foreign authors.
Background: Introduced in 2010, the sub-discipline of gerontologic biostatistics (GBS) was conceptualized to address the specific challenges in analyzing data from research studies involving older adults. However, the evolving technological landscape has catalyzed data science and statistical advancements since the original GBS publication, greatly expanding the scope of gerontologic research. There is a need to describe how these advancements enhance the analysis of multi-modal data and complex phenotypes that are hallmarks of gerontologic research. Methods: This paper introduces GBS 2.0, an updated and expanded set of analytical methods reflective of the practice of gerontologic biostatistics in contemporary and future research. Results: GBS 2.0 topics and relevant software resources include cutting-edge methods in experimental design; analytical techniques that include adaptations of machine learning, quantifying deep phenotypic measurements, high-dimensional -omics analysis; the integration of information from multiple studies, and strategies to foster reproducibility, replicability, and open science. Discussion: The methodological topics presented here seek to update and expan
Proposed as blanket structural materials for fusion power reactors, reduced activation ferritic/martensitic (RAFM) steel undergoes volume expanding and contracting in a cyclic mode under service environment. Particularly, being subjected to significant fluxes of fusion neutrons RAFM steel suffers considerable local volume variations in the radiation damage involved regions. It is necessary to study the structure properties of the alloying elements in contraction and expansion states. In this paper we studied local substitution structures of thirteen alloying elements Al, Co, Cr, Cu, Mn, Mo, Nb, Ni, Si, Ta, Ti, V, and W in bcc Fe and calculated their substitutional energies in the volume variation range from -1.0% to 1.0%. From the structure relaxation results of the first five neighbor shells around the substitutional atom we find the relaxation in each neighbor shell keeps approximately uniform within the volume variation from -1.0% to 1.0% except those of Mn and the relaxation of the fifth neighbor shell is stronger than that of the third and forth, indicating that the lattice distortion due to the substitution atom is easier to spread in <111> direction than in other direc
We present the results of processing the effects of the powerful Gamma Ray Burst GRB221009A captured by the charged particle detectors (electrostatic analyzers and solid-state detectors) onboard spacecraft at different points in the heliosphere on October 9, 2022. To follow the GRB221009A propagation through the heliosphere we used the electron and proton flux measurements from solar missions Solar Orbiter and STEREO-A; Earth magnetosphere and the solar wind missions THEMIS and Wind; meteorological satellites POES15, POES19, MetOp3; and MAVEN - a NASA mission orbiting Mars. GRB221009A had a structure of four bursts: less intense Pulse 1 - the triggering impulse - was detected by gamma-ray observatories at 131659 UT (near the Earth); the most intense Pulses 2 and 3 were detected on board all the spacecraft from the list, and Pulse 4 detected in more than 500 s after Pulse 1. Due to their different scientific objectives, the spacecraft, which data was used in this study, were separated by more than 1 AU (Solar Orbiter and MAVEN). This enabled tracking GRB221009A as it was propagating across the heliosphere. STEREO-A was the first to register Pulse 2 and 3 of the GRB, almost 100 secon
This paper investigates the Twitter interaction patterns of journals from the Science Citation Index (SCI) of Master Journal List (MJL). A total of 953,253 tweets extracted from 857 journal accounts, were analyzed in this study. Findings indicate that SCI journals interacted more with each other but much less with journals from other citation indices. The network structure of the communication graph resembled a tight crowd network, with Nature journals playing a major part. Information sources such as news portals and scientific organizations were mentioned more in tweets, than academic journal Twitter accounts. Journals with high journal impact factors (JIFs) were found to be prominent hubs in the communication graph. Differences were found between the Twitter usage of SCI journals with Humanities and Social Sciences (HSS) journals.
Labeling or classifying time series is a persistent challenge in the physical sciences, where expert annotations are scarce, costly, and often inconsistent. Yet robust labeling is essential to enable machine learning models for understanding, prediction, and forecasting. We present the \textit{Clustering and Indexation Pipeline with Human Evaluation for Recognition} (CIPHER), a framework designed to accelerate large-scale labeling of complex time series in physics. CIPHER integrates \textit{indexable Symbolic Aggregate approXimation} (iSAX) for interpretable compression and indexing, density-based clustering (HDBSCAN) to group recurring phenomena, and a human-in-the-loop step for efficient expert validation. Representative samples are labeled by domain scientists, and these annotations are propagated across clusters to yield systematic, scalable classifications. We evaluate CIPHER on the task of classifying solar wind phenomena in OMNI data, a central challenge in space weather research, showing that the framework recovers meaningful phenomena such as coronal mass ejections and stream interaction regions. Beyond this case study, CIPHER highlights a general strategy for combining sy
Social Network Analysis is a way of studying agents embedded in contexts. In about 1998, physicists discovered social networks as representations of complex systems. Small-world and scale-free networks are the paradigmatic models of this Network Science. Relying on various models and mechanisms of socio-cultural processes, an identity model is developed and calibrated in a case study of Social Network Science. This research domain results from the union of Social Network Analysis and Network Science. A unique dataset of 25,760 scholarly articles from one century of research (1916-2012) is created. Clustering this set of publications, five subdomains are detected and analyzed in terms of authorship, citation, and word usage structures and dynamics. The scaling hypothesis of percolation theory is formulated for socio-cultural systems, namely that power-law size distributions like Lotka's, Bradford's, and Zipf's Law mean that the described identity resides at the phase transition between the stability and change of meaning. In this case, it can be diagnosed using bivariate scaling laws and Abbott's heuristic of fractal distinctions. Identities are not dichotomies but dualities of soci
Artificial intelligence (AI) models trained using medical images for clinical tasks often exhibit bias in the form of disparities in performance between subgroups. Since not all sources of biases in real-world medical imaging data are easily identifiable, it is challenging to comprehensively assess how those biases are encoded in models, and how capable bias mitigation methods are at ameliorating performance disparities. In this article, we introduce a novel analysis framework for systematically and objectively investigating the impact of biases in medical images on AI models. We developed and tested this framework for conducting controlled in silico trials to assess bias in medical imaging AI using a tool for generating synthetic magnetic resonance images with known disease effects and sources of bias. The feasibility is showcased by using three counterfactual bias scenarios to measure the impact of simulated bias effects on a convolutional neural network (CNN) classifier and the efficacy of three bias mitigation strategies. The analysis revealed that the simulated biases resulted in expected subgroup performance disparities when the CNN was trained on the synthetic datasets. More