This paper takes James David Forbes' Encyclopaedia Britannica entry, Dissertation Sixth, as a lens to examine physics as a cognitive, practical, and social, enterprise. Forbes wrote this survey of eighteenth- and nineteenth-century mathematical and physical sciences, in 1852-6, when British "physics" was at a pivotal point in its history, situated between a discipline identified by its mathematical methods - originating in France - and one identified by its university laboratory institutions. Contemporary encyclopaedias provided a nexus for publishers, the book trade, readers, and men of science, in the formation of physics as a field. Forbes was both a witness, whose account of the progress of physics or natural philosophy can be explored at face value, and an agent, who exploited the opportunity offered by the Encyclopaedia Britannica in the mid nineteenth-century to enrol the broadly educated public, and scientific collective, illuminating the connection between the definition of physics and its forms of social practice. Forbes used the terms "physics" and "natural philosophy" interchangeably. He portrayed the field as progressed by the natural genius of great men, who curated t
The current flow of high accuracy astrophysical data, among which are the Cosmic Microwave Background (CMB) measurements by the Planck satellite, offers an unprecedented opportunity to constrain the inflationary theory. This is however a challenging project given the size of the inflationary landscape which contains hundreds of different scenarios. Given that there is currently no observational evidence for primordial non-Gaussianities, isocurvature perturbations or any other non-minimal extension of the inflationary paradigm, a reasonable approach is to consider the simplest models first, namely the slow-roll single field models with minimal kinetic terms. This still leaves us with a very populated landscape, the exploration of which requires new and efficient strategies. It has been customary to tackle this problem by means of approximate model independent methods while a more ambitious alternative is to study the inflationary scenarios one by one. We have developed the publicly available runtime library ASPIC to implement this last approach. The ASPIC code provides all routines needed to quickly derive reheating-consistent observable predictions within this class of scenarios. A
We quantify radial mixing in exoplanet hosts and explore links between birth environment, orbital evolution, planetary architecture, and Galactic habitability. We constructed a homogeneous catalogue by cross-matching the Encyclopaedia of Exoplanetary Systems with Gaia DR3 astrometry and infrared photometry from 2MASS and AllWISE. Stellar orbits were integrated using Galpy. Stellar birth radii were inferred by combining Galactic chemical enrichment models with the generalised additive model introduced in Paper I. Giant-planet hosts preferentially trace inner-Galaxy birth sites, whereas brown-dwarf hosts span a broader, less localised range of radial displacements. Rocky-only systems show smaller radial excursions and less centrally concentrated birth radii, while rocky+giant systems are intermediate, retaining a stronger link to inner-disc birth environments than rocky-only systems. We also find that outward-migrators host more compact outer detected companions than inward-migrators, with non-migrators in between. This trend remains tentative because of heterogeneous detection biases. Giant-planet hosts retain a strong connection to metal-rich inner-Galaxy birth environments, wherea
This paper introduces a dataset of enriched geographic coordinates retrieved from Diderot and d'Alembert's eighteenth-century Encyclopedie. Automatically recovering geographic coordinates from historical texts is a complex task, as they are expressed in a variety of ways and with varying levels of precision. To improve retrieval of coordinates from similar digitized early modern texts, we have created a gold standard dataset, trained models, published the resulting inferred and normalized coordinate data, and experimented applying these models to new texts. From 74,000 total articles in each of the digitized versions of the Encyclopedie from ARTFL and ENCCRE, we examined 15,278 geographical entries, manually identifying 4,798 containing coordinates, and 10,480 with descriptive but non-numerical references. Leveraging our gold standard annotations, we trained transformer-based models to retrieve and normalize coordinates. The pipeline presented here combines a classifier to identify coordinate-bearing entries and a second model for retrieval, tested across encoder-decoder and decoder architectures. Cross-validation yielded an 86% EM score. On an out-of-domain eighteenth-century Trev
Population III (or Pop. III) stars, the first stellar generation built up from metal-free primordial gas, first started to form at redshifts z ~ 30. They formed primarily in small dark matter halos with masses of a few million solar masses. The cooling of the gas in these halos was dominated on all scales by molecular hydrogen. Current theoretical models indicate that Pop. III stars typically formed in small clusters with a logarithmically flat mass function due to widespread fragmentation in the protostellar accretion disks around these primordial stars. Massive Pop. III stars are thought to have played a pivotal role in shaping the early Universe, as their feedback regulates subsequent star formation, although the immediate effects of this feedback remain uncertain. Direct detection of Pop. III stars is challenging, but our chances of detecting at least a few Pop. III supernovae within the next decade are brighter. Indirect approaches based on stellar archaeology or gravitational wave detections offer promising constraints. Current observations suggest that most massive Pop. III stars ended their lives as core-collapse supernovae rather than pair-instability supernovae, offering
We study the problem of generating interesting integer sequences with a combinatorial interpretation. For this we introduce a two-step approach. In the first step, we generate first-order logic sentences which define some combinatorial objects, e.g., undirected graphs, permutations, matchings etc. In the second step, we use algorithms for lifted first-order model counting to generate integer sequences that count the objects encoded by the first-order logic formulas generated in the first step. For instance, if the first-order sentence defines permutations then the generated integer sequence is the sequence of factorial numbers $n!$. We demonstrate that our approach is able to generate interesting new sequences by showing that a non-negligible fraction of the automatically generated sequences can actually be found in the Online Encyclopaedia of Integer Sequences (OEIS) while generating many other similar sequences which are not present in OEIS and which are potentially interesting. A key technical contribution of our work is the method for generation of first-order logic sentences which is able to drastically prune the space of sentences by discarding large fraction of sentences whi
The dataset focuses on Wikipedia users and contains information about demographic and socioeconomic characteristics of the respondents and their activity on Wikipedia. The data was collected using a questionnaire available online between June and July 2023. The link to the questionnaire was distributed via a banner published in 8 languages on the Wikipedia page. Filling out the questionnaire was voluntary and not incentivised in any way. The survey includes 200 questions about: what people were doing on Wikipedia before clicking the link to the questionnaire; how they use Wikipedia as readers (``professional'' and ``personal'' uses); their opinion on the quality, the thematic coverage, the importance of the encyclopaedia; the making of Wikipedia (how they think it is made, if they have ever contributed and how); their social, sport, artistic and cultural activities, both online and offline; their socio-economic characteristics including political beliefs, and trust propensities. More than 200 000 people opened the questionnaire, 100 332 started to answer, and constitute our dataset, and 10 576 finished it. Among other themes identified by future researchers, the dataset can be usef
Spanning two decades, the Encyclopaedia of DNA Elements (ENCODE) is a collaborative research project that aims to identify all the functional elements in the human and mouse genomes. To best serve the scientific community, all data generated by the consortium is shared through a web-portal (https://www.encodeproject.org/) with no access restrictions. The fourth and final phase of the project added a diverse set of new samples (including those associated with human disease), and a wide range of new assays aimed at detection, characterization and validation of functional genomic elements. The ENCODE data portal hosts results from over 23,000 functional genomics experiments, over 800 functional elements characterization experiments (including in vivo transgenic enhancer assays, reporter assays and CRISPR screens) along with over 60,000 results of computational and integrative analyses (including imputations, predictions and genome annotations). The ENCODE Data Coordination Center (DCC) is responsible for development and maintenance of the data portal, along with the implementation and utilisation of the ENCODE uniform processing pipelines to generate uniformly processed data. Here we
We investigate the implications for inflation of the detection of B-modes polarization in the Cosmic Microwave Background (CMB) by BICEP2. We show that the hypothesis of primordial origin of the measurement is only favored by the first four bandpowers, while the others would prefer unreasonably large values of the tensor-to-scalar ratio. Using only those four bandpowers, we carry out a complete analysis in the cosmological and inflationary slow-roll parameter space using the BICEP2 polarization measurements alone and extract the Bayesian evidences and complexities for all the Encyclopaedia Inflationaris models. This allows us to determine the most probable and simplest BICEP2 inflationary scenarios. Although this list contains the simplest monomial potentials, it also includes many other scenarios, suggesting that focusing model building efforts on large field models only is unjustified at this stage. We demonstrate that the sets of inflationary models preferred by Planck alone and BICEP2 alone are almost disjoint, indicating a clear tension between the two data sets. We address this tension with a Bayesian measure of compatibility between BICEP2 and Planck. We find that for models
Wikis can be considered as public domain knowledge sharing system. They provide opportunity for those who may not have the privilege to publish their thoughts through the traditional methods. They are one of the fastest growing systems of online encyclopaedia. In this study, we consider the importance of wikis as a way of creating, sharing and improving public knowledge. We identify some of the problems associated with wikis to include, (a) identification of the identities of information and its creator (b) accuracy of information (c) justification of the credibility of authors (d) vandalism of quality of information (e) weak control over the contents. A solution to some of these problems is sought through the use of an annotation model. The model assumes that contributions in wikis can be seen as annotation to the initial document. It proposed a systematic control of contributors and contributions to the initiative and the keeping of records of what existed and what was done to initial documents. We believe that with this model, analysis can be done on the progress of wiki initiatives. We assumed that using this model, wikis can be better used for creation and sharing of knowledge
In this paper an open-domain factoid question answering system for Polish, RAFAEL, is presented. The system goes beyond finding an answering sentence; it also extracts a single string, corresponding to the required entity. Herein the focus is placed on different approaches to entity recognition, essential for retrieving information matching question constraints. Apart from traditional approach, including named entity recognition (NER) solutions, a novel technique, called Deep Entity Recognition (DeepER), is introduced and implemented. It allows a comprehensive search of all forms of entity references matching a given WordNet synset (e.g. an impressionist), based on a previously assembled entity library. It has been created by analysing the first sentences of encyclopaedia entries and disambiguation and redirect pages. DeepER also provides automatic evaluation, which makes possible numerous experiments, including over a thousand questions from a quiz TV show answered on the grounds of Polish Wikipedia. The final results of a manual evaluation on a separate question set show that the strength of DeepER approach lies in its ability to answer questions that demand answers beyond the tr