共找到 20 条结果
We present a two-stage pipeline for AI-assisted improvement of published algorithm implementations. In the first stage, a large language model with research capabilities identifies recently published algorithms satisfying explicit experimental criteria. In the second stage, Claude Code is given a prompt to reproduce the reported baseline and then iterate an improvement process. We apply this pipeline to published algorithm implementations spanning multiple research domains. Claude Code reported that all eleven experiments yielded improvements. Each improvement could be achieved within a single working day. We analyse the human contributions that remain indispensable, including selecting the target, verifying experimental validity, assessing novelty and impact, providing computational resources, and writing with appropriate AI-use disclosure. Finally, we discuss implications for peer review and academic publishing.
We study the problem of reconstructing tabular data from aggregate statistics, in which the attacker aims to identify interesting claims about the sensitive data that can be verified with 100% certainty given the aggregates. Successful attempts in prior work have conducted studies in settings where the set of published statistics is rich enough that entire datasets can be reconstructed with certainty. In our work, we instead focus on the regime where many possible datasets match the published statistics, making it impossible to reconstruct the entire private dataset perfectly (i.e., when approaches in prior work fail). We propose the problem of partial data reconstruction, in which the goal of the adversary is to instead output a $\textit{subset}$ of rows and/or columns that are $\textit{guaranteed to be correct}$. We introduce a novel integer programming approach that first $\textbf{generates}$ a set of claims and then $\textbf{verifies}$ whether each claim holds for all possible datasets consistent with the published aggregates. We evaluate our approach on the housing-level microdata from the U.S. Decennial Census release, demonstrating that privacy violations can still persist e
Proper citation is of great importance in academic writing for it enables knowledge accumulation and maintains academic integrity. However, citing properly is not an easy task. For published scientific entities, the ever-growing academic publications and over-familiarity of terms easily lead to missing citations. To deal with this situation, we design a special method Citation Recommendation for Published Scientific Entity (CRPSE) based on the cooccurrences between published scientific entities and in-text citations in the same sentences from previous researchers. Experimental outcomes show the effectiveness of our method in recommending the source papers for published scientific entities. We further conduct a statistical analysis on missing citations among papers published in prestigious computer science conferences in 2020. In the 12,278 papers collected, 475 published scientific entities of computer science and mathematics are found to have missing citations. Many entities mentioned without citations are found to be well-accepted research results. On a median basis, the papers proposing these published scientific entities with missing citations were published 8 years ago, which
In this Note, we provide the comments on the paper [A note on the hit problem for the polynomial algebra in the case of odd primes and its application] which was published in the RACSAM [Rev. Real. Acad. Cienc. Exactas. Fis. Nat. Ser. A-Mat. 118, 22 (2024)]. We will show that some of the results in this paper are not new and others are false.
A plethora of scientific software packages are published in repositories, e.g., Zenodo and figshare. These software packages are crucial for the reproducibility of published research. As an additional route to scholarly knowledge graph construction, we propose an approach for automated extraction of machine actionable (structured) scholarly knowledge from published software packages by static analysis of their (meta)data and contents (in particular scripts in languages such as Python). The approach can be summarized as follows. First, we extract metadata information (software description, programming languages, related references) from software packages by leveraging the Software Metadata Extraction Framework (SOMEF) and the GitHub API. Second, we analyze the extracted metadata to find the research articles associated with the corresponding software repository. Third, for software contained in published packages, we create and analyze the Abstract Syntax Tree (AST) representation to extract information about the procedures performed on data. Fourth, we search the extracted information in the full text of related articles to constrain the extracted information to scholarly knowledge
Academic publishers claim that they add value to scholarly communications by coordinating reviews and contributing and enhancing text during publication. These contributions come at a considerable cost: U.S. academic libraries paid $1.7 billion for serial subscriptions in 2008 alone. Library budgets, in contrast, are flat and not able to keep pace with serial price inflation. We have investigated the publishers' value proposition by conducting a comparative study of pre-print papers and their final published counterparts. This comparison had two working assumptions: 1) if the publishers' argument is valid, the text of a pre-print paper should vary measurably from its corresponding final published version, and 2) by applying standard similarity measures, we should be able to detect and quantify such differences. Our analysis revealed that the text contents of the scientific papers generally changed very little from their pre-print to final published versions. These findings contribute empirical indicators to discussions of the added value of commercial publishers and therefore should influence libraries' economic decisions regarding access to scholarly publications.
The problem with the identification of Malaysian scholarly journals lies in the lack of a current and complete listing of journals published in Malaysia. As a result, librarians are deprived of a tool that can be used for journal selection and identification of gaps in their serials collection. This study describes the audit carried out on scholarly journals, with the objectives (a) to trace and characterized scholarly journal titles published in Malaysia, and (b) to determine their visibility in international and national indexing databases. A total of 464 titles were traced and their yearly trends, publisher and publishing characteristics, bibliometrics and indexation in national, international and subject-based indexes were described.
For-profit editors such as Elsevier and Springer have been subject to sustained criticism from academics and university libraries, including calls to boycott, and discontinued subscriptions. Mathematicians have played a particularly active role in this critique, and have endeavored to imagine new publication practices and create new journals. This motivates the monitoring of the share of articles published by different editors. I used data from MathSciNet over the period 2000-2017, and focused on the 100 journals with highest citations per article. Within this category, the share of articles published by Elsevier and Springer has steadily increased over this period, from about a third to almost half of the total.
Explainably estimating confidence in published scholarly work offers opportunity for faster and more robust scientific progress. We develop a synthetic prediction market to assess the credibility of published claims in the social and behavioral sciences literature. We demonstrate our system and detail our findings using a collection of known replication projects. We suggest that this work lays the foundation for a research agenda that creatively uses AI for peer review.
The problem of (non)random distribution of points on the sphere is investigated. Published procedures for obtaining preferred direction and preferred plane for points on the sphere (in the sky) are discussed. It is shown that the published methods are incorrect, and, as a consequence, the results obtained by these methods cannot be considered to be significant. The correct methods and their applications on real data will be presented in other papers of this set of papers.
Comment on six papers published by M.A. El-Hakiem and his co-workers in International Communications in Heat and Mass Transfer, Journal of Magnetism and Magnetic Materials and Heat and Mass Transfer
In the year 1598 Philipp Uffenbach published a printed diptych sundial, which is a forerunner of Franz Ritters horizantal sundial. Uffenbach's sundial contains apart from the usual information on a sundial ascending signs of the zodiac, several brigthest stars, an almucantar and most important the oldest gnomonic world map known so far. The sundial is constructed for the polar height of 50 1/6 degrees, the height of Frankfurt/Main the town of his citizenship.
The article gives an overview on early cosmic-ray work, published in German in the period from around 1910 to about 1940.
Comments on six papers published by S.P. Anjali Devi and R. Kandasamy in Heat and Mass Transfer, ZAMM, Mechanics Research Communications, International Communications in Heat and Mass Transfer, Communications in Numerical Methods in Engineering, Journal of Computational and Applied Mechanics In conclusion all the above papers are of very low quality, written without care and are partly or completely wrong.
Proxima Centauri has become the subject of intense study since the radial-velocity discovery by Anglada-Escudé et al. 2016 of a planet orbiting this nearby M-dwarf every ~ 11.2 days. If Proxima Centauri b transits its host star, independent confirmation of its existence is possible, and its mass and radius can be measured in units of the stellar host mass and radius. To date, there have been three independent claims of possible transit-like event detections in light curve observations obtained by the MOST satellite (in 2014-15), the BSST telescope in Antarctica (in 2016), and the Las Campanas Observatory (in 2016). The claimed possible detections are tentative, due in part to the variability intrinsic to the host star, and in the case of the ground-based observations, also due to the limited duration of the light curve observations. Here, we present preliminary results from an extensive photometric monitoring campaign of Proxima Centauri, using telescopes around the globe and spanning from 2006 to 2017, comprising a total of 329 observations. Considering our data that coincide directly and/or phased with the previously published tentative transit detections, we are unable to indepe
This study examines the effect of article processing charge (APC) waivers on the participation of Ukrainian researchers in fully Gold Open Access (Gold OA) journals published by the five largest academic publishers - Elsevier, SAGE, Springer Nature, Taylor & Francis, and Wiley - during the period 2019-2024. These publishers were selected because, in response to the full-scale war launched against Ukraine in 2022, all five introduced emergency 100% APC-waiver policies for Ukrainian authors. Using bibliometric data from the Web of Science Core Collection, the study analyses publication trends in Ukrainian-authored articles in fully Gold OA journals of these publishers before and after 2022. The results show a marked post-2022 increase in Ukraine's Gold OA output, particularly in journals published by Springer Nature and Elsevier. Disciplinary and publisher-specific patterns are evident, with especially strong growth in the medical and applied sciences. The findings underscore the potential of targeted support measures during times of crisis, while also illustrating the inherent limitations of APC-based publishing models in fostering equitable scholarly communication.
This report is the first of two publications of a joint Working Group of the International Mathematical Union (IMU) and the International Council of Industrial and Applied Mathematics (ICIAM). In it, we shall analyze the current state of publishing in the mathematical sciences and explain the resulting problems. Our second publication will offer concrete recommendations, guidelines, and best practices for researchers, policymakers, and evaluators of mathematical research. It will explain how to detect and counteract attempts to game bibliometric measures, empowering the community to reclaim control over research evaluation and drive necessary change.
These recommendations were formulated by the authors in close collaboration with the IMU Committee on Publishing (chaired by Ilka Agricola) and have been endorsed by the Executive Committee of the IMU and the Board of the ICIAM in May/June 2025.
The domination of scientific publishing in the Global North by major commercial publishers is harmful to science. We need the most powerful members of the research community, funders, governments and Universities, to lead the drive to re-communalise publishing to serve science not the market.
Purpose: This paper introduces the concept of "Agentic Publication," a novel LLM-driven framework designed to complement traditional scientific publishing by transforming papers into interactive knowledge systems that address challenges created by exponential growth in scientific literature. Design/methodology/approach: Our architecture integrates structured data (knowledge graphs, metadata) with unstructured content (text, multimedia) through retrieval-augmented generation and multi-agent verification. The system provides interfaces for humans and artificial agents, offering narrative explanations alongside machine-readable outputs. Implementation leverages vector databases for semantic search, knowledge graphs for structured reasoning, and collaborative verification agents. Findings: Our proof-of-concept demonstration showcases multilingual interaction, API accessibility, continuous knowledge flow, and structured knowledge representation. The framework enables dynamic updating of knowledge, synthesis of new findings, and customizable detail levels. Originality: The Agentic Publication represents a transformative approach to scientific communication by creating responsive knowledg