In 2006 and 2016, the University of Pennsylvania denied any ties to slavery. In 2017, a group of undergraduate researchers, led by Professor Kathleen Brown, investigated this claim. Initial research, focused on 18th century faculty and trustees who owned slaves, revealed deep connections between the university's history and the institution of slavery. These findings, and discussions amongst the researchers shaped the Penn and Slavery Project's goal of redefining complicity beyond ownership. Breanna Moore's contributions in PSP's second semester expanded the project's focus to include generational wealth gaps. In 2018, VanJessica Gladney served as the PSP's Public History Fellow and spread the project outreach in the greater Philadelphia area. That year, the PSP team began to design an augmented reality app as a Digital Interruption and an attempt to display the truth about Penn's history on its campus. Unfortunately, PSP faced delays due to COVID 19. Despite setbacks, the project persisted, engaging with activists and the wider community to confront historical injustices and modern inequalities.
Improved computation of the dielectric function considering excitonic effects and long wavelength is performed and compared with the nearly free electron band approximation, similarly with the Penn's model case. New expressions for the real and imaginary part of the dielectric function are presented and the real part compared with the Penn's result. The obtained functions satisfy the Kramers-Krönig relations, in contrast with earlier results in the literature. In addition, our improved dielectric function presents a coeficient of 2/3 for small gap approximation (different from the value of 1 in the original Penn model) is very close to the value 0.62 obtained in [Can. J. Phys.53,(1975) p.2549] from pure numerical procedures. The obtained dielectric function also is used in a rough and stimative analysis of the metal-insulator transition in molecular hydrogen being the critical densities determined near the experimental values for the hydrogen coming from other approach. The approximated expressions and critical values are given and the usefulness of the rough methods involved in the determination of the critical points briefly discussed.
We present the first parsing results on the Penn-Helsinki Parsed Corpus of Early Modern English (PPCEME), a 1.9 million word treebank that is an important resource for research in syntactic change. We describe key features of PPCEME that make it challenging for parsing, including a larger and more varied set of function tags than in the Penn Treebank. We present results for this corpus using a modified version of the Berkeley Neural Parser and the approach to function tag recovery of Gabbard et al (2006). Despite its simplicity, this approach works surprisingly well, suggesting it is possible to recover the original structure with sufficient accuracy to support linguistic applications (e.g., searching for syntactic structures of interest). However, for a subset of function tags (e.g., the tag indicating direct speech), additional work is needed, and we discuss some further limits of this approach. The resulting parser will be used to parse Early English Books Online, a 1.1 billion word corpus whose utility for the study of syntactic change will be greatly increased with the addition of accurate parse trees.
Agent memory failures are silent: an LLM-based agent can produce a fluent response even when it fails to extract, retain, or retrieve the information needed across sessions. The write-manage-read loop describes the external pipeline of these systems but leaves open which internal computations implement each stage. Tracing feature circuits across the Qwen-3 family (0.6B--14B) and two memory frameworks (mem0 and A-MEM), we report two mechanistic findings and one deliverable. First, control is detectable before content: routing circuitry is causally active at 0.6B, while content circuitry produces no detectable signal until 4B, exposing a deployment regime where small models route memory decisions before they can reliably extract or ground the underlying facts. Second, the shared hub is recruited, not created: Write and Read converge on a late-layer hub that already exists in the base model as a context-grounding substrate, and memory framing recruits a memory-specific functional direction on this substrate rather than building one of its own. Both findings transfer across mem0 and A-MEM, indicating that the underlying computations are properties of the base model rather than of any p
In this paper, we present empirical and theoretical evidence against a central but largely implicit assumption in circuit and sheaf discovery (CSD), which we term the Functional Anisotropy Hypothesis: the idea that functions in large language models (LLMs) are localised to a unique or near-unique internal mechanism. We show that a single LLM task can instead be supported by multiple, structurally distinct circuits or sheaves that are simultaneously faithful, sparse, and complete. To systematically uncover such competing mechanisms, we introduce Overlap-Aware Sheaf Repulsion, a method that augments the CSD objective with an explicit penalty on structural overlap across multiple discovery runs, enabling the discovery of circuits or sheaves with strong task performance but minimal shared structure across a plethora of common CSD benchmarks. We find that this phenomenon becomes increasingly pronounced as the number of discovered sheaves grows and persists robustly across major CSD methods. We further identify an ultra-sparse three-edge sheaf and show that none of its edges is individually indispensable, undermining even weakened notions of canonical or essential components. To explain
In many real-world settings, institutions can and do adjust the consequences attached to algorithmic classification decisions, such as the size of fines, sentence lengths, or benefit levels. We refer to these consequences as the stakes associated with classification. These stakes can give rise to behavioral responses to classification, as people adjust their actions in anticipation of how they will be classified. Much of the algorithmic fairness literature evaluates classification outcomes while holding behavior fixed, treating behavioral differences across groups as exogenous features of the environment. Under this assumption, the stakes of classification play no role in shaping outcomes. We revisit classic impossibility results in algorithmic fairness in a setting where people respond strategically to classification. We show that, in this environment, the well-known incompatibility between error-rate balance and predictive parity disappears, but only by potentially introducing a qualitatively different form of unequal treatment. Concretely, we construct a two-stage design in which a classifier first standardizes its statistical performance across groups, and then adjusts stakes s
The Dynamic Eclipse Broadcast (DEB) Initiative citizen science program observed coronal visible continuum brightness during the 2024 April 8 total eclipse from locations across North America. We present results from 11 DEB sites spanning 2700 km of distance and showing 49 minutes of evolution. The coronal brightness radial profiles from these telescopes are tightly correlated from 1.2 to 4.0 solar radii and comparable to published photometric coronal intensities. The coronal flattening parameter is measured from 1.4 to 2.8 solar radii. A Ludendorff index of 0.0761 +/- 0.0007 is computed but the extrapolation techniques used by some to calculate this index are shown to disagree with this direct measurement, and alternate structure parameters are suggested. Measured radial velocities are compared with an MHD model of the corona during the eclipse from Y. Li et al. (2026). A polar downflow is measured with an average radial velocity of -37 +/- 3 km s-1 and a deceleration of 14 +/- 3 m s-2 at a speed and position which agrees with the model. The predicted mixture of outflows and downflows at low heights is seen, as well as outflows in two western regions of the corona. The fastest obse
Large language models often solve complex reasoning tasks more effectively with Chain-of-Thought (CoT), but at the cost of long, low-bandwidth token sequences. Humans, by contrast, often reason softly by maintaining a distribution over plausible next steps. Motivated by this, we propose Multiplex Thinking, a stochastic soft reasoning mechanism that, at each thinking step, samples K candidate tokens and aggregates their embeddings into a single continuous multiplex token. This preserves the vocabulary embedding prior and the sampling dynamics of standard discrete generation, while inducing a tractable probability distribution over multiplex rollouts. Consequently, multiplex trajectories can be directly optimized with on-policy reinforcement learning (RL). Importantly, Multiplex Thinking is self-adaptive: when the model is confident, the multiplex token is nearly discrete and behaves like standard CoT; when it is uncertain, it compactly represents multiple plausible next steps without increasing sequence length. Across challenging math reasoning benchmarks, Multiplex Thinking consistently outperforms strong discrete CoT and RL baselines from Pass@1 through Pass@1024, while producing
In Titan's atmosphere, the chemistry of small hydrocarbons and nitriles represent an important link from molecular species to the ubiquitous organic haze that gives Titan its characteristic yellow color. Here we present a new search for two previously undetected molecules, triacetylene (C$_{6}$H$_{2}$) and the gas phase dicyanoacetylene (C$_{4}$N$_{2}$), using the Echelon-Cross-Echelle Spectrograph (EXES) instrument aboard the SOFIA (Stratospheric Observatory For Infrared Astronomy) aircraft. We do not detect these two molecules but determine upper limits for their mixing ratios and column abundances. We find the $3σ$ upper limits on the uniform volume mixing ratio (VMR) above 100 km for C$_{6}$H$_{2}$ to be $4.3\times10^{-11}$ which is lower than the photochemical model predictions. This new upper limit suggests that the growth of linear molecules is inhibited. We also put a strict upper limit on the uniform VMR for gas phase C$_{4}$N$_{2}$ above 125 km to be $1.0\times10^{-10}$. This upper limit is well below the saturation mixing ratio at this altitude for C$_{4}$N$_{2}$ and greatly limits the feasibility of C$_{4}$N$_{2}$ forming ice from condensation.
Reinforcement Learning with Human Feedback (RLHF) has been the dominant approach for improving the reasoning capabilities of Large Language Models (LLMs). Recently, Reinforcement Learning with Verifiable Rewards (RLVR) has simplified this paradigm by replacing the reward and value models with rule-based verifiers. A prominent example is Group Relative Policy Optimization (GRPO). However, GRPO inherently suffers from a length bias, since the same advantage is uniformly assigned to all tokens of a response. As a result, longer responses distribute the reward over more tokens and thus contribute disproportionately to gradient updates. Several variants, such as DAPO and Dr. GRPO, modify the token-level aggregation of the loss, yet these methods remain heuristic and offer limited interpretability regarding their implicit token preferences. In this work, we explore the possibility of allowing the model to learn its own token preference during optimization. We unify existing frameworks under a single formulation and introduce a learnable parameter $λ$ that adaptively controls token-level weighting. We use $λ$-GRPO to denote our method, and we find that $λ$-GRPO achieves consistent improve