Deploying robots at scale demands robustness to the long tail of everyday situations. The countless variations in scene layout, object geometry, and task specifications that characterize real environments are vast and underrepresented in existing robot benchmarks. Measuring this level of generalization requires infrastructure at a scale and diversity that physical evaluation alone cannot provide. We introduce MolmoSpaces, a fully open ecosystem to support large-scale benchmarking of robot policies. MolmoSpaces consists of over 230k diverse indoor environments, ranging from handcrafted household scenes to procedurally generated multiroom houses, populated with 130k richly annotated object assets, including 48k manipulable objects with 42M stable grasps. Crucially, these environments are simulator-agnostic, supporting popular options such as MuJoCo, Isaac, and ManiSkill. The ecosystem supports the full spectrum of embodied tasks: static and mobile manipulation, navigation, and multiroom long-horizon tasks requiring coordinated perception, planning, and interaction across entire indoor environments. We also design MolmoSpaces-Bench, a benchmark suite of 8 tasks in which robots interac
Photonic quantum technologies require efficient sources of pure single photons. We present a heralded single-photon source based on spontaneous parametric down-conversion in a monolithic cavity optimized for high spectral and spatial purity. The source heralds single photons at a wavelength of 1540 nm and a spectral bandwidth of 168 MHz, with a maximum heralding efficiency of 70% including all transmission and detection losses, while keeping the multi-photon contamination below 3%. The cavity enhancement predominantly generates photons into the central cavity mode, with a theoretical upper bound on the spectral purity of 79.4% arising from nonzero overlap with adjacent cavity modes. Spectral isolation of the central cavity mode with an etalon yields an increased measured spectral purity of (96.2 $\pm$ 2.7)%.
Many advancements in optics have relied on the tight-binding approximation, which simplifies the description and prediction of complex system behaviors. This approximation describes the dynamics of the total light field by examining the coupling between the guided modes of individual single-mode substructures -- also known as coupled mode theory. However, the underlying assumption, that the guided modes of individual waveguides form an orthogonal basis, breaks down when waveguides are brought into close proximity or when larger arrays are considered. In this work, we systematically analyze the consequences of this non-orthogonality and show that it leads to a generalized eigenvalue problem involving an overlap matrix, causing a fundamental mismatch between the standard TB model and solutions of the paraxial wave equation. To resolve this issue, we introduce a modified TB framework based on the Löwdin orthogonalization, which constructs an orthonormal basis from the non-orthogonal guided modes while minimally altering their physical shape and preserving their symmetry properties. The resulting Löwdin-TB method restores the standard eigenvalue problem and yields excellent agreement w
The concept of quantum tokens dates back alongside quantum cryptography to Stephen Wiesner's seminal work in 1983[1]. Already this initial work proposes society-relevant applications such as secure quantum banknotes, which can be exchanged between a bank and a customer. This quantum currency is based on various physical states that can be easily verified but is protected from being copied by the fundamental quantum laws. Four decades later, these ideas have flourished in the field of quantum information, and the concept of quantum banknotes has not only adopted many varying names, such as quantum money, quantum coins, quantum-digital payments, and quantum tokens, but also reached its first experimental demonstrations. In this perspective article, we discuss the current state-of-the-art of quantum tokens in the field of quantum information, as well as their future perspectives. We present a number of physical realizations of quantum tokens with integrated quantum memories and their applicability scenarios in detail. Finally, we discuss how quantum tokens fit into the information security ecosystem and consider their relationship to post-quantum cryptography.
Web agents--autonomous systems that navigate and execute tasks on the web on behalf of users--have the potential to transform how people interact with the digital world. However, the most capable web agents today rely on proprietary models with undisclosed training data and recipes, limiting scientific understanding, reproducibility, and community-driven progress. We believe agents for the open web should be built in the open. To this end, we introduce (1) MolmoWebMix, a large and diverse mixture of browser task demonstrations and web-GUI perception data and (2) MolmoWeb, a family of fully open multimodal web agents. Specifically, MolmoWebMix combines over 100K synthetic task trajectories from multiple complementary generation pipelines with 30K+ human demonstrations, atomic web-skill trajectories, and GUI perception data, including referring expression grounding and screenshot question answering. MolmoWeb agents operate as instruction-conditioned visual-language action policies: given a task instruction and a webpage screenshot, they predict the next browser action, requiring no access to HTML, accessibility trees, or specialized APIs. Available in 4B and 8B size, on browser-use b
We propose a nested Wolter-I mirror design for a neutron condenser, which is based on established X-ray telescope technology. We demonstrate through simulations that it can increase the flux density at the ESS imaging instrument ODIN by up to two orders of magnitude. Experimental measurements of reflectivity and figure errors on a prototype mirror element confirm the technical feasibility of the approach. Then, we discuss design strategies for an imaging objective to fully exploit the condenser specifications while achieving spatial resolutions comparable to those of X-ray micro-CT instruments. Analytically, we show that for monochromatic beams suitable solutions exist employing arrays of hundreds of identical objectives, realized either as compound refractive lenses (CRLs) or Fresnel zone plates (FZPs). To mitigate the inherent chromatic aberration of these optics, each individual objective could be replaced by an achromatic FZP/CRL combination. Key optical properties of the resulting microscope are estimated. This novel full-field microscopy concept for highly divergent, polychromatic neutron beams has the potential to improve temporal and spatial resolution for large samples and
We prove a duality statement on modules over KH-theory in the stable motivic homotopy category whose dualizing object is given by G-theory, over any quasi-excellent scheme of characteristic zero.
Conventional robot programming methods are complex and time-consuming for users. In recent years, alternative approaches such as mixed reality have been explored to address these challenges and optimize robot programming. While the findings of the mixed reality robot programming methods are convincing, most existing methods rely on gesture interaction for robot programming. Since controller-based interactions have proven to be more reliable, this paper examines three controller-based programming methods within a mixed reality scenario: 1) Classical Jogging, where the user positions the robot's end effector using the controller's thumbsticks, 2) Direct Control, where the controller's position and orientation directly corresponds to the end effector's, and 3) Gripper Control, where the controller is enhanced with a 3D-printed gripper attachment to grasp and release objects. A within-subjects study (n = 30) was conducted to compare these methods. The findings indicate that the Gripper Control condition outperforms the others in terms of task completion time, user experience, mental demand, and task performance, while also being the preferred method. Therefore, it demonstrates promisin
Accurate performance estimation of experimentally demonstrated quantum memories is key to understand the nuances in their deployment in photonic quantum networks. While several software packages allow for accessible quantum simulation, they often do not account for the loss and noise in physical devices. We present a framework for modeling ensemble-based atomic quantum memories using the quantum channel formalism. We provide a Kraus matrix representation of several experimentally implemented state-of-the art quantum memories and give an overview of their most important performance metrics. To showcase the applicability of this approach, we implement a memory-assisted quantum token protocol within our simulation framework. Our digital twin model is readily extensible to other memory implementations and easily compatible with existing frameworks for performance simulation of experimental quantum networks.
We address the problem of survival regression modelling with multivariate responses and nonlinear covariate effects. Our model extends the proportional hazards model by introducing several weakly-parametric elements: the marginal baseline hazard functions are expressed as piecewise constants, association is modelled with copulas, and nonlinear covariate effects are handled by a single-index structure using a spline. The model permits a full likelihood approach to inference, making it possible to obtain individual-level survival or hazard function estimates. Performance of the new model is evaluated through simulation studies and application to the Busselton health study data. The results suggest that the proposed method can capture nonlinear covariate effects well, and that there is benefit to modeling the association between the correlated responses.
We investigate quantum superposition effects in two-dimensional quantum walks of identical particles with different statistics under particle exchange, starting from various different initial configurations. To characterize interparticle correlation dynamics, we focus on joint properties such as two-particle coincidence probabilities and the spread velocity of the interparticle distance. Regarding spatial modes as an environment for the particles internal degrees of freedom, we study the role played by the particle statistics using standard entanglement witnesses, showing that particles possessing fermionic statistics are more resistant to thermalize with their environment. We analyze the presence of multipartite entanglement in the system's degrees of freedom through the Quantum Fisher Information, revealing that fermionic states generated during the walk are better suited to perform quantum metrology tasks. Finally, we discuss the potential for implementing this model using integrated photonic circuits by exploiting $N$-partite entanglement between individual photons.
End-to-end autonomous driving systems promise stronger performance through unified optimization of perception, motion forecasting, and planning. However, vision-based approaches face fundamental limitations in adverse weather conditions, partial occlusions, and precise velocity estimation - critical challenges in safety-sensitive scenarios where accurate motion understanding and long-horizon trajectory prediction are essential for collision avoidance. To address these limitations, we propose SpaRC-AD, a query-based end-to-end camera-radar fusion framework for planning-oriented autonomous driving. Through sparse 3D feature alignment, and doppler-based velocity estimation, we achieve strong 3D scene representations for refinement of agent anchors, map polylines and motion modelling. Our method achieves strong improvements over the state-of-the-art vision-only baselines across multiple autonomous driving tasks, including 3D detection (+4.8% mAP), multi-object tracking (+8.3% AMOTA), online mapping (+1.8% mAP), motion prediction (-4.0% mADE), and trajectory planning (-0.1m L2 and -9% TPC). We achieve both spatial coherence and temporal consistency on multiple challenging benchmarks, in
A prevailing view in robot learning is that simulation alone is not enough; effective sim-to-real transfer is widely believed to require at least some real-world data collection or task-specific fine-tuning to bridge the gap between simulated and physical environments. We challenge that assumption. With sufficiently large-scale and diverse simulated synthetic training data, we show that zero-shot transfer to the real world is not only possible, but effective for both static and mobile manipulation. We introduce MolmoBot-Engine, a fully open-source pipeline for procedural data generation across robots, tasks, and diverse simulated environments in MolmoSpaces. With it, we release MolmoBot-Data, a dataset of 1.8 million expert trajectories for articulated object manipulation and pick-and-place tasks. We train three policy classes: MolmoBot, a Molmo2-based multi-frame vision-language model with a flow-matching action head; MolmoBot-Pi0, which replicates the $π_0$ architecture to enable direct comparison; and MolmoBot-SPOC, a lightweight policy suitable for edge deployment and amenable to RL fine-tuning. We evaluate on two robotic platforms: the Franka FR3 for tabletop manipulation task
Interfacing light from solid-state single-photon sources with scalable and robust room-temperature quantum memories has been a long-standing challenge in photonic quantum information technologies due to inherent noise processes and time-scale mismatches between the operating conditions of solid-state and atomic systems. Here, we demonstrate on-demand storage and retrieval of single photons from a semiconductor quantum dot device in a room-temperature atomic vapor memory. A deterministically fabricated InGaAs quantum dot light source emits single photons at the wavelength of the cesium D1 line at 895\,nm which exhibit an inhomogeneously broadened linewidth of 5.1(7)\,GHz and are subsequently stored in a low-noise ladder-type cesium vapor memory. We show control over the interaction between the single photons and the atomic vapor, allowing for variable retrieval times of up to 19.8(3)\,ns at an internal efficiency of $η_\mathrm{int}=0.6(1)\%$. Our results significantly expand the application space of both room-temperature vapor memories and semiconductor quantum dots in future quantum network architectures.
Large language models (LLMs) have recently transformed natural language processing, enabling machines to generate human-like text and engage in meaningful conversations. This development necessitates speed, efficiency, and accessibility in LLM inference as the computational and memory requirements of these systems grow exponentially. Meanwhile, advancements in computing and memory capabilities are lagging behind, exacerbated by the discontinuation of Moore's law. With LLMs exceeding the capacity of single GPUs, they require complex, expert-level configurations for parallel processing. Memory accesses become significantly more expensive than computation, posing a challenge for efficient scaling, known as the memory wall. Here, compute-in-memory (CIM) technologies offer a promising solution for accelerating AI inference by directly performing analog computations in memory, potentially reducing latency and power consumption. By closely integrating memory and compute elements, CIM eliminates the von Neumann bottleneck, reducing data movement and improving energy efficiency. This survey paper provides an overview and analysis of transformer-based models, reviewing various CIM architectu
On-demand storage and retrieval of quantum information in coherent light-matter interfaces is a key requirement for future quantum networking and quantum communication applications. Alkali vapor memories offer scalable and robust high-bandwidth storage at high repetition rates which makes them a natural fit to interface with solid-state single-photon sources. Here, we experimentally realize a room-temperature ladder-type atomic vapor memory that operates on the Cs D1 line. We provide a detailed experimental characterization and demonstration of on-demand storage and retrieval of weak coherent laser pulses with 0.06 photons per pulse at a high signal-to-noise ratio of SNR$=830(80)$. The memory achieves a maximum internal storage efficiency of $η_{\text{int}}=15(1)\%$ and an estimated $1/e$-storage time of $τ_{\mathrm{s}}\approx32\,$ns. Benchmark properties for the storage of single photons from inhomogeneously broadened state-of-the-art solid-state emitters are estimated from the performance of the memory. Together with the immediate availability of high-quality InGaAs quantum dots emitting at 895\,nm, these results provide clear prospects for the development of a heterogeneous on-d
Rapid-turn slow-roll inflationary trajectories have been shown to be an attractor in two-field models, provided the turn rate is near constant and larger than the slow-roll parameters. These trajectories can produce primordial spectra consistent with current observations on CMB scales. We present the generalized consistency condition for sustained rapid-turn inflationary trajectory with two fields, arbitrary field-space metric and potential valid for any value of the turn rate. This has to be supplemented by a second condition to ensure slow roll evolution. Both conditions together constitute a tool to identify inflationary trajectories with arbitrary values of the turning rate without having to solve the equations of motion. We present a Python package for the numerical identification of regions in field-space and parameter space that allow for rapid-turn trajectories.
Efficient optical quantum memories are a milestone required for several quantum technologies including repeater-based quantum key distribution and on-demand multi-photon generation. We present an efficiency optimization of an optical electromagnetically induced transparency (EIT) memory experiment in a warm cesium vapor using a genetic algorithm and analyze the resulting waveforms. The control pulse is represented either as a Gaussian or free-form pulse, and the results from the optimization are compared. We see an improvement factor of 3(7)\% when using optimized free-form pulses. By limiting the allowed pulse energy in a solution, we show an energy-based optimization giving a 30% reduction in energy, with minimal efficiency loss.
We study the problem of estimating the means of well-separated mixtures when an adversary may add arbitrary outliers. While strong guarantees are available when the outlier fraction is significantly smaller than the minimum mixing weight, much less is known when outliers may crowd out low-weight clusters - a setting we refer to as list-decodable mixture learning (LD-ML). In this case, adversarial outliers can simulate additional spurious mixture components. Hence, if all means of the mixture must be recovered up to a small error in the output list, the list size needs to be larger than the number of (true) components. We propose an algorithm that obtains order-optimal error guarantees for each mixture mean with a minimal list-size overhead, significantly improving upon list-decodable mean estimation, the only existing method that is applicable for LD-ML. Although improvements are observed even when the mixture is non-separated, our algorithm achieves particularly strong guarantees when the mixture is separated: it can leverage the mixture structure to partially cluster the samples before carefully iterating a base learner for list-decodable mean estimation at different scales.
We present the implementation and performance analysis of a portable, rack-mounted standalone warm vapor quantum memory system, that also includes the laser package, control electronics and data processing hardware. The optical memory is based on long-lived hyperfine ground states of Cesium which are connected to an excited state via the $D_1$ line at 895 nm in a $Λ$-configuration. The memory is operated with weak coherent pulses containing on average $<1$ photons per pulse. The long-term stability of the memory efficiency and storage fidelity is demonstrated at the single-photon level together with operation in a non-laboratory environment.