共找到 20 条结果
Spanning trees of complete bipartite graphs exhibit a rich interaction between degree sequences and graph structure. In this paper, we obtain lower bounds on the number of isomorphism classes of spanning trees in $K_{a,b}, 2 \leq a \leq b$ in terms of $P_a(a+b-1)$ and $P_b(a+b-1)$ where $P_k(m)$ is the number of integer partitions of $m$ of length $k$. Furthermore, we obtain an upper bound in terms of $a$ and $b$.
Flaw reporting for deployed AI systems is fundamental to identifying system failures and improving AI safety. Yet the AI reporting ecosystem is fragmented: researchers who identify flaws often do not know what or where to report, and groups who receive reports rarely share them with other relevant stakeholders. As a result, good-faith reporters duplicate effort by submitting many different forms, and recipients lack standardized, triage-ready information. We audit 12 reporting systems published by AI developers, cybersecurity groups, and AI flaw aggregators, identifying five recurring design challenges spanning discoverability, scope, information collection, coordination, and guidance for strict-liability cases. Building on this analysis and feedback from 49 experts across 32 organizations representing developers, security researchers, and ecosystem coordinators, we introduce FLARE-AI, an open-source AI flaw reporting system designed for interoperability with existing systems. FLARE-AI streamlines flaw report creation by collecting triage-relevant information through conditional logic and early classification, then enables optional dissemination of standardized, machine-readable re
Given an explicit presentation of a reflection group of rank two (or any rank two group for that matter), we give a simple procedure for calculating all its systems of imprimitivity, when viewed as a matrix group over the quaternions. This is applied to all the reflection groups, in particular the quaternionic reflection groups, thereby unifying a number of results and ideas in the literature. For example, a primitive complex reflection group of rank two has either uncountably many quaternionic systems of imprimitivity (3 cases) or none (16 cases).
Improving single-thread performance remains a critical challenge in modern processor design, as conventional approaches such as deeper speculation, wider pipelines, and complex out-of-order execution face diminishing returns. This work introduces SAHM-State-Aware Heterogeneous Multicore-a novel architecture that targets performance gains by exploiting fine-grained, time-varying behavioral diversity in single-threaded workloads. Through empirical characterization of performance counter data, we define 16 distinct behavioral states representing different microarchitectural demands. Rather than over-provisioning a monolithic core with all optimizations, SAHM uses a set of specialized cores tailored to specific states and migrates threads at runtime based on detected behavior. This design enables composable microarchitectural enhancements without incurring prohibitive area, power, or complexity costs. We evaluate SAHM in both single-threaded and multiprogrammed scenarios, demonstrating its ability to maintain core utilization while improving overall performance through intelligent state-driven scheduling. Experimental results show opportunity for 17% speed up in realistic scenarios. Th
The evaluation of new microprocessor designs is constrained by slow, cycle-accurate simulators that rely on unrepresentative benchmark traces. This paper introduces a novel deep learning framework for high-fidelity, ``in-the-wild'' simulation on production hardware. Our core contribution is a DL model trained on microarchitecture-independent features to predict cycle-level performance for hypothetical processor designs. This unique approach allows the model to be deployed on existing silicon to evaluate future hardware. We propose a complete system featuring a lightweight hardware trace collector and a principled sampling strategy to minimize user impact. This system achieves a simulation speed of 5 MIPS on a commodity GPU, imposing a mere 0.1% performance overhead. Furthermore, our co-designed Neutrino on-chip accelerator improves performance by 85x over the GPU. We demonstrate that this framework enables accurate performance analysis and large-scale hardware A/B testing on a massive scale using real-world applications.
The reflection subgroups of a reflection group have a natural lattice structure given by the reflections that they contain. By considering the conjugation action orbits of the reflection subgroups for a given root line, we are able to give an essentially combinatorial way to calculate the lattice of all the normal reflection subgroups of a given (finite irreducible) reflection group, and natural generators for them. Moreover, we observe that every complex reflection group is a normal subgroup of the unique maximal reflection group which shares its collineation group. Hence, we are able to present the Shephard-Todd classification of the complex reflection groups as a collection of maximal reflection groups, together with appropriate (collineation preserving) normal reflection subgroups. We investigate the quotients by the normal reflection subgroups, which are known to be reflection groups. We also consider the action of the collineation group on some appropriate small systems of lines, and how these results extend to quaternionic reflection groups. Some novel techniques are introduced, including the notion of a "hidden reflection", a combinatorial-geometric description of the refle
If a (weighted) spherical design is defined as an integration (cubature) rule for a unitarily invariant space P of polynomials (on the sphere), then any unitary image of it is also such a spherical design. It therefore follows that such spherical designs are determined by their Gramian (Gram matrix). We outline a general method to obtain such a characterisation as the minima of a function of the Gramian, which we call a potential. This characterisation can be used for the numerical and analytic construction of spherical designs. When the space P of polynomials is not irreducible under the action of the unitary group, then the potential is not unique. In several cases of interest, e.g., spherical t-designs and half-designs, we use this flexibility to provide potentials with a very simple form. We then use our results to develop certain aspects of the theory of real and complex spherical designs for unitarily invariant polynomial spaces.
Constraint management is a central challenge in modern control systems. A solution is the Reference Governor (RG), which is an add-on strategy to pre-stabilized feedback control systems to enforce state and input constraints by shaping the reference command. While robust formulations of RG exist for linear systems, their extension to nonlinear systems is often computationally intractable. This paper develops a scenario-based robust RG formulation for nonlinear systems and investigates its parallel implementation on multi-core CPUs and CUDA-enabled GPUs. We analyze the computational structure of the algorithm, identify parallelization opportunities, and implement the resulting schemes on modern parallel hardware. Benchmarking on a nonlinear hydrogen fuel cell model demonstrates order-of-magnitude speedups (by as much as three orders of magnitude) compared to sequential implementations.
We give an elementary classification and presentation of the finite quaternionic reflection groups of rank two, based on the notion of a``reflection system''. This simplifies the existing classification, which is shown to be incomplete, e.g., there exist four imprimitive quaternionic reflection groups of order 192 with 22 reflections which are not isomorphic (one of which was previously unknown).
We consider the primitive quaternionic reflection groups of type P for H^2 that are obtained from Blichfeldt's collineation groups for C^4.These are seen to be intimately related to the maximal set of five quaternionic mutually unbiased bases (MUBs) in H2 , for which they are symmetries. From these groups, we construct other interesting sets of lines that they fix, including a new quaternionic spherical 3-design of 16 lines in H^2 with angles {1/5,3/5}, which meets the special bound. Some interesting consequences of this investigation include finding imprimitive quaternionic reflection groups with several systems of imprimitivity, and finding a nontrivial reducible subgroup which has a continuous family of eigenvectors.
We give a holomorphic quartic polynomial in the overlap variables whose zeros on the torus are precisely the Weyl-Heisenberg SICs (symmetric informationally complete positive operator valued measures). By way of comparison, all the other known systems of equations that determine a Weyl-Heisenberg SIC involve variables and their complex conjugates. We also give a related interesting result about the powers of the projective Fourier transform of the group G = Z d x Z d .
Recent breakthroughs in large language models (LLMs) have centered around a handful of data-rich languages. What does it take to broaden access to breakthroughs beyond first-class citizen languages? Our work introduces Aya, a massively multilingual generative language model that follows instructions in 101 languages of which over 50% are considered as lower-resourced. Aya outperforms mT0 and BLOOMZ on the majority of tasks while covering double the number of languages. We introduce extensive new evaluation suites that broaden the state-of-art for multilingual eval across 99 languages -- including discriminative and generative tasks, human evaluation, and simulated win rates that cover both held-out tasks and in-distribution performance. Furthermore, we conduct detailed investigations on the optimal finetuning mixture composition, data pruning, as well as the toxicity, bias, and safety of our models. We open-source our instruction datasets and our model at https://hf.co/CohereForAI/aya-101
General-purpose artificial intelligence (AI) systems are built on massive swathes of public web data, assembled into corpora such as C4, RefinedWeb, and Dolma. To our knowledge, we conduct the first, large-scale, longitudinal audit of the consent protocols for the web domains underlying AI training corpora. Our audit of 14,000 web domains provides an expansive view of crawlable web data and how codified data use preferences are changing over time. We observe a proliferation of AI-specific clauses to limit use, acute differences in restrictions on AI developers, as well as general inconsistencies between websites' expressed intentions in their Terms of Service and their robots.txt. We diagnose these as symptoms of ineffective web protocols, not designed to cope with the widespread re-purposing of the internet for AI. Our longitudinal analyses show that in a single year (2023-2024) there has been a rapid crescendo of data restrictions from web sources, rendering ~5%+ of all tokens in C4, or 28%+ of the most actively maintained, critical sources in C4, fully restricted from use. For Terms of Service crawling restrictions, a full 45% of C4 is now restricted. If respected or enforced, t
The decomposition of the polynomials on the quaternionic unit sphere in $\Hd$ into irreducible modules under the action of the quaternionic unitary (symplectic) group and quaternionic scalar multiplication has been studied by several authors. Typically, these abstract decompositions into ``quaternionic spherical harmonics'' specify the irreducible representations involved and their multiplicities. The elementary constructive approach taken here gives an orthogonal direct sum of irreducibles, which can be described by some low-dimensional subspaces, to which commuting linear operators $L$ and $R$ are applied. These operators map harmonic polynomials to harmonic polynomials, and zonal polynomials to zonal polynomials. We give explicit formulas for the relevant ``zonal polynomials'' and describe the symmetries, dimensions, and ``complexity'' of the subspaces involved. Possible applications include the construction and analysis of desirable sets of points in quaternionic space, such as equiangular lines, lattices and spherical designs (cubature rules).
We give some new explicit examples of putatively optimal projective spherical designs. i.e., ones for which there is numerical evidence that they are of minimal size. These form continuous families, and so have little apparent symmetry in general, which requires the introduction of new techniques for their construction. New examples of interest include an 11-point spherical (3, 3)-design for R 3 , and a 12-point spherical (2, 2)-design for R 4 given by four Mercedes-Benz frames that lie on equi-isoclinic planes. We also give results of an extensive numerical study to determine the nature of the real algebraic variety of optimal projective real spherical designs, and in particular when it is a single point (a unique design) or corresponds to an infinite family of designs.
As new machine learning methods demand larger training datasets, researchers and developers face significant challenges in dataset management. Although ethics reviews, documentation, and checklists have been established, it remains uncertain whether consistent dataset management practices exist across the community. This lack of a comprehensive overview hinders our ability to diagnose and address fundamental tensions and ethical issues related to managing large datasets. We present a systematic review of datasets published at the NeurIPS Datasets and Benchmarks track, focusing on four key aspects: provenance, distribution, ethical disclosure, and licensing. Our findings reveal that dataset provenance is often unclear due to ambiguous filtering and curation processes. Additionally, a variety of sites are used for dataset hosting, but only a few offer structured metadata and version control. These inconsistencies underscore the urgent need for standardized data infrastructures for the publication and management of datasets.
The following is a response to the US Department of Commerce's Request for Information (RFI) regarding AI and Open Government Data Assets. First, we commend the Department for its initiative in seeking public insights on the organization and sharing of data. To facilitate scientific discovery and advance AI development, it is crucial for all data producers, including the Department of Commerce and other governmental entities, to prioritize the quality of their data corpora. Ensuring data is accessible, scalable, and secure is essential for harnessing its full potential. In our response, we outline best practices and key considerations for AI and the Department of Commerce's Open Government Data Assets.
We give a simple presentation of the six quaternionic equiangular lines in $\mathbb{H}^2$ as an orbit of the primitive quaternionic reflection group of order 720 (which is isomorphic to 2.A_6 the double cover of $A_6)$. Other orbits of this group are also seen to give optimal spherical designs (packings) of 10, 15 and 20 lines in $\mathbb{H}^2$, with angles { 1/3, 2/3 }, { 1/4, 5/8 } and { 0, 1/3, 2/3 }, respectively. We consider the origins of this reflection group as one of Blichfeldt's "finite collineation groups" for lines in $\mathbb{C}^4$, and general methods for finding nice systems of quaternionic lines.
Independent evaluation and red teaming are critical for identifying the risks posed by generative AI systems. However, the terms of service and enforcement strategies used by prominent AI companies to deter model misuse have disincentives on good faith safety evaluations. This causes some researchers to fear that conducting such research or releasing their findings will result in account suspensions or legal reprisal. Although some companies offer researcher access programs, they are an inadequate substitute for independent research access, as they have limited community representation, receive inadequate funding, and lack independence from corporate incentives. We propose that major AI developers commit to providing a legal and technical safe harbor, indemnifying public interest safety research and protecting it from the threat of account suspensions or legal reprisal. These proposals emerged from our collective experience conducting safety, privacy, and trustworthiness research on generative AI systems, where norms and incentives could be better aligned with public interests, without exacerbating model misuse. We believe these commitments are a necessary step towards more inclusi
Continuous superradiance using a narrow optical transition has the potential to improve the short-term stability of state-of-the-art optical clocks. Even though pulsed superradiant emission on a mHz linewidth clock transition has been shown, true continuous operation, without Fourier limitation, has turned out to be extremely challenging. The trade-off between maintaining a high atomic flux while minimizing decoherence effects presents a significant obstacle. Here, we discuss the design of a machine that could overcome this problem by combining a high-flux continuous beam of ultra cold strontium atoms with a bowtie cavity for the generation of superradiant lasing. To evaluate the feasibility of our design, we present simulation results for continuous high-efficiency cooling, loading, and pumping to the upper lasing state inside the bowtie cavity. We then present two different models for stimulating the generated superradiant field by taking into account position-dependent shifts, collisional decoherence, light shifts, and atom loss. Finally, we estimate a laser linewidth of less than 100 mHz, limited by atom number fluctuations, and resulting in an output power of hundreds of fW.