共找到 20 条结果
Sound is an essential sensing element for many organisms in nature, and multiple species have evolved organic structures that create complex acoustic scattering and dispersion phenomena to emit and perceive sound unambiguously. To date, it has not proven possible to design artificial scattering structures that rival the performance of those found in organic structures. Contrarily, most sound manipulation relies on active transduction in fluid media rather than relying on passive scattering principles, as are often found in nature. In this work, we utilize computational morphogenesis to synthesize complex energy-efficient wavelength-sized single-material scattering structures that passively decompose radiated sound into its spatio-spectral components. Specifically, we tailor an acoustic rainbow structure with "above unity" efficiency and an acoustic wavelength-splitter. Our work paves the way for a new frontier in sound-field engineering, with potential applications in transduction, bionics, energy harvesting, communications and sensing.
Avatar creation from human images allows users to customize their digital figures in different styles. Existing rendering systems like Bitmoji, MetaHuman, and Google Cartoonset provide expressive rendering systems that serve as excellent design tools for users. However, twenty-plus parameters, some including hundreds of options, must be tuned to achieve ideal results. Thus it is challenging for users to create the perfect avatar. A machine learning model could be trained to predict avatars from images, however the annotators who label pairwise training data have the same difficulty as users, causing high label noise. In addition, each new rendering system or version update requires thousands of new training pairs. In this paper, we propose a Tag-based annotation method for avatar creation. Compared to direct annotation of labels, the proposed method: produces higher annotator agreements, causes machine learning to generates more consistent predictions, and only requires a marginal cost to add new rendering systems.
We have analyzed all preprints on ChatGPT (N=501) and selected the most influential preprints (according to Altmetric) about ChatGPT across scientific disciplines to provide the most discussed research results about ChatGPT. We prompted ChatGPT to create a structured review article based on them. The results are surprisingly promising, suggesting that the future of creating review articles can lie in ChatGPT.
Fractures are a critical process in how materials wear, weaken, and fail whose unpredictable behavior can have dire consequences. While the behavior of smooth cracks in ideal materials is well understood, it is assumed that for real, heterogeneous systems, fracture propagation is complex, generating rough fracture surfaces that are highly sensitive to specific details of the medium. Here we show how fracture roughness and material heterogeneity are inextricably connected via a simple framework. Studying hydraulic fractures in brittle hydrogels that have been supplemented with microbeads or glycerol to create controlled material heterogeneity, we show that the morphology of the crack surface depends solely on one parameter: the probability to perturb the front above a critical size to produce a step-like instability. This probability scales linearly with the number density, and as heterogeneity size to the $5/2$ power. The ensuing behavior is universal and is captured by the 1D ballistic propagation and annihilation of steps along the singular fracture front.
After small forcing, almost every strongness embedding is the lift of a strongness embedding in the ground model. Consequently, small forcing creates neither strong nor Woodin cardinals.
In a recent article R. T. Cahill claims that the cosmological model based on his "new physics of a dynamical 3-space" resolves the CMB-BBN Lithium-7 and Helium-4 abundance anomalies. In this note it is shown that this conclusion is wrong, resulting from a misunderstanding. In fact, primordial nucleosynthesis in this non-standard cosmological model exacerbates the LIthium-7 problem and creates new problems for primordial Helium-4 and Deuterium.
In this paper, we prove that depth with nonlinearity creates no bad local minima in a type of arbitrarily deep ResNets with arbitrary nonlinear activation functions, in the sense that the values of all local minima are no worse than the global minimum value of corresponding classical machine-learning models, and are guaranteed to further improve via residual representations. As a result, this paper provides an affirmative answer to an open question stated in a paper in the conference on Neural Information Processing Systems 2018. This paper advances the optimization theory of deep learning only for ResNets and not for other network architectures.
In deep learning, \textit{depth}, as well as \textit{nonlinearity}, create non-convex loss surfaces. Then, does depth alone create bad local minima? In this paper, we prove that without nonlinearity, depth alone does not create bad local minima, although it induces non-convex loss surface. Using this insight, we greatly simplify a recently proposed proof to show that all of the local minima of feedforward deep linear neural networks are global minima. Our theoretical results generalize previous results with fewer assumptions, and this analysis provides a method to show similar results beyond square loss in deep linear models.
The paper defends the thesis that analysis of time meaning in a context of philosophy of physical and mathematical natural sciences and philosophical anthropology allows to clear basis of human being and to construct special model of general understanding of time as a creation of nature or creation of human. Regulations on discretization and virtual nature of cultural interaction, mutual tension of limits of cultural and historical process allow connecting philosophy of the nature and philosophical anthropology with system of categories (energy, weight, distance, etc.). It finds application both in the physical and mathematical sphere and in the field of humanitarian studies. We can make a conclusion that neither nature nor human create the time. Time is an imaginary phenomenon connecting human activity and natural processes in the limits of human consciousness.
The magnetic nature of the formation of solar active regions lies at the heart of understanding solar activity and, in particular, solar eruptions. A widespread model, used in many theoretical studies, simulations and the interpretation of observations is that the basic structure of an active region is created by the emergence of a large tube of pre-twisted magnetic field. Despite plausible reasons and the availability of various proxies suggesting the veracity of this model, there has not yet been any direct observational evidence of the emergence of large twisted magnetic flux tubes. Thus, the fundamental question, "are active regions formed by large twisted flux tubes?" has remained open. In this work, we answer this question in the affirmative and provide direct evidence to support this. We do this by investigating a robust topological quantity, called magnetic winding, in solar observations. This quantity, combined with other signatures that are currently available, provides the first direct evidence that large twisted flux tubes do emerge to create active regions.
The charging and dissolution of mineral surfaces in contact with flowing liquids are ubiquitous in nature, as most minerals in water spontaneously acquire charge and dissolve. Mineral dissolution has been studied extensively under equilibrium conditions, even though non-equilibrium phenomena are pervasive and substantially affect the mineral-water interface. Here we demonstrate using interface-specific spectroscopy that liquid flow along a calcium fluoride surface creates a reversible, spatial charge gradient, with decreasing surface charge downstream of the flow. The surface charge gradient can be quantitatively accounted for by a reaction-diffusion-advection model, which reveals that the charge gradient results from a delicate interplay between diffusion, advection, dissolution, and desorption/adsorption. The underlying mechanism is expected to be valid for a wide variety of systems, including groundwater flows in nature and microfluidic systems.
Voice design from natural language aims to generate speaker timbres directly from free-form textual descriptions, allowing users to create voices tailored to specific roles, personalities, and emotions. Such controllable voice creation benefits a wide range of downstream applications-including storytelling, game dubbing, role-play agents, and conversational assistants, making it a significant task for modern Text-to-Speech models. However, existing models are largely trained on carefully recorded studio data, which produces speech that is clean and well-articulated, yet lacks the lived-in qualities of real human voices. To address these limitations, we present MOSS-VoiceGenerator, an open-source instruction-driven voice generation model that creates new timbres directly from natural language prompts. Motivated by the hypothesis that exposure to real-world acoustic variation produces more perceptually natural voices, we train on large-scale expressive speech data sourced from cinematic content. Subjective preference studies demonstrate its superiority in overall performance, instruction-following, and naturalness compared to other voice design models.
In this paper, we illustrate how a Schrödinger cat state created via a matter-wave interferometer can be viewed as the simplest quantum-gravity setup where we can treat both matter and gravity on an equal footing at a perturbative level. Here we treat Einstein's theory of general relativity using an effective field theory approach, quantising the massless spin-2 graviton in the presence of a quantum spatial superposition of matter that creates a matter-wave interferometer in the non-relativistic limit. We show that due to the matter-graviton coupling the graviton vacuum is displaced analogous to the coherent state. We study the contrast/overlap between the coherent states of the left and right superpositions in the matter-wave interferometer. We also study the entanglement between matter and the graviton in this setup and relate it to a gravitational contrast, or the overlap of the quantum geometries led by the coherent states. In the appendix, we provide an example of a time-dependent harmonic oscillator and study the contrast/overlap of such coherent states of the graviton.
We present a system that uses LLMs as a tool in the development of Constructed Languages -- ConLangs, which we call IASC (Interactive Agentic System for ConLangs). The system is modular in that it creates each of the components -- phonology, morphology and syntax, lexicon, orthography, and grammatical handbook, using module-specific sets of prompts. The approach is agentic in that various modules allow for refining the output given automatically-generated commentary on a previous step. Our main goals are twofold. First, we aim to provide tools that facilitate an engaging and enjoyable experience in creating artificially constructed languages. Second, the focus of this paper is on using our ConLang framework as a novel way to explore what LLMs 'know' about language -- not what they know about any particular language or encyclopedic facts, but how much they know about and understand language and linguistic concepts. In the experiments, we particularly focus on the morphosyntax module and show that there is a fairly wide gulf in capabilities both among different LLMs and among different linguistic specifications, with it being notably easier for systems to deal with more typologically
A boundary equilibrium bifurcation (BEB) in a hybrid dynamical system occurs when a regular equilibrium collides with a switching surface in phase space. This causes a transition to a pseudo-equilibrium embedded within the switching surface, but limit cycles (LCs) and other invariant sets can also be created and the nature of these is not well understood for systems with more than two dimensions. This work treats two codimension-two scenarios in hybrid systems of any number of dimensions, where the number of small-amplitude limit cycles bifurcating from a BEB changes. The first scenario involves a limit cycle (LC) with a Floquet multiplier $1$ and for nearby parameter values the BEB creates a pair of limit cycles. The second scenario involves a limit cycle with a Floquet multiplier $-1$ and for nearby parameter values the BEB creates a period-doubled solution. Both scenarios are unfolded in a general setting, showing that typical two-parameter bifurcation diagrams have a curve of saddle-node or period-doubling bifurcations emanating transversally from a curve of BEBs at the codimension-two point. The results are illustrated with three-dimensional examples and an eight-dimensional a
Designers often engage with video to gain rich, temporal insights about the context of users, collaboratively analyzing it to gather ideas, challenge assumptions, and foster empathy. To capture the full visual context of users and their situations, designers are adopting 360$^\circ$ video, providing richer, more multi-layered insights. Unfortunately, the spherical nature of 360$^\circ$ video means designers cannot create tangible video artifacts such as storyboards for collaborative analysis. To overcome this limitation, we created Tangi, a web-based tool that converts 360$^\circ$ images into tangible 360$^\circ$ video artifacts, that enable designers to embody and share their insights. Our evaluation with nine experienced designers demonstrates that the artifacts Tangi creates enable tangible interactions found in collaborative workshops and introduce two new capabilities: spatial orientation within 360$^\circ$ environments and linking specific details to the broader 360$^\circ$ context. Since Tangi is an open-source tool, designers can immediately leverage 360$^\circ$ video in collaborative workshops.
Although existing technology cannot yet directly produce fields at the Schwinger level, experimental facilities can already explore strong-field QED phenomena by taking advantage of the Lorentz boost of energetic electron beams. Recent studies show that QED cascades can create electron-positron pairs at sufficiently high density to exhibit collective plasma effects. Signatures of the collective pair plasma effects can appear in exquisite detail through plasma-induced frequency upshifts and chirps in the laser spectrum. Maximizing the magnitude of the QED plasma signature demands high pair density and low pair energy, which suits the configuration of colliding an over $10^{18}{Jm^{-3}}$ energy-density electron beam with a $10^{22}\mathrm{-}10^{23}{Wcm^{-2}}$ intensity laser pulse. The collision creates pairs that have a large plasma frequency, made even larger as they slow down or reverse direction due to both the radiation reaction and laser pressure. This paper explains at a tutorial level the key properties of the QED cascades and laser frequency upshift, and at the same time finds the minimum parameters that can be used to produce observable QED plasma.
An experienced human Observer reading a document -- such as a crime report -- creates a succinct plot-like $\textit{``Working Memory''}$ comprising different actors, their prototypical roles and states at any point, their evolution over time based on their interactions, and even a map of missing Semantic parts anticipating them in the future. $\textit{An equivalent AI Observer currently does not exist}$. We introduce the $\textbf{[G]}$enerative $\textbf{[S]}$emantic $\textbf{[W]}$orkspace (GSW) -- comprising an $\textit{``Operator''}$ and a $\textit{``Reconciler''}$ -- that leverages advancements in LLMs to create a generative-style Semantic framework, as opposed to a traditionally predefined set of lexicon labels. Given a text segment $C_n$ that describes an ongoing situation, the $\textit{Operator}$ instantiates actor-centric Semantic maps (termed ``Workspace instance'' $\mathcal{W}_n$). The $\textit{Reconciler}$ resolves differences between $\mathcal{W}_n$ and a ``Working memory'' $\mathcal{M}_n^*$ to generate the updated $\mathcal{M}_{n+1}^*$. GSW outperforms well-known baselines on several tasks ($\sim 94\%$ vs. FST, GLEN, BertSRL - multi-sentence Semantics extraction, $\sim 1
An in-in framework under Schwinger pair creating fields in strong-field quantum electrodynamics is formulated using in-out propagators in coordinate space, that have first-quantized or worldline representation. The framework is derived to all orders in the background field coupling from both the Bogoliubov coefficient method and Schwinger-Keldysh closed-time path formalism. In-out matrix elements in pair creating fields are readily handled using first-quantized methods, and the approach we develop serves to facilitate the evaluation of in-in observables in pair creating backgrounds. We find that in-in augmentations to the in-out partition function and or propagator amount to the insertion of a non-local interaction term that sandwiches a function that receives contributions from singularities and critical points in complex Schwinger propertime. Furthermore, we show the resummation of the in-in partition function leading to vacuum non-persistence that en-route gives an exact first-quantized definition of creating $N$-pairs.
A "match cut" is a common video editing technique where a pair of shots that have a similar composition transition fluidly from one to another. Although match cuts are often visual, certain match cuts involve the fluid transition of audio, where sounds from different sources merge into one indistinguishable transition between two shots. In this paper, we explore the ability to automatically find and create "audio match cuts" within videos and movies. We create a self-supervised audio representation for audio match cutting and develop a coarse-to-fine audio match pipeline that recommends matching shots and creates the blended audio. We further annotate a dataset for the proposed audio match cut task and compare the ability of multiple audio representations to find audio match cut candidates. Finally, we evaluate multiple methods to blend two matching audio candidates with the goal of creating a smooth transition. Project page and examples are available at: https://denfed.github.io/audiomatchcut/