共找到 20 条结果
Artificial intelligence (AI) chatbots (e.g., ChatGPT) can communicate in strikingly humanlike ways. This has prompted many chatbot users to attribute psychological properties, including consciousness, to these systems. However, there is little scientific evidence that current AI chatbots are conscious. How, then, should we understand people's consciousness attributions to chatbots? Are they merely metaphorical claims, or do they express genuine beliefs? If these attributions lack evidential support, are users epistemically blameworthy for making them, or might they be epistemically innocent, yielding significant benefits otherwise unattainable? This paper offers a conceptual analysis of consciousness attributions to AI chatbots and develops a multidimensional taxonomy of the attitudes they may express, ranging from non-doxastic stances (e.g., pretence) to different forms of belief, including delusions. This taxonomy helps avoid conflations by showing that linguistically identical attributions can reflect importantly different attitudes and degrees of epistemic commitment to the proposition that chatbots are conscious. The taxonomy also provides a framework for empirical studies to
Multimodal learning methods with targeted unimodal learning objectives have exhibited their superior efficacy in alleviating the imbalanced multimodal learning problem. However, in this paper, we identify the previously ignored gradient conflict between multimodal and unimodal learning objectives, potentially misleading the unimodal encoder optimization. To well diminish these conflicts, we observe the discrepancy between multimodal loss and unimodal loss, where both gradient magnitude and covariance of the easier-to-learn multimodal loss are smaller than the unimodal one. With this property, we analyze Pareto integration under our multimodal scenario and propose MMPareto algorithm, which could ensure a final gradient with direction that is common to all learning objectives and enhanced magnitude to improve generalization, providing innocent unimodal assistance. Finally, experiments across multiple types of modalities and frameworks with dense cross-modal interaction indicate our superior and extendable method performance. Our method is also expected to facilitate multi-task cases with a clear discrepancy in task difficulty, demonstrating its ideal scalability. The source code and
In asynchronous games, Melli{è}s proved that innocent strategies are positional: their behaviour only depends on the position, not the temporal order used to reach it. This insightful result shaped our understanding of the link between dynamic (i.e. game) and static (i.e. relational) semantics. In this paper, we investigate the positionality of innocent strategies in the traditional setting of Hyland-Ong-Nickau-Coquand pointer games. We show that though innocent strategies are not positional, total finite innocent strategies still enjoy a key consequence of positionality, namely positional injectivity: they are entirely determined by their positions. Unfortunately, this does not hold in general: we show a counterexample if finiteness and totality are lifted. For finite partial strategies we leave the problem open; we show however the partial result that two strategies with the same positions must have the same P-views of maximal length.
Seeking a general framework for reasoning about and comparing programming languages, we derive a new view of Milner's CCS. We construct a category E of plays, and a subcategory V of views. We argue that presheaves on V adequately represent innocent strategies, in the sense of game semantics. We then equip innocent strategies with a simple notion of interaction. This results in an interpretation of CCS. Based on this, we propose a notion of interactive equivalence for innocent strategies, which is close in spirit to Beffara's interpretation of testing equivalences in concurrency theory. In this framework we prove that the analogues of fair and must testing equivalences coincide, while they differ in the standard setting.
Although the HO/N games are fully abstract for PCF, the traditional notion of innocence (which underpins these games) is not satisfactory for such language features as non-determinism and probabilistic branching, in that there are stateless terms that are not innocent. Based on a category of P-visible plays with a notion of embedding as morphisms, we propose a natural generalisation by viewing innocent strategies as sheaves over (a site of) plays, echoing a slogan of Hirschowitz and Pous. Our approach gives rise to fully complete game models in each of the three cases of deterministic, nondeterministic and probabilistic branching. To our knowledge, in the second and third cases, ours are the first such factorisation-free constructions.
Critical pair analysis provides a convenient and computable criterion of confluence, which is a fundamental property in rewriting theory, for a wide variety of rewriting systems. Bonchi et al. showed validity of critical pair analysis for rewriting on string diagrams in symmetric monoidal categories. This work aims at automation of critical pair analysis for string diagram rewriting, and develops an algorithm that implements the core part of critical pair analysis. The algorithm enumerates all critical pairs of a given left-connected string diagram rewriting system, and it can be realised by concrete manipulation of hypergraphs. We prove correctness and exhaustiveness of the algorithm, for string diagrams in symmetric monoidal categories without a Frobenius structure.
High-resolution RGB imagery acquired from low-altitude UAV surveys was processed through a modular pipeline incorporating transformer-based semantic segmentation, connected-component vegetation extraction, fine-grained species classification using a ConvNeXt architecture, and grid-based dominance scoring at 2x2m resolution. The framework targeted two ecologically significant halophytic grasses, Spartina maritima and Puccinellia maritima, and was trained using a curated and manually annotated UAV imagery, along with biodiversity imagery sourced from publicly accessible datasets. In order to identify these plants from the imagery, our segmentation yielded reliable species masks (mean IoU = 0.56; pixel-level accuracy = 0.96), while object-level classification achieved very good discrimination (F1 = 0.99). Dominance estimates closely matched quadrat-based field surveys, with mean absolute differences below 8%, preserving fine-scale spatial structure under realistic survey conditions. The developed system, named EcoVision, establishes a practical foundation for scalable, high-resolution salt marsh monitoring, demonstrating how AI-driven workflows can translate pixel-level predictions in
Climate change is intensifying human heat exposure, particularly in densely built urban centers of the Global South. Low-cost construction materials and high thermal-mass surfaces further exacerbate this risk. Yet scalable methods for assessing such heat-relevant building attributes remain scarce. We propose a machine learning framework that fuses openly available unmanned aerial vehicle (UAV) and street-view (SV) imagery via a coupled global context vision transformer (CGCViT) to learn heat-relevant representations of urban structures. Thermal infrared (TIR) measurements from HotSat-1 are used to quantify the relationship between building attributes and heat-associated health risks. Our dual-modality cross-view learning approach outperforms the best single-modality models by up to $9.3\%$, demonstrating that UAV and SV imagery provide valuable complementary perspectives on urban structures. The presence of vegetation surrounding buildings (versus no vegetation), brighter roofing (versus darker roofing), and roofing made of concrete, clay, or wood (versus metal or tarpaulin) are all significantly associated with lower HotSat-1 TIR values. Deployed across the city of Dar es Salaam,
Urban flooding is a growing climate change-related hazard in rapidly expanding African cities, where inadequate waste management often blocks drainage systems and amplifies flood risks. This study introduces an AI-powered urban waste mapping workflow that leverages openly available aerial and street-view imagery to detect municipal solid waste at high resolution. Applied in Dar es Salaam, Tanzania, our approach reveals spatial waste patterns linked to informal settlements and socio-economic factors. Waste accumulation in waterways was found to be up to three times higher than in adjacent urban areas, highlighting critical hotspots for climate-exacerbated flooding. Unlike traditional manual mapping methods, this scalable AI approach allows city-wide monitoring and prioritization of interventions. Crucially, our collaboration with local partners ensured culturally and contextually relevant data labeling, reflecting real-world reuse practices for solid waste. The results offer actionable insights for urban planning, climate adaptation, and sustainable waste management in flood-prone urban areas.
Several sources of flexibility in transmission and, especially, distribution networks are being unlocked by advances in information and communication technologies, aggregators, and new flexibility markets. However, maximizing benefits for both transmission and distribution system operators in a coordinated way requires new algorithms, modeling tools, and modernization of regulatory frameworks. Such approaches must account for uncertainties, the physical and operational constraints of flexibility providers and the grid itself, constraints on information exchange, and scalability, including computational requirements and time constraints. Given the diverse contexts and jurisdictions around the world, there is no single recipe for achieving coordination, but important trends and shared challenges are emerging. This paper surveys the complexities of coordination from technical, market, and technological perspectives, and outlines current practices, proposed approaches, and future research directions to effectively manage, coordinate, model, and leverage flexibility across voltage levels.
Acquiring and annotating surgical data is often resource-intensive, ethical constraining, and requiring significant expert involvement. While generative AI models like text-to-image can alleviate data scarcity, incorporating spatial annotations, such as segmentation masks, is crucial for precision-driven surgical applications, simulation, and education. This study introduces both a novel task and method, SimGen, for Simultaneous Image and Mask Generation. SimGen is a diffusion model based on the DDPM framework and Residual U-Net, designed to jointly generate high-fidelity surgical images and their corresponding segmentation masks. The model leverages cross-correlation priors to capture dependencies between continuous image and discrete mask distributions. Additionally, a Canonical Fibonacci Lattice (CFL) is employed to enhance class separability and uniformity in the RGB space of the masks. SimGen delivers high-fidelity images and accurate segmentation masks, outperforming baselines across six public datasets assessed on image and semantic inception distance metrics. Ablation study shows that the CFL improves mask quality and spatial separation. Downstream experiments suggest gener
As AI systems become more integrated into daily life, the need for safer and more reliable moderation has never been greater. Large Language Models (LLMs) have demonstrated remarkable capabilities, surpassing earlier models in complexity and performance. Their evaluation across diverse tasks has consistently showcased their potential, enabling the development of adaptive and personalized agents. However, despite these advancements, LLMs remain prone to errors, particularly in areas requiring nuanced moral reasoning. They struggle with detecting implicit hate, offensive language, and gender biases due to the subjective and context-dependent nature of these issues. Moreover, their reliance on training data can inadvertently reinforce societal biases, leading to inconsistencies and ethical concerns in their outputs. To explore the limitations of LLMs in this role, we developed an experimental framework based on state-of-the-art (SOTA) models to assess human emotions and offensive behaviors. The framework introduces a unified benchmark dataset encompassing 49 distinct categories spanning the wide spectrum of human emotions, offensive and hateful text, and gender and racial biases. Furt
Surgical future prediction, driven by real-time AI analysis of surgical video, is critical for operating room safety and efficiency. It provides actionable insights into upcoming events, their timing, and risks-enabling better resource allocation, timely instrument readiness, and early warnings for complications (e.g., bleeding, bile duct injury). Despite this need, current surgical AI research focuses on understanding what is happening rather than predicting future events. Existing methods target specific tasks in isolation, lacking unified approaches that span both short-term (action triplets, events) and long-term horizons (remaining surgery duration, phase transitions). These methods rely on coarse-grained supervision while fine-grained surgical action triplets and steps remain underexplored. Furthermore, methods based only on future feature prediction struggle to generalize across different surgical contexts and procedures. We address these limits by reframing surgical future prediction as state-change learning. Rather than forecasting raw observations, our approach classifies state transitions between current and future timesteps. We introduce SurgFUTR, implementing this thro
Intraoperative adverse events (IAEs), such as bleeding or thermal injury, can lead to severe postoperative complications if undetected. However, their rarity results in highly imbalanced datasets, posing challenges for AI-based detection and severity quantification. We propose BetaMixer, a novel deep learning model that addresses these challenges through a Beta distribution-based mixing approach, converting discrete IAE severity scores into continuous values for precise severity regression (0-5 scale). BetaMixer employs Beta distribution-based sampling to enhance underrepresented classes and regularizes intermediate embeddings to maintain a structured feature space. A generative approach aligns the feature space with sampled IAE severity, enabling robust classification and severity regression via a transformer. Evaluated on the MultiBypass140 dataset, which we extended with IAE labels, BetaMixer achieves a weighted F1 score of 0.76, recall of 0.81, PPV of 0.73, and NPV of 0.84, demonstrating strong performance on imbalanced data. By integrating Beta distribution-based sampling, feature mixing, and generative modeling, BetaMixer offers a robust solution for IAE detection and quantif
We explore the roles of the three competitors, namely, gravity, turbulence, and magnetic fields, in controlling star formation (SF) within dense, massive clumps identified in the ATLASGAL survey. By examining scaling relations, virial parameters, and turbulent energy spectra, we evaluate the dynamical state of these clumps. We observe a weak velocity dispersion-size relation (sigma proportional to L^0.11), which is much shallower than the classical Larson-like relations, suggesting that turbulence does not mainly drive internal dynamics. The turbulent energy spectrum, E(k) proportional to k^-1.21, is also less steep than what is expected for both incompressible and compressible turbulence. We equally observe a decreasing trend in the virial parameter with increasing mass (alpha_vir proportional to M^-0.37), indicating that more massive clumps are increasingly gravitationally bound. These trends indicate an increasing relative dominance of gravity over turbulence at smaller scales, aligning with multiscale collapse scenarios; however, the absolute energy balance remains unquantifiable with the current data. Although magnetic fields are not directly measured, their potential influenc
Accurate tool tracking is essential for the success of computer-assisted intervention. Previous efforts often modeled tool trajectories rigidly, overlooking the dynamic nature of surgical procedures, especially tracking scenarios like out-of-body and out-of-camera views. Addressing this limitation, the new CholecTrack20 dataset provides detailed labels that account for multiple tool trajectories in three perspectives: (1) intraoperative, (2) intracorporeal, and (3) visibility, representing the different types of temporal duration of tool tracks. These fine-grained labels enhance tracking flexibility but also increase the task complexity. Re-identifying tools after occlusion or re-insertion into the body remains challenging due to high visual similarity, especially among tools of the same category. This work recognizes the critical role of the tool operators in distinguishing tool track instances, especially those belonging to the same tool category. The operators' information are however not explicitly captured in surgical videos. We therefore propose SurgiTrack, a novel deep learning method that leverages YOLOv7 for precise tool detection and employs an attention mechanism to mode
A universal theory of linear instabilities in swirling flows, occurring in both natural settings and industrial applications, is formulated. The theory encompasses a wide range of open and confined flows, including spiral isothermal flows and baroclinic flows driven by radial temperature gradients and natural gravity in rotating fluids. By employing short-wavelength local analysis, the theory generalizes previous findings from numerical simulations and linear stability analyses of specific swirling flows, such as spiral Couette flow, spiral Poiseuille flow, and baroclinic Couette flow. A general criterion, extending and unifying existing criteria for instability to both centrifugal and shear-driven perturbations in swirling flows is derived, taking into account viscosity and thermal diffusion, and guiding experimental and numerical investigations in the otherwise inaccessible parameter regimes.
Although tractor services are increasingly used in the Ejura-Sekyedumase Municipality, access remains uneven, and some smallholder maize farmers still rely on labour-intensive production methods. This study investigated the profitability effects and adoption determinants of tractor service utilisation in Ejura-Sekyedumase Municipality, Ashanti Region, Ghana. Cross-sectional data were collected from 359 farmers using multi-stage proportionate random sampling. A multivariate probit model identified simultaneous adoption determinants across ploughing and shelling services. Propensity score matching (PSM) estimated a positive profitability effect of GHS 471 per acre after controlling for observable selection bias. Users achieved a net profit margin of 29.87 percent and a return-to-cost ratio of 42.60 percent, compared with 25.10 percent and 33.51 percent for non-users, respectively. Farming experience and fertiliser intensity positively predicted adoption, while farm size exerted a consistent negative influence; this may suggest supply-side availability constraints, including potential fleet-capacity limitations. The negative FBO result may reflect labour-sharing functions that substit
Glitch activity refers to the mean increase in pulsar spin frequency per year due to rotational glitches. It is an important tool for studying super-nuclear matter using neutron star interiors as templates. Glitch events are typically observed in the spin frequency ($ν$) and frequency derivative ($\dotν$) of pulsars. The rate of glitch recurrence decreases as the pulsar ages, and the activity parameter is usually measured by linear regression of cumulative glitches over a given period. This method is effective for pulsars with multiple regular glitch events. However, due to the scarcity of glitch events and the difficulty of monitoring all known pulsars, only a few have multiple records of glitch events. This limits the use of the activity parameter in studying neutron star interiors with multiple pulsars. In this study, we examined the relationship between the activity parameters and pulsar spin parameters (spin frequency, frequency derivative, and pulsar characteristic age). We found that a quadratic function provides a better fit for the relationship between activity parameters and spin parameters than the commonly used linear functions. Using this information, we were able to e
An analytical theory is presented for linear, local, short-wavelength instabilities in swirling flows, in which axial shear, differential rotation, radial thermal stratification, viscosity, and thermal diffusivity are all taken into account. A geometrical optics approach is applied to the Navier-Stokes equations, coupled with the energy equation, leading to a set of amplitude transport equations. From these, a dispersion relation is derived, capturing two distinct types of instability: a stationary centrifugal instability and an oscillatory, visco-diffusive McIntyre instability. Instability regions corresponding to different axial or azimuthal wavenumbers are found to possess envelopes in the plane of physical parameters, which are explicitly determined using the discriminants of polynomials. As these envelopes are shown to bound the union of instability regions associated with particular wavenumbers, it is concluded that the envelopes correspond to curves of critical values of physical parameters, thereby providing compact, closed-form criteria for the onset of instability. The derived analytical criteria are validated for swirling flows modelled by a cylindrical, differentially r