共找到 20 条结果
International humanitarian law protects civilians from direct attack unless and for such time as they take direct part in hostilities, with the ICRC's 2009 Interpretive Guidance operationalising this rule through a three-criterion cumulative test. This paper argues that AI-mediated civilian cyber operations challenge the direct causation element of this test in a structurally specific way: when a civilian deploys an autonomous multi-agent cyber system of the kind recently demonstrated in offensive AI research, the "one causal step" standard fails because harm is produced by system-generated decisions made after human disengagement, and the integral-part requirement does not extend because it presupposes downstream human contributors whose conduct can be independently classified. The framework therefore defaults to treating such deployments as indirect participation, in tension with its purpose of capturing civilians who personally take part in hostilities. Beyond the doctrinal analysis, this paper identifies goal-specification granularity as the property on which the integral-part test's concreteness component implicitly turns, classifies AI-mediated operations along a five-level s
The number of active shooter incidents in the US has been increasing alarmingly. It is imperative for the government as well as the public to understand these events. Though both analytic and agent-based models have been proposed for studying active shooter incidents, there are only a few analytic models in the literature, and none incorporate civilian resistance. This article analytically investigates the survival probability of a civilian during an active shooter incident when he can hide, fight, or run, depending on whether or not he is in a closed arena and whether or not he is armed. The key findings are (i) the civilian's chance of survival decreases over time, irrespective of his action; (ii) there are desperate situations in which a civilian should fight back even if he is unarmed; (iii) carrying a firearm does not always increase the civilian's chance of survival during an active shooter incident; (iv) carrying a firearm makes "resistance" more likely to be the optimal action of the civilian; (v) "hide" might be the best action even if the civilian is armed; (vi) more armed civilians might increase a civilian's chance of survival, but it is not always the case.
After the end of World War II, the commitment to confine scientific activities in universities and research institutions to peaceful and civilian purposes has entered, in the form of {\it Civil Clauses}, the charters of many research institutions and universities. In the wake of recent world events, the relevance and scope of such Civil Clauses has been questioned in reports issued by some governments and by the EU Commission, a development that opens the door to a possible blurring of the distinction between peaceful and military research. This paper documents the reflections stimulated by a panel discussion on this issue recently organized by the Science4Peace Forum. We review the adoptions of Civil Clauses in research organizations and institutions in various countries, present evidence of the challenges that are emerging to such Civil Clauses, and collect arguments in favour of maintaining the purely civilian and peaceful focus of public (non-military) research.
In this article, we explore how the escalating victimization of civilians during civil wars is mirrored in the fragmented distribution of territorial control, focusing on the Colombian armed conflict. Through an exhaustive characterization of the topology of bipartite and projected networks of municipalities, we describe changes in territorial configurations across different periods between 1978 and 2007. By employing stochastic block models for count data, we show that, during periods dominated by a small set of actors, the networks adopt a centralized node periphery structure, whereas during times of widespread conflict, areas of influence overlap in complex ways. Our findings also suggest the existence of cohesive municipal communities shaped by both geographic proximity and affinities between armed structures, as well as internally dispersed groups with a high likelihood of interaction. As the spatial distribution shifts toward a more fragmented arrangement, the average interaction intensity between communities predicted by the stochastic block model approaches that within communities, indicating a weakening of modular structure and increased inter community connectivity.
During large-scale crises disrupting cellular and Internet infrastructure, civilians lack reliable methods for communication, aid coordination, and access to trustworthy information. This paper presents a unified emergency communication system integrating a low-power, long-range network with a crisis-oriented smartphone application, enabling decentralized and off-grid civilian communication. Unlike previous solutions separating physical layer resilience from user layer usability, our design merges these aspects into a cohesive crisis-tailored framework. The system is evaluated in two dimensions: communication performance and application functionality. Field experiments in urban Zürich demonstrate that the 868 MHz band, using the LongFast configuration, achieves a communication range of up to 1.2 km with 92% Packet Delivery Ratio, validating network robustness under real-world infrastructure degraded conditions. In parallel, a purpose-built mobile application featuring peer-to-peer messaging, identity verification, and community moderation was evaluated through a requirements-based analysis.
Prior work has demonstrated that incorporating well-known quantum tunnelling (QT) probability into neural network models effectively captures important nuances of human perception, particularly in the recognition of ambiguous objects and sentiment analysis. In this paper, we employ novel QT-based neural networks and assess their effectiveness in distinguishing customised CIFAR-format images of military and civilian vehicles, as well as sentiment, using a proprietary military-specific vocabulary. We suggest that QT-based models can enhance multimodal AI applications in battlefield scenarios, particularly within human-operated drone warfare contexts, imbuing AI with certain traits of human reasoning.
This report describes trade-offs in the design of international governance arrangements for civilian artificial intelligence (AI) and presents one approach in detail. This approach represents the extension of a standards, licensing, and liability regime to the global level. We propose that states establish an International AI Organization (IAIO) to certify state jurisdictions (not firms or AI projects) for compliance with international oversight standards. States can give force to these international standards by adopting regulations prohibiting the import of goods whose supply chains embody AI from non-IAIO-certified jurisdictions. This borrows attributes from models of existing international organizations, such as the International Civilian Aviation Organization (ICAO), the International Maritime Organization (IMO), and the Financial Action Task Force (FATF). States can also adopt multilateral controls on the export of AI product inputs, such as specialized hardware, to non-certified jurisdictions. Indeed, both the import and export standards could be required for certification. As international actors reach consensus on risks of and minimum standards for advanced AI, a jurisdict
Unidentified Anomalous Phenomena (UAP) have historically been stigmatized and regarded as pseudoscience due to a general lack of robust evidence. Recently, however, the subject has gained interest among astronomers and the military. This review explores how astronomers can enhance our understanding of these enigmatic phenomena by focusing on empirical tests of specific hypotheses (e.g. the hypothesis of extraterrestrial visitations) rather than solely collecting and categorizing data. We compare the investigation of UAP to the process of calibration and interpretations of astronomical discoveries and propose a toy model involving a network of neuro-interface extraterrestrial probes to model exotic UAP. This model aids in predicting probe signatures and behaviour, improving detection methods, and addressing ethical concerns in UAP research.
Military intelligence is underutilized in the study of civil war violence. Declassified records are hard to acquire and difficult to explore with the standard econometrics toolbox. I investigate a contemporary government database of civilians targeted during the Vietnam War. The data are detailed, with up to 45 attributes recorded for 73,712 individual civilian suspects. I employ an unsupervised machine learning approach of cleaning, variable selection, dimensionality reduction, and clustering. I find support for a simplifying typology of civilian targeting that distinguishes different kinds of suspects and different kinds targeting methods. The typology is robust, successfully clustering both government actors and rebel departments into groups that mirror their known functions. The exercise highlights methods for dealing with high dimensional found conflict data. It also illustrates how aggregating measures of political violence masks a complex underlying empirical data generating process as well as a complex institutional reporting process.
This chapter explores moral responsibility for civilian harms by human-artificial intelligence (AI) teams. Although militaries may have some bad apples responsible for war crimes and some mad apples unable to be responsible for their actions during a conflict, increasingly militaries may 'cook' their good apples by putting them in untenable decision-making environments through the processes of replacing human decision-making with AI determinations in war making. Responsibility for civilian harm in human-AI military teams may be contested, risking operators becoming detached, being extreme moral witnesses, becoming moral crumple zones or suffering moral injury from being part of larger human-AI systems authorised by the state. Acknowledging military ethics, human factors and AI work to date as well as critical case studies, this chapter offers new mechanisms to map out conditions for moral responsibility in human-AI teams. These include: 1) new decision responsibility prompts for critical decision method in a cognitive task analysis, and 2) applying an AI workplace health and safety framework for identifying cognitive and psychological risks relevant to attributions of moral respons
We perform the largest known computational analysis of Canadian news narratives about police-involved deaths, spanning 4,000 articles from the last quarter-century. We develop a novel computational model, PerspectiveGap, grounded in prior sociological work on media representation of policing. We find that reporting on police-involved deaths on average features perspectives from state bureaucrats at a rate nearly three times as much as perspectives from other members of the public, including relatives, community members, eyewitnesses, lawyers representing the family, or civil liberties groups. A considerable fraction of articles contain no points of view from civilian actors, though civilian representation has increased in recent years. Qualitatively, we find that state bureaucrats' accounts of these deaths tend to be clinical and procedural, while civilian discourse carries considerably more emotional valence. The PerspectiveGap framework developed here can be contextualized to other jurisdictions, offering a scalable approach for analyzing how media systems construct narratives around policing and accountability.
Effective crisis response requires spatially grounded communication that bridges linguistic guidance of civilians with the physical environment, accounting for structural bottlenecks, evolving threats, and agent-specific contexts. Yet, current NLP research in crisis communication remains mainly limited to static, text-only classification settings, overlooking the critical communicative role of AI operators in dynamic, embodied scenarios. We address this gap with a novel benchmarking framework for evaluating Vision-Language Models (VLMs) tasked with guiding civilian agents through simulated evacuations. We test two communication strategies (narrowcast vs. broadcast), two environment representations (visual vs. graph-based), and two threat behaviors (static vs. moving) across nine maps of varying structural complexity. Our results show that Narrowcast consistently reduces civilian Fail rates compared to Broadcast across all difficulty levels. Guidance quality depends heavily on how the VLM operator represents the world: the visual modality drives performance, while adding an adjacency graph is model-dependent and often harmful. Moving threats raise Fail rates across all conditions as
Inferring racial discrimination in police use of force -- the average causal effect of civilian race on use of force -- requires two assumptions about policing prior to potential use of force: that officers do not discriminate in whom they would stop (no discrimination in stops) and that, conditional on patrol context, the probability that an encounter is with a minority rather than a white civilian does not vary across encounters (no bias in encounters). As Knox et al. (2020) show, violations of the first can mask racial disparity in force. Whether it reflects discrimination in force also depends on the second. Existing sensitivity analyses address one assumption at a time. We develop a framework that varies both sequentially and apply it to NYPD Stop, Question, and Frisk data (2003--2013). Under plausible levels of discrimination in stops, we find substantial racial disparity in force. However, the conclusion that this disparity reflects discrimination is fragile to modest departures from no bias in encounters that census-based calibration suggests are demographically feasible. By jointly addressing both confounding channels, the framework reveals how they interact in ways that s
This paper examines how the Joint Institute for Nuclear Research (JINR), an international organization formally committed to peaceful science, is deeply embedded in an ecosystem of military-industrial enterprises in the city of Dubna in Russia, contributing to training specialists and developing technologies used in Russia's military operations, including attacks on civilian facilities in Ukraine. It also shows how JINR collaborates with scientific institutions on the Ukrainian territories occupied by Russia, legitimizing the occupation and exposing international partners to legal and ethical risks. Despite these ties, JINR maintains broad international collaborations, allowing its scientists and engineers to access advanced technologies and indirectly support Russia's military capabilities, highlighting the need for greater awareness in the global scientific community and coordinated sanctions enforcement.
This paper analyzes the macroeconomic consequences of military spending and militarization within a dynamic growth framework. Building on a Keynesian goods-market model, we examine how the allocation of government expenditure between civilian and military sectors affects capital accumulation and technological progress. Military spending generates opposing effects: it stimulates aggregate demand and may support innovation through defense-related research, but it also crowds out civilian investment and creates structural rigidities. We formalize these mechanisms in a stylized endogenous-growth model in which productivity depends on the degree of militarization, producing a non-linear relationship between the military burden and long-run growth. Calibrated simulations show that moderate levels of military spending can temporarily support growth, whereas excessive militarization reduces long-run development. We further illustrate the asymmetric growth costs of conflict using a simple two-country war simulation between an advanced economy and a sanctioned middle-income economy.
Emergency vehicle (EV) response time is a critical determinant of survival outcomes, yet deployed signal preemption strategies remain reactive and uncontrollable. We propose a return-conditioned framework for emergency corridor optimization based on the Decision Transformer (DT). By casting corridor optimization as offline, return-conditioned sequence modeling, our approach (1) eliminates online environment interaction during policy learning, (2) enables dispatch-level urgency control through a single target-return scalar, and (3) extends to multi-agent settings via a Multi-Agent Decision Transformer (MADT) with graph attention for spatial coordination. On the LightSim simulator, DT reduces average EV travel time by 37.7% relative to fixed-timing preemption on a 4x4 grid (88.6 s vs. 142.3 s), achieving the lowest civilian delay (11.3 s/veh) and fewest EV stops (1.2) among all methods, including online RL baselines that require environment interaction. MADT further improves on larger grids, overtaking DT with 45.2% reduction on 8x8 via graph-attention coordination. Return conditioning produces a smooth dispatch interface: varying the target return from 100 to -400 trades EV travel t
As military organisations consider integrating large language models (LLMs) into command and control (C2) systems for planning and decision support, understanding their behavioural tendencies is critical. This study develops a benchmarking framework for evaluating aspects of legal and moral risk in targeting behaviour by comparing LLMs acting as agents in multi-turn simulated conflict. We introduce four metrics grounded in International Humanitarian Law (IHL) and military doctrine: Civilian Target Rate (CTR) and Dual-use Target Rate (DTR) assess compliance with legal targeting principles, while Mean and Max Simulated Non-combatant Casualty Value (SNCV) quantify tolerance for civilian harm. We evaluate three frontier models, GPT-4o, Gemini-2.5, and LLaMA-3.1, through 90 multi-agent, multi-turn crisis simulations across three geographic regions. Our findings reveal that off-the-shelf LLMs exhibit concerning and unpredictable targeting behaviour in simulated conflict environments. All models violated the IHL principle of distinction by targeting civilian objects, with breach rates ranging from 16.7% to 66.7%. Harm tolerance escalated through crisis simulations with MeanSNCV increasing
This paper examines the dual-use challenges of foundation models and the consequent risks they pose for international security. As artificial intelligence (AI) models are increasingly tested and deployed across both civilian and military sectors, distinguishing between these uses becomes more complex, potentially leading to misunderstandings and unintended escalations among states. The broad capabilities of foundation models lower the cost of repurposing civilian models for military uses, making it difficult to discern another state's intentions behind developing and deploying these models. As military capabilities are increasingly augmented by AI, this discernment is crucial in evaluating the extent to which a state poses a military threat. Consequently, the ability to distinguish between military and civilian applications of these models is key to averting potential military escalations. The paper analyzes this issue through four critical factors in the development cycle of foundation models: model inputs, capabilities, system use cases, and system deployment. This framework helps elucidate the points at which ambiguity between civilian and military applications may arise, leadin
Which factors determine AI's propensity to support military intervention? While the use of AI in high-stakes decision-making is growing exponentially, we still lack systematic analysis of the key drivers embedded in these models. This paper conducts a conjoint experiment in which large language models (LLMs) from leading providers (OpenAI, Anthropic, Google) are asked to decide on military intervention across 128 vignettes, with each vignette run 10 times. This design enables a systematic assessment of AI decision-making in military contexts. The results are remarkably consistent across models: all models place substantial weight on the probability of success and domestic support, prioritizing these factors over civilian casualties, economic shock, or international sanctions. The paper then tests whether LLMs are sensitive to context by introducing different motivations for intervention. The scoring is indeed context-dependent; however, probability of victory remains the most important factor in all scenarios. Finally, the paper evaluates numerical sensitivity and finds that models display some responsiveness to the scale of civilian casualties but no detectable sensitivity to the
Violence is commonly linked with large urban areas, and as a social phenomenon, it is presumed to scale super-linearly with population size. This study explores the hypothesis that smaller, isolated cities in Africa may experience a heightened intensity of violence against civilians. It aims to investigate the correlation between the risk of experiencing violence with a city's size and its geographical isolation. Over a 20-year period, the incidence of civilian casualties has been analysed to assess lethality in relation to varying degrees of isolation and city sizes. African cities are categorised by isolation (number of highway connections) and centrality (the estimated frequency of journeys). Findings suggest that violence against civilians exhibits a sub-linear pattern, with larger cities witnessing fewer casualties per 100,000 inhabitants. Remarkably, individuals in isolated cities face a quadrupled risk of a casualty compared to those in more connected cities.