Standard count models such as the Poisson and Negative Binomial models often fail to capture the large proportion of zero claims commonly observed in insurance data. To address such issue of excessive zeros, zero-inflated and hurdle models introduce additional parameters that explicitly account for excess zeros, thereby improving the joint representation of zero and positive claim outcomes. These models have further been extended with random effects to accommodate longitudinal dependence and unobserved heterogeneity. However, their consistency with fundamental probabilistic principles in insurance, particularly stochastic monotonicity, has not been formally examined. This paper provides a rigorous analysis showing that standard counting random-effect models for excessive zeros may violate this property, leading to inconsistencies in posterior credibility. We then propose new classes of counting random-effect models that both accommodate excessive zeros and ensure stochastic monotonicity, thereby providing fair and theoretically coherent credibility adjustments as claim histories evolve.
Sycophancy in language models is typically studied as excessive agreement or validation, while explicit praise and flattery have received comparatively little attention. We argue that sycophantic praise is a distinct alignment problem that cannot be reliably measured using current methods. We introduce a parameterized framework that measures whether praise is excessive relative to contribution quality and expected user ability. We show that our framework substantially outperforms generic LLM judges in agreement with human annotations, and that sycophantic praise occurs far more frequently in social and interpretive domains than in objective reasoning settings. Together, these findings position praise calibration as a distinct alignment challenge.
Recent reasoning large language models (LLMs), such as OpenAI o1 and DeepSeek-R1, exhibit strong performance on complex tasks through test-time inference scaling. However, prior studies have shown that these models often incur significant computational costs due to excessive reasoning, such as frequent switching between reasoning trajectories (e.g., underthinking) or redundant reasoning on simple questions (e.g., overthinking). In this work, we expose a novel threat: adversarial inputs can be crafted to exploit excessive reasoning behaviors and substantially increase computational overhead without compromising model utility. Therefore, we propose a novel loss framework consisting of three components: (1) Priority Cross-Entropy Loss, a modification of the standard cross-entropy objective that emphasizes key tokens by leveraging the autoregressive nature of LMs; (2) Excessive Reasoning Loss, which encourages the model to initiate additional reasoning paths during inference; and (3) Delayed Termination Loss, which is designed to extend the reasoning process and defer the generation of final outputs. We optimize and evaluate our attack for the GSM8K and ORCA datasets on DeepSeek-R1-Dis
Excessive use of smartphones is a worldwide known issue. In this study, we proposed a notification-based intervention approach to reduce smartphone overuse without making the user feel any annoyance or irritation. Most of the work in this field tried to reduce smartphone overuse by making smartphone use more difficult for the user. In our user study (n = 109), we found that 19.3% of the participants are unwilling to use any usage-limiting application because a) they do not want their smartphone activities to get restricted or b) those applications are annoying. Following that, we devised a hypothesis to minimize smartphone usage among undergraduates. Finally, we designed a prototype for Android, "App Usage Monitor," and conducted a 3-week experiment through which we found proof of concept for our hypothesis. In our prototype, we combined techniques such as nudge and visualization to increase self-awareness among the user by leveraging notifications.
We classify protocols of entanglement distribution as excessive and non-excessive ones. In a non-excessive protocol, the gain of entanglement is bounded by the amount of entanglement being communicated between the remote parties, while excessive protocols violate such bound. We first present examples of excessive protocols that achieve a significant entanglement gain. Next we consider their use in noisy scenarios, showing that they improve entanglement achieved in other ways and for some situations excessive distribution is the only possibility of gaining entanglement.
We propose a new best-of-both-worlds algorithm for bandits with variably delayed feedback. In contrast to prior work, which required prior knowledge of the maximal delay $d_{\mathrm{max}}$ and had a linear dependence of the regret on it, our algorithm can tolerate arbitrary excessive delays up to order $T$ (where $T$ is the time horizon). The algorithm is based on three technical innovations, which may all be of independent interest: (1) We introduce the first implicit exploration scheme that works in best-of-both-worlds setting. (2) We introduce the first control of distribution drift that does not rely on boundedness of delays. The control is based on the implicit exploration scheme and adaptive skipping of observations with excessive delays. (3) We introduce a procedure relating standard regret with drifted regret that does not rely on boundedness of delays. At the conceptual level, we demonstrate that complexity of best-of-both-worlds bandits with delayed feedback is characterized by the amount of information missing at the time of decision making (measured by the number of outstanding observations) rather than the time that the information is missing (measured by the delays).
Given two positive integers l and m, with l \le m, an [l,m]-covering of a graph G is a set M of matchings of G whose union is the edge set of G and such that l \le |L| \le m for every matching L of M. An [l,m]-covering M of G is an excessive [l,m]-factorization of G if the cardinality of M is as small as possible. The number of matchings in an excessive [l,m]-factorization of G (or \infty, if G does not admit an excessive [l,m]-factorization) is a graph parameter called the excessive [l,m]-index of G and denoted by χ'[l,m](G). In this paper we study such parameter. Our main result is a general formula for the excessive [l,m]-index of a graph G in terms of other graph parameters. Furthermore, we give a polynomial time algorithm which computes χ'[l,m](G) and outputs an excessive [l,m]-factorization of G, whenever the latter exists.
The excessive [m]-index of a graph G is the minimum number of matchings of size m needed to cover the edge-set of G. We call a graph G [m]-coverable if its excessive [m]-index is finite. Obviously the excessive [1]-index is |E(G)| for all graphs and it is an easy task the computation of the excessive [2]-index for a [2]-coverable graph. The case m=3 is completely solved by Cariolaro and Fu in 2009. In this paper we prove a general formula to compute the excessive [4]-index of a tree and we conjecture a possible generalization for any value of m. Furthermore, we prove that such a formula does not work for the excessive [4]-index of an arbitrary graph.
A diffusion spider is a strong Markov process with continuous paths taking values on a graph with one vertex and a finite number of edges (of infinite length). An example is Walsh's Brownian spider where the process on each edge behaves as Brownian motion. We calculate the density of the resolvent kernel in terms of the characteristics of the underlying diffusion. Excessive functions are studied via the Martin boundary theory. The main result is an explicit expression for the representing measure of a given excessive function. These results are used to solve optimal stopping problems for diffusion spiders.
The principle of smooth fit is probably the most used tool to find solutions to optimal stopping problems of one-dimensional diffusions. It is important, e.g., in financial mathematical applications to understand in which kind of models and problems smooth fit can fail. In this paper we connect - in case of one-dimensional diffusions - the validity of smooth fit and the differentiability of excessive functions. The basic tool to derive the results is the representation theory of excessive functions; in particular, the Riesz and Martin representations. It is seen that the differentiability may not hold in case the speed measure of the diffusion or the representing measure of the excessive function has atoms. As an example, we study optimal stopping of sticky Brownian motion. It is known that the validity of the smooth fit in this case depends on the value of the discounting parameter (when the other parameters are fixed). We decompose the size of the jump in the derivative of the value function in two factors. The first one is due to the atom of the representing measure and the second one due to the atom of the speed measure.
This paper explores transverse coordinates for the purpose of orbitally stabilizing periodic motions of nonlinear, control-affine dynamical systems. It is shown that the dynamics of any (minimal or excessive) set of transverse coordinates, which are defined in terms of a particular parameterization of the motion and a strictly state-dependent projection operator recovering the parameterizing variable, admits a (transverse) linearization along the target motion, with explicit expressions stated. Special focus is then placed on a generic excessive set of orthogonal coordinates, revealing a certain limitation of the "excessive" transverse linearization for the purpose of control design. To overcome this limitation, a linear comparison system is introduced, and conditions are stated for when the asymptotic stability of its origin corresponds to the asymptotic stability of the origin of linearized transverse dynamics. This allows for the construction of feedback controllers utilizing this comparison system which, when implemented on the dynamical system, renders the desired motion asymptotically stable in the orbital sense.
Underwater robots in shallow waters usually suffer from strong wave forces, which may frequently exceed robot's control constraints. Learning-based controllers are suitable for disturbance rejection control, but the excessive disturbances heavily affect the state transition in Markov Decision Process (MDP) or Partially Observable Markov Decision Process (POMDP). Also, pure learning procedures on targeted system may encounter damaging exploratory actions or unpredictable system variations, and training exclusively on a prior model usually cannot address model mismatch from the targeted system. In this paper, we propose a transfer learning framework that adapts a control policy for excessive disturbance rejection of an underwater robot under dynamics model mismatch. A modular network of learning policies is applied, composed of a Generalized Control Policy (GCP) and an Online Disturbance Identification Model (ODI). GCP is first trained over a wide array of disturbance waveforms. ODI then learns to use past states and actions of the system to predict the disturbance waveforms which are provided as input to GCP (along with the system state). A transfer reinforcement learning algorithm
Pressure ulcers (PU) are known to be a high-cost disease with a risk of severe morbidity. This work evaluates a new clinical strategy based on an innovative medical device (Tongue Display Unit-TDU) that implements perceptive supplementation in order to reduce prolonged excessive pressure, recognized as one of the main causes of PU. A randomized, controlled, parallel-group trial was carried out with 12 subjects with spinal cord injuries (SCI). Subjects were assigned to the control (without TDU, n=6) or intervention (with TDU, n=5) group. Each subject took part in two sessions, during which the subject, seated on a pressure map sensor, watched a movie for one hour. The TDU was activated during the second session of the intervention group. Intention-to-treat analysis showed that the improvement in adequate weight shifting between the two sessions was higher in the intervention group (0.84 [0.24; 0.89]) than in the control group (0.01 [-0.01; 0.09]; p=0.004) and that the ratio of prolonged excessive pressure between the two sessions was lower in the intervention group (0.74 [0.37; 1.92]) than in the control group (1.72 [1.32; 2.56]; p=0.06). The pressure map sensor was evaluated as bei
There has been tremendous recent progress on equilibrium-finding algorithms for zero-sum imperfect-information extensive-form games, but there has been a puzzling gap between theory and practice. First-order methods have significantly better theoretical convergence rates than any counterfactual-regret minimization (CFR) variant. Despite this, CFR variants have been favored in practice. Experiments with first-order methods have only been conducted on small- and medium-sized games because those methods are complicated to implement in this setting, and because CFR variants have been enhanced extensively for over a decade they perform well in practice. In this paper we show that a particular first-order method, a state-of-the-art variant of the excessive gap technique---instantiated with the dilated entropy distance function---can efficiently solve large real-world problems competitively with CFR and its variants. We show this on large endgames encountered by the Libratus poker AI, which recently beat top human poker specialist professionals at no-limit Texas hold'em. We show experimental results on our variant of the excessive gap technique as well as a prior version. We introduce a n
Despite their impressive performance, deep neural networks exhibit striking failures on out-of-distribution inputs. One core idea of adversarial example research is to reveal neural network errors under such distribution shifts. We decompose these errors into two complementary sources: sensitivity and invariance. We show deep networks are not only too sensitive to task-irrelevant changes of their input, as is well-known from epsilon-adversarial examples, but are also too invariant to a wide range of task-relevant changes, thus making vast regions in input space vulnerable to adversarial attacks. We show such excessive invariance occurs across various tasks and architecture types. On MNIST and ImageNet one can manipulate the class-specific content of almost any image without changing the hidden activations. We identify an insufficiency of the standard cross-entropy loss as a reason for these failures. Further, we extend this objective based on an information-theoretic analysis so it encourages the model to consider all task-dependent features in its decision. This provides the first approach tailored explicitly to overcome excessive invariance and resulting vulnerabilities.
Let $\mathfrak X$ be a Hunt process on a locally compact space $X$ such that the set $\mathcal E_{\mathfrak X}$ of its Borel measurable excessive functions separates points, every function in $\mathcal E_{\mathfrak X}$ is the supremum of its continuous minorants in $\mathcal E_{\mathfrak X}$ and there are strictly positive continuous functions $v,w\in\mathcal E_{\mathfrak X}$ such that $v/w$ vanishes at infinity. A numerical function $u\ge 0$ on $X$ is said to be nearly hyperharmonic, if $\int^\ast u\circ X_{τ_V}\,dP^x\le u(x)$ for all $x\in X$ and relatively compact open neighborhoods $V$ of $x$, where $τ_V$ denotes the exit time of $V$. For every such function $u$, its lower semicontinous regularization $\hat u$ is excessive. The main purpose of the paper is to give a short, complete and understandable proof for the statement that every Borel measurable nearly hyperharmonic function on $X$ is the infimum of its majorants in $E_{\mathfrak X}$. The major novelties of our approach are the following: 1. A quick reduction to the special case, where starting at $x\in X$ with $u(x)<\infty$ the expected number of times the process $\mathfrak X$ visits the set of points $y\in X$, where
Three regions of excessive flux of cosmic rays with energies of the order of PeV are found in the experimental data of the EAS MSU array at a confidence level greater than 4σ. For two of them, there are similar regions in the experimental data of the EAS-1000 Prototype array. One of the interesting features of the regions is the absence of supernova remnants in their vicinities, traditionally considered as the main sources of Galactic cosmic rays, but the presence of isolated pulsars, some of which are able to accelerate heavy nuclei up to energies close to PeV. In our opinion, this favors the assumption that isolated pulsars are able to contribute to the flux of Galactic cosmic rays more than is usually assumed.
A digraph $G$ is \emph{$k$-geodetic} if for any pair $u,v \in V(G)$ there is at most one $u,v$-walk of length not exceeding $k$. The order of a $k$-geodetic digraph with minimum out-degree $d$ is bounded below by the directed Moore bound $M(d,k) = 1 + d + d^2+ \cdots +d^k$. It is known that the Moore bound cannot be achieved for $d,k \geq 2$. A $k$-geodetic digraph with minimum degree $d$ and order one greater than the Moore bound has \emph{excess one}. In this paper we prove a conjecture that no excess one digraphs exist for $d,k \geq 2$, thus complementing the result of Bannai and Ito on the non-existence of undirected graphs with excess one.
Many years have passed since the conception of the quintessential method of shortcut to adiabaticity known as counterdiabatic driving (or transitionless quantum driving). Yet, this method appears to be energetically cost-free and thus continually challenges the task of quantifying the amount of energy it demands to be accomplished. This paper proposes that the energy cost of controlling a closed quantum system using the counterdiabatic method can also be assessed using the instantaneous excess work during the process and related quantities, as the time-averaged excess work. Starting from the Mandelstam-Tamm bound for driven dynamics, we have shown that the speed-up of counterdiabatic driving is linked with the spreading of energy between the eigenstates of the total Hamiltonian, which is necessarily accompanied by transitions between these eigenstates. Nonetheless, although excess work can be used to quantify energetically these transitions, it is well known that the excess work is zero throughout the entire process under counterdiabatic driving. To recover the excess work as an energetic cost quantifier for counterdiabatic driving, we will propose a different interpretation of the
Multi-task learning (MTL) considers learning a joint model for multiple tasks by optimizing a convex combination of all task losses. To solve the optimization problem, existing methods use an adaptive weight updating scheme, where task weights are dynamically adjusted based on their respective losses to prioritize difficult tasks. However, these algorithms face a great challenge whenever label noise is present, in which case excessive weights tend to be assigned to noisy tasks that have relatively large Bayes optimal errors, thereby overshadowing other tasks and causing performance to drop across the board. To overcome this limitation, we propose Multi-Task Learning with Excess Risks (ExcessMTL), an excess risk-based task balancing method that updates the task weights by their distances to convergence instead. Intuitively, ExcessMTL assigns higher weights to worse-trained tasks that are further from convergence. To estimate the excess risks, we develop an efficient and accurate method with Taylor approximation. Theoretically, we show that our proposed algorithm achieves convergence guarantees and Pareto stationarity. Empirically, we evaluate our algorithm on various MTL benchmarks