Although increasingly used for research, electronic health records (EHR) often lack gold-standard assessment of key data elements. Linking EHRs to other data sources with higher-quality measurements can improve statistical inference, but such analyses must account for selection bias if the linked data source arises from a non-probability sample. We propose a set of novel estimators targeting the average treatment effect (ATE) that combine information from binary outcomes measured with error in a large, population-representative EHR database with gold-standard outcomes obtained from a smaller validation sample subject to selection bias. We evaluate our approach in extensive simulations and an analysis of data from the Adult Changes in Thought (ACT) study, a longitudinal study of incident dementia in a cohort of Kaiser Permanente Washington members with linked EHR data. For a subset of deceased ACT participants who consented to brain autopsy prior to death, gold-standard measures of Alzheimer's disease neuropathology are available. Our proposed estimators reduced bias and improved efficiency for the ATE, facilitating valid inference with EHR data when key data elements are ascertaine
This paper is motivated by basic complexity and probability questions about permanents of random matrices over finite fields, and in particular, about properties separating the permanent and the determinant. Fix $q = p^m$ some power of an odd prime, and let $k \leq n$ both be growing. For a uniformly random $n \times k$ matrix $A$ over $\mathbb{F}_q$, we study the probability that all $k \times k$ submatrices of $A$ have zero permanent; namely that $A$ does not have full "permanental rank". When $k = n$, this is simply the probability that a random square matrix over $\mathbb{F}_q$ has zero permanent, which we do not understand. We believe that the probability in this case is $\frac{1}{q} + o(1)$, which would be in contrast to the case of the determinant, where the answer is $\frac{1}{q} + Ω_q(1)$. Our main result is that when $k$ is $O(\sqrt{n})$, the probability that a random $n \times k$ matrix does not have full permanental rank is essentially the same as the probability that the matrix has a $0$ column, namely $(1 +o(1)) \frac{k}{q^n}$. In contrast, for determinantal (standard) rank the analogous probability is $Θ(\frac{q^k}{q^n})$. At the core of our result are some basic lin
The permanent of an $n \times n$ matrix $M = (m_{ij})$ is defined as $\mathrm{per}(M) = \sum_{σ\in S_n} \prod_{i=1}^n m_{i,σ(i)}$, where $S_n$ denotes the symmetric group on $\{1,2,\ldots,n\}$. The permanental polynomial of $M$, is defined by $ψ(M;x) = \mathrm{per}(xI_n - M)$. We study two fundamental variants: the Laplacian permanental polynomial $ψ(L(G);x)$ and signless Laplacian permanental polynomial $ψ(Q(G);x)$ of a graph $G$. A graph is said to be {determined} by its (signless) Laplacian permanental polynomial if no other non-isomorphic graph shares the same polynomial. A graph is combinedly determined when isomorphism is guaranteed by the equality of both polynomials. Characterizing which graphs are determined by their(signless) Laplacian permanental polynomials is an interesting problem. This paper investigates the permanental characterization problem for several families of starlike graphs, including: spider graphs (tree), coconut tree, perfect binary tree, corona product of $C_m$ and $K_n$, and $\bar K_n$ for various values of $m$ and $n$. We establish which of these graphs are determined by their Laplacian or signless Laplacian permanental polynomials, and which require
The rank of an n x n matrix A is equal to the size of its largest square submatrix with a nonzero determinant, and it can be computed in O(n^2.37) time. Analogously, the size of the largest square submatrix with nonzero permanent is defined as the permanental rank. Computing the permanent or the coefficients of the permanental polynomial is #P-complete. The permanental nullity is defined as the multiplicity of zero as a root of the permanental polynomial. We establish a permanental analog of the rank-nullity theorem, showing that the sum of the permanental rank and the permanental nullity equals n for symmetric nonnegative matrices, positive semidefinite matrices, and adjacency matrices of balanced signed graphs. Using this theorem, we can compute the permanental nullity for symmetric nonnegative matrices and adjacency matrices of balanced signed graphs in polynomial time. For symmetric matrices with entries in {0, plus or minus 1}, we also provide a complete characterization of when the permanental rank-nullity identity holds.
Let $Y$ be a symmetric Borel right process with locally compact state space $T\subseteq R^{1}$ and potential densities $u(x,y)$ with respect to some $σ$-finite measure on $T$. Let $g$ and $f$ be finite excessive functions for $ Y$. Set $$ u_{g, f}(x,y)= u(x,y)+g(x)f(y),\qquad x,y\in T.$$ In this paper we take $Y$ to be a symmetric Lévy process, or a diffusion, that is killed at the end of an independent exponential time or the first time it hits 0. Under general smoothness conditions on $g$, $f$, $u$ and points $d\in T$, laws of the iterated logarithm are found for $X_{k/2} =\{X_{k/2}(t), t\in T \}$, a $k/2-$permanental process with kernel $ \{u_{g, f}(x,y),x,y\in T \}$, of the following form: For all integers $k\geq 1$, $$\limsup_{x \to 0}\frac{| X_{k/2}( d+x)- X_{k/2} (d)|}{ \left( 2 σ^{2}\left(x\right)\log\log 1/x\right)^{1/2}}= \left( 2 X _{k/2} (d)\right)^{1/2}, \qquad a.s. ,$$ where, $$σ^2(x)=u(d+x,d+x)+u(x,x)-2u(d+x,x).$$ Using these limit theorems and the Eisenbaum Kaspi Isomorphism Theorem, laws of the iterated logarithm are found for the local times of certain Markov processes with potential densities that have the form of $ \{u_{g, f}(x,y),x,y\in T \}$ or are slight modi
As a variant of the Ulam's vertex reconstruction conjecture and the Harary's edge reconstruction conjecture, Cvetković and Schwenk posed independently the following problem: Can the characteristic polynomial of a simple graph $G$ with vertex set $V$ be reconstructed from the characteristic polynomials of all subgraphs in $\{G-v|v\in V\}$ for $|V|\geq 3$? This problem is still open. A natural problem is: Can the characteristic polynomial of a simple graph $G$ with edge set $E$ be reconstructed from the characteristic polynomials of all subgraphs in $\{G-e|e\in E\}$? In this paper, we prove that if $|V| eq |E|$, then the characteristic polynomial of $G$ can be reconstructed from the characteristic polynomials of all subgraphs in $\{G-uv, G-u-v|uv\in E\}$, and the similar result holds for the permanental polynomial of $G$. We also prove that the Laplacian (resp. signless Laplacian) characteristic polynomial of $G$ can be reconstructed from the Laplacian (resp. signless Laplacian) characteristic polynomials of all subgraphs in $\{G-e|e\in E\}$ (resp. if $|V| eq |E|$).
Let ${\rm Mat}_n(\mathbb{F})$ denote the set of square $n\times n$ matrices over a field $\mathbb{F}$ of characteristic different from two. The permanental rank ${\rm prk}\,(A)$ of a matrix $A \in{\rm Mat}_{n}(\mathbb{F})$ is the size of the maximal square submatrix in $A$ with nonzero permanent. By $Λ^{k}$ and $Λ^{\leq k}$ we denote the subsets of matrices $A \in {\rm Mat}_{n}(\mathbb{F})$ with ${\rm prk}\,(A) = k$ and ${\rm prk}\,(A) \leq k$, respectively. In this paper for each $1 \leq k \leq n-1$ we obtain a complete characterization of linear maps $T: {\rm Mat}_{n}(\mathbb{F}) \to {\rm Mat}_{n}(\mathbb{F})$ satisfying $T(Λ^{\leq k}) = Λ^{\leq k}$ or bijective linear maps satisfying $T(Λ^{\leq k}) \subseteq Λ^{\leq k}$. Moreover, we show that if $\mathbb{F}$ is an infinite field, then $Λ^{k}$ is Zariski dense in $Λ^{\leq k}$ and apply this to describe such bijective linear maps satisfying $T(Λ^{k}) \subseteq Λ^{k}$.
作者:Eric J. Rose, Erica E. M. Moodie, Susan Shortreed
Data-driven methods for personalizing treatment assignment have garnered much attention from clinicians and researchers. Dynamic treatment regimes formalize this through a sequence of decision rules that map individual patient characteristics to a recommended treatment. Observational studies are commonly used for estimating dynamic treatment regimes due to the potentially prohibitive costs of conducting sequential multiple assignment randomized trials. However, estimating a dynamic treatment regime from observational data can lead to bias in the estimated regime due to unmeasured confounding. Sensitivity analyses are useful for assessing how robust the conclusions of the study are to a potential unmeasured confounder. A Monte Carlo sensitivity analysis is a probabilistic approach that involves positing and sampling from distributions for the parameters governing the bias. We propose a method for performing a Monte Carlo sensitivity analysis of the bias due to unmeasured confounding in the estimation of dynamic treatment regimes. We demonstrate the performance of the proposed procedure with a simulation study and apply it to an observational study examining tailoring the use of anti