This paper presents an approach for surgical phase recognition using video data, aiming to provide a comprehensive understanding of surgical procedures for automated workflow analysis. The advent of robotic surgery, digitized operating rooms, and the generation of vast amounts of data have opened doors for the application of machine learning and computer vision in the analysis of surgical videos. Among these advancements, Surgical Phase Recognition(SPR) stands out as an emerging technology that has the potential to recognize and assess the ongoing surgical scenario, summarize the surgery, evaluate surgical skills, offer surgical decision support, and facilitate medical training. In this paper, we analyse and evaluate both frame-based and video clipping-based phase recognition on thoracic surgery dataset consisting of 11 classes of phases. Specifically, we utilize ImageNet ViT for image-based classification and VideoMAE as the baseline model for video-based classification. We show that Masked Video Distillation(MVD) exhibits superior performance, achieving a top-1 accuracy of 72.9%, compared to 52.31% achieved by ImageNet ViT. These findings underscore the efficacy of video-based cl
Purpose: We investigated the utilization of privacy-preserving, locally-deployed, open-source Large Language Models (LLMs) to extract diagnostic information from free-text cardiovascular magnetic resonance (CMR) reports. Materials and Methods: We evaluated nine open-source LLMs on their ability to identify diagnoses and classify patients into various cardiac diagnostic categories based on descriptive findings in 109 clinical CMR reports. Performance was quantified using standard classification metrics including accuracy, precision, recall, and F1 score. We also employed confusion matrices to examine patterns of misclassification across models. Results: Most open-source LLMs demonstrated exceptional performance in classifying reports into different diagnostic categories. Google's Gemma2 model achieved the highest average F1 score of 0.98, followed by Qwen2.5:32B and DeepseekR1-32B with F1 scores of 0.96 and 0.95, respectively. All other evaluated models attained average scores above 0.93, with Mistral and DeepseekR1-7B being the only exceptions. The top four LLMs outperformed our board-certified cardiologist (F1 score of 0.94) across all evaluation metrics in analyzing CMR reports.
One of the most important issue in soft tissue modeling is to assess the quality of the simulations. A validation protocol is presented based on two CT scans of the patient acquired before and after cranio-maxillofacial surgery. The actual bones repositioning realized during the intervention are accurately measured and reproduced. A evaluation of the soft tissue deformation is then computed using a finite element model of the face. The simulations are therefore compared, qualitatively and quantitatively, with the actual outcome of the surgery. This protocol enable to rigorously evaluate different modeling methods, and to assess the clinical relevance of soft tissue simulation in maxillofacial surgery.
BACKGROUND: Clinical factors influence surgery duration. This study also investigated non-clinical effects. METHODS: 22 months of data about thoracic operations in a large hospital in China were reviewed. Linear and nonlinear regression models were used to predict the duration of the operations. Interactions among predictors were also considered. RESULTS: Surgery duration decreased with the number of operations a surgeon performed in a day (P<0.001). Also, it was found that surgery duration decreased with the number of operations allocated to an OR as long as there were no more than four surgeries per day in the OR (P<0.001), but increased with the number of operations if it was more than four (P<0.01). The duration of surgery was affected by its position in a sequence of surgeries performed by a surgeon. In addition, surgeons exhibited different patterns of the effects of surgery type for surgeries in different positions in the day. CONCLUSIONS: Surgery duration was affected not only by clinical effects but also some non-clinical effects. Scheduling and allocation decisions significantly influenced surgery duration.
Poor adaptation of orbital implants remains a major contributor to postoperative complications and revision surgery. Although preformed orbital plates are widely used to reduce cost and operative time compared with customized implants, surgeons currently lack publicly available tools and standardized metrics to quantitatively compare plate fit across vendors, sizes, and patient anatomy. We developed SlicerOrbitSurgerySim, an open-source extension for the 3D Slicer platform that enables interactive virtual registration, evaluation, and comparison of multiple preformed orbital plates in a patient-specific virtual planning environment. The software generates reproducible quantitative plate-to-orbit distance metrics and visualization tools that support both patient-specific planning and population-level statistical analysis of plate adaptability. By facilitating objective comparison of implant designs and placement strategies, this tool aims to improve preoperative decision-making, reduce intraoperative plate modification, and promote collaborative research and surgical education. Pilot studies, sample datasets, and detailed tutorials are provided to support testing, transparency, and
For a nullhomologous Legendrian knot in a closed contact 3-manifold Y we consider a contact structure obtained by positive rational contact surgery. We prove that in this situation the Heegaard Floer contact invariant of Y is mapped by a surgery cobordism to the contact invariant of the result of contact surgery. In addition we characterize the spin-c structure on the cobordism that induces the relevant map. As a consequence we determine necessary and sufficient conditions for the nonvanishing of the contact invariant after rational surgery when Y is the standard 3-sphere, generalizing previous results of Lisca-Stipsicz and Golla. In fact our methods allow direct calculation of the contact invariant in terms of the rational surgery mapping cone of Ozsváth and Szabó. The proof involves a construction called reducible open book surgery, which reduces in special cases to the capping-off construction studied by Baldwin.
We exhibit an explicit short basis of the Stickelberger ideal of cyclotomic fields of any conductor $m$, i.e., a basis containing only short elements. By definition, an element of $\mathbb{Z}[G_m]$, where $G_m$ denotes the Galois group of the field, is called short whenever it writes as $\sum_{σ\in G_m} \varepsilon_σσ$ with all $\varepsilon_σ\in\{0,1\}$. One ingredient for building such a basis consists in picking wisely generators $α_m(b)$ in a large family of short elements. As a direct practical consequence, we deduce from this short basis an explicit upper bound on the relative class number, that is valid for any conductor. This basis also has several concrete applications, in particular for the cryptanalysis of the Shortest Vector Problem on Ideal lattices.
This is a short proof of Ledoit-Péché's RIE formula for covariance matrices. The proof is based on the Stein formula, which gives a very simple way to derive the result. One of the advantages of this approach is that it shows that the only really needed hypothesis, for the machinery to work, is that the mean of the eigenvalues of the true covariance matrix and the largest of them have the same order.
This paper presents simulations of the impact of tongue surgery on tongue movements and on speech articulation. For this, a 3D biomechanical Finite Element (FE) model of the tongue is used. Muscles are represented within the FE structure by specific subsets of elements. The tongue model is inserted in the upper airways including jaw, palate and pharyngeal walls. Two examples of tongue surgery, which are quite common in the treatment of cancers of the oral cavity are modelled: hemiglossectomy and large resection of the mouth floor. Three kinds of reconstruction are also modelled, assuming flaps with a low, medium or high stiffnesses. The impact of the surgery without any reconstruction and with the three different reconstructions is quantitatively measured and compared during simulated speech production sequences. More precisely, differences in global 3D tongue shape and in velocity patterns during tongue displacements are evaluated.
It is proved the equivalence of the compatibility condition of [A. Ramos, J. Phys. A 44 (2011) 342001, Phys. Lett. A 376 (2012) 3499] with a condition found in [Yadav et al., Ann. Phys. 359 (2015) 46]. The link of Shape Invariance with the existence of a Potential Algebra is reinforced for the rationally extended Shape Invariant potentials. Some examples on X1 and Xl Jacobi and Laguerre cases are given.
The mechanisms by which blast pressure waves cause mild to moderate traumatic brain injury (mTBI) are an open question. Possibilities include acceleration of the head, direct passage of the blast wave via the cranium, and propagation of the blast wave to the brain via a thoracic mechanism. The hypothesis that the blast pressure wave reaches the brain via a thoracic mechanism is considered in light of ballistic and blast pressure wave research. Ballistic pressure waves, caused by penetrating ballistic projectiles or ballistic impacts to body armor, can only reach the brain via an internal mechanism and have been shown to cause cerebral effects. Similar effects have been documented when a blast pressure wave has been applied to the whole body or focused on the thorax in animal models. While vagotomy reduces apnea and bradycardia due to ballistic or blast pressure waves, it does not eliminate neural damage in the brain, suggesting that the pressure wave directly affects the brain cells via a thoracic mechanism. An experiment is proposed which isolates the thoracic mechanism from cranial mechanisms of mTBI due to blast wave exposure. Results have implications for evaluating risk of mTB
Understanding the causes and patient impacts of surgical adverse events will help improve systems and operational practices to avoid incidents in the future. We analyzed the adverse events data related to robotic systems and instruments used in minimally invasive surgery, reported to the U.S. FDA MAUDE database from January 2000 to December 2013. We determined the number of events reported per procedure and per surgical specialty, the most common types of device malfunctions and their impact on patients, and the causes for catastrophic events such as major complications, patient injuries, and deaths. During the study period, 144 deaths (1.4% of the 10,624 reports), 1,391 patient injuries (13.1%), and 8,061 device malfunctions (75.9%) were reported. The numbers of injury and death events per procedure have stayed relatively constant since 2007 (mean = 83.4, 95% CI, 74.2-92.7). Surgical specialties, for which robots are extensively used, such as gynecology and urology, had lower number of injuries, deaths, and conversions per procedure than more complex surgeries, such as cardiothoracic and head and neck (106.3 vs. 232.9, Risk Ratio = 2.2, 95% CI, 1.9-2.6). Device and instrument malf
Video-assisted thoracic surgery (VATS) is a minimally invasive approach for treating early-stage non-small-cell lung cancer. Optimal trocar placement during VATS ensures comprehensive access to the thoracic cavity, provides a panoramic endoscopic view, and prevents instrument crowding. While established principles such as the Baseball Diamond Principle (BDP) and Triangle Target Principle (TTP) exist, surgeons mainly rely on experience and patient-specific anatomy for trocar placement, potentially leading to sub-optimal surgical plans that increase operative time and fatigue. To address this, we present the first virtual reality (VR)-based pre-operative planning tool with tailored data visualization and interaction designs for efficient and optimal VATS trocar placement, following the established surgical principles and consultation with an experienced surgeon. In our preliminary study, we demonstrate the system's application in right upper lung lobectomy, a common thoracic procedure typically using three trocars. A preliminary user study of our system indicates it is efficient, robust, and user-friendly for planning optimal trocar placement, with a great promise for clinical applic
A historical record of a seismic tsunami is identified in the Irish annals for October 720 (all dates herein CE). It is contained in the earliest stratum of the annals, which survives in the form of a handful of iterated scribal copies of the foundational text of the tradition. This was compiled by the contemporary observation of noteworthy events for the years c. 563-740 at the monastery of Iona in the Scottish Hebrides. The 720 event is close outside the 2$σ$ radiocarbon terminus ante quem date ranges for tsunami deposits identified at Dury Voe (530-660) and Basta Voe (430-650) in the Shetland Isles, and is identified as a candidate progenitor. The possibility of the existence of associated tsunami deposits in Scotland or on the north coast of Ireland is highlighted.
Phosphorus (P) is considered to be one of the key elements for life, making it an important element to look for in the abundance analysis of spectra of stellar systems. Yet, there exists only a handful of spectroscopic studies to estimate the P abundances and investigate its trend across a range of metallicities. We have observed full HK band spectra at a spectral resolving power of R=45,000 with IGRINS instrument. Abundances are determined using SME in combination with 1D MARCS stellar atmosphere models. The investigated sample of stars have reliable stellar parameters estimated using optical FIES spectra (GILD; Jönsson et al. in prep.). In order to determine the P abundances from the 16482.92 Angstrom P line, we take special care of the CO($ν=7-4$) blend. We determine the C, N, O abundances from atomic carbon and a range of non-blended molecular lines (CO, CN, OH) which are aplenty in the H band region of K giant stars, assuring an appropriate modelling of the blending CO($ν=7-4$) line. We present [P/Fe] vs [Fe/H] trend for 38 K giant stars in the metallicity range of -1.2 dex $<$ [Fe/H] $<$ 0.4 dex. We find that our trend matches well with the compiled literature sample of
Short text classi cation is a method for classifying short sentence with prede ned labels. However, short text is limited in shortness in text length that leads to a challenging problem of sparse features. Most of existing methods treat each short sentences as independently and identically distributed (IID), local context only in the sentence itself is focused and the relational information between sentences are lost. To overcome these limitations, we propose a PathWalk model that combine the strength of graph networks and short sentences to solve the sparseness of short text. Experimental results on four different available datasets show that our PathWalk method achieves the state-of-the-art results, demonstrating the efficiency and robustness of graph networks for short text classification.
Foundation models are a promising path toward general-purpose and user-friendly robots. The prevalent approach involves training a generalist policy that, like a reinforcement learning policy, uses observations to output actions. Although this approach has seen much success, several concerns arise when considering deployment and end-user interaction with these systems. In particular, the lack of modularity between tasks means that when model weights are updated (e.g., when a user provides feedback), the behavior in other, unrelated tasks may be affected. This can negatively impact the system's interpretability and usability. We present an alternative approach to the design of robot foundation models, Diffusion for Policy Parameters (DPP), which generates stand-alone, task-specific policies. Since these policies are detached from the foundation model, they are updated only when a user wants, either through feedback or personalization, allowing them to gain a high degree of familiarity with that policy. We demonstrate a proof-of-concept of DPP in simulation then discuss its limitations and the future of interpretable foundation models.
The short text matching task employs a model to determine whether two short texts have the same semantic meaning or intent. Existing short text matching models usually rely on the content of short texts which are lack information or missing some key clues. Therefore, the short texts need external knowledge to complete their semantic meaning. To address this issue, we propose a new short text matching framework for introducing external knowledge to enhance the short text contextual representation. In detail, we apply a self-attention mechanism to enrich short text representation with external contexts. Experiments on two Chinese datasets and one English dataset demonstrate that our framework outperforms the state-of-the-art short text matching models.
Eccentric planets may spend a significant portion of their orbits at large distances from their host stars, where low temperatures can cause atmospheric CO2 to condense out onto the surface, similar to the polar ice caps on Mars. The radiative effects on the climates of these planets throughout their orbits would depend on the wavelength-dependent albedo of surface CO2 ice that may accumulate at or near apoastron and vary according to the spectral energy distribution of the host star. To explore these possible effects, we incorporated a CO2 ice-albedo parameterization into a one-dimensional energy balance climate model. With the inclusion of this parameterization, our simulations demonstrated that F-dwarf planets require 29% more orbit-averaged flux to thaw out of global water ice cover compared with simulations that solely use a traditional pure water ice-albedo parameterization. When no eccentricity is assumed, and host stars are varied, F-dwarf planets with higher bond albedos relative to their M-dwarf planet counterparts require 30% more orbit-averaged flux to exit a water snowball state. Additionally, the intense heat experienced at periastron aids eccentric planets in exiting
Background and aim: Image registration and alignment are the main limitations of augmented reality-based knee replacement surgery. This research aims to decrease the registration error, eliminate outcomes that are trapped in local minima to improve the alignment problems, handle the occlusion, and maximize the overlapping parts. Methodology: markerless image registration method was used for Augmented reality-based knee replacement surgery to guide and visualize the surgical operation. While weight least square algorithm was used to enhance stereo camera-based tracking by filling border occlusion in right to left direction and non-border occlusion from left to right direction. Results: This study has improved video precision to 0.57 mm~0.61 mm alignment error. Furthermore, with the use of bidirectional points, for example, forwards and backwards directional cloud point, the iteration on image registration was decreased. This has led to improve the processing time as well. The processing time of video frames was improved to 7.4~11.74 fps. Conclusions: It seems clear that this proposed system has focused on overcoming the misalignment difficulty caused by movement of patient and enhan