The paper proposes a method to align surgical videos using kinematic data.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Automated surgical workflow analysis and understanding can assist surgeons to standardize procedures and enhance post-surgical assessment and indexing, as well as, interventional monitoring. Computer-assisted interventional (CAI) systems based on video can perform workflow estimation through surgical instruments' recog…
Paper proposes an active learning method for surgical workflow recognition using long-range temporal dependency.
Vision algorithms capable of interpreting scenes from a real-time video stream are necessary for computer-assisted surgery systems to achieve context-aware behavior. In laparoscopic procedures one particular algorithm needed for such systems is the identification of surgical phases, for which the current state of the a…
The paper uses facial keypoints to estimate post-surgical pain intensity.
Method detects surgical deviations during laparoscopic rectopexy.
Surgeons normally need surgical scissors and tissue grippers to cut through a deformable surgical tissue. The cutting accuracy depends on the skills to manipulate these two tools. Such skills are part of basic surgical skills training as in the Fundamentals of Laparoscopic Surgery. The gripper is used to pinch a point …
Convolutional neural networks improve surgical skill evaluation.
PHASE predicts surgical complications from physiological signals.
Algorithm detects surgical site infections from hospital data.
We aim to create a framework for transfer learning using latent factor models to learn the dependence structure between a larger source dataset and a target dataset. The methodology is motivated by our goal of building a risk-assessment model for surgery patients, using both institutional and national surgical outcomes…
Evaluating surgeon skill has predominantly been a subjective task. Development of objective methods for surgical skill assessment are of increased interest. Recently, with technological advances such as robotic-assisted minimally invasive surgery (RMIS), new opportunities for objective and automated assessment framewor…
Paper introduces OTR for efficient offline RL in surgical robotics.
We give a criterion under which a solution g(t) of the Kahler-Ricci flow contracts exceptional divisors on a compact manifold and can be uniquely continued on a new manifold. As t tends to the singular time T from each direction, we prove convergence of g(t) in the sense of Gromov-Hausdorff and smooth convergence away …
Optimization framework for reconstructing missing mandible segments.
Despite spectacular advances in defining invariants for simply connected smooth and symplectic 4-dimensional manifolds and the discovery of effective surgical techniques, we still have been unable to classify simply connected smooth manifolds up to diffeomorphism. In these notes, adapted from six lectures given at the …
A large fraction of the electronic health records consists of clinical measurements collected over time, such as blood tests, which provide important information about the health status of a patient. These sequences of clinical measurements are naturally represented as time series, characterized by multiple variables a…
The study uses transfer learning to compare surgical outcomes across racial/ethnic subgroups.
Ground-A-Video edits videos without training, preserving intended changes.
We augment linear Support Vector Machine (SVM) classifiers by adding three important features: (i) we introduce a regularization constraint to induce a sparse classifier; (ii) we devise a method that partitions the positive class into clusters and selects a sparse SVM classifier for each cluster; and (iii) we develop a…
It is well-known that any pair of closed orientable 3-manifolds are related by a finite sequence of Dehn surgeries on knots. Furthermore Kawauchi showed that such knots can be taken to be hyperbolic. In this article, we consider the minimal length of such sequences connecting a pair of 3-manifolds, in particular, a pai…
CB-GLNs learn video data's complex dependencies via graph representation.
RaMViD uses diffusion models for video prediction and infilling.
Current deep learning results on video generation are limited while there are only a few first results on video prediction and no relevant significant results on video completion. This is due to the severe ill-posedness inherent in these three problems. In this paper, we focus on human action videos, and propose a gene…
LumièreNet creates lecture videos from audio narration.
Bayesian models can be tricked into believing false data.
Paper defends against adversarial videos by detecting and reducing imperceptible perturbations.
Paper improves video categorization using temporal coherence.
Improves video search by balancing text and visual modalities.
A new video prediction model treats videos as continuous processes, reducing sampling steps and improving efficiency.
UDVD uses deep learning to denoise videos without supervision.
Proposes TDNs for learning complex video structures.
SummaryNet automates video summarisation using deep learning.
Paper proposes SMFN for high-res spherical video super-resolution.
Generative model produces high-fidelity video samples.
Paper proposes new principles and framework for AVC learning from user-generated videos.
The extension of image generation to video generation turns out to be a very difficult task, since the temporal dimension of videos introduces an extra challenge during the generation process. Besides, due to the limitation of memory and training stability, the generation becomes increasingly challenging with the incre…
Recent advances in deep generative models have lead to remarkable progress in synthesizing high quality images. Following their successful application in image processing and representation learning, an important next step is to consider videos. Learning generative models of video is a much harder task, requiring a mod…
Lower-dimensional video discriminators improve GAN performance.
System automates discovery and classification of training videos for career progression.
With the growth of user-generated content, we observe the constant rise of the number of companies, such as search engines, content aggregators, etc., that operate with tremendous amounts of web content not being the services hosting it. Thus, aiming to locate the most important content and promote it to the users, the…
Paper proposes graph-based separable transforms for video coding.
We prove that a general complex Monge-Ampère flow on a Hermitian manifold can be run from an arbitrary initial condition with zero Lelong number at all points. Using this property, we confirm a conjecture of Tosatti-Weinkove: the Chern-Ricci flow performs a canonical surgical contraction. Finally, we study a generaliza…
Jointly trains images and videos using residual vectors.
Paper proposes detecting video manipulation using stream descriptors.
Finding compact representation of videos is an essential component in almost every problem related to video processing or understanding. In this paper, we propose a generative model to learn compact latent codes that can efficiently represent and reconstruct a video sequence from its missing or under-sampled measuremen…
Paper introduces adversarial lossy compression for video artifacts reduction.
Advanced video classification systems decode video frames to derive the necessary texture and motion representations for ingestion and analysis by spatio-temporal deep convolutional neural networks (CNNs). However, when considering visual Internet-of-Things applications, surveillance systems and semantic crawlers of la…