AI systems that explain their decisions can be monitored for harmful intentions.
arXiv research
A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.
Trend · papers per month
Simple online monitor detects unsafe LLM outputs.
SafeML monitors ML systems for safety and security risks.
Real time large scale streaming data pose major challenges to forecasting, in particular defying the presence of human experts to perform the corresponding analysis. We present here a class of models and methods used to develop an automated, scalable and versatile system for large scale forecasting oriented towards saf…
System predicts ice formation to improve road safety.
Paper monitors DNN accuracy to enhance trustworthiness.
Online monitoring system for safety classifiers with shift detection and conformal adaptation
Online monitor detects classifier drift and adapts predictions.
For using neural networks in safety critical domains, it is important to know if a decision made by a neural network is supported by prior similarities in training. We propose runtime neuron activation pattern monitoring - after the standard training process, one creates a monitor by feeding the training data to the ne…
Proposes real-time risk monitoring for machine learning systems under unknown shifts.
AI monitors social distancing and masks at manufacturing plants.
The paper proposes an on-line monitoring framework for continuous real-time safety/security in learning-based control systems (specifically application to a unmanned ground vehicle). We monitor validity of mappings from sensor inputs to actuator commands, controller-focused anomaly detection (CFAM), and from actuator c…
While current machine learning models have impressive performance over a wide range of applications, their large size and complexity render them unsuitable for tasks such as remote monitoring on edge devices with limited storage and computational power. A naive approach to resolve this on the model level is to use simp…
This research tackles monitoring machine learning algorithms post-deployment, addressing performativity issues.
The failure of a complex and safety critical industrial asset can have extremely high consequences. Close monitoring for early detection of abnormal system conditions is therefore required. Data-driven solutions to this problem have been limited for two reasons: First, safety critical assets are designed and maintained…
Although aviation accidents are rare, safety incidents occur more frequently and require a careful analysis to detect and mitigate risks in a timely manner. Analyzing safety incidents using operational data and producing event-based explanations is invaluable to airline companies as well as to governing organizations s…
New approach extracts AI model representations for steering and monitoring.
Novel method detects faults in helicopter transmissions using healthy data only.
In multimodal traffic monitoring, we gather traffic statistics for distinct transportation modes, such as pedestrians, cars and bicycles, in order to analyze and improve people's daily mobility in terms of safety and convenience. On account of its robustness to bad light and adverse weather conditions, and inherent spe…
Can engineering neural networks be approached in a disciplined way similar to how engineers build software for civil aircraft? We present nn-dependability-kit, an open-source toolbox to support safety engineering of neural networks for autonomous driving systems. The rationale behind nn-dependability-kit is to consider…
This paper reviews traditional and modern methods for detecting structural damage using vibrations.
Efficient classifier with uncertainty bounds for safety-critical applications.
Fine particulate matter (PM) is one of the criteria air pollutants regulated by the Environmental Protection Agency in the United States. There is strong evidence that ambient exposure to (PM) increases risk of mortality and hospitalization. Large scale epidemiological studies on the health effects of P…
In a world of global trading, maritime safety, security and efficiency are crucial issues. We propose a multi-task deep learning framework for vessel monitoring using Automatic Identification System (AIS) data streams. We combine recurrent neural networks with latent variable modeling and an embedding of AIS messages t…
New framework uses OR to ensure AI systems make safe decisions.
Deep learning identifies precipitation clouds from all-sky camera data.
Despite the tremendous advances that have been made in the last decade on developing useful machine-learning applications, their wider adoption has been hindered by the lack of strong assurance guarantees that can be made about their behavior. In this paper, we consider how formal verification techniques developed for …
Transformer model predicts train axle vibrations for safer maintenance.
Using raw sensor data to model and train networks for Human Activity Recognition can be used in many different applications, from fitness tracking to safety monitoring applications. These models can be easily extended to be trained with different data sources for increased accuracies or an extension of classifications …
Visual object detection is a computer vision-based artificial intelligence (AI) technique which has many practical applications (e.g., fire hazard monitoring). However, due to privacy concerns and the high cost of transmitting video data, it is highly challenging to build object detection models on centrally stored lar…
Measures policy-violating content prevalence with ML-assisted sampling and LLM labeling.
Gaussian Process Regression improves damage assessment in structural health monitoring.
Air traffic control is a real-time safety-critical decision making process in highly dynamic and stochastic environments. In today's aviation practice, a human air traffic controller monitors and directs many aircraft flying through its designated airspace sector. With the fast growing air traffic complexity in traditi…
RLVR maintains safety while improving reasoning capabilities in LLMs.
Cyber-physical systems (CPS) greatly benefit by using machine learning components that can handle the uncertainty and variability of the real-world. Typical components such as deep neural networks, however, introduce new types of hazards that may impact system safety. The system behavior depends on data that are availa…
Conformal Test Martingales can be 'blind' to significant changes in data distribution.
In recent years, deep learning methods have outperformed other methods in image recognition. This has fostered imagination of potential application of deep learning technology including safety relevant applications like the interpretation of medical images or autonomous driving. The passage from assistance of a human d…
Framework predicts remaining useful life of DSH subsystems under unknown failure modes.
How can we find patterns and anomalies in a tensor, or multi-dimensional array, in an efficient and directly interpretable way? How can we do this in an online environment, where a new tensor arrives each time step? Finding patterns and anomalies in a tensor is a crucial problem with many applications, including buildi…
Approves updates to machine learning models in healthcare based on accumulating data.
Autonomous vehicles rely on machine learning to solve challenging tasks in perception and motion planning. However, automotive software safety standards have not fully evolved to address the challenges of machine learning safety such as interpretability, verification, and performance limitations. In this paper, we revi…
Reward hacking exploits misspecified rewards, affecting agent capabilities and true performance.
Fine-tuning LLMs improves capability but harms safety, study finds.
RAGuard improves safety in LLMs for offshore wind maintenance.
Survey of algorithms for testing AI-driven CPS safety.
Safe imitation learning with a safety layer for flexible training.
Deep learning methods are widely regarded as indispensable when it comes to designing perception pipelines for autonomous agents such as robots, drones or automated vehicles. The main reasons, however, for deep learning not being used for autonomous agents at large scale already are safety concerns. Deep learning appro…
This paper formalizes AI safety using hypothesis testing in GenAI.