Python library for boosting statistical relational models.
problem Expressing learning and inference problems in statistical relational models.
method Adapting scikit-learn interface for boosted statistical relational models.
result Provides examples for using srlearn.
ParaMonte::Python streamlines Bayesian data analysis with fast Monte Carlo and MCMC routines.
problem Efficiently sampling posterior distributions in Bayesian modeling and data science.
method Serial and MPI-parallelized Markov Chain Monte Carlo (MCMC) routines.
result Automated model calibration and uncertainty quantification in Bayesian analysis.
NoMoPy models noise as HMM/FHMM in Python.
problem Modeling noise in data.
method Approximate and exact EM algorithms, cross-validation, confidence region estimation.
result Validated on example problems.
BARMPy offers a Python package for Bayesian Additive Regression Models.
problem Making complex Bayesian models accessible to machine learning practitioners.
method Object-oriented design compatible with SciKit-Learn, documentation and tutorial provided.
result Ease of use and compatibility with existing machine learning tools.
PyHHMM is a Python library for HHMMs with advanced features.
problem Handling heterogeneous observation models and missing data in HMMs.
method Object-oriented Python implementation with advanced features.
result PyHHMM supports a heterogeneous observation model and missing data inference.
OpenML-Python API simplifies access to OpenML for Python users.
problem Limited access to OpenML for Python users.
method Developed a Python API (OpenML-Python) to integrate OpenML with Python-based tools.
result Facilitates easy access to OpenML's datasets, tasks, and experiments.
AutoGMM automates Gaussian mixture modeling in Python.
problem Automatic clustering of complex data with uncertainty-aware grouping.
method Strategic initialization using an agglomerative Mahalanobis heuristic, parallelized model selection by information criteria.
result Strong out-of-the-box performance on classic benchmarks and real datasets.
Cyanure offers efficient solvers for linear model learning in Python, C++, and more.
problem Efficiently solving empirical risk minimization problems for linear models.
method Stochastic variance-reduced optimization with acceleration mechanisms.
result Handles a wide range of loss and regularization functions.
Python package ajdmom simplifies moment formula derivation for jump diffusions.
problem Deriving moment formulae for complex jump diffusion processes.
method Automatically generates closed-form expressions and derivatives for any order of moments.
result Enhances usability and usability of affine jump diffusion models.
metric-learn simplifies metric learning in Python.
problem Performing distance metric learning efficiently.
method Unified scikit-learn compatible interface for supervised and weakly-supervised metric learning.
result Unified interface for cross-validation and model selection.
DPPy offers Python tools for sampling DPPs.
problem Sampling from DPPs is challenging.
method Gathers exact and approximate sampling algorithms for finite and continuous DPPs.
result DPPy provides Python tools for DPP sampling.
Python package for functional data analysis.
problem Handling and analysis of functional data.
method Comprehensive tools for representation, preprocessing, and exploratory analysis of functional data.
result Scikit-fda package provides a comprehensive set of tools for functional data analysis.
Python toolbox uncovers causal relationships from data.
problem Discovering causal relationships from observational data.
method End-to-end approach using algorithms from 'Bnlearn' and 'Pcalg', including pairwise causal discovery.
result Recovery of direct dependencies and causal relationships.
DoubleML is a Python library for causal inference using machine learning.
problem Estimating causal parameters in complex models with machine learning.
method Double machine learning framework for valid statistical inference.
result High flexibility and easy extension for various model specifications.
dalex simplifies model exploration and fairness for Python developers.
problem Model black-box nature and risks of discrimination, lack of reproducibility, and data drift.
method Model-agnostic interface for interactive model exploration.
result Enhances model transparency and accountability through interactive explainability and fairness.
PyVBMC speeds up Bayesian inference for expensive models in Python.
problem Efficient Bayesian inference for computationally expensive models.
method Variational Bayesian Monte Carlo (VBMC) algorithm.
result PyVBMC provides a flexible and efficient method for parameter estimation and model assessment.
Autoconj automates conjugacy exploitation in Python without a DSL.
problem Time-consuming and error-prone conjugacy derivations.
method Operates on Python functions, supports any PPL.
result Accelerates development of inference algorithms.
Python models predict stock sentiment for market-beating returns.
problem Predicting public sentiment for stock trading.
method Crowd-sourced labeled data, trained and evaluated various models.
result Best models predict market-beating returns from public sentiment.
GridPyM handles grid diagrams for knot theory.
problem Handling grid diagrams for knot theory.
method Generates and simplifies grids, models local transformations.
result Models local transformations between grid diagrams.
arfpy simplifies data generation with adversarial random forests.
problem Synthesizing new data that matches given data.
method Adversarial Random Forests (ARF) integrated into a Python package.
result Effective and user-friendly data generation for various fields.
CausalML simplifies causal inference methods in Python.
problem Combining causal inference and machine learning.
method Collection of causal inference methods in Python.
result Makes causal inference methods accessible in Python.
RobPy offers robust statistical methods in Python.
problem Lack of robust statistical methods in Python.
method Built on NumPy, SciPy, and scikit-learn, RobPy includes robust tools for various statistical tasks.
result RobPy enables more users to perform robust data analysis in Python.
GraSPy simplifies graph analysis in Python.
problem Analyzing and understanding graphs.
method Scikit-learn compliant API for statistical inference and machine learning.
result Flexible algorithms for graph statistics.
A Python package for GLHMM, a flexible HMM framework.
problem Handling diverse HMM applications in neuroscience.
method Stochastic variational inference for large datasets.
result Enables statistical testing and out-of-sample prediction.
TrueLearn Python library for personalized educational recommendations.
problem Building educational recommendation systems with humanly-intuitive user representations.
method Online learning Bayesian models and open learner concept.
result Library includes models and representations for user control and interpretability.
Python library for conformal prediction, licensed under MIT.
problem Improving prediction accuracy with uncertainty quantification.
method Conformal prediction framework implemented in Python.
result Stable API and algorithms for conformal prediction.
Tangent automates derivatives in Python, improving expressiveness and performance.
problem Efficiently calculating derivatives for complex models in Python.
method Source-code transformation for dynamically typed array programming.
result Demonstrates improved expressiveness and performance in automatic differentiation.
Python package for fast simulation-based inference.
problem Intractable likelihood functions in Bayesian inference.
method Uses neural networks as surrogate models for Bayesian inference.
result Highly efficient and user-friendly for constructing SBI estimators.
This document serves to complement our website which was developed with the aim of exposing the students to Gaussian Processes (GPs). GPs are non-parametric Bayesian regression models that are largely used by statisticians and geospatial data scientists for modeling spatial data. Several open source libraries spanning …
Pythae is a Python library for benchmarking VAE models.
problem Improving variational autoencoders for various tasks.
method Unified implementation and framework for 19 generative autoencoder models.
result Benchmarking 19 VAE models across multiple tasks.
Comprisk simplifies competing-risks analysis in Python.
problem Analyzing medical time-to-event data with competing risks.
method A scikit-learn-compatible toolkit for competing-risks survival analysis.
result Comprisk provides a unified API for various competing-risks methods.
XDeep interprets deep neural networks for practitioners and researchers.
problem Understanding and interpreting deep neural networks.
method Post-hoc interpretation algorithms integrated into XDeep.
result XDeep provides local and global explanations for deep models.
Automatic differentiation (AD) is an essential primitive for machine learning programming systems. Tangent is a new library that performs AD using source code transformation (SCT) in Python. It takes numeric functions written in a syntactic subset of Python and NumPy as input, and generates new Python functions which c…
SurvLIMEpy is a Python package for computing feature importance in survival analysis.
problem Computing feature importance in survival analysis models.
method Parallelized implementation of the SurvLIME algorithm for various survival models.
result SurvLIMEpy accurately captures feature importance in survival analysis models.
A Python tool assesses fairness, accountability, and transparency in AI decisions.
problem Lack of regulation and certification for AI-driven decisions.
method Developed an open-source Python toolbox to analyze fairness, accountability, and transparency aspects of machine learning.
result Automatically reports fairness, accountability, and transparency aspects of AI decisions to stakeholders.
FairLangProc simplifies fairness in NLP models for Python users.
problem Addressing bias in NLP models for decision-making contexts.
method Develops a Python package for implementing fairness metrics and algorithms.
result Promotes the use of bias mitigation techniques in NLP.
MKLpy simplifies Multiple Kernel Learning in Python.
problem Learning optimal kernel functions from data.
method Python-based framework for Multiple Kernel Learning algorithms.
result Maximizes usability and simplifies development of novel solutions.
PyOD offers scalable outlier detection for multivariate data.
problem Scalable outlier detection for multivariate data.
method Wide range of outlier detection algorithms, including ensembles and neural networks.
result Robust and scalable outlier detection.
hyppo simplifies multivariate hypothesis testing in Python.
problem Inconsistent multivariate hypothesis testing interfaces in Python.
method Unified library for multivariate testing procedures.
result Easy-to-use and flexible for future extensions.
PySAD offers a unified Python framework for efficient streaming anomaly detection.
problem Efficient anomaly detection in streaming data with strict constraints.
method Unified architecture with 17+ streaming algorithms, specialized components, and support for multiple learning paradigms.
result PySAD enables real-time processing with bounded memory and is compatible with other Python frameworks.
Python library for causal discovery from observational data.
problem Revealing causal relations from observational data.
method Comprehensive collection of causal discovery methods in Python.
result Ease of use for non-specialists and modular building blocks for developers.
Factor Engine simplifies financial factor computation and analysis in Python.
problem Efficient computation and analysis of financial factors.
method Modular, extensible Python library with decorators, integrates with data science ecosystem.
result Mispricing factors computed by Factor Engine and Stata implementation are highly similar.
combo library simplifies model combination for various machine learning tasks.
problem Facilitating model combination in machine learning.
method Easy-to-use Python toolkit for aggregating models and scores.
result Unified and consistent way to combine models from multiple libraries.
Python package cegpy models processes with asymmetries.
problem Leveraging CEGs for processes with structural asymmetries.
method Developed cegpy, a Python package for CEGs with Bayesian model selection and probability propagation.
result First CEG package in any language that can model symmetric and asymmetric structures.
Distiller simplifies DNN compression research with a Python package.
problem Efficiently compressing deep neural networks.
method Open-source Python package with DNN compression algorithms.
result Facilitates new research and learning tasks in DNN compression.
Pykg2vec simplifies knowledge graph embedding research.
problem Learning representations of entities and relations in knowledge graphs.
method Flexible and modular software architecture implementing 16 state-of-the-art algorithms.
result Accelerates research in knowledge graph representation learning.
UncertaintyPlayground simplifies uncertainty estimation in Python.
problem Uncertainty estimation in supervised learning tasks.
method Sparse and Variational Gaussian Process Regressions for normally distributed outcomes, Mixed Density Networks for mixed distributions.
result Fast and simplified uncertainty estimation through Python library.
PyDTS analyzes survival data with discrete intervals and competing risks.
problem Discrete-time survival analysis with competing risks and optional penalization.
method Regularized estimation methods, model evaluation metrics, variable screening tools, and simulation module.
result Supports research and development in discrete-time survival analysis.