This review discusses challenges and solutions for AI in chemical engineering.
problem Challenges in applying classical machine learning to chemical engineering data.
method Identifying four data characteristics and discussing their applications and solutions.
result Current research extends data science and machine learning to handle chemical engineering data challenges.
Based on 46 in-depth interviews with scientists, engineers, and CEOs, this document presents a list of concrete machine research problems, progress on which would directly benefit tech ventures in East Africa.
This essay examines how what is considered to be artificial intelligence (AI) has changed over time and come to intersect with the expertise of the author. Initially, AI developed on a separate trajectory, both topically and institutionally, from pattern recognition, neural information processing, decision and control …
GAN-based data augmentation can perpetuate biases in synthetic data.
problem Biases in synthetic data generated by GANs.
method Used a dataset of engineering researchers' head-shots to demonstrate how GANs can reinforce and amplify biases.
result GAN-based data augmentation can amplify biases in synthetic data.
Factor Engine simplifies financial factor computation and analysis in Python.
problem Efficient computation and analysis of financial factors.
method Modular, extensible Python library with decorators, integrates with data science ecosystem.
result Mispricing factors computed by Factor Engine and Stata implementation are highly similar.
Advocates for user-friendly RL problem descriptions to improve usability and generalization.
problem Usability and generalization challenges in RL for non-engineers.
method Development of user-friendly description languages for RL problems.
result Improved ability of RL algorithms to generalize to new problems.
Deep learning models enhance engineering design automation.
problem Design optimization and customization across various industries.
method Leveraging deep neural networks, GANs, VAEs, and DRL for design synthesis.
result Recent advances in DGMs show promise in structural optimization, materials design, and shape synthesis.
Paper proposes verifier engineering for improving foundation models.
problem Challenges in providing effective supervision signals for foundation models.
method Leverages automated verifiers to perform verification tasks and deliver feedback.
result Verifier engineering can enhance foundation models' capabilities.
Survey of ART neural networks for engineering applications.
problem Understanding and utilizing ART neural networks for various machine learning tasks.
method Comprehensive review of classic and modern ART models, describing learning dynamics and engineering properties.
result Compilation of ART models and their properties for engineering applications.
These notes were originally written for the Stochastic Analysis Seminar in the Department of Operations Research and Financial Engineering at Princeton University, in February of 2011. The seminar was attended and supported by members of the Research Training Group, with the author being partially supported by NSF gran…
ABM automates feature engineering and variable selection for loss-based models.
problem Improving model performance through better feature engineering and variable selection.
method ABM uses group and fused lasso regularization to automatically select cutting points and variables.
result ABM integrates feature engineering, variable selection, and model training.
This paper summarizes AI methods for detecting credit card fraud.
problem Detecting credit card fraud from millions of transactions.
method Rule-based and AI approaches, addressing imbalanced datasets, real-time scenarios, and feature engineering.
result Summarizes state-of-the-art AI methods for fraud detection.
Paper integrates ML with physics models for engineering and environmental challenges.
problem Complex science and engineering problems require new methodologies combining physics-based models and ML.
method Structured overview of integrating physics-based models with ML techniques.
result Taxonomy of existing techniques and potential research gaps identified.
Unified strategy for efficient data compression and model estimation.
problem Limited interactive exploration and data interaction in linear model development and deployment.
method Conditionally sufficient statistics for optimal data compression and estimation of linear models.
result Linear models can be estimated from compressed data without loss of parameters or covariances.
Bayesian optimization simplifies bioprocess engineering experiments.
problem Complex biological systems and experimental uncertainty.
method Adapts classical Bayesian optimization for bioprocess engineering.
result Provides accessible introduction to Bayesian optimization for practitioners.
Distiller simplifies DNN compression research with a Python package.
problem Efficiently compressing deep neural networks.
method Open-source Python package with DNN compression algorithms.
result Facilitates new research and learning tasks in DNN compression.
Machine learning identified 13 key equations for distillation column dynamics.
problem Identify governing laws for complex engineered systems.
method Sparse Identification of Non-Linear Dynamics (SINDy) applied to distillation column data.
result Reduced 1000s of equations to 13 interpretable terms.
DeepONet accelerates nuclear DT inference with high accuracy and efficiency.
problem Real-time prediction and model evaluation in nuclear systems.
method Deep Neural Operator (DeepONet) for surrogate modeling.
result DeepONet outperforms traditional ML methods in accuracy and speed.
The rise of Big Data has led to new demands for Machine Learning (ML) systems to learn complex models with millions to billions of parameters, that promise adequate capacity to digest massive datasets and offer powerful predictive analytics thereupon. In order to run ML algorithms at such scales, on a distributed clust…
A new model improves click-through rate prediction for recommendation systems.
problem Improving accuracy of click-through rate prediction in recommendation systems.
method Combines traditional feature engineering with deep neural networks to automate feature combinations.
result The model (FNFM) outperforms current deep learning feature combination models.
Research identifies risks in selecting project managers for civil engineering projects.
problem Lack of awareness of project manager selection criteria and associated risks.
method Combined ANP-FMEA approach for risk analysis.
result ANP-FMEA model identifies more significant risks than traditional FMEA.
Paper uses neural nets for financial optimization problems.
problem Financial optimization and derivative pricing problems.
method Neural networks and deep reinforcement learning for solving PDEs and dynamic optimization.
result Efficient resolution of nonlinear PDEs and dynamic optimization in finance.
VEST automates feature engineering for time series forecasting.
problem Challenges in time series forecasting with improved performance.
method VEST combines auto-regression with statistical summarization of recent past dynamics.
result VEST significantly improves forecasting performance.
Paper proposes a method to monitor research topic evolution.
problem Difficulty in tracking research topic diffusion and evolution.
method Deep Non-negative Autoencoder with information divergence measurement.
result Identifies evolution of research topics and discovers topic diffusions.
FeatureEnVi aids in feature engineering with visual analytics.
problem Insufficient support for feature engineering in visual analytics tools.
method Stepwise selection and semi-automatic extraction approaches.
result Extracts heavily engineered features evaluated by multiple metrics.
Alpha-GPT mines new trading signals with human-AI interaction.
problem Mining new alphas for effective trading signals.
method Human-AI interaction and prompt engineering algorithmic framework.
result Demonstrates Alpha-GPT's effectiveness in generating creative, insightful, and effective alphas.
Researchers can reverse-engineer deep ReLU networks from outputs.
problem Recovering a deep ReLU network from its outputs.
method Dissecting region boundaries to identify neuron states and weights.
result Weights and neuron arrangement of a deep ReLU network can be recovered.
This research categorizes AMM designs for secure token exchanges.
problem Designing AMMs for cryptoeconomic systems can lead to financial risks and inefficiencies.
method Developed an AMM taxonomy and proposed three archetypes.
result AMM archetypes meet key requirements for token issuance and exchange.
AI agent improves performance attribution analysis with high accuracy.
problem Improving accuracy in performance attribution analysis.
method Leveraging large language models and advanced prompt engineering techniques.
result Achieves accuracy rates exceeding 93% in analyzing performance drivers.
Chess engines Stockfish and LCZero differ in their approach to solving endgame puzzles.
problem Comparing machine and human chess problem-solving abilities.
method Used Plaskett's Puzzle to compare Stockfish and LCZero's performance.
result Stockfish outperforms LCZero on the puzzle.
Recent advances in artificial intelligence have been driven by the presence of increasingly realistic and complex simulated environments. However, many of the existing environments provide either unrealistic visuals, inaccurate physics, low task complexity, restricted agent perspective, or a limited capacity for intera…
Recently, link prediction has attracted more attentions from various disciplines such as computer science, bioinformatics and economics. In this problem, unknown links between nodes are discovered based on numerous information such as network topology, profile information and user generated contents. Most of the previo…
Quant 4.0 uses AI to automate, explain, and incorporate knowledge in investment.
problem Limitations of deep learning in quant investment.
method Automated AI, Explainable AI, Knowledge-driven AI.
result Improves investment decision-making through automation, interpretability, and prior knowledge integration.
Researchers simulate and estimate a market model with a matching engine to understand its impact on order submission and management.
problem The impact of a matching engine on the modeling of order submission and management in financial markets.
method Simulation of a 10-variate Hawkes process with rules for different order types, including limit orders, to compare model parameters with the original order generating process.
result Practical considerations, not directly related to model specification, can significantly distort the true model specification in an asynchronous trading environment.
While deep learning models have achieved state-of-the-art accuracies for many prediction tasks, understanding these models remains a challenge. Despite the recent interest in developing visual tools to help users interpret deep learning models, the complexity and wide variety of models deployed in industry, and the lar…
Study improves mortality prediction in ICU patients using feature engineering and 1D CNN.
problem Improving mortality prediction in ICU patients with high-dimensional, imbalanced, and missing data.
method Feature engineering, 1D Convolutional Neural Network (1D CNN), traditional machine learning algorithms.
result Best AUC of 0.848 achieved with 1D CNN model.
Model predicts stock market trends for better investment decisions.
problem Identifying optimal times to buy and sell stocks.
method XGBoost machine learning model using time series data and feature engineering.
result Model accurately predicts stock market trends and their endpoints.
Due to recent explosion of text data, researchers have been overwhelmed by ever-increasing volume of articles produced by different research communities. Various scholarly search websites, citation recommendation engines, and research databases have been created to simplify the text search tasks. However, it is still d…
This research optimizes fluid-dynamic designs using deep learning and active learning.
problem Expensive and limited empirical design verification for fluid dynamics.
method Applied a deep learning architecture to predict and optimize fluid dynamics performance.
result Reduced the number of required simulation data points from ~8000 to 625.
Study proposes using auxiliary classification to improve unsupervised anomaly detection.
problem Challenging anomaly detection in high-dimensional data.
method Use of an auxiliary classification task to extract features from unlabelled data by supervised learning.
result Our feature learning approach yields best anomaly detection performance.
Ancestry improves genealogy search results by ranking diverse record types.
problem Ranking diverse genealogy records equitably from various sources.
method Customized Coordinate Ascent, Stochastic Search, Normalized Cumulative Entropy.
result Demonstrated effectiveness of algorithms in improving relevance and diversity.
TgNN improves neural network accuracy for subsurface flow modeling.
problem Improving accuracy of neural network predictions for subsurface flow.
method Theory-guided Neural Network (TgNN) trained with data and physical constraints.
result TgNN achieves higher accuracy and better generalizability than ANN models.
Surveying machine learning for solving graph optimization problems.
problem Solving combinatorial optimization problems on graphs requires algorithmic engineering.
method Surveying machine learning approaches for graph optimization.
result Machine learning offers new ways to solve graph optimization problems.
Survey of AutoML methods and their performance.
problem Building high-quality DL systems requires human expertise.
method Reviews AutoML methods including data preparation, feature engineering, hyperparameter optimization, and neural architecture search.
result Summarizes the performance of NAS algorithms on CIFAR-10 and ImageNet datasets.
Machine Learning algorithms are increasingly being used in recent years due to their flexibility in model fitting and increased predictive performance. However, the complexity of the models makes them hard for the data analyst to interpret the results and explain them without additional tools. This has led to much rese…
Meta-learning helps models learn quickly from few samples.
problem Deep learning requires many samples, which are hard to get.
method Meta-learning optimizes models to adapt quickly to new tasks.
result Meta-learning can improve model efficiency and adaptability.
Study examines scientific research on Bitcoin across various disciplines.
problem Limited time but large number of papers on Bitcoin.
method Bibliometric analysis of papers indexed in Web of Science Core Collection.
result Observes research clusters, emerging topics, and leading scholars.
Proposes Genetic Thompson Sampling for multi-armed bandits, improving performance in nonstationary settings.
problem Improving sequential decision making tasks of online learning agents using multi-armed bandits.
method Integrates genetic principles into Thompson Sampling for multi-armed bandits.
result Significantly outperforms baselines in nonstationary settings.