New adaptive questionnaire identifies optimal product design without precise preference estimates.
problem Identifying the most profitable product design from unknown consumer preferences.
method Integrates engineering feasibility and cost models with adaptive discrete-choice questionnaire to directly determine the optimal design.
result The optimal design can be determined without accurate preference estimation, leveraging engineering knowledge.
Robo-advisors estimate clients' risk aversion using interactive questionnaires.
problem Estimating risk aversion of non-expert clients using adaptive questionnaires.
method Model risk aversion with cost functions and spectral risk measures. Use inverse reinforcement learning to design questions maximizing distinguishing power.
result Designing questions by maximizing distinguishing power achieves satisfactory accuracy in learning risk aversion with fewer than 50 questions.
A framework uses preprocessing to improve psychiatric questionnaire predictions while maintaining interpretability.
problem Weak predictive accuracy and limited interpretability of psychiatric questionnaires.
method Two-stage method: stable preprocessing followed by a linear mapping.
result REFINE outperforms other interpretable approaches in psychiatric and non-psychiatric prediction tasks.
CrowdMI uses crowdsourcing to impute missing data, achieving results similar to complex models.
problem Missing data imputation using human crowdsourcing.
method Replicating a multiple imputation framework with multiple crowdworkers completing a survey.
result Valid imputations for both qualitative and quantitative missing data comparable to complex models.
Study explores factors influencing saving behavior among Dhaka employees.
problem Factors influencing saving behavior among Dhaka employees.
method Quantitative approach with cross-sectional survey design, structured questionnaire, descriptive statistics, reliability analysis, regression analysis.
result Only financial management practices had a significant positive relationship with saving behavior.
The study analyzes how large language models form and express investor risk profiles.
problem Understanding how large language models (LLMs) form and express investor risk profiles.
method Examined three LLMs (GPT, Gemini, and Llama) and assessed their responses to a standardized risk questionnaire under varying prompts.
result LLMs generally form long-term investment profiles, but they exhibit different risk tolerance levels.
New method detects careless responding in long surveys.
problem Careless responding in long surveys threatens internal validity.
method Detects a changepoint in combined measurements of carelessness.
result Highly accurate in identifying carelessness onset.
An efficient algorithm selects the correct number of latent dimensions in multidimensional probit models.
problem Determining the correct number of latent dimensions in multidimensional probit graded response models.
method Adaptive Bayesian dimension selection framework using cumulative ordered spike-and-slab (COSS) prior and Albert--Chib latent response augmentation.
result The proposed method accurately recovers latent structures and avoids repeated model fitting.
MOAI evaluates indoor airflow's impact on COVID-19 transmission.
problem Understanding indoor airflow's role in COVID-19 transmission.
method Developed a privacy-preserving app and model to evaluate risk exposure.
result Quantified factors contributing to higher or lower contamination in settings.
GPT-4 assesses its confidence in answering USMLE questions with and without feedback.
problem Understanding AI's performance in healthcare applications, especially in sensitive areas like medical education.
method Used a prompting technique to evaluate GPT-4's confidence scores before and after answering USMLE questions, categorized into with and without feedback.
result Feedback influences relative confidence but doesn't consistently increase or decrease it.
The main aim of this paper is to inspect the properties of survey based on households inflation expectations, conducted by Reserve Bank of India. It is theorized that the respondents answers are exaggerated by extreme response bias. Latent class analysis has been hailed as a promising technique for studying measurement…
TabSODA improves imputation of surveys with skips and ordinal data.
problem Handling structural skips and ordinal responses in survey data.
method TabSODA uses an Elucidated Diffusion Model with skip pattern detection and ordinal awareness.
result TabSODA reduces ordinal missing-at-random (MACE) by up to 23.7% and improves categorical accuracy by up to 9%.
In many countries information on expectations collected through consumer confidence surveys are used in macroeconomic policy formulation. Unfortunately, before doing so, the consistency of responses is often not taken into account, leading to biases creeping in and affecting the reliability of the indices hence created…
Ordinal data is omnipresent in almost all multiuser-generated feedback - questionnaires, preferences etc. This paper investigates modelling of ordinal data with Gaussian restricted Boltzmann machines (RBMs). In particular, we present the model architecture, learning and inference procedures for both vector-variate and …
Study predicts onset of type II diabetes using survey data and machine learning.
problem Early diagnosis of type II diabetes from patient data.
method Developed an ensemble classifier using five classification algorithms.
result Ensemble model had an AUC of 0.834, indicating high performance.
Archetypal analysis helps understand binary data sets.
problem Explaining binary questionnaire data.
method Using archetypal analysis for binary observations.
result The approach contributes to understanding binary data sets.
New algorithm estimates intrinsic dimension of discrete datasets.
problem Inaccuracies in using continuous methods for discrete datasets.
method Introduced an algorithm to infer intrinsic dimension of discrete spaces.
result Demonstrated accuracy on benchmark datasets and found a small intrinsic dimension in a metagenomic dataset.
Machine learning aids in shark detection at Muizenberg Beach.
problem Improving shark spotting efficiency using automated methods.
method Defined desirable properties, selected mathematical techniques, and partially implemented model.
result Extract useful information from shark images despite geometric transformations.
Models predict player motivation from game data.
problem Predict player motivation from gameplay data.
method Collected gameplay data and survey responses, used SVM for inference.
result Models accurately predict player motivation (92%-94%).
Survey shows users value usability over functionality in process discovery tools.
problem Users prioritize usability over functional aspects in process discovery tools.
method A survey was conducted with 66 respondents to gather feedback on process discovery tools.
result Users prefer usability over functionality in process discovery tools.
Survey of techniques for diagnosing pediatric sleep apnea from inexpensive data.
problem Diagnosing pediatric sleep apnea from limited and variable data.
method Exploratory data analysis using correlation networks, Mapper, SVD; supervised and unsupervised learning techniques.
result Analysis of various learning techniques applied to pediatric sleep apnea data.
Method completes mixed matrix from complex surveys with heterogeneous missingness.
problem Recovering a mixed dataframe matrix from complex survey sampling with different missingness patterns.
method Two-stage procedure: logistic regression for missingness modeling, and weighted log-likelihood maximization with low-rank constraint.
result The proposed method achieves sublinear convergence and shows superior performance compared to existing methods.
Researchers study how teachers' advising relationships influence their perceptions of satisfaction and students, not policy influence.
problem Understanding the relationship between teachers' advising relationships and their perceptions of satisfaction and students.
method Proposed a novel joint model of network and item responses (JNIRM) with correlated latent variables.
result Teachers' advising relationships contribute more to satisfaction and students than to influence over educational policies.
The aim of this paper is to get an overview of the online buyer profile, and also some key aspects in the way the online shopping is conducted. In this project we conducted a quantitative research, consisting of a questionnaire based survey. For data processing and interpretation we used SPSS statistical software and E…
Model predicts increased social unrest during COVID-19 using social media data.
problem Detecting rising conflict potential in societies during pandemics.
method Neural implicit motive pattern recognition from social media texts.
result Significant increase in conflict indicators during the pandemic.
Early diagnosis is important for type 2 diabetes (T2D) to improve patient prognosis, prevent complications and reduce long-term treatment costs. We present a novel risk profiling approach based exclusively on health expenditure data that is available to Belgian mutual health insurers. We used expenditure data related t…
A new personality-based recommender system tackles data sparsity without feedback.
problem Data sparsity without common feedback among users.
method Implicitly identifying users' personality type and incorporating it with personal interests and knowledge level.
result The model's effectiveness, especially in data sparsity situations, demonstrated on a real-world dataset.
Study reveals LLM personas have two distinct components: frame-robust aggregated traits and frame-dependent geometric features.
problem Evaluation of LLM personas via psychometric questionnaires discards within-instance correlation structure.
method Constructed within-instance correlation matrices from IPIP-50 responses and analyzed geometry on SPD manifolds under manipulated question orderings.
result Persona expression comprises two dissociable components: aggregated features (Big Five scores) and geometric features (SPD manifold).
Client appraisal improves efficiency in microfinance banks in Adamawa State.
problem Increasing loan defaults and losses in microfinance institutions.
method Survey method with primary and secondary data collection, multi-stage sampling, questionnaires, descriptive and inferential statistics.
result Client appraisal positively affects efficiency and productivity.
Study shows financial literacy, social capital, and financial tech positively impact financial inclusion of Indonesian students.
problem Financial literacy, social capital, and financial technology's impact on financial inclusion of Indonesian students.
method Quantitative research using questionnaires distributed to 100 students from 7 private colleges in Tangerang, Indonesia.
result Financial literacy, social capital, and financial technology have a positive and significant influence on financial inclusion.
Tangles improve clustering in various datasets.
problem Clustering diverse datasets efficiently and accurately.
method Tangles aggregate cuts to identify dense structures, leading to soft cluster characterization.
result Tangle framework generates hierarchical soft dendrograms for cluster exploration.
New dataset from clinicians improves sepsis prediction models.
problem Circularity in previous sepsis prediction models.
method Developed an independent dataset from clinical judgments, avoiding circularity.
result Achieved state-of-the-art AUROC scores.
The paper proposes a test to determine the number of latent classes in ordinal categorical data.
problem Determining the correct number of latent classes in latent class models with ordinal categorical data.
method The test statistic centers the largest singular value of a normalized residual matrix by a simple sample-size adjustment.
result The test statistic converges to zero under the null hypothesis and exceeds a fixed positive constant under an under-fitted alternative.
New estimators improve Rasch model item parameter estimation for sparse data.
problem Estimating item parameters in sparse Rasch model data.
method Random pairing maximum likelihood estimator (RP-MLE) and its bootstrapped variant (MRP-MLE).
result RP-MLE and MRP-MLE are minimax optimal and provide precise item parameter estimates.
Study predicts online procrastination using machine learning.
problem Predicting procrastination in eLearning to prevent drop-outs.
method Comparison of multiple machine learning models with subjective and objective predictors.
result Models with objective predictors outperform those with subjective predictors.
This paper optimizes survey questions to reduce bias and select key tokens for QoE analysis.
problem Reducing bias in user surveys and selecting informative tokens for Quality of Experience analysis.
method Randomized question order and greedy submodular maximization for selecting tokens.
result Randomizing token order can significantly reduce bias, and a subset of 30% tokens captures 94% of the information.
Paper presents a new validation method for simulation workflow.
problem Validation of simulation workflow is challenging and domain-dependent.
method Empirical learning-based validation procedure using AHP and machine learning.
result Validation procedure is semi-automated and efficient.
The present research aims to highlight the main factors influencing the development of entrepreneurial innovation in a rural environment and to perform an empirical study with the purpose of assessing the main problems in rural development. The research performed is mostly of a quantitative nature, being based on the u…
New Bayesian method for sparse multidimensional item response theory.
problem Sparse interpretable explanations for questionnaire data.
method Bayesian EM algorithm for sparse factor loadings.
result Reliable recovery of factor dimensionality and latent structure.
Novel approach models life events using causal discovery and survival analysis.
problem Modeling life event choices and occurrence from a probabilistic perspective.
method Bi-level problem formulation: causal discovery for life events graph, survival analysis for time-to-event modeling.
result Identification of causal relationships and factors influencing transition rates between life events.
Sparse GFA identifies disease factors in FTD subgroups.
problem Heterogeneity in neurological disorders hinders understanding and treatment.
method Sparse Group Factor Analysis (GFA) with regularised horseshoe priors.
result Identified latent disease factors differentially expressed in FTD subgroups.
The paper presents algorithms to select important data prototypes with weights.
problem Mining meaningful data prototypes from complex distributions.
method Develops algorithms with theoretical guarantees to select and weight data prototypes.
result Our framework provides a unified approach for selecting and weighting prototypes.
Machine learning detectors can be accurate but introduce bias.
problem The accuracy of machine learning detectors affects the reliability of measured constructs.
method Examined the impact of detector accuracy on estimated correlations between constructs and phenomena.
result The expected correlation between phenomena decreases as detector accuracy decreases.
Automated process links oral health to systemic conditions using machine learning.
problem Correlating oral health with systemic health conditions.
method Intraoral fluorescent biomarker imaging, machine learning segmentation, and clinical examination.
result Machine learning classifier achieved AUC of 0.677, indicating a learned association between disease signatures in images and periodontal disease.
Developed predictive models for improving programming course performance.
problem Improving student performance in programming courses.
method Used M5P Decision Tree and Linear Regression Classifier on structured questionnaire data.
result Variable-based LRC model produced the best model with least evaluation metrics.
Bluetooth data predicts depression severity, showing 18.8% extra variance.
problem Predicting depressive symptom severity using Bluetooth data.
method Extracted 49 Bluetooth features from NBDC data, used linear mixed-effect and hierarchical Bayesian linear regression models.
result Hierarchical Bayesian model achieved best prediction metrics (R2=0.526, RMSE=3.891).
Study evaluates human capital's role in Moroccan hotels' value creation.
problem Limited research on human capital's role in value creation.
method Linear regression analysis using hotel annual reports and a human capital scale.
result Human capital positively impacts value creation in Moroccan hotels.
The abstract covers various aspects of eBusiness and eGovernment, including digital currencies, m-government services, gender inclusivity, eLearning, export performance, SME digitalization, and banking customer behavior.
problem Various challenges and opportunities in eBusiness and eGovernment.
method Critical review, UTAUT model with perceived risk theory, GAD approach, inductive research paradigm, one-on-one interviews, survey questionnaires, convenience sampling.
result Impediments to eLearning uptake, gender inclusivity in e-procurement, export performance of manufacturing firms, SME digitalization impact, measuring and modeling framework for Internet banking customers.