Research
On-device research index

arXiv research

A locally-built, LLM-digested index of recent arXiv papers in quant finance, geometry/topology, and statistical ML — keyword search served straight from SQLite on this machine.

168,695 papers · 148 categories

Trend · papers per month

265177102 · Oct 201919922001200920172026
48 results for California Community Colleges

Study shows GDP and CPI predict CCC funding, highlighting need for economic forecasting.

problem Challenges in aligning CCC funding with DEI initiatives.
method Quantitative correlational design, analyzing 30 years of economic data.
result Strong positive correlation between GDP growth and CCC funding levels, and between CPI and funding levels.

This paper explores integration and contagion among US metropolitan housing markets. The analysis applies Federal Housing Finance Agency (FHFA) house price repeat sales indexes from 384 metropolitan areas to estimate a multi-factor model of U.S. housing market integration. It then identifies statistical jumps in metrop…

2011-10-18abs ↗pdf ↗

Paper proposes efficient multivariate spatial Fay-Herriot models using variational autoencoders.

problem Estimating population characteristics in small areas with limited data.
method Integrates multivariate spatial Fay-Herriot model with variational autoencoders to leverage spatial structure efficiently.
result Significant computational efficiency improvements for high-dimensional datasets.

Geospatial framework assesses climate risks for California's banking and exposed sectors.

problem Evaluating climate risks on banking and exposed sectors in California.
method Integrates hazard mapping, exposure analysis, and scenario-based financial risk assessment.
result Framework supports portfolio monitoring and institutional readiness under new standards.

Differential Calculus is a staple of the college mathematics major's diet. Eventually one becomes tired of the same routine, and wishes for a more diverse meal. The college math major may seek to generalize applications of the derivative that involve functions of more than one variable, and thus enjoy a course on Multi…

2009-09-30abs ↗pdf ↗

The paper critiques ε-fairness, showing it can lead to unfair outcomes and proposes a utility-based approach.

problem The limitations of probabilistic fairness metrics in real-world contexts.
method Utility-based approach to measure fairness, addressing the issue of unavailable data on false negatives.
result A utility-based approach uncovers necessary actions to achieve true fairness, contrasting with traditional probability-based evaluations.

The paper sorts big data by revealed preferences, improving consumer and policy decisions.

problem Sorting diverse consumer preferences for big data objects like colleges.
method Endogenous weighting of revealed preferences, considering spillover effects.
result Consistent steady-state solution to counterbalance equilibrium.

Predicts academic risk in college students using interpretable machine learning.

problem Predicting academic risk from high-dimensional, unbalanced student data.
method Binary classification task using LightGBM model and Shapley value.
result 8 predictors for academic risk identified, including quality of academic partners and dormitory study atmosphere.

New benchmark for earthquake forecasting models shows current neural point processes are not yet suitable.

problem Lack of a modern benchmark for evaluating neural point process models in earthquake forecasting.
method Curated and standardized earthquake catalog, evaluation protocols, and datasets.
result None of the tested NPPs outperformed the classical ETAS model.

Modeling student behaviors and multiple predictions for early intervention.

problem Predicting student outcomes and interactions among multiple tasks.
method Proposes a variant of LSTM and soft-attention mechanism for heterogeneous behaviors, and co-attention mechanism for task interactions.
result Demonstrated effectiveness in predicting student outcomes and interactions.

Paper introduces RPWithPrior for efficient label differential privacy in regression.

problem Protecting user privacy in regression tasks with minimal accuracy loss.
method Modeling responses as continuous random variables, avoiding discretization; estimating optimal intervals for randomized responses.
result RPWithPrior algorithm guarantees ε-label differential privacy and outperforms existing methods.

The study uses Hidden Markov Models to analyze student enrollment patterns and academic performance.

problem Limited understanding of how enrollment patterns affect academic performance.
method Applied Hidden Markov Models to categorize enrollment strategies and compare academic outcomes.
result Mixed enrollment strategies lead to better academic performance, especially during part-time semesters.

1. Translated by Thomas E. Cecil, Department of Mathematics and Computer Science, College of the Holy Cross, Worcester, MA 01610, USA; E-mail address: cecil@mathcs.holycross.edu 2. Typed by Wenjiao Yan, School of Mathematical Sciences, Laboratory of Mathematics and Complex Systems, Beijing Normal University, Beijing 10…

2011-12-13abs ↗pdf ↗

Study shows human advisors use context to improve student outcomes in algorithm-assisted advising.

problem How human advisors use context to guide interventions in algorithm-assisted advising.
method Mixed-methods approach combining quantitative and qualitative data from a randomized controlled trial.
result 2 out of 3 interventions by advisors were plausibly 'expertly targeted' to students using non-algorithmic context.

Study shows financial literacy, social capital, and financial tech positively impact financial inclusion of Indonesian students.

problem Financial literacy, social capital, and financial technology's impact on financial inclusion of Indonesian students.
method Quantitative research using questionnaires distributed to 100 students from 7 private colleges in Tangerang, Indonesia.
result Financial literacy, social capital, and financial technology have a positive and significant influence on financial inclusion.

Time series classification models have been garnering significant importance in the research community. However, not much research has been done on generating adversarial samples for these models. These adversarial samples can become a security concern. In this paper, we propose utilizing an adversarial transformation …

2019-02-27abs ↗pdf ↗

Paper proposes a sparse synthetic control method to select important predictors.

problem Choosing and weighting predictors affects synthetic control estimator performance.
method Sparse synthetic control procedure that penalizes predictors, derived in a linear factor model.
result Sparse synthetic control achieves lower bias and better post-treatment performance.

This paper determines the minimal degree sequence for two compact rational knots, namely the trefoil and figure-eight knots. We find explicit projections with the minimal degree sequence of each knot. This is done by modifying a non-compact rational minimal-degree parameterization of the trefoil and figure-eight knots …

2011-11-14abs ↗pdf ↗

We present a new approach for mitigating unfairness in learned classifiers. In particular, we focus on binary classification tasks over individuals from two populations, where, as our criterion for fairness, we wish to achieve similar false positive rates in both populations, and similar false negative rates in both po…

2017-06-30abs ↗pdf ↗

Study predicts academic achievement using students' support networks.

problem Predicting academic achievement in college students.
method Decision tree and random forest algorithms applied to Ties data.
result Different types of support are important for different demographics and genders.

Study compares geostatistical and machine learning models for PM2.5 prediction.

problem Improving accuracy of hourly PM2.5 maps across California.
method Traditional geostatistical methods (kriging, land use regression) and machine learning models (neural networks, random forests, support vector machines) were evaluated.
result Ensemble model enhanced predictive accuracy of PM2.5 concentration by correcting PurpleAir data bias.

In May 2015, a conference entitled "Groups, Geometry, and 3-manifolds" was held at the University of California, Berkeley. The organizers asked participants to suggest problems and open questions, related in some way to the subject of the conference. These have been collected here, roughly divided by topic. The name (o…

2015-12-15abs ↗pdf ↗

Transfer learning improves highway traffic forecasting using graph neural networks.

problem Lack of historical data for traffic forecasting on large highway networks.
method Developed a transfer learning approach for DCRNN, a graph neural network for highway forecasting.
result TL-DCRNN can forecast traffic on unseen regions of the highway network with high accuracy.

This paper introduces EQShapelets (EarthQuake Shapelets) a time-series shape-based approach embedded in machine learning to autonomously detect earthquakes. It promises to overcome the challenges in the field of seismology related to automated detection and cataloging of earthquakes. EQShapelets are amplitude and phase…

2019-11-20abs ↗pdf ↗

Paper proposes transforming ATN to attack multivariate time series models.

problem Generating adversarial samples for multivariate time series classification models.
method Proposes using a distilled model as a surrogate to mimic attacked models and applies 1-NN DTW and FCN attacks.
result Both models were susceptible to attacks on all 18 datasets.

This manuscript served as lecture notes for a mini-course in the 2016 Southern California Geometric Analysis Seminar Winter School. The goal is to give a quick introduction to Kahler geometry by describing the recent resolution of Tian's three influential properness conjectures in joint work with T. Darvas. These resul…

2018-07-02abs ↗pdf ↗

Hybrid approach protects privacy while analyzing smart meter data.

problem Privacy concerns in AMI data analysis under CPUC regulations.
method Anonymization, differential privacy, federated learning, synthetic data, cryptography.
result Comprehensive privacy-preserving analytics framework for AMI data.

In these notes, I will sketch a new approach to Khovanov homology of knots and links based on counting the solutions of certain elliptic partial differential equations in four and five dimensions. The equations are formulated on four and five-dimensional manifolds with boundary, with a rather subtle boundary condition …

2011-08-15abs ↗pdf ↗

The web contains a vast corpus of HTML tables. They can be used to provide direct answers to many web queries. We focus on answering two classes of queries with those tables: those seeking lists of entities (e.g., `cities in california') and those seeking superlative entities (e.g., `largest city in california'). The m…

2020-01-10abs ↗pdf ↗

Belief networks are a new, potentially important, class of knowledge-based models. ARCO1, currently under development at the Atlantic Richfield Company (ARCO) and the University of Southern California (USC), is the most advanced reported implementation of these models in a financial forecasting setting. ARCO1's underly…

2013-03-20abs ↗pdf ↗

The paper shows how demographic data can lead to biased predictions, proposing 'Affirmative Information' as a solution.

problem Bias in predictions due to demographic data.
method Characterization of error types and conditions leading to disparate impact.
result Demographic variables in data can lead to biased predictions, with higher average outcomes receiving higher false positive rates.

We propose a categorical data synthesizer with a quantifiable disclosure risk. Our algorithm, named Perturbed Gibbs Sampler, can handle high-dimensional categorical data that are often intractable to represent as contingency tables. The algorithm extends a multiple imputation strategy for fully synthetic data by utiliz…

2013-12-18abs ↗pdf ↗

Study shows houses appreciated more during pandemic due to speculation, not just price uncertainty.

problem Impact of COVID-19 on house prices and speculation.
method Quasi-experimental design, unit-level matching, multivariate difference-in-difference regression.
result Properties listed for sale appreciated an additional 1% per month after pandemic onset, with an excess annual growth of 12.7 percentage points.