Goto

Collaborating Authors

 Country


A General Weighting Theory for Ensemble Learning: Beyond Variance Reduction via Spectral and Geometric Structure

arXiv.org Machine Learning

Ensemble learning is traditionally justified as a variance-reduction strategy, explaining its strong performance for unstable predictors such as decision trees. This explanation, however, does not account for ensembles constructed from intrinsically stable estimators-including smoothing splines, kernel ridge regression, Gaussian process regression, and other regularized reproducing kernel Hilbert space (RKHS) methods whose variance is already tightly controlled by regularization and spectral shrinkage. This paper develops a general weighting theory for ensemble learning that moves beyond classical variance-reduction arguments. We formalize ensembles as linear operators acting on a hypothesis space and endow the space of weighting sequences with geometric and spectral constraints. Within this framework, we derive a refined bias-variance approximation decomposition showing how non-uniform, structured weights can outperform uniform averaging by reshaping approximation geometry and redistributing spectral complexity, even when variance reduction is negligible. Our main results provide conditions under which structured weighting provably dominates uniform ensembles, and show that optimal weights arise as solutions to constrained quadratic programs. Classical averaging, stacking, and recently proposed Fibonacci-based ensembles appear as special cases of this unified theory, which further accommodates geometric, sub-exponential, and heavy-tailed weighting laws. Overall, the work establishes a principled foundation for structure-driven ensemble learning, explaining why ensembles remain effective for smooth, low-variance base learners and setting the stage for distribution-adaptive and dynamically evolving weighting schemes developed in subsequent work.


A review of NMF, PLSA, LBA, EMA, and LCA with a focus on the identifiability issue

arXiv.org Machine Learning

Across fields such as machine learning, social science, geography, considerable attention has been given to models that factorize a nonnegative matrix into the product of two or three matrices, subject to nonnegative or row-sum-to-1 constraints. Although these models are to a large extend similar or even equivalent, they are presented under different names, and their similarity is not well known. This paper highlights similarities among five popular models, latent budget analysis (LBA), latent class analysis (LCA), end-member analysis (EMA), probabilistic latent semantic analysis (PLSA), and nonnegative matrix factorization (NMF). We focus on an essential issue-identifiability-of these models and prove that the solution of LBA, EMA, LCA, PLSA is unique if and only if the solution of NMF is unique. We also provide a brief review for algorithms of these models. We illustrate the models with a time budget dataset from social science, and end the paper with a discussion of closely related models such as archetypal analysis.


Predicting Mycotoxin Contamination in Irish Oats Using Deep and Transfer Learning

arXiv.org Machine Learning

Mycotoxin contamination poses a significant risk to cereal crop quality, food safety, and agricultural productivity. Accurate prediction of mycotoxin levels can support early intervention strategies and reduce economic losses. This study investigates the use of neural networks and transfer learning models to predict mycotoxin contamination in Irish oat crops as a multi-response prediction task. Our dataset comprises oat samples collected in Ireland, containing a mix of environmental, agronomic, and geographical predictors. Five modelling approaches were evaluated: a baseline multilayer perceptron (MLP), an MLP with pre-training, and three transfer learning models; TabPFN, TabNet, and FT-Transformer. Model performance was evaluated using regression (RMSE, $R^2$) and classification (AUC, F1) metrics, with results reported per toxin and on average. Additionally, permutation-based variable importance analysis was conducted to identify the most influential predictors across both prediction tasks. The transfer learning approach TabPFN provided the overall best performance, followed by the baseline MLP. Our variable importance analysis revealed that weather history patterns in the 90-day pre-harvest period were the most important predictors, alongside seed moisture content.


Sampling with Shielded Langevin Monte Carlo Using Navigation Potentials

arXiv.org Machine Learning

We introduce shielded Langevin Monte Carlo (LMC), a constrained sampler inspired by navigation functions, capable of sampling from unnormalized target distributions defined over punctured supports. In other words, this approach samples from non-convex spaces defined as convex sets with convex holes. This defines a novel and challenging problem in constrained sampling. To do so, the sampler incorporates a combination of a spatially adaptive temperature and a repulsive drift to ensure that samples remain within the feasible region. Experiments on a 2D Gaussian mixture and multiple-input multiple-output (MIMO) symbol detection showcase the advantages of the proposed shielded LMC in contrast to unconstrained cases.


Towards Unsupervised Causal Representation Learning via Latent Additive Noise Model Causal Autoencoders

arXiv.org Machine Learning

Unsupervised representation learning seeks to recover latent generative factors, yet standard methods relying on statistical independence often fail to capture causal dependencies. A central challenge is identifiability: as established in disentangled representation learning and nonlinear ICA literature, disentangling causal variables from observational data is impossible without supervision, auxiliary signals, or strong inductive biases. In this work, we propose the Latent Additive Noise Model Causal Autoencoder (LANCA) to operationalize the Additive Noise Model (ANM) as a strong inductive bias for unsupervised discovery. Theoretically, we prove that while the ANM constraint does not guarantee unique identifiability in the general mixing case, it resolves component-wise indeterminacy by restricting the admissible transformations from arbitrary diffeo-morphisms to the affine class. Methodologically, arguing that the stochastic encoding inherent to V AEs obscures the structural residuals required for latent causal discovery, LANCA employs a deterministic Wasserstein Auto-Encoder (W AE) coupled with a differentiable ANM Layer. This architecture transforms residual independence from a passive assumption into an explicit optimization objective. Empirically, LANCA outperforms state-of-the-art baselines on synthetic physics benchmarks (Pendulum, Flow), and on photorealistic environments (CANDLE), where it demonstrates superior robustness to spurious correlations arising from complex background scenes.


How to watch the LG CES 2026 press conference

Engadget

The Korean electronics giant will showcase new TVs and plenty of AI-powered products when it kicks off press day in Las Vegas. For years, LG has kicked off CES press day with the first event of the morning -- and 2026 will be no different. The Korea-based corporation is theming its presentation as Innovation in Tune with You, and -- if it follows the template of past presentations -- it will highlight both the consumer electronics and large appliance sides of its mammoth global businesses. Like nearly all tech-centric events these days, expect AI to be the binding theme of the LG presentation at CES 2026 . Just be aware that, like Apple, LG has its own customized abbreviation for AI: Affectionate Intelligence.


'Stranger Things' fans review-bomb 'woke' coming-out scene in show's final season

FOX News

This material may not be published, broadcast, rewritten, or redistributed. Quotes displayed in real-time or delayed by at least 15 minutes. Market data provided by Factset . Powered and implemented by FactSet Digital Solutions . Mutual Fund and ETF data provided by Refinitiv Lipper .


Zelenskyy denies Russian claim of Ukrainian strike on Putin residence

Al Jazeera

Could Ukraine hold a presidential election right now? Will Europe use frozen Russian assets to fund war? How can Ukraine rebuild China ties? 'Ukraine is running out of men, money and time' Volodymyr Zelenskyy quickly denied a claim by Moscow that his country's military launched a drone attack on Vladimir Putin's residence in the city of Novgorod. The Ukrainian leader accused Russia of trying to derail peace talks a day after Zelenskyy met with US President Donald Trump.


Baby spider monkeys rescued in Texas

Popular Science

Animal traffickers face up to 20 years in prison and a $250,000 fine. Breakthroughs, discoveries, and DIY tips sent every weekday. It should go without saying, but please don't smuggle spider monkeys. While responding to a human trafficking case earlier this year, United States Border Patrol agents in Laredo, Texas, found two of these tiny primates . The driver failed to yield and fled the scene, leading officers to respond.


700Credit data breach exposes SSNs of 5.8M consumers

FOX News

U.S.-based fintech company 700Credit confirms 2025 cybersecurity incident affecting 5.8 million consumers after hackers accessed data through third-party vendor compromise that went undetected.