AITopics | overparameterisation

Collaborating Authors

overparameterisation

Information about AI from the News, Publications, and Conferences

Automatic Classification – Tagging and Summarization – Customizable Filtering and Analysis

If you are looking for an answer to the question What is Artificial Intelligence? and you only have a minute, then here's the definition the Association for the Advancement of Artificial Intelligence offers on its home page: "the scientific understanding of the mechanisms underlying thought and intelligent behavior and their embodiment in machines."

However, if you are fortunate enough to have more than a minute, then please get ready to embark upon an exciting journey exploring AI (but beware, it could last a lifetime) …

Tilting the Odds at the Lottery: the Interplay of Overparameterisation and Curricula in Neural Networks

Mannelli, Stefano Sarao, Ivashinka, Yaraslau, Saxe, Andrew, Saglietti, Luca

arXiv.org Machine LearningJun-3-2024

A wide range of empirical and theoretical works have shown that overparameterisation can amplify the performance of neural networks. According to the lottery ticket hypothesis, overparameterised networks have an increased chance of containing a sub-network that is well-initialised to solve the task at hand. A more parsimonious approach, inspired by animal learning, consists in guiding the learner towards solving the task by curating the order of the examples, i.e. providing a curriculum. However, this learning strategy seems to be hardly beneficial in deep learning applications. In this work, we undertake an analytical study that connects curriculum learning and overparameterisation. In particular, we investigate their interplay in the online learning setting for a 2-layer network in the XOR-like Gaussian Mixture problem. Our results show that a high degree of overparameterisation -- while simplifying the problem -- can limit the benefit from curricula, providing a theoretical account of the ineffectiveness of curricula in deep learning.

curriculum, neuron, overparameterisation, (16 more...)

arXiv.org Machine Learning

2406.01589

Country:

Europe > Austria > Vienna (0.14)
Africa > Middle East > Tunisia > Ben Arous Governorate > Ben Arous (0.05)
Europe > United Kingdom > England > Greater London > London (0.04)
Europe > Italy > Lombardy > Milan (0.04)

Genre: Research Report > New Finding (1.00)

Industry: Education (0.66)

Technology: Information Technology > Artificial Intelligence > Machine Learning > Neural Networks > Deep Learning (0.75)

Add feedback

Learning Compact Neural Networks with Deep Overparameterised Multitask Learning

Ren, Shen, Shi, Haosen

arXiv.org Artificial IntelligenceAug-25-2023

The left and right singular vectors are trained with all task losses, and the diagonal matrices are trained using taskspecific Compact neural network offers many benefits for losses. Our design is mainly inspired by analytical real-world applications. However, it is usually studies on overparameterised networks for MTL [Lampinen challenging to train the compact neural networks and Ganguli, 2018] that the training/test error dynamics depends with small parameter sizes and low computational on the time-evolving alignment of the network parameters costs to achieve the same or better model performance to the singular vectors of the training data, and a quantifiable compared to more complex and powerful task alignment describing the transfer benefits among architecture. This is particularly true for multitask multiple tasks depends on the singular values and input feature learning, with different tasks competing for resources.

artificial intelligence, machine learning, matrix, (15 more...)

arXiv.org Artificial Intelligence

2308.133

Country: Asia > Singapore (0.04)

Genre: Research Report (0.50)

Technology: Information Technology > Artificial Intelligence > Machine Learning > Neural Networks (1.00)

Add feedback