AITopics | training summary

Collaborating Authors

training summary

Information about AI from the News, Publications, and Conferences

Automatic Classification – Tagging and Summarization – Customizable Filtering and Analysis

If you are looking for an answer to the question What is Artificial Intelligence? and you only have a minute, then here's the definition the Association for the Advancement of Artificial Intelligence offers on its home page: "the scientific understanding of the mechanisms underlying thought and intelligent behavior and their embodiment in machines."

However, if you are fortunate enough to have more than a minute, then please get ready to embark upon an exciting journey exploring AI (but beware, it could last a lifetime) …

Lexical Repetitions Lead to Rote Learning: Unveiling the Impact of Lexical Overlap in Train and Test Reference Summaries

Choubey, Prafulla Kumar, Fabbri, Alexander R., Xiong, Caiming, Wu, Chien-Sheng

arXiv.org Artificial IntelligenceNov-15-2023

Ideal summarization models should generalize to novel summary-worthy content without remembering reference training summaries by rote. However, a single average performance score on the entire test set is inadequate in determining such model competencies. We propose a fine-grained evaluation protocol by partitioning a test set based on the lexical similarity of reference test summaries with training summaries. We observe up to a 5x (1.2x) difference in ROUGE-2 (entity recall) scores between the subsets with the lowest and highest similarity. Next, we show that such training repetitions also make a model vulnerable to rote learning, reproducing data artifacts such as factual errors, especially when reference test summaries are lexically close to training summaries. Consequently, we propose to limit lexical repetitions in training summaries during both supervised fine-tuning and likelihood calibration stages to improve the performance on novel test cases while retaining average performance. Our automatic and human evaluations on novel test subsets and recent news articles show that limiting lexical repetitions in training summaries can prevent rote learning and improve generalization.

dataset, subset, training summary, (14 more...)

arXiv.org Artificial Intelligence

2311.09458

Country:

Europe > Russia (0.05)
Asia > Russia (0.05)
Europe > Spain > Catalonia > Barcelona Province > Barcelona (0.04)
(8 more...)

Genre: Research Report (0.82)

Industry: Government (0.46)

Technology:

Information Technology > Artificial Intelligence > Natural Language (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Memory-Based Learning > Rote Learning (0.81)

Add feedback

How to Build a Poisson Hidden Markov Model Using Python and Statsmodels

#artificialintelligenceDec-6-2021, 06:35:52 GMT

A Poisson Hidden Markov Model is a mixture of two regression models: A Poisson regression model which is visible and a Markov model which is ‘hidden’.

matrix, poisson hidden markov model, statsmodel, (14 more...)

#artificialintelligence

Technology: Information Technology > Artificial Intelligence > Machine Learning > Learning Graphical Models > Undirected Networks > Markov Models (1.00)

Add feedback

Image Classifier: Deployed on Heroku Using FastAI, Flask, and Node JS

#artificialintelligenceAug-24-2021, 07:25:12 GMT

The code below is a boilerplate of image classification models seen elsewhere and has been retooled specifically for this dataset. For the dataset, I have built a web scraper using Beautiful Soup to download the images of top 50 dog breeds as reported in American Kennel Club. In total, there were 5000 images, 100 images per breed, allowing us to maintain the same distrubtion of training and validation dataset between classes. I have chosen ResNet34 over ResNet101, ResNet50 and ResNet18 as the model architecture here because of its optimal performance metrics (speed and accuracy). To faciltate model generalization, default data augmentation was applied to the training dataset using a batch size of 8. I have used a batch size of 8 here because of an'out of memory' error when 32 or 64 were used in AWS SageMaker notebook instance.

dataset, deployed, image classifier, (8 more...)

#artificialintelligence

Technology: Information Technology > Artificial Intelligence > Machine Learning (0.99)

Add feedback