Goto

Collaborating Authors

 Government


Generalization v.s. Memorization: Tracing Language Models' Capabilities Back to Pretraining Data

arXiv.org Artificial Intelligence

Despite the proven utility of large language models (LLMs) in real-world applications, there remains a lack of understanding regarding how they leverage their large-scale pretraining text corpora to achieve such capabilities. In this work, we investigate the interplay between generalization and memorization in pretrained LLMs at scale, through a comprehensive $n$-gram analysis of their training data. Our experiments focus on three general task types: translation, question-answering, and multiple-choice reasoning. With various sizes of open-source LLMs and their pretraining corpora, we observe that as the model size increases, the task-relevant $n$-gram pair data becomes increasingly important, leading to improved task performance, decreased memorization, stronger generalization, and emergent abilities. Our results support the hypothesis that LLMs' capabilities emerge from a delicate balance of memorization and generalization with sufficient task-related pretraining data, and point the way to larger-scale analyses that could further improve our understanding of these models.


Implementing Fairness: the view from a FairDream

arXiv.org Artificial Intelligence

In this paper, we propose an experimental investigation of the problem of AI fairness in classification. We train an AI model and develop our own fairness package FairDream to detect inequalities and then to correct for them, using income prediction as a case study. Our experiments show that it is a property of FairDream to fulfill fairness objectives which are conditional on the ground truth (Equalized Odds), even when the algorithm is set the task of equalizing positives across groups (Demographic Parity). While this may be seen as an anomaly, we explain this property by comparing our approach with a closely related fairness method (GridSearch), which can enforce Demographic Parity at the expense of Equalized Odds. We grant that a fairness metric conditioned on true labels does not give a sufficient criterion to reach fairness, but we argue that it gives us at least a necessary condition to implement Demographic Parity cautiously. We also explain why neither Equal Calibration nor Equal Precision stand as relevant fairness criteria in classification. Addressing their limitations to warn the decision-maker for any disadvantaging rate, Equalized Odds avoids the peril of strict conservatism, while keeping away the utopia of a whole redistribution of resources through algorithms.


Text Style Transfer: An Introductory Overview

arXiv.org Artificial Intelligence

Text Style Transfer (TST) is a pivotal task in natural language generation to manipulate text style attributes while preserving style-independent content. The attributes targeted in TST can vary widely, including politeness, authorship, mitigation of offensive language, modification of feelings, and adjustment of text formality. TST has become a widely researched topic with substantial advancements in recent years. This paper provides an introductory overview of TST, addressing its challenges, existing approaches, datasets, evaluation measures, subtasks, and applications. This fundamental overview improves understanding of the background and fundamentals of text style transfer.


A Measure for Level of Autonomy Based on Observable System Behavior

arXiv.org Artificial Intelligence

Contemporary artificial intelligence systems are pivotal in enhancing human efficiency and safety across various domains. One such domain is autonomous systems, especially in automotive and defense use cases. Artificial intelligence brings learning and enhanced decision-making to autonomy system goal-oriented behaviors and human independence. However, the lack of clear understanding of autonomy system capabilities hampers human-machine or machine-machine interaction and interdiction. This necessitates varying degrees of human involvement for safety, accountability, and explainability purposes. Yet, measuring the level autonomous capability in an autonomous system presents a challenge. Two scales of measurement exist, yet measuring autonomy presupposes a variety of elements not available in the wild. This is why existing measures for level of autonomy are operationalized only during design or test and evaluation phases. No measure for level of autonomy based on observed system behavior exists at this time. To address this, we outline a potential measure for predicting level of autonomy using observable actions. We also present an algorithm incorporating the proposed measure. The measure and algorithm have significance to researchers and practitioners interested in a method to blind compare autonomous systems at runtime. Defense-based implementations are likewise possible because counter-autonomy depends on robust identification of autonomous systems.


Decentralized Federated Anomaly Detection in Smart Grids: A P2P Gossip Approach

arXiv.org Artificial Intelligence

Decentralized Federated Anomaly Detection in Smart Grids: A P2P Gossip Approach Muhammad Akbar Husnoo a,, Adnan Anwar a, Md Enamul Haque b and Abdun Naser Mahmood c a Centre for Cyber Resilience and Trust (CREST), Deakin University, 75 Pigdons Rd, Waurn Ponds, 3216, Victoria, Australia b Centre for Smart Power and Energy Research (CSPER)), Deakin University, 75 Pigdons Rd, Waurn Ponds, 3216, Victoria, Australia c Department of Computer Science & IT, Latrobe University, Plenty Rd, Bundoora, 3086, Victoria, AustraliaA R T I C L E I N F OKeywords: Anomaly Detection Decentralized Federated Learning (DFL) Cyberattack Internet of Things (Io T) Smart Grid A B S T R A C T Amidst escalating concerns regarding security and privacy within the Smart Grid domain, the need for robust intrusion detection mechanisms in critical energy infrastructure has surged in recent times. To address the challenges posed by privacy preservation and decentralized power zones with distinct data ownership, Federated Learning (FL) has emerged as a promising privacy-preserving solution which facilitates collaborative training of attack detection models without necessitating the sharing of raw data. However, FL presents several implementation limitations in the power system domain due to its heavy reliance on a centralized aggregator and the risks of privacy leakage during model update transmission. In response to the technical bottlenecks, this paper introduces a novel decentralized federated anomaly detection scheme based on two main gossip protocols namely Random Walk and Epidemic. Our findings indicate that the Random Walk protocol exhibits superior performance compared to the Epidemic protocol, highlighting its efficacy in decentralized federated learning environments. Experimental validation of the proposed framework utilizing publicly available industrial control systems datasets demonstrates superior attack detection accuracy while safeguarding data confidentiality and mitigating the impact of communication latency and stragglers. Moreover, a notable 35% improvement in training time against conventional FL highlights the efficacy and robustness of our decentralized learning approach.1.


Hyperspectral Unmixing Under Endmember Variability: A Variational Inference Framework

arXiv.org Artificial Intelligence

This work proposes a variational inference (VI) framework for hyperspectral unmixing in the presence of endmember variability (HU-EV). An EV-accounted noisy linear mixture model (LMM) is considered, and the presence of outliers is also incorporated into the model. Following the marginalized maximum likelihood (MML) principle, a VI algorithmic structure is designed for probabilistic inference for HU-EV. Specifically, a patch-wise static endmember assumption is employed to exploit spatial smoothness and to try to overcome the ill-posed nature of the HU-EV problem. The design facilitates lightweight, continuous optimization-based updates under a variety of endmember priors. Some of the priors, such as the Beta prior, were previously used under computationally heavy, sampling-based probabilistic HU-EV methods. The effectiveness of the proposed framework is demonstrated through synthetic, semi-real, and real-data experiments.


Operationalizing a Threat Model for Red-Teaming Large Language Models (LLMs)

arXiv.org Artificial Intelligence

Creating secure and resilient applications with large language models (LLM) requires anticipating, adjusting to, and countering unforeseen threats. Red-teaming has emerged as a critical technique for identifying vulnerabilities in real-world LLM implementations. This paper presents a detailed threat model and provides a systematization of knowledge (SoK) of red-teaming attacks on LLMs. We develop a taxonomy of attacks based on the stages of the LLM development and deployment process and extract various insights from previous research. In addition, we compile methods for defense and practical red-teaming strategies for practitioners. By delineating prominent attack motifs and shedding light on various entry points, this paper provides a framework for improving the security and robustness of LLM-based systems.


Visual Geo-Localization from images

arXiv.org Artificial Intelligence

Algorithms process this data to pinpoint exact coordinates[11][12]. Geo-localization is important for organizing and analyzing large volumes of imagery data, as demonstrated by systems like the US Geological Survey (USGS), which classify and locate satellite and drone images to streamline data collection and analysis. Social media platforms like Instagram use geo-localization to tag photos with specific locations, enabling users to explore location-based content[11]. Despite its significance, many images and videos lack geo-localization data, particularly those collected in the past or by devices without GPS capabilities[12].


Mapping the Technological Future: A Topic, Sentiment, and Emotion Analysis in Social Media Discourse

arXiv.org Artificial Intelligence

People worldwide are currently confronted with a number of technological challenges, which act as a potent source of uncertainty. The uncertainty arising from the volatility and unpredictability of technology (such as AI) and its potential consequences is widely discussed on social media. This study uses BERTopic modelling along with sentiment and emotion analysis on 1.5 million tweets from 2021 to 2023 to identify anticipated tech-driven futures and capture the emotions communicated by 400 key opinion leaders (KOLs). Findings indicate positive sentiment significantly outweighs negative, with a prevailing dominance of positive anticipatory emotions. Specifically, the 'Hope' score is approximately 10.33\% higher than the median 'Anxiety' score. KOLs emphasize 'Optimism' and benefits over 'Pessimism' and challenges. The study emphasizes the important role KOLs play in shaping future visions through anticipatory discourse and emotional tone during times of technological uncertainty.


Israel defense minister says country will 'settle the score' after Houthi drone attack on Tel Aviv

FOX News

Israel's defense minister struck an ominous tone Friday after an Iranian-made drone fired by Houthi rebels in Yemen struck Tel Aviv, telling Israeli media that Jerusalem would "settle the score." "I held an operational situation assessment this morning to review the steps required to strengthen our defense arrays in light of events overnight, as well as the intelligence and operational activities required against those responsible for the attack," Israeli Minister of Defense Yoav Gallant said in a statement. "The year 2024 is marked by war. We must be prepared for every scenario and every arena." Israeli Minister of Defense Yoav Gallant sits with defense officials after a Yemen-based Houthi drone strike on Tel Aviv July 19, 2024.