Overview
Review for NeurIPS paper: Understanding Deep Architecture with Reasoning Layer
The analysis connects underlying algorithm property and the performance of the deep learning models. In the learning theory analysis, the local Rademacher complexity technique is utilized to obtain tighter bound, which enables to reveal trade-off corresponding to the number of layers. The theoretical findings are justified from numerical experiments. This paper deals with a new problem setting and gives a nice first step. Although its problem setting is quite simple, it is expected that this kind of study will open up a new direction of researches.
Entity Linking using LLMs for Automated Product Carbon Footprint Estimation
Castle, Steffen, Schneider, Julian Moreno, Hennig, Leonhard, Rehm, Georg
Growing concerns about climate change and sustainability are driving manufacturers to take significant steps toward reducing their carbon footprints. For these manufacturers, a first step towards this goal is to identify the environmental impact of the individual components of their products. We propose a system leveraging large language models (LLMs) to automatically map components from manufacturer Bills of Materials (BOMs) to Life Cycle Assessment (LCA) database entries by using LLMs to expand on available component information. Our approach reduces the need for manual data processing, paving the way for more accessible sustainability practices.
Corporate Greenwashing Detection in Text - a Survey
Calamai, Tom, Balalau, Oana, Guenedal, Thรฉo Le, Suchanek, Fabian M.
This increased awareness has translated into guidelines, laws, and investments, such as the European Green Deal [84] or the Inflation Reduction Act in the US [106]. Many companies have used the financial incentives offered by states, and the guidelines and legislation to make significant steps towards sustainability [109]. At the same time, this growing attention also generated an advertising opportunity for companies that aim to promote themselves as environmentally aware and responsible. Indeed, some companies have been found to deliberately manipulate their data and statistics to appear more environment-friendly. The Diesel Scandal around the Volkswagen car company is a prominent example [116]. However, such cases are not the norm. More commonly, companies avoid outright data manipulation but present themselves in a misleadingly positive light regarding their environmental impact - a practice called greenwashing.
Large Language Models as Proxies for Theories of Human Linguistic Cognition
Ziv, Imry, Lan, Nur, Chemla, Emmanuel, Katzir, Roni
We consider the possible role of current large language models (LLMs) in the study of human linguistic cognition. We focus on the use of such models as proxies for theories of cognition that are relatively linguistically-neutral in their representations and learning but differ from current LLMs in key ways. We illustrate this potential use of LLMs as proxies for theories of cognition in the context of two kinds of questions: (a) whether the target theory accounts for the acquisition of a given pattern from a given corpus; and (b) whether the target theory makes a given typologically-attested pattern easier to acquire than another, typologically-unattested pattern. For each of the two questions we show, building on recent literature, how current LLMs can potentially be of help, but we note that at present this help is quite limited.
A Survey of In-Context Reinforcement Learning
Moeini, Amir, Wang, Jiuqi, Beck, Jacob, Blaser, Ethan, Whiteson, Shimon, Chandra, Rohan, Zhang, Shangtong
Reinforcement learning (RL) agents typically optimize their policies by performing expensive backward passes to update their network parameters. However, some agents can solve new tasks without updating any parameters by simply conditioning on additional context such as their action-observation histories. This paper surveys work on such behavior, known as in-context reinforcement learning.
Out-of-Distribution Detection on Graphs: A Survey
Cai, Tingyi, Jiang, Yunliang, Liu, Yixin, Li, Ming, Huang, Changqin, Pan, Shirui
Graph machine learning has witnessed rapid growth, driving advancements across diverse domains. However, the in-distribution assumption, where training and testing data share the same distribution, often breaks in real-world scenarios, leading to degraded model performance under distribution shifts. This challenge has catalyzed interest in graph out-of-distribution (GOOD) detection, which focuses on identifying graph data that deviates from the distribution seen during training, thereby enhancing model robustness. In this paper, we provide a rigorous definition of GOOD detection and systematically categorize existing methods into four types: enhancement-based, reconstruction-based, information propagation-based, and classification-based approaches. We analyze the principles and mechanisms of each approach and clarify the distinctions between GOOD detection and related fields, such as graph anomaly detection, outlier detection, and GOOD generalization. Beyond methodology, we discuss practical applications and theoretical foundations, highlighting the unique challenges posed by graph data. Finally, we discuss the primary challenges and propose future directions to advance this emerging field. The repository of this survey is available at https://github.com/ca1man-2022/Awesome-GOOD-Detection.
Mathematical reasoning and the computer
Computers have already changed the way that humans do mathematics: they enable us to compute efficiently. But will they soon be helping us to reason? And will they one day start reasoning themselves? We give an overview of recent developments in neural networks, computer theorem provers and large language models.
Vision-Language Models for Edge Networks: A Comprehensive Survey
Sharshar, Ahmed, Khan, Latif U., Ullah, Waseem, Guizani, Mohsen
Vision Large Language Models (VLMs) combine visual understanding with natural language processing, enabling tasks like image captioning, visual question answering, and video analysis. While VLMs show impressive capabilities across domains such as autonomous vehicles, smart surveillance, and healthcare, their deployment on resource-constrained edge devices remains challenging due to processing power, memory, and energy limitations. This survey explores recent advancements in optimizing VLMs for edge environments, focusing on model compression techniques, including pruning, quantization, knowledge distillation, and specialized hardware solutions that enhance efficiency. We provide a detailed discussion of efficient training and fine-tuning methods, edge deployment challenges, and privacy considerations. Additionally, we discuss the diverse applications of lightweight VLMs across healthcare, environmental monitoring, and autonomous systems, illustrating their growing impact. By highlighting key design strategies, current challenges, and offering recommendations for future directions, this survey aims to inspire further research into the practical deployment of VLMs, ultimately making advanced AI accessible in resource-limited settings.
Enhancing Video Understanding: Deep Neural Networks for Spatiotemporal Analysis
Fadaei, Amir Hosein, Dehaqani, Mohammad-Reza A.
It's no secret that video has become the primary way we share information online. That's why there's been a surge in demand for algorithms that can analyze and understand video content. It's a trend going to continue as video continues to dominate the digital landscape. These algorithms will extract and classify related features from the video and will use them to describe the events and objects in the video. Deep neural networks have displayed encouraging outcomes in the realm of feature extraction and video description. This paper will explore the spatiotemporal features found in videos and recent advancements in deep neural networks in video understanding. We will review some of the main trends in video understanding models and their structural design, the main problems, and some offered solutions in this topic. We will also review and compare significant video understanding and action recognition datasets.
GENERator: A Long-Context Generative Genomic Foundation Model
Wu, Wei, Li, Qiuyi, Li, Mingyang, Fu, Kun, Feng, Fuli, Ye, Jieping, Xiong, Hui, Wang, Zheng
Advancements in DNA sequencing technologies have significantly improved our ability to decode genomic sequences. However, the prediction and interpretation of these sequences remain challenging due to the intricate nature of genetic material. Large language models (LLMs) have introduced new opportunities for biological sequence analysis. Recent developments in genomic language models have underscored the potential of LLMs in deciphering DNA sequences. Nonetheless, existing models often face limitations in robustness and application scope, primarily due to constraints in model structure and training data scale. To address these limitations, we present GENERator, a generative genomic foundation model featuring a context length of 98k base pairs (bp) and 1.2B parameters. Trained on an expansive dataset comprising 386B bp of eukaryotic DNA, the GENERator demonstrates state-of-the-art performance across both established and newly proposed benchmarks. The model adheres to the central dogma of molecular biology, accurately generating protein-coding sequences that translate into proteins structurally analogous to known families. It also shows significant promise in sequence optimization, particularly through the prompt-responsive generation of promoter sequences with specific activity profiles. These capabilities position the GENERator as a pivotal tool for genomic research and biotechnological advancement, enhancing our ability to interpret and predict complex biological systems and enabling precise genomic interventions.