Africa
Spatioformer: A Geo-encoded Transformer for Large-Scale Plant Species Richness Prediction
Guo, Yiqing, Mokany, Karel, Levick, Shaun R., Yang, Jinyan, Moghadam, Peyman
Earth observation data have shown promise in predicting species richness of vascular plants ($\alpha$-diversity), but extending this approach to large spatial scales is challenging because geographically distant regions may exhibit different compositions of plant species ($\beta$-diversity), resulting in a location-dependent relationship between richness and spectral measurements. In order to handle such geolocation dependency, we propose Spatioformer, where a novel geolocation encoder is coupled with the transformer model to encode geolocation context into remote sensing imagery. The Spatioformer model compares favourably to state-of-the-art models in richness predictions on a large-scale ground-truth richness dataset (HAVPlot) that consists of 68,170 in-situ richness samples covering diverse landscapes across Australia. The results demonstrate that geolocational information is advantageous in predicting species richness from satellite observations over large spatial scales. With Spatioformer, plant species richness maps over Australia are compiled from Landsat archive for the years from 2015 to 2023. The richness maps produced in this study reveal the spatiotemporal dynamics of plant species richness in Australia, providing supporting evidence to inform effective planning and policy development for plant diversity conservation. Regions of high richness prediction uncertainties are identified, highlighting the need for future in-situ surveys to be conducted in these areas to enhance the prediction accuracy.
Correlating Variational Autoencoders Natively For Multi-View Imputation
Orme, Ella S. C., Evangelou, Marina, Paquet, Ulrich
This is mirrored in correlation between the latent spaces of separate variational autoencoders (VAEs) trained on each data-view. A multi-view VAE approach is proposed that incorporates a joint prior with a non-zero correlation structure between the latent spaces of the VAEs. By enforcing such correlation structure, more strongly correlated latent spaces are uncovered. Using conditional distributions to move between these latent spaces, missing views can be imputed and used for downstream analysis. Learning this correlation structure involves maintaining validity of the prior distribution, as well as a successful parameterization that allows end-to-end learning.
How AI Is Being Used to Respond to Natural Disasters in Cities
The number of people living in urban areas has tripled in the last 50 years, meaning when a major natural disaster such as an earthquake strikes a city, more lives are in danger. Meanwhile, the strength and frequency of extreme weather events has increased--a trend set to continue as the climate warms. That is spurring efforts around the world to develop a new generation of earthquake monitoring and climate forecasting systems to make detecting and responding to disasters quicker, cheaper, and more accurate than ever. On Nov. 6, at the Barcelona Supercomputing Center in Spain, the Global Initiative on Resilience to Natural Hazards through AI Solutions will meet for the first time. The new United Nations initiative aims to guide governments, organizations, and communities in using AI for disaster management.
The Artificial State
"Jacob Javits of New York is the first United States senator to become fully automated," the Chicago Tribune announced in 1962 from the Republican state convention in Buffalo, where an electronic Javits spat out slips of paper with answers to questions about everything from Cuba's missiles ("a serious threat") to the Cubs' prospects (dim). Javits also harbors thoughts on medical care for the elderly, Berlin, the communist menace," and more than a hundred other subjects, the Tribune reported after an interview with the machine. Javits may have been the first automated American politician, but he wasn't the last. Since the nineteen-sixties, much of American public life has become automated, driven by computers and predictive algorithms that can do the political work of rallying support, running campaigns, communicating with constituents, and even crafting policy. In that same stretch of time, the proportion of Americans who say that they trust the U.S. government to do what is right ...
The Shipwreck Detective
The wreck was like a bug on the wall, a jumbly shape splayed on the abyssal plain. It was noticed by a team of autonomous-underwater-vehicle operators on board a subsea exploration vessel, working at an undisclosed location in the Atlantic Ocean, about a thousand miles from the nearest shore. The analysts belonged to a small private company that specializes in deep-sea search operations; I have been asked not to name it. They were looking for something else. In the past decade, the company has helped to transform the exploration of the seabed by deploying fleets of A.U.V.s--underwater drones--which cruise in formation, mapping large areas of the ocean floor with high-definition imagery.
Reshaping UAV-Enabled Communications with Omnidirectional Multi-Rotor Aerial Vehicles
Licea, Daniel Bonilla, Silano, Giuseppe, Hammouti, Hajar El, Ghogho, Mounir, Saska, Martin
A new class of Multi-Rotor Aerial Vehicles (MRAVs), known as omnidirectional MRAVs (o-MRAVs), has attracted significant interest in the robotics community. These MRAVs have the unique capability of independently controlling their 3D position and 3D orientation. In the context of aerial communication networks, this translates into the ability to control the position and orientation of the antenna mounted on the MRAV without any additional devices tasked for antenna orientation. This additional Degrees of Freedom (DoF) adds a new dimension to aerial communication systems, creating various research opportunities in communications-aware trajectory planning and positioning. This paper presents this new class of MRAVs and discusses use cases in areas such as physical layer security and optical communications. Furthermore, the benefits of these MRAVs are illustrated with realistic simulation scenarios. Finally, new research problems and opportunities introduced by this advanced robotics technology are discussed.
SibylSat: Using SAT as an Oracle to Perform a Greedy Search on TOHTN Planning
Quenard, Gaspard, Pellier, Damier, Fiorino, Humbert
This paper presents SibylSat, a novel SAT-based method designed to efficiently solve totally-ordered HTN problems (TOHTN). In contrast to prevailing SAT-based HTN planners that employ a breadth-first search strategy, SibylSat adopts a greedy search approach, enabling it to identify promising decompositions for expansion. The selection process is facilitated by a heuristic derived from solving a relaxed problem, which is also expressed as a SAT problem. Our experimental evaluations demonstrate that SibylSat outperforms existing SAT-based TOHTN approaches in terms of both runtime and plan quality on most of the IPC benchmarks, while also solving a larger number of problems.
Social Support Detection from Social Media Texts
Ahani, Zahra, Tash, Moein Shahiki, Balouchzahi, Fazlourrahman, Ramos, Luis, Sidorov, Grigori, Gelbukh, Alexander
Social support, conveyed through a multitude of interactions and platforms such as social media, plays a pivotal role in fostering a sense of belonging, aiding resilience in the face of challenges, and enhancing overall well-being. This paper introduces Social Support Detection (SSD) as a Natural language processing (NLP) task aimed at identifying supportive interactions within online communities. The study presents the task of Social Support Detection (SSD) in three subtasks: two binary classification tasks and one multiclass task, with labels detailed in the dataset section. We conducted experiments on a dataset comprising 10,000 YouTube comments. Traditional machine learning models were employed, utilizing various feature combinations that encompass linguistic, psycholinguistic, emotional, and sentiment information. Additionally, we experimented with neural network-based models using various word embeddings to enhance the performance of our models across these subtasks.The results reveal a prevalence of group-oriented support in online dialogues, reflecting broader societal patterns. The findings demonstrate the effectiveness of integrating psycholinguistic, emotional, and sentiment features with n-grams in detecting social support and distinguishing whether it is directed toward an individual or a group. The best results for different subtasks across all experiments range from 0.72 to 0.82.
Toward Robust Incomplete Multimodal Sentiment Analysis via Hierarchical Representation Learning
Li, Mingcheng, Yang, Dingkang, Liu, Yang, Wang, Shunli, Chen, Jiawei, Wang, Shuaibing, Wei, Jinjie, Jiang, Yue, Xu, Qingyao, Hou, Xiaolu, Sun, Mingyang, Qian, Ziyun, Kou, Dongliang, Zhang, Lihua
Multimodal Sentiment Analysis (MSA) is an important research area that aims to understand and recognize human sentiment through multiple modalities. The complementary information provided by multimodal fusion promotes better sentiment analysis compared to utilizing only a single modality. Nevertheless, in real-world applications, many unavoidable factors may lead to situations of uncertain modality missing, thus hindering the effectiveness of multimodal modeling and degrading the model's performance. To this end, we propose a Hierarchical Representation Learning Framework (HRLF) for the MSA task under uncertain missing modalities. Specifically, we propose a fine-grained representation factorization module that sufficiently extracts valuable sentiment information by factorizing modality into sentiment-relevant and modality-specific representations through crossmodal translation and sentiment semantic reconstruction. Moreover, a hierarchical mutual information maximization mechanism is introduced to incrementally maximize the mutual information between multi-scale representations to align and reconstruct the high-level semantics in the representations. Ultimately, we propose a hierarchical adversarial learning mechanism that further aligns and adapts the latent distribution of sentiment-relevant representations to produce robust joint multimodal representations. Comprehensive experiments on three datasets demonstrate that HRLF significantly improves MSA performance under uncertain modality missing cases.
Detecting Student Disengagement in Online Classes Using Deep Learning: A Review
Mohamed, Ahmed, Ali, Mostafa, Ahmed, Shahd, Hani, Nouran, Hisham, Mohammed, Mahmoud, Meram
Student disengagement in online learning has become a critical challenge, particularly post-pandemic. This review explores deep learning techniques used to detect disengagement, emphasizing computer vision and affective computing as effective approaches. We examine recent studies focusing on facial expressions, eye movements, and posture to assess student attention, along with non-face-based indicators like mouse activity. A systematic review of 38 selected studies outlines the indicators, methods, and models employed in this field, providing insights for future research on real-time engagement monitoring in online classrooms