Goto

Collaborating Authors

 Africa


UK says it thwarted Houthis' drone attack in the Red Sea

Al Jazeera

A UK vessel shot down a Houthi drone in the Red Sea, the United Kingdom's Ministry of Defence has said, as tensions in the Middle East soar amid the ongoing war in Gaza. "Yesterday HMS Diamond successfully repelled a drone attack from the Iranian-backed Houthis in the Red Sea," read a statement from the ministry published on Sunday on X. "Diamond destroyed a drone targeting her, with no injuries or damage sustained to Diamond or her crew," it added. There was no immediate comment from the Houthis. The Yemen-based group previously pledged to target Israel-linked vessels in the region as part of an effort to pressure the country's government to end its bombardment of Gaza and allow more humanitarian aid supplies into the coastal Palestinian enclave. Gaza has been under heavy bombardment by Israeli forces since October 7, when Hamas fighters stormed communities in southern Israel, killing at least 1,139 people and taking about 240 others captive, according to Israeli officials.


Three US service members killed in Jordan drone attack, Biden says

Al Jazeera

Three US service members have been killed and "many" others wounded during an unmanned aerial drone attack on US forces stationed in northeastern Jordan near the Syrian border, President Joe Biden has said, blaming Iran-backed groups for the attack. The United States military said in a statement that at least 25 people were injured. "While we are still gathering the facts of this attack, we know it was carried out by radical Iran-backed militant groups operating in Syria and Iraq," Biden said in a statement on Sunday. Biden said the US "will hold all those responsible to account at a time and in a manner [of] our choosing." Jordanian state television quoted Muhannad Mubaidin, a spokesperson for Jordan's government, as saying the attack happened outside of the kingdom across the border in Syria. There was no immediate comment from Iran.


3 American troops killed, 25 injured in attack on Jordan base near Syria border

FOX News

Fox News reporter Stephanie Bennett has more on the rising tension in the Middle East on'Fox News Live.' Three U.S. service members were killed and 25 others were injured in a drone attack on an outpost in northeast Jordan near the Syria border, U.S. Central Command confirmed on Sunday. "On Jan. 28, three U.S. service members were killed and 25 injured from a one-way attack UAS that impacted at a base in northeast Jordan, near the Syria border. As a matter of respect for the families and in accordance with DoD policy, the identities of the servicemembers will be withheld until 24 hours after their next of kin have been notified," CENTCOM said. "Updates will be provided as they become available," it added. The White House says President Biden was briefed Sunday morning by Defense Secretary Lloyd Austin, National Security Advisor Jake Sullivan, and Principal Deputy National Security Advisor Jon Finer about the attack, which marked a significant escalation in the Middle East as the first time American troops have been killed by enemy fire in the region since the Israel-Hamas war began.


Importance-Aware Adaptive Dataset Distillation

arXiv.org Artificial Intelligence

Herein, we propose a novel dataset distillation method for constructing small informative datasets that preserve the information of the large original datasets. The development of deep learning models is enabled by the availability of large-scale datasets. Despite unprecedented success, large-scale datasets considerably increase the storage and transmission costs, resulting in a cumbersome model training process. Moreover, using raw data for training raises privacy and copyright concerns. To address these issues, a new task named dataset distillation has been introduced, aiming to synthesize a compact dataset that retains the essential information from the large original dataset. State-of-the-art (SOTA) dataset distillation methods have been proposed by matching gradients or network parameters obtained during training on real and synthetic datasets. The contribution of different network parameters to the distillation process varies, and uniformly treating them leads to degraded distillation performance. Based on this observation, we propose an importance-aware adaptive dataset distillation (IADD) method that can improve distillation performance by automatically assigning importance weights to different network parameters during distillation, thereby synthesizing more robust distilled datasets. IADD demonstrates superior performance over other SOTA dataset distillation methods based on parameter matching on multiple benchmark datasets and outperforms them in terms of cross-architecture generalization. In addition, the analysis of self-adaptive weights demonstrates the effectiveness of IADD. Furthermore, the effectiveness of IADD is validated in a real-world medical application such as COVID-19 detection.


UnMASKed: Quantifying Gender Biases in Masked Language Models through Linguistically Informed Job Market Prompts

arXiv.org Artificial Intelligence

Language models (LMs) have become pivotal in the realm of technological advancements. While their capabilities are vast and transformative, they often include societal biases encoded in the human-produced datasets used for their training. This research delves into the inherent biases present in masked language models (MLMs), with a specific focus on gender biases. This study evaluated six prominent models: BERT, RoBERTa, DistilBERT, BERT-multilingual, XLM-RoBERTa, and DistilBERT-multilingual. The methodology employed a novel dataset, bifurcated into two subsets: one containing prompts that encouraged models to generate subject pronouns in English, and the other requiring models to return the probabilities of verbs, adverbs, and adjectives linked to the prompts' gender pronouns. The analysis reveals stereotypical gender alignment of all models, with multilingual variants showing comparatively reduced biases.


Real-time object detection and robotic manipulation for agriculture using a YOLO-based learning approach

arXiv.org Artificial Intelligence

The optimisation of crop harvesting processes for commonly cultivated crops is of great importance in the aim of agricultural industrialisation. Nowadays, the utilisation of machine vision has enabled the automated identification of crops, leading to the enhancement of harvesting efficiency, but challenges still exist. This study presents a new framework that combines two separate architectures of convolutional neural networks (CNNs) in order to simultaneously accomplish the tasks of crop detection and harvesting (robotic manipulation) inside a simulated environment. Crop images in the simulated environment are subjected to random rotations, cropping, brightness, and contrast adjustments to create augmented images for dataset generation. The you only look once algorithmic framework is employed with traditional rectangular bounding boxes for crop localization. The proposed method subsequently utilises the acquired image data via a visual geometry group model in order to reveal the grasping positions for the robotic manipulation.


YODA: Teacher-Student Progressive Learning for Language Models

arXiv.org Artificial Intelligence

Although large language models (LLMs) have demonstrated adeptness in a range of tasks, they still lag behind human learning efficiency. This disparity is often linked to the inherent human capacity to learn from basic examples, gradually generalize and handle more complex problems, and refine their skills with continuous feedback. Inspired by this, this paper introduces YODA, a novel teacher-student progressive learning framework that emulates the teacher-student education process to improve the efficacy of model fine-tuning. The framework operates on an interactive \textit{basic-generalized-harder} loop. The teacher agent provides tailored feedback on the student's answers, and systematically organizes the education process. This process unfolds by teaching the student basic examples, reinforcing understanding through generalized questions, and then enhancing learning by posing questions with progressively enhanced complexity. With the teacher's guidance, the student learns to iteratively refine its answer with feedback, and forms a robust and comprehensive understanding of the posed questions. The systematic procedural data, which reflects the progressive learning process of humans, is then utilized for model training. Taking math reasoning as a testbed, experiments show that training LLaMA2 with data from YODA improves SFT with significant performance gain (+17.01\% on GSM8K and +9.98\% on MATH). In addition, we find that training with curriculum learning further improves learning robustness.


Deep Learning for Gamma-Ray Bursts: A data driven event framework for X/Gamma-Ray analysis in space telescopes

arXiv.org Artificial Intelligence

The HERMES (High Energy Rapid Modular Ensemble of Satellites) Pathfinder mission serves as an in-orbit demonstration of a constellation of nanosatellites whose primary scientific purpose is to discover intense high-energy transients, such as gamma-ray bursts, across a broad energy range (few keV to few MeV) with unparalleled temporal precision and exact localisation. By 2024, the first constellation of six nanosatellites is expected to be launched. To fully exploit satellite data and allow faint astronomical events to emerge, a precise estimation of satellite background count rates is required to determine whether the event is statistically valid or not. The dynamics of the background are related to the satellite's orbital information, which varies in the order of minutes, potentially hiding long transient events. This work introduces two main contributions I have brought ahead; first a novel background estimator is presented that could potentially be fitted to any type of X/Gamma-ray satellite space telescope, capable of capturing long-term dynamics and accurate enough to detect faint transients. This estimator is built using a Neural Network and tested on data from the Fermi Gamma-ray Space Telescope's Gamma Burst Monitor (GBM). As a second objective, it is employed a trigger algorithm, called FOCuS (Functional Online CUSUM), to extract events from the background using the background estimator. The resulting framework, DeepGRB, can identify astronomical events that are both present and absent from the Fermi-GBM catalog. The analysis of the discovered events reveals the strengths and weaknesses of the framework.


Hyperedge Interaction-aware Hypergraph Neural Network

arXiv.org Artificial Intelligence

Hypergraphs provide an effective modeling approach for modeling high-order relationships in many real-world datasets. To capture such complex relationships, several hypergraph neural networks have been proposed for learning hypergraph structure, which propagate information from nodes to hyperedges and then from hyperedges back to nodes. However, most existing methods focus on information propagation between hyperedges and nodes, neglecting the interactions among hyperedges themselves. In this paper, we propose HeIHNN, a hyperedge interaction-aware hypergraph neural network, which captures the interactions among hyperedges during the convolution process and introduce a novel mechanism to enhance information flow between hyperedges and nodes. Specifically, HeIHNN integrates the interactions between hyperedges into the hypergraph convolution by constructing a three-stage information propagation process. After propagating information from nodes to hyperedges, we introduce a hyperedge-level convolution to update the hyperedge embeddings. Finally, the embeddings that capture rich information from the interaction among hyperedges will be utilized to update the node embeddings. Additionally, we introduce a hyperedge outlier removal mechanism in the information propagation stages between nodes and hyperedges, which dynamically adjusts the hypergraph structure using the learned embeddings, effectively removing outliers. Extensive experiments conducted on real-world datasets show the competitive performance of HeIHNN compared with state-of-the-art methods.


Can AI Assistants Know What They Don't Know?

arXiv.org Artificial Intelligence

Recently, AI assistants based on large language models (LLMs) show surprising performance in many tasks, such as dialogue, solving math problems, writing code, and using tools. Although LLMs possess intensive world knowledge, they still make factual errors when facing some knowledge intensive tasks, like open-domain question answering. These untruthful responses from the AI assistant may cause significant risks in practical applications. We believe that an AI assistant's refusal to answer questions it does not know is a crucial method for reducing hallucinations and making the assistant truthful. Therefore, in this paper, we ask the question "Can AI assistants know what they don't know and express them through natural language?" To answer this question, we construct a model-specific "I don't know" (Idk) dataset for an assistant, which contains its known and unknown questions, based on existing open-domain question answering datasets. Then we align the assistant with its corresponding Idk dataset and observe whether it can refuse to answer its unknown questions after alignment. Experimental results show that after alignment with Idk datasets, the assistant can refuse to answer most its unknown questions. For questions they attempt to answer, the accuracy is significantly higher than before the alignment.