Oceania
LLaMA-E: Empowering E-commerce Authoring with Multi-Aspect Instruction Following
Shi, Kaize, Sun, Xueyao, Wang, Dingxian, Fu, Yinlin, Xu, Guandong, Li, Qing
E-commerce authoring involves creating attractive, abundant, and targeted promotional content to drive product sales. The emergence of large language models (LLMs) introduces an innovative paradigm, offering a unified solution to address various authoring tasks within this scenario. However, mainstream LLMs trained on general corpora with common sense knowledge reveal limitations in fitting complex and personalized features unique to e-commerce products and customers. Furthermore, LLMs like GPT-3.5 necessitate remote accessibility, raising concerns about safeguarding voluminous customer privacy data during transmission. This paper proposes the LLaMA-E, the unified and customized instruction-following language models focusing on diverse e-commerce authoring tasks. Specifically, the domain experts create the seed instruction set from the tasks of ads generation, query-enhanced product title rewriting, product classification, purchase intent speculation, and general Q&A. These tasks enable the models to comprehensively understand precise e-commerce authoring knowledge by interleaving features covering typical service aspects of customers, sellers, and platforms. The GPT-3.5 is introduced as a teacher model, which expands the seed instructions to form a training set for the LLaMA-E models with various scales. The experimental results show that the proposed LLaMA-E models achieve state-of-the-art results in quantitative and qualitative evaluations, also exhibiting the advantage in zero-shot scenes. To the best of our knowledge, this study is the first to serve the LLMs to specific e-commerce authoring scenarios.
Intrinsic Motivation via Surprise Memory
Le, Hung, Do, Kien, Nguyen, Dung, Venkatesh, Svetha
We present a new computing model for intrinsic rewards in reinforcement learning that addresses the limitations of existing surprise-driven explorations. The reward is the novelty of the surprise rather than the surprise norm. We estimate the surprise novelty as retrieval errors of a memory network wherein the memory stores and reconstructs surprises. Our surprise memory (SM) augments the capability of surprise-based intrinsic motivators, maintaining the agent's interest in exciting exploration while reducing unwanted attraction to unpredictable or noisy observations. Our experiments demonstrate that the SM combined with various surprise predictors exhibits efficient exploring behaviors and significantly boosts the final performance in sparse reward environments, including Noisy-TV, navigation and challenging Atari games.
Neuro-Symbolic RDF and Description Logic Reasoners: The State-Of-The-Art and Challenges
Singh, Gunjan, Bhatia, Sumit, Mutharaju, Raghava
Ontologies are used in various domains, with RDF and OWL being prominent standards for ontology development. RDF is favored for its simplicity and flexibility, while OWL enables detailed domain knowledge representation. However, as ontologies grow larger and more expressive, reasoning complexity increases, and traditional reasoners struggle to perform efficiently. Despite optimization efforts, scalability remains an issue. Additionally, advancements in automated knowledge base construction have created large and expressive ontologies that are often noisy and inconsistent, posing further challenges for conventional reasoners. To address these challenges, researchers have explored neuro-symbolic approaches that combine neural networks' learning capabilities with symbolic systems' reasoning abilities. In this chapter,we provide an overview of the existing literature in the field of neuro-symbolic deductive reasoning supported by RDF(S), the description logics EL and ALC, and OWL 2 RL, discussing the techniques employed, the tasks they address, and other relevant efforts in this area.
Feature Matching Data Synthesis for Non-IID Federated Learning
Li, Zijian, Sun, Yuchang, Shao, Jiawei, Mao, Yuyi, Wang, Jessie Hui, Zhang, Jun
Federated learning (FL) has emerged as a privacy-preserving paradigm that trains neural networks on edge devices without collecting data at a central server. However, FL encounters an inherent challenge in dealing with non-independent and identically distributed (non-IID) data among devices. To address this challenge, this paper proposes a hard feature matching data synthesis (HFMDS) method to share auxiliary data besides local models. Specifically, synthetic data are generated by learning the essential class-relevant features of real samples and discarding the redundant features, which helps to effectively tackle the non-IID issue. For better privacy preservation, we propose a hard feature augmentation method to transfer real features towards the decision boundary, with which the synthetic data not only improve the model generalization but also erase the information of real features. By integrating the proposed HFMDS method with FL, we present a novel FL framework with data augmentation to relieve data heterogeneity. The theoretical analysis highlights the effectiveness of our proposed data synthesis method in solving the non-IID challenge. Simulation results further demonstrate that our proposed HFMDS-FL algorithm outperforms the baselines in terms of accuracy, privacy preservation, and computational cost on various benchmark datasets.
How Does the Inner Geometry of Soft Actuators Modulate the Dynamic and Hysteretic Response?
Libby, Jacqueline, Somwanshi, Aniket A., Stancati, Federico, Tyagi, Gayatri, Mehrdad, Sarmad, Rizzo, JohnRoss, Atashzar, S. Farokh
This paper investigates the influence of the internal geometrical structure of soft pneu-nets on the dynamic response and hysteresis of the actuators. The research findings indicate that by strategically manipulating the stress distribution within soft robots, it is possible to enhance the dynamic response while reducing hysteresis. The study utilizes the Finite Element Method (FEM) and includes experimental validation through markerless motion tracking of the soft robot. In particular, the study examines actuator bending angles up to 500% strain while achieving 95% accuracy in predicting the bending angle. The results demonstrate that the particular design with the minimum air chamber width in the center significantly improves both high- and low-frequency hysteresis behavior by 21.5% while also enhancing dynamic response by 60% to 112% across various frequencies and peak-to-peak pressures. Consequently, the paper evaluates the effectiveness of "mechanically programming" stress distribution and distributed energy storage within soft robots to maximize their dynamic performance, offering direct benefits for control.
A Survey on Graph Neural Networks for Time Series: Forecasting, Classification, Imputation, and Anomaly Detection
Jin, Ming, Koh, Huan Yee, Wen, Qingsong, Zambon, Daniele, Alippi, Cesare, Webb, Geoffrey I., King, Irwin, Pan, Shirui
Time series are the primary data type used to record dynamic system measurements and generated in great volume by both physical sensors and online processes (virtual sensors). Time series analytics is therefore crucial to unlocking the wealth of information implicit in available data. With the recent advancements in graph neural networks (GNNs), there has been a surge in GNN-based approaches for time series analysis. These approaches can explicitly model inter-temporal and inter-variable relationships, which traditional and other deep neural network-based methods struggle to do. In this survey, we provide a comprehensive review of graph neural networks for time series analysis (GNN4TS), encompassing four fundamental dimensions: forecasting, classification, anomaly detection, and imputation. Our aim is to guide designers and practitioners to understand, build applications, and advance research of GNN4TS. At first, we provide a comprehensive task-oriented taxonomy of GNN4TS. Then, we present and discuss representative research works and introduce mainstream applications of GNN4TS. A comprehensive discussion of potential future research directions completes the survey. This survey, for the first time, brings together a vast array of knowledge on GNN-based time series research, highlighting foundations, practical applications, and opportunities of graph neural networks for time series analysis.
Google says AI systems should be able to mine publishers' work unless companies opt out
Publishers should be able to opt out of having their works mined by generative artificial intelligence systems, according to Google, but the company has not said how such a system would work. The call for a fair use exception for AI systems is a view the company has expressed to the Australian government in the past, but the notion of an opt-out option for publishers is a new argument from Google. When asked how such a system would work, a spokesperson pointed to a recent blog post by Google where the company said it wanted a discussion around creating a community-developed web standard similar to the robots.txt Google's comments come as news companies such as News Corp have already reportedly been initiating conversations with AI companies about payment for scraping news articles. Toby Murray, associate professor at the University of Melbourne's computing and information systems school, said Google's proposal would put the onus on content creators to specify whether AI systems could absorb their content or not, but he indicated existing licensing schemes such as Creative Commons already allowed creators to mark how their works can be used.
Spotify's new AI 'DJ' expands to 50 countries
The beta version of Spotify's AI-enhanced DJ feature is coming to 50 new countries, after soft-launching in the US and Canada back in February. In recent months, it's rolled out in the UK and Ireland, but now the robotic Wolfman Jack is headed to more countries in Europe, Asia and Africa, in addition to Australia and New Zealand. There's a caveat, but it depends on some initial understanding of what this tool actually does. The Spotify DJ is available to premium subscription members and provides algorithmic recommendations of what to listen to, just like any music streaming app. However, these recommendations are accompanied by AI-generated DJ commentary on what you're listening to. The DJ, based on Spotify's Xavier Jernigan, only speaks English, no matter where you live.
Path Signatures for Diversity in Probabilistic Trajectory Optimisation
Barcelos, Lucas, Lai, Tin, Oliveira, Rafael, Borges, Paulo, Ramos, Fabio
Abstract-- Motion planning can be cast as a trajectory optimisation problem where a cost is minimised as a function of the trajectory being generated. In complex environments with several obstacles and complicated geometry, this optimisation problem is usually difficult to solve and prone to local minima. However, recent advancements in computing hardware allow for parallel trajectory optimisation where multiple solutions are obtained simultaneously, each initialised from a different starting point. Unfortunately, without a strategy preventing two solutions to collapse on each other, naive parallel optimisation can suffer from mode collapse diminishing the efficiency of the approach and the likelihood of finding a global solution. In this paper we leverage on recent advances in the theory of rough paths to devise an algorithm for parallel trajectory optimisation that promotes diversity over the range of solutions, therefore avoiding mode collapses and achieving better global properties. These can be roughly divided into two main paradigms: sampling-based and trajectory optimisation algorithms. Sampling-based planning [2] is a class of planners with Trajectory optimisation is one of the key tools in robotic probabilistically complete and asymptotically optimal guarantees motion, used to find control signals or paths in obstaclecluttered [3]. These approaches decompose the planning problem environments that allow the robot to perform into a series of sequential decision-making problems with desired tasks. These trajectories can represent a variety of a tree-based [4] or graph-based [5], [6] approach.
Why Data Science Projects Fail
Data Science is a modern Data Intelligence practice, which is the core of many businesses and helps businesses build smart strategies around to deal with businesses challenges more efficiently. Data Science practice also helps in automating business processes using the algorithm, and it has several other benefits, which also deliver in a non-profitable framework. In regards to data science, three key components primarily influence the effective outcome of a data science project. Those are 1.Availability of Data 2.Algorithm 3.Processing power or infrastructure