Goto

Collaborating Authors

 Asia


Improved Neural Machine Translation with SMT Features

AAAI Conferences

Neural machine translation (NMT) conducts end-to-end translation with a source language encoder and a target language decoder, making promising translation performance. However, as a newly emerged approach, the method has some limitations. An NMT system usually has to apply a vocabulary of certain size to avoid the time-consuming training and decoding, thus it causes a serious out-of-vocabulary problem. Furthermore, the decoder lacks a mechanism to guarantee all the source words to be translated and usually favors short translations, resulting in fluent but inadequate translations. In order to solve the above problems, we incorporate statistical machine translation (SMT) features, such as a translation model and an n-gram language model, with the NMT model under the log-linear framework. Our experiments show that the proposed method significantly improves the translation quality of the state-ofthe-art NMT system on Chinese-to-English translation tasks. Our method produces a gain of up to 2.33 BLEU score on NIST open test sets.


Inferring a Personalized Next Point-of-Interest Recommendation Model with Latent Behavior Patterns

AAAI Conferences

In this paper, we address the problem of personalized next Point-of-interest (POI) recommendation which has become an important and very challenging task in location-based social networks (LBSNs), but not well studied yet. With the conjecture that, under different contextual scenario, human exhibits distinct mobility patterns, we attempt here to jointly model the next POI recommendation under the influence of user's latent behavior pattern. We propose to adopt a third-rank tensor to model the successive check-in behaviors. By incorporating softmax function to fuse the personalized Markov chain with latent pattern, we furnish a Bayesian Personalized Ranking (BPR) approach and derive the optimization criterion accordingly. Expectation Maximization (EM) is then used to estimate the model parameters. Extensive experiments on two large-scale LBSNs datasets demonstrate the significant improvements of our model over several state-of-the-art methods.


Community-Based Question Answering via Heterogeneous Social Network Learning

AAAI Conferences

Community-based question answering (cQA) sites have accumulated vast amount of questions and corresponding crowdsourced answers over time. How to efficiently share the underlying information and knowledge from reliable (usually highly-reputable) answerers has become an increasingly popular research topic. A major challenge in cQA tasks is the accurate matching of high-quality answers w.r.t given questions. Many of traditional approaches likely recommend corresponding answers merely depending on the content similarity between questions and answers, therefore suffer from the sparsity bottleneck of cQA data. In this paper, we propose a novel framework which encodes not only the contents of question-answer(Q-A) but also the social interaction cues in the community to boost the cQA tasks. More specifically, our framework collaboratively utilizes the rich interaction among questions, answers and answerers to learn the relative quality rank of different answers w.r.t a same question. Moreover, the information in heterogeneous social networks is comprehensively employed to enhance the quality of question-answering (QA) matching by our deep random walk learning framework. Extensive experiments on a large-scale dataset from a real world cQA site show that leveraging the heterogeneous social information indeed achieves better performance than other state-of-the-art cQA methods.


From Tweets to Wellness: Wellness Event Detection from Twitter Streams

AAAI Conferences

Social media platforms have become the most popular means for users to share what is happening around them. The abundance and growing usage of social media has resulted in a large repository of users' social posts, which provides a stethoscope for inferring individuals' lifestyle and wellness. As users' social accounts implicitly reflect their habits, preferences, and feelings, it is feasible for us to monitor and understand the wellness of users by harvesting social media data towards a healthier lifestyle. As a first step towards accomplishing this goal, we propose to automatically extract wellness events from users' published social contents. Existing approaches for event extraction are not applicable to personal wellness events due to its domain nature characterized by plenty of noise and variety in data, insufficient samples, and inter-relation among events.To tackle these problems, we propose an optimization learning framework that utilizes the content information of microblogging messages as well as the relations between event categories. By imposing a sparse constraint on the learning model, we also tackle the problems arising from noise and variation in microblogging texts. Experimental results on a real-world dataset from Twitter have demonstrated the superior performance of our framework.


Social Role-Aware Emotion Contagion in Image Social Networks

AAAI Conferences

Psychological theories suggest that emotion represents the state of mind and instinctive responses of oneโ€™s cognitive system (Cannon 1927). Emotions are a complex state of feeling that results in physical and psychological changes that influence our behavior. In this paper, we study an interesting problem of emotion contagion in social networks. In particular, by employing an image social network (Flickr) as the basis of our study, we try to unveil how usersโ€™ emotional statuses influence each other and how usersโ€™ positions in the social network affect their influential strength on emotion. We develop a probabilistic framework to formalize the problem into a role-aware contagion model. The model is able to predict usersโ€™ emotional statuses based on their historical emotional statuses and social structures. Experiments on a large Flickr dataset show that the proposed model significantly outperforms (+31% in terms of F1-score) several alternative methods in predicting usersโ€™ emotional status. We also discover several intriguing phenomena. For example, the probability that a user feels happy is roughly linear to the number of friends who are also happy; but taking a closer look, the happiness probability is superlinear to the number of happy friends who act as opinion leaders (Page et al. 1999) in the network and sublinear in the number of happy friends who span structural holes (Burt 2001). This offers a new opportunity to understand the underlying mechanism of emotional contagion in online social networks.


Little Is Much: Bridging Cross-Platform Behaviors through Overlapped Crowds

AAAI Conferences

People often use multiple platforms to fulfill their different information needs. With the ultimate goal of serving people intelligently, a fundamental way is to get comprehensive understanding about user needs. How to organically integrate and bridge cross-platform information in a human-centric way is important. Existing transfer learning assumes either fully-overlapped or non-overlapped among the users. However, the real case is the users of different platforms are partially overlapped. The number of overlapped users is often small and the explicitly known overlapped users is even less due to the lacking of unified ID for a user across different platforms. In this paper, we propose a novel semi-supervised transfer learning method to address the problem of cross-platform behavior prediction, called XPTrans. To alleviate the sparsity issue, it fully exploits the small number of overlapped crowds to optimally bridge a user's behaviors in different platforms. Extensive experiments across two real social networks show that XPTrans significantly outperforms the state-of-the-art. We demonstrate that by fully exploiting 26% overlapped users, XPTrans can predict the behaviors of non-overlapped users with the same accuracy as overlapped users, which means the small overlapped crowds can successfully bridge the information across different platforms.


Chinese scientists built a 'robot goddess', then made it subservient and insecure

#artificialintelligence

An ultra-realistic robot was unveiled last week by researchers from the University of Science and Technology in China (USTC). Jia Jia, as the female robot has been named, is apparently capable of basic communication, interaction with nearby people, and natural facial expressions. Unfortunately, many of her pre-programmed interactions appear to be highly stereotypical.


In Japan, an artificial intelligence has been appointed creative director Springwise

#artificialintelligence

Weird Of The Week: This is part of a series of articles that looks at some of the most bizarre and niche business ideas we see here at Springwise. Advertising and media are often at the forefront of new technology, and we have already seen augmented reality platforms showing content in the real world and a virtual reality advertising network for brands. Now an artificial intelligence robot, AI-CD?, developed by Japanese advertising and marketing agency McCann Japan, is set to work on providing new creative direction for commercials. The AI will give input on projects, mining and analyzing creative databases of adverts to find the best commercials for products and messages. But the robot is also being treated as somewhat part of the team at McCann, taking the title of "creative director" and attending the opening ceremony for new company employees.


Nvidia Puts The Accelerator To The Metal With Pascal

#artificialintelligence

The revolution in GPU computing started with games, and spread to the HPC centers of the world eight years ago with the first "Fermi" Tesla accelerators from Nvidia. But hyperscalers and their deep learning algorithms are driving the architecture of the "Pascal" GPUs and the Tesla accelerators that Nvidia unveiled today at the GPU Technical Conference in its hometown of San Jose. Not only did the hyperscalers and their AI efforts help drive the Pascal architecture, but they will be the first companies to get their hands on all of the Tesla P100 accelerators based on the Pascal GP100 GPU that Nvidia can manufacture, long before they become generally available in early 2017 through server partners who make hybrid CPU-GPU systems. As was the case with the prior generations of GPU compute engines, Nvidia will eventually offer multiple versions of the Pascal GPU for specific workloads and use cases, but Nvidia has made the big bet and created its high-end GP100 variant of Pascal and making other big bets at the same time, such as moving to a 16 nanometer FinFET process from chip fab partner Taiwan Semiconductor Manufacturing Corp and adding in High Bandwidth Memory from memory partner Samsung at the same time. Jen-Hsun Huang, co-founder and CEO at Nvidia, said during his opening keynote that Nvidia has a rule about how many big bets it can make.


Intel: Facing A Real Threat

#artificialintelligence

Shares of Intel (NASDAQ:INTC) have been trading along the 100-day moving average after bouncing off the low in February, as analysts raised concerns about PC and notebook sales during the first-quarter 2016. Intel will report its first-quarter 2016 earnings after the market close on April 19. Investors will be closely watching the results from the Client Computing Group, or CCG, Intel's mobile and PC business, and the Data Center Group, or DCG, as about 88.9% of their revenues last year came from these two groups. Less than a week ahead of the earnings report, Pacific Crest warned it expects Intel to report first-quarter earnings below the midpoint of guidance, and to lower its guidance for the full-year as well. Wall Street opinions about Intel's data center outlook are mixed, according to Barron's.