Africa
He got Facebook hooked on AI. Now he can't fix its misinformation addiction
The Cambridge Analytica scandal would kick off Facebook's largest publicity crisis ever. It compounded fears that the algorithms that determine what people see on the platform were amplifying fake news and hate speech, and that Russian hackers had weaponized them to try to sway the election in Trump's favor. Millions began deleting the app; employees left in protest; the company's market capitalization plunged by more than $100 billion after its July earnings call. In the ensuing months, Mark Zuckerberg began his own apologizing. He apologized for not taking "a broad enough view" of Facebook's responsibilities, and for his mistakes as a CEO. Internally, Sheryl Sandberg, the chief operating officer, kicked off a two-year civil rights audit to recommend ways the company could prevent the use of its platform to undermine democracy. Finally, Mike Schroepfer, Facebook's chief technology officer, asked Quiñonero to start a team with a directive that was a little vague: to examine the societal impact of the company's algorithms. The group named itself the Society and AI Lab (SAIL); last year it combined with another team working on issues of data privacy to form Responsible AI. Quiñonero was a natural pick for the job. He, as much as anybody, was the one responsible for Facebook's position as an AI powerhouse. In his six years at Facebook, he'd created some of the first algorithms for targeting users with content precisely tailored to their interests, and then he'd diffused those algorithms across the company. Now his mandate would be to make them less harmful. Facebook has consistently pointed to the efforts by Quiñonero and others as it seeks to repair its reputation. It regularly trots out various leaders to speak to the media about the ongoing reforms.
Automated Fact-Checking for Assisting Human Fact-Checkers
Nakov, Preslav, Corney, David, Hasanain, Maram, Alam, Firoj, Elsayed, Tamer, Barrón-Cedeño, Alberto, Papotti, Paolo, Shaar, Shaden, Martino, Giovanni Da San
The reporting and analysis of current events around the globe has expanded from professional, editor-lead journalism all the way to citizen journalism. Politicians and other key players enjoy direct access to their audiences through social media, bypassing the filters of official cables or traditional media. However, the multiple advantages of free speech and direct communication are dimmed by the misuse of the media to spread inaccurate or misleading claims. These phenomena have led to the modern incarnation of the fact-checker -- a professional whose main aim is to examine claims using available evidence to assess their veracity. As in other text forensics tasks, the amount of information available makes the work of the fact-checker more difficult. With this in mind, starting from the perspective of the professional fact-checker, we survey the available intelligent technologies that can support the human expert in the different steps of her fact-checking endeavor. These include identifying claims worth fact-checking; detecting relevant previously fact-checked claims; retrieving relevant evidence to fact-check a claim; and actually verifying a claim. In each case, we pay attention to the challenges in future work and the potential impact on real-world fact-checking.
Image Segmentation Methods for Non-destructive testing Applications
Guerrout, EL-Hachemi, Mahiou, Ramdane, Boukabene, Randa, Ouali, Assia
In this paper, we present new image segmentation methods based on hidden Markov random fields (HMRFs) and cuckoo search (CS) variants. HMRFs model the segmentation problem as a minimization of an energy function. CS algorithm is one of the recent powerful optimization techniques. Therefore, five variants of the CS algorithm are used to compute a solution. Through tests, we conduct a study to choose the CS variant with parameters that give good results (execution time and quality of segmentation). CS variants are evaluated and compared with non-destructive testing (NDT) images using a misclassification error (ME) criterion.
Spectral Temporal Graph Neural Network for Multivariate Time-series Forecasting
Cao, Defu, Wang, Yujing, Duan, Juanyong, Zhang, Ce, Zhu, Xia, Huang, Conguri, Tong, Yunhai, Xu, Bixiong, Bai, Jing, Tong, Jie, Zhang, Qi
Multivariate time-series forecasting plays a crucial role in many real-world applications. It is a challenging problem as one needs to consider both intra-series temporal correlations and inter-series correlations simultaneously. Recently, there have been multiple works trying to capture both correlations, but most, if not all of them only capture temporal correlations in the time domain and resort to pre-defined priors as inter-series relationships. In this paper, we propose Spectral Temporal Graph Neural Network (StemGNN) to further improve the accuracy of multivariate time-series forecasting. StemGNN captures inter-series correlations and temporal dependencies \textit{jointly} in the \textit{spectral domain}. It combines Graph Fourier Transform (GFT) which models inter-series correlations and Discrete Fourier Transform (DFT) which models temporal dependencies in an end-to-end framework. After passing through GFT and DFT, the spectral representations hold clear patterns and can be predicted effectively by convolution and sequential learning modules. Moreover, StemGNN learns inter-series correlations automatically from the data without using pre-defined priors. We conduct extensive experiments on ten real-world datasets to demonstrate the effectiveness of StemGNN. Code is available at https://github.com/microsoft/StemGNN/
Biggest influencers in future cities in Q4 2020: The top individuals to follow
GlobalData research has found the top influencers in future cities based on their performance online and on social media.Using research from GlobalData's Influencer platform, Verdict has named ten of the most influential people and companies in digital construction on Twitter during Q4 2020. Ronald Van Loon is a principal analyst and CEO of the Intelligent World, an influencer network connecting businesses and experts with new tech, artificial intelligence (AI), analytics, and data enthusiasts. He is a recognised thought leader in technologies such as AI, the internet of things (IoT), machine learning, and 5G, among others. Loon is an advisory board member at Simplilearn, an education management company and has also served as director of Advertisement, an information technology and services company. Glen Gilmore is the founding faculty for digital marketing programmes at the Rutgers University School of Business.
Probabilistic Surface Friction Estimation Based on Visual and Haptic Measurements
Le, Tran Nguyen, Verdoja, Francesco, Abu-Dakka, Fares J., Kyrki, Ville
Accurately modeling local surface properties of objects is crucial to many robotic applications, from grasping to material recognition. Surface properties like friction are however difficult to estimate, as visual observation of the object does not convey enough information over these properties. In contrast, haptic exploration is time consuming as it only provides information relevant to the explored parts of the object. In this work, we propose a joint visuo-haptic object model that enables the estimation of surface friction coefficient over an entire object by exploiting the correlation of visual and haptic information, together with a limited haptic exploration by a robotic arm. We demonstrate the validity of the proposed method by showing its ability to estimate varying friction coefficients on a range of real multi-material objects. Furthermore, we illustrate how the estimated friction coefficients can improve grasping success rate by guiding a grasp planner toward high friction areas.
Conceptual capacity and effective complexity of neural networks
Szymanski, Lech, McCane, Brendan, Atkinson, Craig
We propose a complexity measure of a neural network mapping function based on the diversity of the set of tangent spaces from different inputs. Treating each tangent space as a linear PAC concept we use an entropy-based measure of the bundle of concepts in order to estimate the conceptual capacity of the network. The theoretical maximal capacity of a ReLU network is equivalent to the number of its neurons. In practice however, due to correlations between neuron activities within the network, the actual capacity can be remarkably small, even for very big networks. Empirical evaluations show that this new measure is correlated with the complexity of the mapping function and thus the generalisation capabilities of the corresponding network. It captures the effective, as oppose to the theoretical, complexity of the network function. We also showcase some uses of the proposed measure for analysis and comparison of trained neural network models.
Private Cross-Silo Federated Learning for Extracting Vaccine Adverse Event Mentions
Kanani, Pallika, Marathe, Virendra J., Peterson, Daniel, Harpaz, Rave, Bright, Steve
Federated Learning (FL) is quickly becoming a goto distributed training paradigm for users to jointly train a global model without physically sharing their data. Users can indirectly contribute to, and directly benefit from a much larger aggregate data corpus used to train the global model. However, literature on successful application of FL in real-world problem settings is somewhat sparse. In this paper, we describe our experience applying a FL based solution to the Named Entity Recognition (NER) task for an adverse event detection application in the context of mass scale vaccination programs. We present a comprehensive empirical analysis of various dimensions of benefits gained with FL based training. Furthermore, we investigate effects of tighter Differential Privacy (DP) constraints in highly sensitive settings where federation users must enforce Local DP to ensure strict privacy guarantees. We show that local DP can severely cripple the global model's prediction accuracy, thus dis-incentivizing users from participating in the federation. In response, we demonstrate how recent innovation on personalization methods can help significantly recover the lost accuracy. We focus our analysis on the Federated Fine-Tuning algorithm, FedFT, and prove that it is not PAC Identifiable, thus making it even more attractive for FL-based training.
Automating the GDPR Compliance Assessment for Cross-border Personal Data Transfers in Android Applications
Guamán, Danny S., Ferrer, Xavier, del Alamo, Jose M., Such, Jose
Abstract-- The General Data Protection Regulation (GDPR) aims to ensure that all personal data processing activities are fair and transparent for the European Union (EU) citizens, regardless of whether these are carried out within the EU or anywhere else. To this end, it sets strict requirements to transfer personal data outside the EU. However, checking these requirements is a daunting task for supervisory authorities, particularly in the mobile app domain due to the huge number of apps available and their dynamic nature. In this paper, we propose a fully automated method to assess compliance of mobile apps with the GDPR requirements for cross-border personal data transfers. We have applied the method to the top-free 10,080 apps from the Google Play Store. The results reveal that there is still a very significant gap between what app providers and third-party recipients do in practice and what is intended by the GDPR. A substantial 56% of analysed apps are potentially non-compliant with the GDPR cross-border transfer requirements. THE distributed nature of today's digital systems and services across the world [1], or shared between chains of thirdparty not only facilitates the collection of personal data service providers [6], even without the app developer's from individuals anywhere, but also their transfer to different knowledge [7]. Second, apps are distributed through countries around the world [1]. This raises potential global stores, enabling app providers to easily reach markets risks to the privacy of individuals, as the organizations and users beyond its country of residence. In this sending and receiving personal data can be subject to different context, there is a need for constant vigilance by the various data protection laws and, therefore, may not offer an stakeholders, including app developers, supervisory equivalent level of protection.
Evidence-Based Policy Learning
Spiess, Jann, Syrgkanis, Vasilis
The past years have seen seen the development and deployment of machine-learning algorithms to estimate personalized treatment-assignment policies from randomized controlled trials. Yet such algorithms for the assignment of treatment typically optimize expected outcomes without taking into account that treatment assignments are frequently subject to hypothesis testing. In this article, we explicitly take significance testing of the effect of treatment-assignment policies into account, and consider assignments that optimize the probability of finding a subset of individuals with a statistically significant positive treatment effect. We provide an efficient implementation using decision trees, and demonstrate its gain over selecting subsets based on positive (estimated) treatment effects. Compared to standard tree-based regression and classification tools, this approach tends to yield substantially higher power in detecting subgroups with positive treatment effects. INTRODUCTION Recent years have seen the development of machine-learning algorithms that estimate heterogeneous causal effects from randomized controlled trials. While the estimation of average effects - for example, how effective a vaccine is overall, whether a conditional cash transfer reduces poverty, or which ad leads to more clicks - can inform the decision whether to deploy a treatment or not, heterogeneous treatment effect estimation allows us to decide who should get treated. These algorithms aim to maximize realized outcomes, and thus focus on assigning treatment to individuals with positive (estimated) treatment effects. Yet in practice, the deployment of assignment policies often only happens after passing a test that the assignment produces a positive net effect relative to some status quo. For example, a drug manufacturer may have to demonstrate that the drug is effective on the target population by submitting a hypothesis test to the FDA for approval.