AITopics | disdain

Collaborating Authors

disdain

Information about AI from the News, Publications, and Conferences

Automatic Classification – Tagging and Summarization – Customizable Filtering and Analysis

If you are looking for an answer to the question What is Artificial Intelligence? and you only have a minute, then here's the definition the Association for the Advancement of Artificial Intelligence offers on its home page: "the scientific understanding of the mechanisms underlying thought and intelligent behavior and their embodiment in machines."

However, if you are fortunate enough to have more than a minute, then please get ready to embark upon an exciting journey exploring AI (but beware, it could last a lifetime) …

Learning more skills through optimistic exploration

Strouse, DJ, Baumli, Kate, Warde-Farley, David, Mnih, Vlad, Hansen, Steven

arXiv.org Artificial IntelligenceJul-29-2021

Unsupervised skill learning objectives (Gregor et al., 2016, Eysenbach et al., 2018) allow agents to learn rich repertoires of behavior in the absence of extrinsic rewards. They work by simultaneously training a policy to produce distinguishable latent-conditioned trajectories, and a discriminator to evaluate distinguishability by trying to infer latents from trajectories. The hope is for the agent to explore and master the environment by encouraging each skill (latent) to reliably reach different states. However, an inherent exploration problem lingers: when a novel state is actually encountered, the discriminator will necessarily not have seen enough training data to produce accurate and confident skill classifications, leading to low intrinsic reward for the agent and effective penalization of the sort of exploration needed to actually maximize the objective. To combat this inherent pessimism towards exploration, we derive an information gain auxiliary objective that involves training an ensemble of discriminators and rewarding the policy for their disagreement. Our objective directly estimates the epistemic uncertainty that comes from the discriminator not having seen enough training examples, thus providing an intrinsic reward more tailored to the true objective compared to pseudocount-based methods (Burda et al., 2019). We call this exploration bonus discriminator disagreement intrinsic reward, or DISDAIN. We demonstrate empirically that DISDAIN improves skill learning both in a tabular grid world (Four Rooms) and the 57 games of the Atari Suite (from pixels). Thus, we encourage researchers to treat pessimism with DISDAIN.

discriminator, disdain, exploration, (13 more...)

arXiv.org Artificial Intelligence

2107.14226

Genre: Research Report (0.82)

Industry: Leisure & Entertainment > Games (0.69)

Technology:

Information Technology > Artificial Intelligence > Machine Learning > Reinforcement Learning (0.94)
Information Technology > Artificial Intelligence > Machine Learning > Inductive Learning (0.68)

Add feedback

Microsoft acquires AI company to make Cortana and bots sound more human

#artificialintelligenceMay-30-2018, 15:37:05 GMT

Microsoft is acquiring conversational AI startup Semantic Machines in an effort to make bots and intelligent assistants like Cortana sound and respond more like humans. Founded in 2014, Semantic Machines uses machine learning to make bots respond in a more natural way to queries. Semantic Machines is led by UC Berkeley professor Dan Klein and former Apple chief speech scientist Larry Gillick. Both are considered pioneers in conversational AI. Microsoft's acquisition will boost the company's Cortana digital assistant, as well as the company's Azure Bot Service that's used by 300,000 developers.

artificial intelligence, chatbot, natural language, (15 more...)

#artificialintelligence

Technology:

Information Technology > Artificial Intelligence > Representation & Reasoning > Personal Assistant Systems (1.00)
Information Technology > Artificial Intelligence > Natural Language > Chatbot (1.00)

Add feedback