Education
Senior Manager, Personalization Machine Learning Science at Wayfair Inc. - Mountain View, CA
Our algorithms tackle a varied & broad spectrum of challenges in the Wayfair marketplace; from empowering suppliers to easily add products to our catalog, to enabling our customers to discover and purchase a vast & diverse assortment of home goods. The Search & Recommendations Machine Learning Science team is looking for an experienced Machine Learning Senior Manager to lead one of Wayfair's recommendation science teams. In this role, you'll be part of the Search & Recommendations leadership. You'll lead a team of 5 to 10 machine learning scientists and machine learning engineers, leveraging customer behavior and Wayfair's content and product data to create a tailored customer experience. We are looking for a candidate who would lead the team to build and implement personalized product recommendations, customer understanding, and personalized content targeting strategies using cutting-edge machine learning and deep learning techniques.
Musk calls for action as AI tech grows stronger, media's smears against Christianity and more top headlines
Subscribe now to get Fox News First in your email. And here's what you need to know to start your day ... 'DANGEROUS RACE' - Elon Musk, futurists call for major action as AI technology grows stronger, smarter. Continue reading โฆ 'HORRENDOUS' COVERAGE - Nashville massacre coverage marked by media's'subtle smears' against Christianity. RED ALERT - China makes serious threats over meeting between Speaker McCarthy, Taiwan's leader. 'PROTECTED HER CHILDREN' - Head of school praised for going'straight for the shooter' during school massacre.
We shouldn't fear ChatGPT in education -- we need to work with it
While ChatGPT's primary purpose is to assist users in generating human-like text, it has also made significant contributions to the field of philosophy. This, in turn, has had a significant impact on educational assessment, enabling educators to evaluate students' critical thinking skills in new and exciting ways. The doomsayers, by contrast, are sceptical. Here at UCC, for example, I've heard more than a few colleagues echoing Socrates, who, in Plato's dialogue Phaedrus (370 BC), expresses similar worries about the invention of โฆ wait for it โฆ writing. "This invention will produce forgetfulness in the minds of those who learn to use it ... you offer your pupils the appearance of wisdom, not true wisdom, for they will read many things without instruction and will therefore seem to know many things, when they are for the most part ignorant and hard to get along with since they are not wise, but only appear wise."
Questions of science: chatting with ChatGPT about complex systems
Crokidakis, Nuno, de Menezes, Marcio Argollo, Cajueiro, Daniel O.
We are currently in a great era for researchers and scientists studying and developing in the field of complex systems. Half of the physics Nobel prize of 2021 was awarded to the physicist Giorgio Parisi for his contributions to the theory of complex systems [9] and the other half to two meteorologists Syukuro Manabe and Klaus Hasselmann to the modeling of the Earth's climate [10]. Parisi has made significant contributions to the literature on complex systems, including areas such as spin glass [11, 12, 13], stochastic resonance [14], surface growth [15], multifractality [16], and bird flocking [17].
Text revision in Scientific Writing Assistance: An Overview
Jourdan, Lรฉane, Boudin, Florian, Dufour, Richard, Hernandez, Nicolas
Writing a scientific article is a challenging task as it is a highly codified genre. Good writing skills are essential to properly convey ideas and results of research work. Since the majority of scientific articles are currently written in English, this exercise is all the more difficult for non-native English speakers as they additionally have to face language issues. This article aims to provide an overview of text revision in writing assistance in the scientific domain. We will examine the specificities of scientific writing, including the format and conventions commonly used in research articles. Additionally, this overview will explore the various types of writing assistance tools available for text revision. Despite the evolution of the technology behind these tools through the years, from rule-based approaches to deep neural-based ones, challenges still exist (tools' accessibility, limited consideration of the context, inexplicit use of discursive information, etc.)
DERA: Enhancing Large Language Model Completions with Dialog-Enabled Resolving Agents
Nair, Varun, Schumacher, Elliot, Tso, Geoffrey, Kannan, Anitha
Large language models (LLMs) have emerged as valuable tools for many natural language understanding tasks. In safety-critical applications such as healthcare, the utility of these models is governed by their ability to generate outputs that are factually accurate and complete. In this work, we present dialog-enabled resolving agents (DERA). DERA is a paradigm made possible by the increased conversational abilities of LLMs, namely GPT-4. It provides a simple, interpretable forum for models to communicate feedback and iteratively improve output. We frame our dialog as a discussion between two agent types - a Researcher, who processes information and identifies crucial problem components, and a Decider, who has the autonomy to integrate the Researcher's information and makes judgments on the final output. We test DERA against three clinically-focused tasks. For medical conversation summarization and care plan generation, DERA shows significant improvement over the base GPT-4 performance in both human expert preference evaluations and quantitative metrics. In a new finding, we also show that GPT-4's performance (70%) on an open-ended version of the MedQA question-answering (QA) dataset (Jin et al. 2021, USMLE) is well above the passing level (60%), with DERA showing similar performance. We release the open-ended MEDQA dataset at https://github.com/curai/curai-research/tree/main/DERA.
Evaluating GPT-3.5 and GPT-4 Models on Brazilian University Admission Exams
Nunes, Desnes, Primi, Ricardo, Pires, Ramon, Lotufo, Roberto, Nogueira, Rodrigo
The present study aims to explore the capabilities of Language Models (LMs) in tackling high-stakes multiple-choice tests, represented here by the Exame Nacional do Ensino M\'edio (ENEM), a multidisciplinary entrance examination widely adopted by Brazilian universities. This exam poses challenging tasks for LMs, since its questions may span into multiple fields of knowledge, requiring understanding of information from diverse domains. For instance, a question may require comprehension of both statistics and biology to be solved. This work analyzed responses generated by GPT-3.5 and GPT-4 models for questions presented in the 2009-2017 exams, as well as for questions of the 2022 exam, which were made public after the training of the models was completed. Furthermore, different prompt strategies were tested, including the use of Chain-of-Thought (CoT) prompts to generate explanations for answers. On the 2022 edition, the best-performing model, GPT-4 with CoT, achieved an accuracy of 87%, largely surpassing GPT-3.5 by 11 points. The code and data used on experiments are available at https://github.com/piresramon/gpt-4-enem.
Student-centric Model of Learning Management System Activity and Academic Performance: from Correlation to Causation
Mandalapu, Varun, Chen, Lujie Karen, Shetty, Sushruta, Chen, Zhiyuan, Gong, Jiaqi
In recent years, there is a lot of interest in modeling students' digital traces in Learning Management System (LMS) to understand students' learning behavior patterns including aspects of meta-cognition and self-regulation, with the ultimate goal to turn those insights into actionable information to support students to improve their learning outcomes. In achieving this goal, however, there are two main issues that need to be addressed given the existing literature. Firstly, most of the current work is course-centered (i.e. models are built from data for a specific course) rather than student-centered; secondly, a vast majority of the models are correlational rather than causal. Those issues make it challenging to identify the most promising actionable factors for intervention at the student level where most of the campus-wide academic support is designed for. In this paper, we explored a student-centric analytical framework for LMS activity data that can provide not only correlational but causal insights mined from observational data. We demonstrated this approach using a dataset of 1651 computing major students at a public university in the US during one semester in the Fall of 2019. This dataset includes students' fine-grained LMS interaction logs and administrative data, e.g. demographics and academic performance. In addition, we expand the repository of LMS behavior indicators to include those that can characterize the time-of-the-day of login (e.g. chronotype). Our analysis showed that student login volume, compared with other login behavior indicators, is both strongly correlated and causally linked to student academic performance, especially among students with low academic performance. We envision that those insights will provide convincing evidence for college student support groups to launch student-centered and targeted interventions that are effective and scalable.
Training Feedforward Neural Networks with Bayesian Hyper-Heuristics
Schreuder, Arnรฉ, Bosman, Anna, Engelbrecht, Andries, Cleghorn, Christopher
The process of training feedforward neural networks (FFNNs) can benefit from an automated process where the best heuristic to train the network is sought out automatically by means of a high-level probabilistic-based heuristic. This research introduces a novel population-based Bayesian hyper-heuristic (BHH) that is used to train feedforward neural networks (FFNNs). The performance of the BHH is compared to that of ten popular low-level heuristics, each with different search behaviours. The chosen heuristic pool consists of classic gradient-based heuristics as well as meta-heuristics (MHs). The empirical process is executed on fourteen datasets consisting of classification and regression problems with varying characteristics. The BHH is shown to be able to train FFNNs well and provide an automated method for finding the best heuristic to train the FFNNs at various stages of the training process.
TempCLR: Temporal Alignment Representation with Contrastive Learning
Yang, Yuncong, Ma, Jiawei, Huang, Shiyuan, Chen, Long, Lin, Xudong, Han, Guangxing, Chang, Shih-Fu
Video representation learning has been successful in video-text pre-training for zero-shot transfer, where each sentence is trained to be close to the paired video clips in a common feature space. For long videos, given a paragraph of description where the sentences describe different segments of the video, by matching all sentence-clip pairs, the paragraph and the full video are aligned implicitly. However, such unit-level comparison may ignore global temporal context, which inevitably limits the generalization ability. In this paper, we propose a contrastive learning framework TempCLR to compare the full video and the paragraph explicitly. As the video/paragraph is formulated as a sequence of clips/sentences, under the constraint of their temporal order, we use dynamic time warping to compute the minimum cumulative cost over sentence-clip pairs as the sequence-level distance. To explore the temporal dynamics, we break the consistency of temporal succession by shuffling video clips w.r.t. temporal granularity. Then, we obtain the representations for clips/sentences, which perceive the temporal information and thus facilitate the sequence alignment. In addition to pre-training on the video and paragraph, our approach can also generalize on the matching between video instances. We evaluate our approach on video retrieval, action step localization, and few-shot action recognition, and achieve consistent performance gain over all three tasks. Detailed ablation studies are provided to justify the approach design.