Technology
Was Israeli PM's Lebanon destruction video a snub to Trump?
Why is Israel still in southern Lebanon? A war to shape Lebanon's future Was Israeli PM's Lebanon destruction video a snub to Trump? NewsFeed Was Israeli PM's Lebanon destruction video a snub to Trump? Hours after US President Donald Trump asked Benjamin Netanyahu to stop destroying buildings in Lebanon as it "makes Israel look bad", the Israeli prime minister published a montage of forces blowing up infrastructure across southern Lebanon.
May skygazing: A blue moon, fading comet, and a lot of meteors
Two full moons in one month occurs about once every 2.5 years. More information Adding us as a Preferred Source in Google by using this link indicates that you would like to see more of our content in Google News results. The Eta Aquariids meteor shower appears in the night sky in Kandalama, Sri Lanka, on May 5, 2024. Unlike most major annual meteor showers, there is no sharp peak for this shower, but rather a broad maximum with good rates that last approximately one week. Breakthroughs, discoveries, and DIY tips sent six days a week.
Adaptive Experimentation When You Can't Experiment
This paper introduces the confounded pure exploration transductive linear bandit (CPET-LB) problem. As a motivating example, often online services cannot directly assign users to specific control or treatment experiences either for business or practical reasons. In these settings, naively comparing treatment and control groups that may result from self-selection can lead to biased estimates of underlying treatment effects. Instead, online services can employ a properly randomized encouragement that incentivizes users toward a specific treatment. Our methodology provides online services with an adaptive experimental design approach for learning the best-performing treatment for such encouragement designs. We consider a more general underlying model captured by a linear structural equation and formulate pure exploration linear bandits in this setting. Though pure exploration has been extensively studied in standard adaptive experimental design settings, we believe this is the first work considering a setting where noise is confounded. Elimination-style algorithms using experimental design methods in combination with a novel finite-time confidence interval on an instrumental variable style estimator are presented with sample complexity upper bounds nearly matching a minimax lower bound. Finally, experiments are conducted that demonstrate the efficacy of our approach.
Neural P 3 M: A Long-Range Interaction Modeling Enhancer for Geometric GNNs
Geometric graph neural networks (GNNs) have emerged as powerful tools for modeling molecular geometry. However, they encounter limitations in effectively capturing long-range interactions in large molecular systems. To address this challenge, we introduce **Neural P$^3$M**, a versatile enhancer of geometric GNNs to expand the scope of their capabilities by incorporating mesh points alongside atoms and reimaging traditional mathematical operations in a trainable manner. Neural P$^3$M exhibits flexibility across a wide range of molecular systems and demonstrates remarkable accuracy in predicting energies and forces, outperforming on benchmarks such as the MD22 dataset. It also achieves an average improvement of 22% on the OE62 dataset while integrating with various architectures.
The Download: the North Pole's future and humanoid data
Plus: Google, Microsoft, Amazon and Meta have all set AI spending records. In the past, getting to the North Pole involved a treacherous trip through ice many meters thick. But last year, a research vessel encountered open water and thin ice, which created an easy passage. It provided a reminder of how quickly the Arctic is changing. Now scientists are digging deep below the seabed to find out if the Arctic Ocean was ever ice-free--and what that could mean for the future of Earth's northernmost waters. Here's what they hope to discover .
SPIQA: A Dataset for Multimodal Question Answering on Scientific Papers
Seeking answers to questions within long scientific research articles is a crucial area of study that aids readers in quickly addressing their inquiries. However, existing question-answering (QA) datasets based on scientific papers are limited in scale and focus solely on textual content. We introduce SPIQA (Scientific Paper Image Question Answering), the first large-scale QA dataset specifically designed to interpret complex figures and tables within the context of scientific research articles across various domains of computer science. Leveraging the breadth of expertise and ability of multimodal large language models (MLLMs) to understand figures, we employ automatic and manual curation to create the dataset. We craft an information-seeking task on interleaved images and text that involves multiple images covering a wide variety of plots, charts, tables, schematic diagrams, and result visualizations. SPIQA comprises 270K questions divided into training, validation, and three different evaluation splits. Through extensive experiments with 12 prominent foundational models, we evaluate the ability of current multimodal systems to comprehend the nuanced aspects of research articles. Additionally, we propose a Chain-of-Thought (CoT) evaluation strategy with in-context retrieval that allows fine-grained, step-by-step assessment and improves model performance. We further explore the upper bounds of performance enhancement with additional textual information, highlighting its promising potential for future research and the dataset's impact on revolutionizing how we interact with scientific literature.
FLAME : Factuality-Aware Alignment for Large Language Models
Alignment is a procedure to fine-tune pre-trained large language models (LLMs) to follow natural language instructions and serve as helpful AI assistants. We have observed, however, that the conventional alignment process fails to enhance the factual accuracy of LLMs, and often leads to the generation of more false facts (i.e.,). In this paper, we study how to make the LLM alignment process more factual, by first identifying factors that lead to hallucination in both alignment steps: supervised fine-tuning (SFT) and reinforcement learning (RL).In particular, we find that training the LLM on new or unfamiliar knowledge can encourage hallucination.This makes SFT less factual as it trains on human-labeled data that may be novel to the LLM. Furthermore, reward functions used in standard RL often inadequately capture factuality and favor longer and more detailed responses, which inadvertently promote hallucination.Based on these observations, we propose, comprised of and through direct preference optimization. Experiments show that our proposed guides LLMs to output more factual responses while maintaining their instruction-following capability.
ChatGPT isn't a mind-reader. Use this prompt for better results
PCWorld explains how vague prompts produce poor results from AI tools like ChatGPT and Gemini, emphasizing the need for specific, detailed requests. The article introduces prompt decomposition, a technique that breaks complex tasks into key variables to create more effective AI prompts. This method helps users guide AI tools more precisely, resulting in higher-quality, less biased outputs for complex tasks. It's never a good idea to hand ChatGPT, Claude, or Gemini big, vague tasks like "draw up a business plan for my new venture" or "act as my personal assistant." Fuzzy prompts like those are sure to yield equally fuzzy results, allowing the AI to make decisions based on its training data and inherent biases, potentially leading you down a path you never intended.