Africa
From Charts to Fair Narratives: Uncovering and Mitigating Geo-Economic Biases in Chart-to-Text
Mahbub, Ridwan, Islam, Mohammed Saidul, Nayeem, Mir Tafseer, Laskar, Md Tahmid Rahman, Rahman, Mizanur, Joty, Shafiq, Hoque, Enamul
Charts are very common for exploring data and communicating insights, but extracting key takeaways from charts and articulating them in natural language can be challenging. The chart-to-text task aims to automate this process by generating textual summaries of charts. While with the rapid advancement of large Vision-Language Models (VLMs), we have witnessed great progress in this domain, little to no attention has been given to potential biases in their outputs. This paper investigates how VLMs can amplify geo-economic biases when generating chart summaries, potentially causing societal harm. Specifically, we conduct a large-scale evaluation of geo-economic biases in VLM-generated chart summaries across 6,000 chart-country pairs from six widely used proprietary and open-source models to understand how a country's economic status influences the sentiment of generated summaries. Our analysis reveals that existing VLMs tend to produce more positive descriptions for high-income countries compared to middle- or low-income countries, even when country attribution is the only variable changed. We also find that models such as GPT-4o-mini, Gemini-1.5-Flash, and Phi-3.5 exhibit varying degrees of bias. We further explore inference-time prompt-based debiasing techniques using positive distractors but find them only partially effective, underscoring the complexity of the issue and the need for more robust debiasing strategies. Our code and dataset are publicly available here.
A deep learning and machine learning approach to predict neonatal death in the context of São Paulo
Raihan, Mohon, Saha, Plabon Kumar, Gupta, Rajan Das, Kabir, A Z M Tahmidul, Tamanna, Afia Anjum, Harun-Ur-Rashid, Md., Salam, Adnan Bin Abdus, Anjum, Md Tanvir, Kabir, A Z M Ahteshamul
Neonatal death is still a concerning reality for underdeveloped and even some developed countries. Worldwide data indicate that 26.693 babies out of 1,000 births die, according to Macro Trades. To reduce this number, early prediction of endangered babies is crucial. Such prediction enables the opportunity to take ample care of the child and mother so that early child death can be avoided. In this context, machine learning was used to determine whether a newborn baby is at risk. To train the predictive model, historical data of 1.4 million newborns was used. Machine learning and deep learning techniques such as logical regression, K-nearest neighbor, random forest classifier, extreme gradient boosting (XGBoost), convolutional neural network, and long short-term memory (LSTM) were implemented using the dataset to identify the most accurate model for predicting neonatal mortality. Among the machine learning algorithms, XGBoost and random forest classifier achieved the best accuracy with 94%, while among the deep learning models, LSTM delivered the highest accuracy with 99%. Therefore, using LSTM appears to be the most suitable approach to predict whether precautionary measures for a child are necessary.
Elon Musk and Sam Altman's AI Feud Gets Nasty
A long-running feud between Elon Musk and Sam Altman spilled out into the open this week as the AI billionaire heavyweights publicly fought over their rival companies. The latest round in the battle between the X CEO and the CEO of OpenAI began when Musk claimed that Apple had been favoring Altman's AI app over his own in the Apple Store rankings. "Apple is behaving in a manner that makes it impossible for any AI company besides OpenAI to reach #1 in the App Store, which is an unequivocal antitrust violation," Musk said on X on Monday evening. "xAI will take immediate legal action," he added, referring to the AI company he leads. "Hey @Apple App Store, why do you refuse to put either X or Grok in your'Must Have' section when X is the #1 news app in the world and Grok is #5 among all apps?" he asked.
Claire's on brink of collapse putting 2,150 jobs at risk
Claire's on brink of collapse putting 2,150 jobs at risk 15 minutes agoShareSaveTom EspinerBusiness reporter, BBC NewsShareSaveEPA Claire's will appoint administrators after struggles with online competition. Fashion accessories chain Claire's is on the brink of collapse after the retailer said it will appoint administrators in the UK and Ireland, putting 2,150 jobs at risk. The company has 278 stores in the UK and 28 in Ireland but has been struggling with falling sales and fierce competition. All the shops will continue trading while administrators at Interpath, once appointed, will "assess options for the company". Interpath chief executive Will Wright, said options include "exploring the possibility of a sale which would secure a future for this well-loved brand". Claire's in the US filed for bankruptcy in the US earlier this month.
Multilingual Diversity Improves Vision-Language Representations
Massive web-crawled image-text datasets lay the foundation for recent progress in multimodal learning. These datasets are designed with the goal of training a model to do well on standard computer vision benchmarks, many of which, however, have been shown to be English-centric (e.g., ImageNet). Consequently, existing data curation techniques gravitate towards using predominantly English image-text pairs and discard many potentially useful non-English samples. Multilingual data is inherently enriching not only because it provides a gateway to learn about culturally salient concepts, but also because it depicts common concepts differently from monolingual data. We thus conduct a systematic study to explore the performance benefits of using more samples of non-English origins with respect to English vision tasks.
Heartificial Intelligence: Exploring Empathy in Language Models
Williams, Victoria, Rosman, Benjamin
Large language models have become increasingly common, used by millions of people worldwide in both professional and personal contexts. As these models continue to advance, they are frequently serving as virtual assistants and companions. In human interactions, effective communication typically involves two types of empathy: cognitive empathy (understanding others' thoughts and emotions) and affective empathy (emotionally sharing others' feelings). In this study, we investigated both cognitive and affective empathy across several small (SLMs) and large (LLMs) language models using standardized psychological tests. Our results revealed that LLMs consistently outperformed humans - including psychology students - on cognitive empathy tasks. However, despite their cognitive strengths, both small and large language models showed significantly lower affective empathy compared to human participants. These findings highlight rapid advancements in language models' ability to simulate cognitive empathy, suggesting strong potential for providing effective virtual companionship and personalized emotional support. Additionally, their high cognitive yet lower affective empathy allows objective and consistent emotional support without running the risk of emotional fatigue or bias.
Grounding Multilingual Multimodal LLMs With Cultural Knowledge
Nyandwi, Jean de Dieu, Song, Yueqi, Khanuja, Simran, Neubig, Graham
Multimodal Large Language Models excel in high-resource settings, but often misinterpret long-tail cultural entities and underperform in low-resource languages. To address this gap, we propose a data-centric approach that directly grounds MLLMs in cultural knowledge. Leveraging a large scale knowledge graph from Wikidata, we collect images that represent culturally significant entities, and generate synthetic multilingual visual question answering data. The resulting dataset, CulturalGround, comprises 22 million high-quality, culturally-rich VQA pairs spanning 42 countries and 39 languages. We train an open-source MLLM CulturalPangea on CulturalGround, interleaving standard multilingual instruction-tuning data to preserve general abilities. CulturalPangea achieves state-of-the-art performance among open models on various culture-focused multilingual multimodal benchmarks, outperforming prior models by an average of 5.0 without degrading results on mainstream vision-language tasks. Our findings show that our targeted, culturally grounded approach could substantially narrow the cultural gap in MLLMs and offer a practical path towards globally inclusive multimodal systems.
Apple's AI Ambitions Leave Big Questions Over Its Climate Goals
Apple's AI Ambitions Leave Big Questions Over Its Climate Goals Halfway to its 2030 net-zero goal, Apple faces slow and hold-out suppliers, a tariffs scramble, and an AI race that could profoundly impact eco-friendly ambitions. Here's a simple question: Is the current top iPhone better for the environment than the top iPhone was five years ago? Let's take the iPhone Pro series. If we're looking at recycled and renewable materials, it's an easy yes. Compare the iPhone 11 Pro, released in September 2019, with the iPhone 16 Pro, released in September 2024, and there has been good progress--from a few smaller components and packaging to now at more than 25 percent of the whole phone.
Charges dropped against teen pilot detained in Antarctica
Charges against an American influencer and teen pilot who has been stranded on a remote island in the Antarctic since June have been dropped. Ethan Guo, 19, is alleged to have illegally landed his plane in Chilean territory after embarking on a solo trip to all seven continents to raise money for cancer research, according to local authorities. They accused him of providing false flight plan information to officials who detained him and opened an investigation. A judge has ordered him to leave the area, pay a $30,000 (£22,332) donation to a children's cancer foundation and is banned from re-entering Chilean territory for three years. Mr Guo made headlines last year when he began an attempt to become the youngest person to fly solo to all seven continents and collect donations for research into childhood cancer.