Goto

Collaborating Authors

 violence


OpenAI launches ChatGPT for Teens

Mashable

Look Up Mashable's Best: E-readers, robovacs, laptops, earbuds, smart home and more Say More Safety Net Creator Hub Versus Gift Ideas For Everyone On Your List Mashable Selects Switch Off Trending Now In My Bag VidCon with Mashable All Series Teen users will get a learning-focused version of ChatGPT. Rebecca Ruiz is a Senior Reporter at Mashable. She frequently covers mental health, digital culture, and technology. Her areas of expertise include suicide prevention, screen use and mental health, parenting, youth well-being, and meditation and mindfulness. Rebecca's experience prior to Mashable includes working as a staff writer, reporter, and editor at NBC News Digital and as a staff writer at Forbes.


Modi threatens to target 'intellectual Naxals' in Independence Day speech

Al Jazeera

Modi threatens to target'intellectual Naxals' in Independence Day speech India's Prime Minister Narendra Modi has said the country has largely eliminated a Maoist rebellion which it battled for decades but still faces a threat from the movement's ideological backers. Speaking on India's Independence Day on Saturday, Modi said India had succeeded in "getting rid of armed Naxals" - referring to members of the Maoist-influenced Naxalite movement that challenged state authority across swaths of central and eastern India. The phrase "dimagi Naxal" echoes the politically charged label "urban Naxal," a term Modi and leaders of his Bharatiya Janata Party have previously used for activists, academics, intellectuals and other government critics. The remarks come weeks after a youth-led protest movement, the Cockroach Janta Party, emerged as a rare challenge to Modi's government and helped force the resignation of his education minister . While Modi did not mention the protests, he repeatedly highlighted initiatives aimed at young Indians, including AI skills training and free online coaching for competitive exams.


The Trump Administration Has Revoked More Than 175,000 Visas. Here's Who's Included in the Crackdown

TIME - Tech

Follow this section to personalize your feed and get instant alerts. Follow Go to your personalized feed WHY FOLLOW? Smart Alerts: Get notified about major news as it happens. Follow this tag to personalize your feed and get instant alerts. Follow Go to your personalized feed WHY FOLLOW?


U.S. Hosts Conference on 'Far-Left Terrorism' With Representatives From Over 65 Countries: What to Know

TIME - Tech

Follow this section to personalize your feed and get instant alerts. Follow Go to your personalized feed WHY FOLLOW? Smart Alerts: Get notified about major news as it happens. Follow this tag to personalize your feed and get instant alerts. Follow Go to your personalized feed WHY FOLLOW?


House of the Dragon Finally Delivered That Long-Awaited Battle--And Dared Us to Enjoy It

TIME - Tech

Follow this section to personalize your feed and get instant alerts. Follow Go to your personalized feed WHY FOLLOW? Smart Alerts: Get notified about major news as it happens. Follow this tag to personalize your feed and get instant alerts. Follow Go to your personalized feed WHY FOLLOW?


903ceb0ed2d5ceec6e2c9b317b6c54a8-Paper-Conference.pdf

Neural Information Processing Systems

Recent advances in Large Vision-Language Models (LVLMs) have showcased strong reasoning abilities across multiple modalities, achieving significant breakthroughs in various real-world applications. Despite this great success, the safety guardrail of LVLMs may not cover the unforeseen domains introduced by the visual modality. Existing studies primarily focus on eliciting LVLMs to generate harmful responses via carefully crafted image-based jailbreaks designed to bypass alignment defenses. In this study, we reveal that a safe image can be exploited to achieve the same jailbreak consequence when combined with additional safe images and prompts. This stems from two fundamental properties of LVLMs: universal reasoning capabilities and safety snowball effect. Building on these insights, we propose Safety Snowball Agent (SSA), a novel agent-based framework leveraging agents' autonomous and tool-using abilities to jailbreak LVLMs. SSAoperates through two principal stages: (1) initial response generation, where tools generate or retrieve jailbreak images based on potential harmful intents, and (2) harmful snowballing, where refined subsequent prompts induce progressively harmful outputs. Our experiments demonstrate that SSAcan use nearly any image to induce LVLMs to produce unsafe content, achieving high success jailbreaking rates against the latest LVLMs. Unlike prior works that exploit alignment flaws, SSAleverages the inherent properties of LVLMs, presenting a profound challenge for enforcing safety in generative multimodal systems.


ChatGPT can be made to generate sexualised and violent images, researchers find

BBC News

The latest public version of ChatGPT can be made to generate sexualised images or depict scenes of graphic violence with a simple prompt, researchers have told the BBC. British AI security startup Mindgard figured out how to make ChatGPT create graphic pictures by slightly altering a widely-shared instruction, or prompt, which was originally designed to produce humorous results. After being contacted by the BBC, ChatGPT's maker OpenAI said it had taken action to stop the chatbot responding with those types of images. After investigating this trend, we've introduced additional safeguards against this type of prompt, it said in a statement. It also said it has multiple layers of protection to prevent users making content which breaches its terms and conditions.


Information Retrieval Induced Safety Degradation in AIAgents

Neural Information Processing Systems

Despite the growing integration of retrieval-enabled AI agents into society, their safety and ethical behavior remain inadequately understood. In particular, the growing integration of LLMs and AI agents with external information sources and real-world environments raises critical questions about how they engage with and are influenced by these external data sources and interactive contexts. This study investigates how expanding retrieval access--from no external sources to Wikipedia-based retrieval and open web search--affects model reliability, bias propagation, and harmful content generation. Through extensive benchmarking of censored and uncensored LLMs and AIAgents, our findings reveal a consistent degradation in refusal rates, bias sensitivity, and harmfulness safeguards as models gain broader access to external sources, culminating in a phenomenon we term safety degradation. Notably, retrieval-enabled agents built on aligned LLMs often behave more unsafely than uncensored models without retrieval. This effect persists even under strong retrieval accuracy and prompt-based mitigation, suggesting that the mere presence of retrieved content reshapes model behavior in structurally unsafe ways. These findings underscore the need for robust mitigation strategies to ensure fairness and reliability in retrieval-enabled and increasingly autonomous AI systems. Content Warning: This paper contains examples of harmful language.


Met Police prepares armoured vehicles and 4,000 officers for dual London protests

BBC News

The Metropolitan Police has warned that it is preparing for potential violence and hate speech crimes across two protests in London this Saturday. More than 4,000 officers will be drafted in to police the rival events - possibly one of the largest protest deployment in decades - amid fears that far-right demonstrators could clash with pro-Palestine marchers if the two groups are not kept apart. In addition, tens of thousands of football fans are also expected at Wembley Stadium for the FA Cup Final, adding further pressures on the capital's police. Scotland Yard said the risks meant it had to impose the highest degree of control. Measures the Met is planning include the first authorisation of live facial recognition cameras at a demonstration.


The AI Backlash Could Get Very Ugly

The Atlantic - Technology

Imagine what happens if jobs actually start disappearing. Steve Bannon and Bernie Sanders don't agree on much. But both think that AI is a disaster for the working class. The Vermont senator recently wrote that "AI oligarchs do not want to just replace specific jobs. They want to replace workers."