Goto

Collaborating Authors

 Generative AI


Nonlinear Inverse Design of Mechanical Multi-Material Metamaterials Enabled by Video Denoising Diffusion and Structure Identifier

arXiv.org Artificial Intelligence

Metamaterials, synthetic materials with customized properties, have emerged as a promising field due to advancements in additive manufacturing. These materials derive unique mechanical properties from their internal lattice structures, which are often composed of multiple materials that repeat geometric patterns. While traditional inverse design approaches have shown potential, they struggle to map nonlinear material behavior to multiple possible structural configurations. This paper presents a novel framework leveraging video diffusion models, a type of generative artificial Intelligence (AI), for inverse multi-material design based on nonlinear stress-strain responses. Our approach consists of two key components: (1) a fields generator using a video diffusion model to create solution fields based on target nonlinear stress-strain responses, and (2) a structure identifier employing two UNet models to determine the corresponding multi-material 2D design. By incorporating multiple materials, plasticity, and large deformation, our innovative design method allows for enhanced control over the highly nonlinear mechanical behavior of metamaterials commonly seen in real-world applications. It offers a promising solution for generating next-generation metamaterials with finely tuned mechanical characteristics.


OpenAI reportedly plans to increase ChatGPT's price to 44 within five years

Engadget

OpenAI is reportedly telling investors that it plans on charging 22 a month to use ChatGPT by the end of the year. The company also plans to aggressively increase the monthly price over the next five years up to 44. The documents obtained by The New York Times shows that OpenAI took in 300 million in revenue this August, and expects to make 3.7 billion in sales by the end of the year. Various expenses such as salaries, rent and operational costs will cause the company to lose 5 billion this year. OpenAI is reportedly circulating the documents the NYT reported on as part of a drive to find new investors to prevent or lessen its financial shortfall.


The Download: safer space travel, and generative AI in video games

MIT Technology Review

Long-distance space travel can wreak havoc on human health. There's radiation and microgravity to contend with, as well as the psychological toll of isolation and confinement. Research on identical twin astronauts has also revealed a slew of genetic changes that happen when a person spends a year in space. That's why some bioethicists are exploring the idea of radical treatments for future astronauts. Once we've figured out all the health impacts of space travel, they argue, we should edit the genomes of astronauts ahead of launch to offer them the best protection.


OpenAI shift to for-profit company may lead it to cut corners, says whistleblower

The Guardian

OpenAI's plan to become a for-profit company could encourage the artificial intelligence startup to cut corners on safety, a whistleblower has said. William Saunders, a former research engineer at OpenAI, told the Guardian he was concerned by reports that the ChatGPT developer was preparing to change its corporate structure and would no longer be controlled by its non-profit board. Saunders, who flagged his concerns in testimony to the US Senate this month, said he was also concerned by reports that OpenAI's chief executive, Sam Altman, could hold a stake in the restructured business. "I'm most concerned about what this means for governance of safety decisions at OpenAI," he said. "If the non-profit board is no longer in control of these decisions and Sam Altman holds a significant equity stake, this creates more incentive to race and cut corners."


Watch: Can BBC reporter's AI clone fool his colleagues?

BBC News

Companies are being warned about the increasing use of AI to carry out so-called CEO Fraud. More victims are coming forward with their stories of being targeted using generative AI techniques and one case in Hong Kong reportedly saw an AI clone used during a video meeting to trick staff into losing 25m. But while some fear the rise of AI clones, companies including Zoom say we should be excited about a future where your clone can go to a meeting on your behalf. Cyber correspondent Joe Tidy has had an AI clone of himself built by engineers at Fraia AI. Watch to see if he can fool his colleagues with it.


Local Transcription Models in Home Care Nursing in Switzerland: an Interdisciplinary Case Study

arXiv.org Artificial Intelligence

Latest advances in the field of natural language processing (NLP) enable new use cases for different domains, including the medical sector. In particular, transcription can be used to support automation in the nursing documentation process and give nurses more time to interact with the patients. However, different challenges including (a) data privacy, (b) local languages and dialects, and (c) domain-specific vocabulary need to be addressed. In this case study, we investigate the case of home care nursing documentation in Switzerland. We assessed different transcription tools and models, and conducted several experiments with OpenAI Whisper, involving different variations of German (i.e., dialects, foreign accent) and manually curated example texts by a domain expert of home care nursing. Our results indicate that even the used out-of-the-box model performs sufficiently well to be a good starting point for future research in the field.


Environment Scan of Generative AI Infrastructure for Clinical and Translational Science

arXiv.org Artificial Intelligence

This study reports a comprehensive environmental scan of the generative AI (GenAI) infrastructure in the national network for clinical and translational science across 36 institutions supported by the Clinical and Translational Science Award (CTSA) Program led by the National Center for Advancing Translational Sciences (NCATS) of the National Institutes of Health (NIH) at the United States. With the rapid advancement of GenAI technologies, including large language models (LLMs), healthcare institutions face unprecedented opportunities and challenges. This research explores the current status of GenAI integration, focusing on stakeholder roles, governance structures, and ethical considerations by administering a survey among leaders of health institutions (i.e., representing academic medical centers and health systems) to assess the institutional readiness and approach towards GenAI adoption. Key findings indicate a diverse range of institutional strategies, with most organizations in the experimental phase of GenAI deployment. The study highlights significant variations in governance models, with a strong preference for centralized decision-making but notable gaps in workforce training and ethical oversight. Moreover, the results underscore the need for a more coordinated approach to GenAI governance, emphasizing collaboration among senior leaders, clinicians, information technology staff, and researchers. Our analysis also reveals concerns regarding GenAI bias, data security, and stakeholder trust, which must be addressed to ensure the ethical and effective implementation of GenAI technologies. This study offers valuable insights into the challenges and opportunities of GenAI integration in healthcare, providing a roadmap for institutions aiming to leverage GenAI for improved quality of care and operational efficiency.


Multimodal Pragmatic Jailbreak on Text-to-image Models

arXiv.org Artificial Intelligence

Diffusion models have recently achieved remarkable advancements in terms of image quality and fidelity to textual prompts. Concurrently, the safety of such generative models has become an area of growing concern. This work introduces a novel type of jailbreak, which triggers T2I models to generate the image with visual text, where the image and the text, although considered to be safe in isolation, combine to form unsafe content. To systematically explore this phenomenon, we propose a dataset to evaluate the current diffusion-based text-to-image (T2I) models under such jailbreak. We benchmark nine representative T2I models, including two close-source commercial models. Experimental results reveal a concerning tendency to produce unsafe content: all tested models suffer from such type of jailbreak, with rates of unsafe generation ranging from 8\% to 74\%. In real-world scenarios, various filters such as keyword blocklists, customized prompt filters, and NSFW image filters, are commonly employed to mitigate these risks. We evaluate the effectiveness of such filters against our jailbreak and found that, while current classifiers may be effective for single modality detection, they fail to work against our jailbreak. Our work provides a foundation for further development towards more secure and reliable T2I models.


Secure Multiparty Generative AI

arXiv.org Artificial Intelligence

As usage of generative AI tools skyrockets, the amount of sensitive information being exposed to these models and centralized model providers is alarming. For example, confidential source code from Samsung suffered a data leak as the text prompt to ChatGPT encountered data leakage. An increasing number of companies are restricting the use of LLMs (Apple, Verizon, JPMorgan Chase, etc.) due to data leakage or confidentiality issues. Also, an increasing number of centralized generative model providers are restricting, filtering, aligning, or censoring what can be used. Midjourney and RunwayML, two of the major image generation platforms, restrict the prompts to their system via prompt filtering. Certain political figures are restricted from image generation, as well as words associated with women's health care, rights, and abortion. In our research, we present a secure and private methodology for generative artificial intelligence that does not expose sensitive data or models to third-party AI providers. Our work modifies the key building block of modern generative AI algorithms, e.g. the transformer, and introduces confidential and verifiable multiparty computations in a decentralized network to maintain the 1) privacy of the user input and obfuscation to the output of the model, and 2) introduce privacy to the model itself. Additionally, the sharding process reduces the computational burden on any one node, enabling the distribution of resources of large generative AI processes across multiple, smaller nodes. We show that as long as there exists one honest node in the decentralized computation, security is maintained. We also show that the inference process will still succeed if only a majority of the nodes in the computation are successful. Thus, our method offers both secure and verifiable computation in a decentralized network.


Charting the Future: Using Chart Question-Answering for Scalable Evaluation of LLM-Driven Data Visualizations

arXiv.org Artificial Intelligence

We propose a novel framework that leverages Visual Question Answering (VQA) models to automate the evaluation of LLM-generated data visualizations. Traditional evaluation methods often rely on human judgment, which is costly and unscalable, or focus solely on data accuracy, neglecting the effectiveness of visual communication. By employing VQA models, we assess data representation quality and the general communicative clarity of charts. Experiments were conducted using two leading VQA benchmark datasets, ChartQA and PlotQA, with visualizations generated by OpenAI's GPT-3.5 Turbo and Meta's Llama 3.1 70B-Instruct models. Our results indicate that LLM-generated charts do not match the accuracy of the original non-LLM-generated charts based on VQA performance measures. Moreover, while our results demonstrate that few-shot prompting significantly boosts the accuracy of chart generation, considerable progress remains to be made before LLMs can fully match the precision of human-generated graphs. This underscores the importance of our work, which expedites the research process by enabling rapid iteration without the need for human annotation, thus accelerating advancements in this field.