AITopics | Mervin, Lewis

Collaborating Authors

Mervin, Lewis

Information about AI from the News, Publications, and Conferences

Automatic Classification – Tagging and Summarization – Customizable Filtering and Analysis

If you are looking for an answer to the question What is Artificial Intelligence? and you only have a minute, then here's the definition the Association for the Advancement of Artificial Intelligence offers on its home page: "the scientific understanding of the mechanisms underlying thought and intelligent behavior and their embodiment in machines."

However, if you are fortunate enough to have more than a minute, then please get ready to embark upon an exciting journey exploring AI (but beware, it could last a lifetime) …

Temporal Distribution Shift in Real-World Pharmaceutical Data: Implications for Uncertainty Quantification in QSAR Models

Friesacher, Hannah Rosa, Svensson, Emma, Winiwarter, Susanne, Mervin, Lewis, Arany, Adam, Engkvist, Ola

arXiv.org Artificial IntelligenceFeb-6-2025

The estimation of uncertainties associated with predictions from quantitative structure-activity relationship (QSAR) models can accelerate the drug discovery process by identifying promising experiments and allowing an efficient allocation of resources. Several computational tools exist that estimate the predictive uncertainty in machine learning models. However, deviations from the i.i.d. setting have been shown to impair the performance of these uncertainty quantification methods. We use a real-world pharmaceutical dataset to address the pressing need for a comprehensive, large-scale evaluation of uncertainty estimation methods in the context of realistic distribution shifts over time. We investigate the performance of several uncertainty estimation methods, including ensemble-based and Bayesian approaches. Furthermore, we use this real-world setting to systematically assess the distribution shifts in label and descriptor space and their impact on the capability of the uncertainty estimation methods. Our study reveals significant shifts over time in both label and descriptor space and a clear connection between the magnitude of the shift and the nature of the assay. Moreover, we show that pronounced distribution shifts impair the performance of popular uncertainty estimation methods used in QSAR models. This work highlights the challenges of identifying uncertainty quantification methods that remain reliable under distribution shifts introduced by real-world data.

artificial intelligence, assay, machine learning, (15 more...)

arXiv.org Artificial Intelligence

2502.03982

Country:

Europe (1.00)
North America > United States > New York (0.14)

Genre: Research Report > New Finding (0.93)

Industry: Health & Medicine > Pharmaceuticals & Biotechnology (1.00)

Technology:

Information Technology > Artificial Intelligence > Machine Learning > Statistical Learning (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks (1.00)
Information Technology > Artificial Intelligence > Representation & Reasoning > Uncertainty > Bayesian Inference (0.66)
(2 more...)

Add feedback

Publishing Neural Networks in Drug Discovery Might Compromise Training Data Privacy

Krüger, Fabian P., Östman, Johan, Mervin, Lewis, Tetko, Igor V., Engkvist, Ola

arXiv.org Artificial IntelligenceOct-22-2024

This study investigates the risks of exposing confidential chemical structures when machine learning models trained on these structures are made publicly available. We use membership inference attacks, a common method to assess privacy that is largely unexplored in the context of drug discovery, to examine neural networks for molecular property prediction in a black-box setting. Our results reveal significant privacy risks across all evaluated datasets and neural network architectures. Combining multiple attacks increases these risks. Molecules from minority classes, often the most valuable in drug discovery, are particularly vulnerable. We also found that representing molecules as graphs and using message-passing neural networks may mitigate these risks. We provide a framework to assess privacy risks of classification models and molecular representations. Our findings highlight the need for careful consideration when sharing neural networks trained on proprietary chemical structures, informing organisations and researchers about the trade-offs between data confidentiality and model openness.

artificial intelligence, machine learning, membership inference attack, (16 more...)

arXiv.org Artificial Intelligence

2410.16975

Country: Europe > Sweden (0.28)

Genre: Research Report > New Finding (1.00)

Industry:

Information Technology > Security & Privacy (1.00)
Health & Medicine > Pharmaceuticals & Biotechnology (1.00)

Technology:

Information Technology > Artificial Intelligence > Machine Learning > Neural Networks (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Performance Analysis > Accuracy (0.96)

Add feedback

Industry-Scale Orchestrated Federated Learning for Drug Discovery

Oldenhof, Martijn, Ács, Gergely, Pejó, Balázs, Schuffenhauer, Ansgar, Holway, Nicholas, Sturm, Noé, Dieckmann, Arne, Fortmeier, Oliver, Boniface, Eric, Mayer, Clément, Gohier, Arnaud, Schmidtke, Peter, Niwayama, Ritsuya, Kopecky, Dieter, Mervin, Lewis, Rathi, Prakash Chandra, Friedrich, Lukas, Formanek, András, Antal, Peter, Rahaman, Jordon, Zalewski, Adam, Heyndrickx, Wouter, Oluoch, Ezron, Stößel, Manuel, Vančo, Michal, Endico, David, Gelus, Fabien, de Boisfossé, Thaïs, Darbier, Adrien, Nicollet, Ashley, Blottière, Matthieu, Telenczuk, Maria, Nguyen, Van Tien, Martinez, Thibaud, Boillet, Camille, Moutet, Kelvin, Picosson, Alexandre, Gasser, Aurélien, Djafar, Inal, Simon, Antoine, Arany, Ádám, Simm, Jaak, Moreau, Yves, Engkvist, Ola, Ceulemans, Hugo, Marini, Camille, Galtier, Mathieu

arXiv.org Artificial IntelligenceDec-12-2022

To apply federated learning to drug discovery we developed a novel platform in the context of European Innovative Medicines Initiative (IMI) project MELLODDY (grant n{\deg}831472), which was comprised of 10 pharmaceutical companies, academic research labs, large industrial companies and startups. The MELLODDY platform was the first industry-scale platform to enable the creation of a global federated model for drug discovery without sharing the confidential data sets of the individual partners. The federated model was trained on the platform by aggregating the gradients of all contributing partners in a cryptographic, secure way following each training iteration. The platform was deployed on an Amazon Web Services (AWS) multi-account architecture running Kubernetes clusters in private subnets. Organisationally, the roles of the different partners were codified as different rights and permissions on the platform and administrated in a decentralized way. The MELLODDY platform generated new scientific discoveries which are described in a companion paper.

artificial intelligence, machine learning, platform, (16 more...)

arXiv.org Artificial Intelligence

2210.08871

Country:

Europe > Germany (0.46)
Europe > France (0.28)

Genre: Research Report (1.00)

Industry: Health & Medicine > Pharmaceuticals & Biotechnology (1.00)

Technology: Information Technology > Artificial Intelligence > Machine Learning > Neural Networks (0.93)

Add feedback