Journal of Chemical Information and Modeling· 2026Q1
Toksisite Dili: Açıklanabilir Bir Yapay Zeka Yaklaşımı
Language of Toxicity: An eXplainable Artificial Intelligence Approach
- 0atıf
- Q1SCImago
- 2026yıl
Kısa özet
Tanımlayıcı içermeyen, dile dayalı yapay zeka modeli, CNN, GRU ve dikkat mekanizmalarını kullanarak küçük ve dengesiz veri kümelerinde ortalama 0.83 AUC ile kimyasal toksisiteyi doğru bir şekilde tahmin ediyor.
Yapay zekâ ile başlık ve abstract'tan üretildi; tam metin okunmaz.
Ana noktalar
- CNN, GRU ve dikkat mekanizmalarını kullanarak toksisite tahmini için tanımlayıcı içermeyen bir yapay zeka modeli geliştirildi.
- Toksik/toksik olmayan kimyasalların kanonik SMILES dizeleri, iki farklı 'dil'de 'kelimeler' olarak ele alındı.
- Sekiz toksisite uç noktası boyunca ortalama 0.83 AUC (aralık 0.70-0.94) elde edildi.
- Model, tahminleri etkileyen ilgili moleküler alt yapıları vurgulayarak yorumlanabilirlik gösteriyor.
Yapay zekâ ile başlık ve abstract'tan üretildi; tam metin okunmaz.
Özet (abstract)
Abstract Toxicity prediction in small molecules represents a fundamental challenge in drug development and chemical safety assessment. Traditional approaches heavily rely on predefined molecular descriptors or fingerprints, potentially limiting the ability to capture complex and nonlinear structure–activity relationships. Here, we present a descriptor-free, language-inspired framework that can be applied to different toxicity prediction tasks within a unified architecture. The model proposed combines a multiscale Convolutional Neural Network (CNN) layer to capture chemical patterns at different scales and a Gated Recurrent Unit (GRU) layer to capture the sequential nature of these patterns. This architecture also exploits an attention mechanism that computes attention weights across the sequence, enabling the model to focus on the most relevant molecular substructures for toxicity prediction. Toxic and nontoxic chemicals, represented by canonical SMILES, are investigated as the words of two languages which have to be discriminated; using eight different end points, the model provided an accurate description of toxicity patterns, with an average Area under the ROC curve (AUC) of 0.83 (min: 0.70, max: 0.94) under repeated cross-validation. The models were trained on relatively small data sets (∼1000 samples) and often strongly imbalanced, two important challenges that highlight the opportunities for future improvement; moreover, the proposed attention-based framework offers a representation of the molecular regions influencing model predictions, providing a basis for future investigations into toxicity-related structural patterns and potentially supporting hypothesis generation in drug design or drug repurposing applications.
Yazarların özeti; kaynağından alınmıştır. Journal of Chemical Information and Modeling, 2026 · DOI ↗
Ücretsiz hesapla devam et
Makaleye Sor ile bu makaleye günde 3 soru ücretsiz; makaleyi kaydet, kaynakçasını al, ilgi alanına göre her gün yeni özetler. Çıkarımlar Premium.
Web'de ücretsiz devam etGoogle ya da Apple hesabınla giriş; kart istemez. Bu makaleye geri dönersin.
Telefonda:
Alan: Hesaplamalı Kuram ve Matematik
Computational Theory and MathematicsComputer Science