Sharing chemical knowledge between human and machine

Research team develops AI tool that translates chemical structures into machine-readable codes

Chemists have long relied on structural formulae to understand the composition and arrangement of chemical compounds. These formulae provide insights into reactions between molecules, synthesis of complex compounds, and potential therapeutic effects of natural substances. However, while these visual representations are intuitive for humans, they’re not easily processed by software. To bridge this gap, a team led by Prof. Christoph Steinbeck and Prof. Achim Zielesny has developed an AI tool called “DECIMER” (Deep Learning for Chemical Image Recognition).

DECIMER transforms images of chemical structural formulae into machine-readable code. This open-source platform automatically identifies and classifies images in scientific articles, converting recognized structural formulae into machine-readable structure codes. For instance, the caffeine molecule’s structural formula becomes the code CN1C=NC2=C1C(=O)N(C(=O)N2C)C. This code can be uploaded into databases and linked with additional molecule information.

The AI tool employs modern AI methods, similar to those used in Large Language Models like ChatGPT. To train DECIMER, the researchers utilized existing machine-readable databases, generating around 450 million structural formulae for training data. Companies and researchers are already using DECIMER to convert structural formulae from patent specifications into databases.

The inspiration for DECIMER arose from the development of AI methods for the ancient Asian board game Go. The chemists were intrigued by the capabilities of AI during the famous Go tournament between human champion Lee Sedol and the AI software “AlphaGo.” Realizing the potential of AI, the researchers aimed to apply these methods to solve complex problems in their field.

The team aspires to make chemical literature from the 1950s onwards machine-readable, preserving knowledge and sharing it with the scientific community. This effort aligns with Prof. Steinbeck’s role as the coordinator of Germany’s National Research Data Infrastructure for Chemistry.

Posted in

Aihub Team

Leave a Comment





News firms seek transparency, collective negotiation over content use by AI makers - letter

News firms seek transparency, collective negotiation over content use by AI makers – letter

White House launches AI-based contest to secure government systems from hacks

White House launches AI-based contest to secure government systems from hacks

Britain appoints tech expert and diplomat to spearhead AI summit

Britain appoints tech expert and diplomat to spearhead AI summit

AI Drafted in War on Online Crimes Against Kids

AI Drafted in War on Online Crimes Against Kids

AI for Disaster Recovery: AI-powered systems for post-disaster recovery and reconstruction.

AI for Disaster Recovery: AI-powered systems for post-disaster recovery and reconstruction.

AI in Drug Repurposing: AI-driven drug discovery for repurposing existing medications.

AI in Drug Repurposing: AI-driven drug discovery for repurposing existing medications.

AI in Augmented Reality: Enhancing AR experiences with AI-generated content and interactions.

AI in Augmented Reality: Enhancing AR experiences with AI-generated content and interactions.

AI in Oil and Gas Exploration: AI applications in seismic data analysis for oil exploration.

AI in Oil and Gas Exploration: AI applications in seismic data analysis for oil exploration.

AI in Podcasting: AI-driven podcast transcription and content recommendation.

AI in Podcasting: AI-driven podcast transcription and content recommendation.

AI in Speech Recognition: Improving speech recognition and transcription with AI algorithms.

AI in Speech Recognition: Improving speech recognition and transcription with AI algorithms.

AI and Blockchain Integration: The potential of combining AI and blockchain technologies.

AI and Blockchain Integration: The potential of combining AI and blockchain technologies.

AI for Wildlife Tracking: AI-enabled tracking systems for studying animal migration and behavior.

AI for Wildlife Tracking: AI-enabled tracking systems for studying animal migration and behavior.

Combating Global Health Crises: The Power of AI in Epidemic Prediction and Prevention

Combating Global Health Crises: The Power of AI in Epidemic Prediction and Prevention

Global cloud market soars again, but AI could pose a risk

Global cloud market soars again, but AI could pose a risk

Interview Mrs.Anita Schjøll Brede

Interview Mrs.Anita Schjøll Brede

Interview with Mr.Jürgen Schmidhuber

Interview with Mr.Jürgen Schmidhuber

Interview with Mr.Fei-Fei Li

Interview with Dr.Fei-Fei Li

AI and Music Composition: The intersection of AI and creativity in composing music.

AI and Music Composition: The intersection of AI and creativity in composing music.

AI in Art Authentication: AI techniques for art forgery detection and provenance verification.

AI in Art Authentication: AI techniques for art forgery detection and provenance verification.

AI for Accessibility: How AI is making technology more accessible for individuals with disabilities.

AI for Accessibility: How AI is making technology more accessible for individuals with disabilities.

AI in Retail Personalization: Customizing shopping experiences with AI-driven recommendations.

AI in Retail Personalization: Customizing shopping experiences with AI-driven recommendations.

AI in Supply Chain Management: AI-driven optimization of supply chain logistics and inventory management.

AI in Supply Chain Management: AI-driven optimization of supply chain logistics and inventory management.

AI in Veterinary Medicine: AI applications for animal health diagnosis and treatment.

AI in Veterinary Medicine: AI applications for animal health diagnosis and treatment.

AI and Genome Sequencing: AI's contribution to accelerating genomic research and precision medicine.

AI and Genome Sequencing: AI’s contribution to accelerating genomic research and precision medicine.

AI and Drone Technology: AI's role in enhancing drone capabilities for various industries.

AI and Drone Technology: AI’s role in enhancing drone capabilities for various industries.

AI in Transportation: Innovations in autonomous vehicles and AI for traffic management.

AI in Transportation: Innovations in autonomous vehicles and AI for traffic management.

AI in Environmental Monitoring: AI applications for monitoring air and water quality.

AI in Environmental Monitoring: AI applications for monitoring air and water quality.

AI in Criminal Justice: AI's impact on crime prevention, offender profiling, and legal analytics.

AI in Criminal Justice: AI’s impact on crime prevention, offender profiling, and legal analytics.

AI for Elderly Care: Enhancing senior care with AI-powered health monitoring and companionship.

AI for Elderly Care: Enhancing senior care with AI-powered health monitoring and companionship.

AI and Disaster Prediction: Predicting natural disasters using AI-based models and algorithms.

AI and Disaster Prediction: Predicting natural disasters using AI-based models and algorithms.