Asset Details
MbrlCatalogueTitleDetail
Do you wish to reserve the book?
Advancing plant metabolic research by using large language models to expand databases and extract labeled data
by
Busta, Lucas
, Johnson, Braidon
, Knapp, Rachel
in
accuracy
/ Application
/ artificial intelligence
/ Automation
/ Data collection
/ Data mining
/ Engineering
/ engineers
/ Enzymes
/ extracts
/ Flowers & plants
/ Genomes
/ image analysis
/ Language
/ languages
/ Large language models
/ Metabolism
/ model validation
/ Natural products
/ pipelines
/ Plant extracts
/ plant metabolism
/ planting
/ Proteins
/ structured data extraction
2025
Hey, we have placed the reservation for you!
By the way, why not check out events that you can attend while you pick your title.
You are currently in the queue to collect this book. You will be notified once it is your turn to collect the book.
Oops! Something went wrong.
Looks like we were not able to place the reservation. Kindly try again later.
Are you sure you want to remove the book from the shelf?
Advancing plant metabolic research by using large language models to expand databases and extract labeled data
by
Busta, Lucas
, Johnson, Braidon
, Knapp, Rachel
in
accuracy
/ Application
/ artificial intelligence
/ Automation
/ Data collection
/ Data mining
/ Engineering
/ engineers
/ Enzymes
/ extracts
/ Flowers & plants
/ Genomes
/ image analysis
/ Language
/ languages
/ Large language models
/ Metabolism
/ model validation
/ Natural products
/ pipelines
/ Plant extracts
/ plant metabolism
/ planting
/ Proteins
/ structured data extraction
2025
Oops! Something went wrong.
While trying to remove the title from your shelf something went wrong :( Kindly try again later!
Do you wish to request the book?
Advancing plant metabolic research by using large language models to expand databases and extract labeled data
by
Busta, Lucas
, Johnson, Braidon
, Knapp, Rachel
in
accuracy
/ Application
/ artificial intelligence
/ Automation
/ Data collection
/ Data mining
/ Engineering
/ engineers
/ Enzymes
/ extracts
/ Flowers & plants
/ Genomes
/ image analysis
/ Language
/ languages
/ Large language models
/ Metabolism
/ model validation
/ Natural products
/ pipelines
/ Plant extracts
/ plant metabolism
/ planting
/ Proteins
/ structured data extraction
2025
Please be aware that the book you have requested cannot be checked out. If you would like to checkout this book, you can reserve another copy
We have requested the book for you!
Your request is successful and it will be processed during the Library working hours. Please check the status of your request in My Requests.
Oops! Something went wrong.
Looks like we were not able to place your request. Kindly try again later.
Advancing plant metabolic research by using large language models to expand databases and extract labeled data
Journal Article
Advancing plant metabolic research by using large language models to expand databases and extract labeled data
2025
Request Book From Autostore
and Choose the Collection Method
Overview
Premise Recently, plant science has seen transformative advances in scalable data collection for sequence and chemical data. These large datasets, combined with machine learning, have demonstrated that conducting plant metabolic research on large scales yields remarkable insights. A key next step in increasing scale has been revealed with the advent of accessible large language models, which, even in their early stages, can distill structured data from the literature. This brings us closer to creating specialized databases that consolidate virtually all published knowledge on a topic. Methods Here, we first test different combinations of prompt engineering techniques and language models in the identification of validated enzyme–product pairs. Next, we evaluate the application of automated prompt engineering and retrieval‐augmented generation to identify compound–species associations. Finally, we build and determine the accuracy of a multimodal language model–based pipeline that transcribes images of tables into machine‐readable formats. Results When tuned for each specific task, these methods perform with high (80–90%) or modest (50%) accuracies for enzyme–product pair identification and table image transcription, but with lower false‐negative rates than previous methods (decreasing from 55% to 40%) for compound–species pair identification. Discussion We enumerate several suggestions for researchers working with language models, among which is the importance of the user's domain‐specific expertise and knowledge.
This website uses cookies to ensure you get the best experience on our website.