Bioinfo-Bench: A Simple Benchmark Framework for LLM Bioinformatics Skills Evaluation
Chen, Q.; Deng, C.
Show abstract
AO_SCPLOWBSTRACTC_SCPLOWLarge Language Models (LLMs) have garnered significant recognition in the life sciences for their capacity to comprehend and utilize knowledge. The contemporary expectation in diverse industries extends beyond employing LLMs merely as chatbots; instead, there is a growing emphasis on harnessing their potential as adept analysts proficient in dissecting intricate issues within these sectors. The realm of bioinformatics is no exception to this trend. In this paper, we introduce BO_SCPLOWIOINFOC_SCPLOW-BO_SCPLOWENCHC_SCPLOW, a novel yet straightforward benchmark framework suite crafted to assess the academic knowledge and data mining capabilities of foundational models in bioinformatics. BO_SCPLOWIOINFOC_SCPLOW-BO_SCPLOWENCHC_SCPLOW systematically gathered data from three distinct perspectives: knowledge acquisition, knowledge analysis, and knowledge application, facilitating a comprehensive examination of LLMs. Our evaluation encompassed prominent models ChatGPT, Llama, and Galactica. The findings revealed that these LLMs excel in knowledge acquisition, drawing heavily upon their training data for retention. However, their proficiency in addressing practical professional queries and conducting nuanced knowledge inference remains constrained. Given these insights, we are poised to delve deeper into this domain, engaging in further extensive research and discourse. It is pertinent to note that project BO_SCPLOWIOINFOC_SCPLOW-BO_SCPLOWENCHC_SCPLOW is currently in progress, and all associated materials will be made publicly accessible.1
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Knowledge Graph-based Thought: a knowledge graph enhanced LLMs framework for pan-cancer question answering 96%
- Smash++: an alignment-free and memory-efficient tool to find genomic rearrangements 94%
- Sequence Compression Benchmark (SCB) database - a comprehensive evaluation of reference-free compressors for FASTA-formatted sequences 94%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- Optimizing biomedical information retrieval with a keyword frequency-driven Prompt Enhancement Strategy 97%
- Relation extraction between bacteria and biotopes from biomedical texts with attention mechanisms and domain-specific contextual representations 94%
- StrongestPath: a Cytoscape application for protein-protein interaction analysis 94%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.