Back to search

Article

Benchmarking Large Language Models on Long-Tail Plant Taxonomic Knowledge with PTTB-600

2026-07-15

Abstract excerpt

Plant taxonomic knowledge contains a long tail of infrequently encountered names, diagnostic characters, and nomenclatural decisions, yet model reliability across this distribution remains unclear. We developed the Chinese-language PTTB-600, comprising 200 general, 300 ordinary specialized, and 100 long-tail fill-in questions, and evaluated 31 large language models (LLMs) or run modes under closed-book conditions...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
8c05e24a-a7ef-50ee-80a0-10ebb9b14b51
DOI
10.20944/preprints202607.1052.v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Benchmarking Large Language Models on Long-Tail Plant Taxonomic Knowledge with PTTB-600DOI 10.20944/preprints202607.1052.v1
Select a neighboring publication to make it the new centre.