← 返回论文检索
NeurIPS 2024PosterAccept (Poster)

Empowering and Assessing the Utility of Large Language Models in Crop Science

Hang Zhang, Jiawei SUN, Renqi Chen, Wei Liu, Zhonghang Yuan, Xinzhe Zheng, Zhefan Wang, Zhiyuan Yang, Hang Yan, Han-Sen Zhong, Xiqing Wang, Wanli Ouyang, Fan Yang, Nanqing Dong

Shanghai Artificial Intelligence Laboratory · Hangzhou dianzi university · Shanghai Jiao Tong University · University of Science and Technology of China · Liaoning University · Hangzhou Dianzi University · AI lab · China Agricultural University · Shanghai AI Lab · Yazhouwan Lab

PDF 由论文原始站点提供,PaperCompass 不保存论文文件。

摘要

Large language models (LLMs) have demonstrated remarkable efficacy across knowledge-intensive tasks. Nevertheless, their untapped potential in crop science presents an opportunity for advancement. To narrow this gap, we introduce CROP, which includes a novel instruction tuning dataset specifically designed to enhance LLMs’ professional capabilities in the crop science sector, along with a benchmark that serves as a comprehensive evaluation of LLMs’ understanding of the domain knowledge. The CROP dataset is curated through a task-oriented and LLM-human integrated pipeline, comprising 210,038 single-turn and 1,871 multi-turn dialogues related to crop science scenarios. The CROP benchmark includes 5,045 multiple-choice questions covering three difficulty levels. Our experiments based on the CROP benchmark demonstrate notable enhancements in crop science-related tasks when LLMs are fine-tuned with the CROP dataset. To the best of our knowledge, CROP dataset is the first-ever instruction tuning dataset in the crop science domain. We anticipate that CROP will accelerate the adoption of LLMs in the domain of crop science, ultimately contributing to global food production.