文章摘要

李菲菲.国家哲学社会科学文献中心高质量数据集建设回顾与前瞻[J].中国图书馆学报,2026,52(3):47~66
国家哲学社会科学文献中心高质量数据集建设回顾与前瞻
Review and Prospect on the Development of High quality Datasets at the National Center for Philosophy and Social Sciences Documentation
投稿时间:2025-12-19  修订日期:2026-05-04
DOI:
中文关键词: 国家哲学社会科学文献中心  高质量数据集  自主知识体系  可信数据空间  知识服务人工智能大模型
英文关键词: National Center for Philosophy and Social Sciences Documentation  High quality datasets  Independent knowledge system  Trusted data space  Knowledge servicesArtificial intelligence large models
基金项目:
作者单位
李菲菲 中国社会科学院图书馆 北京 100732 
摘要点击次数: 450
全文下载次数: 142
中文摘要:
      在数据驱动学术研究范式转型的背景下,AI就绪高质量数据集是构建中国哲学社会科学自主知识体系的基础底座。国家哲学社会科学文献中心作为国家级公益性学术平台,历经十年建设积累了海量多模态学术资源,并成功入选国家高质量数据集建设先行先试名单。本文系统梳理平台发展历程及多模态资源体系;围绕全生命周期可信治理体系,概述平台完成从传统文献资源向可计算、AI就绪数据转型的高质量数据集建设实践。自建数据集可为哲学社会科学领域专用大模型基础底座搭建提供基础,覆盖资源建设、平台运营、国际传播等应用场景。实践证明,该数据集体系在公共普惠服务、学术研究服务、智库资政支撑等领域成效显著。文章从资源迭代升级、人工智能深度融合、协同生态构建、服务国家战略、复合型人才培育五个维度提出未来发展路径,可为哲学社会科学领域高质量数据集长效建设、社科新质生产力培育、国家文化软实力提升提供可复制的实践范式,助力我国加快构建哲学社会科学自主知识体系、提升国际学术话语权。图2。参考文献5。
英文摘要:
Against the backdrop of the paradigm shift toward data driven academic research,AI ready high quality datasets constitute the foundational infrastructure for developing the independent knowledge system of Chinese philosophy and social sciences. As a national public welfare academic platform,the National Center for Philosophy and Social Sciences Documentation has accumulated massive multimodal academic resources over ten years of development and has been included in the pilot list for national high quality datasets. This paper reviews the development history of the platform and elaborates its multimodal resource system consisting of Chinese journals,foreign language resources,ancient books,special subject databases,and achievements of The National Social Science Fund of China. The platform has established a full lifecycle trusted governance system covering data collection,processing,quality inspection,storage,and continuous update,as well as a four level linked quality control mechanism,realizing the transformation from traditional literature to computable and AI ready data. Based on self built datasets,the platform can develop domain specific large language models for philosophy and social sciences,together with four major agent clusters for literature management,academic research,translation,and think tank services,which cover three key application scenarios including resource construction,platform operation,and international academic communication. Practices have verified that this dataset system has achieved remarkable results in inclusive public services,academic research,and think tank policy consultation.
This paper puts forward future development paths from five aspects:resource iteration,in depth AI integration,collaborative ecosystem construction,national strategy support,and interdisciplinary talent cultivation. It provides a replicable practical paradigm for the long term construction of high quality datasets in philosophy and social sciences,the cultivation of new quality productive forces for social sciences,and the improvement of national cultural soft power. Meanwhile,it facilitates the construction of independent knowledge system of Chinese philosophy and social sciences and the enhancement of China's international academic discourse power. 2 figs. 5 refs.
查看全文   查看/发表评论  下载PDF阅读器