1
Return

Large language models for knowledge-centric scientific intelligence: methods, challenges, and lessons from geoscience

delete2026-08-04
delete0
delete
OA
AI
J
Jianhua Ma
Y
Yongzhang Zhou *
L
Luhao He *
X
Xian Liu
DOI:10.1007/s10462-026-11637-zdelete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Large language models (LLMs) are increasingly being explored in geoscience, where scientific knowledge is expressed through specialized terminology, heterogeneous documents, maps, imagery, geospatial structures, and temporally ordered interpretations. This review critically examines the emerging literature on geoscience-oriented LLMs (GeoLLMs), focusing on the tasks, construction strategies, evaluation needs, and unresolved challenges that distinguish them from generic LLM applications. We first synthesize the GeoLLM task landscape, including geological information extraction and semantic normalization, relation modeling and knowledge graph construction, evidence-grounded question answering, multimodal map–image–text reasoning, and high-value applications such as hazard-related information analysis, mineral prospectivity evidence synthesis, and chronostratigraphic interpretation. We then review model construction and adaptation strategies, including geoscience corpus engineering, parameter-efficient tuning, domain-adaptive pretraining, retrieval augmentation, ontology and knowledge-graph grounding, and multimodal representation learning. Across these studies, a consistent theme is that GeoLLM outputs should be assessed not only by linguistic fluency, but also by terminology consistency, evidence traceability, spatial and temporal coherence, multimodal grounding, uncertainty expression, and expert validation. The review further identifies major open challenges, including fragmented data and benchmarks, regional and multilingual terminology variation, scale-aware multimodal reasoning, causal and spatiotemporal consistency, hallucination, overtrust, data privacy, proprietary-model dependence, and deployment governance. By using geoscience as a demanding application domain rather than a universal testbed, this review clarifies where domain-specific LLMs can add value, where current evidence remains limited, and what evaluation and governance practices are needed for reliable scientific use.
Keywords:
Large language models
Geoscience knowledge discovery
Geological text mining
Multimodal earth data integration
Knowledge graph
Geological hazard assessment
Mineral exploration

Journal

Artificial Intelligence Review cover
Artificial Intelligence Review
IF:
13.9
Papers:
6.1K
Citations:
1.9W

Organization

G
Guangzhou Institute of Geochemistry
Scholars:
222
Papers: 75
Citations: 6.1K
C
center for earth environment and earth resources
Scholars:
11
Papers: 5
Citations: 0
Cited Papers

Cited Papers

Citing Papers

Citing Papers