AtomWorld: A Benchmark for Evaluating Spatial Reasoning in Large Language Models on Material Structures
Large language models (LLMs) have shown promising potential in materials science, enabling tasks ranging from knowledge retrieval to property prediction. Existing materials science benchmarks mainly focus on perceptual or knowledge-based tasks, largely ignoring the structure modelling tasks, a core …