REFERENCES

1. Yao, S.; Zhao, J.; Yu, D.; et al. ReAct: synergizing reasoning and acting in language models. In International Conference on Learning Representations (ICLR), 2023. https://openreview.net/forum?id=WE_vluYUL-X. (accessed 2026-09-01).

2. Schick, T.; Dwivedi-Yu, J.; Dessi, R.; et al. Toolformer: language models can teach themselves to use tools. In Advances in Neural Information Processing Systems 36, New Orleans, Louisiana, USA, December 10-16, 2023; Neural Information Processing Systems Foundation, Inc. (NeurIPS): San Diego, California, USA, 2023; pp 68539-51.

3. Curtarolo, S.; Hart, G. L. W.; Nardelli, M. B.; Mingo, N.; Sanvito, S.; Levy, O. The high-throughput highway to computational materials design. Nat. Mater. 2013, 12, 191-201.

4. Lejaeghere, K.; Bihlmayer, G.; Björkman, T.; et al. Reproducibility in density functional theory calculations of solids. Science 2016, 351, aad3000.

5. Liu, Z. K. Thermodynamics and its prediction and CALPHAD modeling: review, state of the art, and perspectives. Calphad 2023, 82, 102580.

6. Meng, Z.; Hao, C.; Li, X. Bridging machine learning and water electrolysis: Concepts, methods, and perspectives. Mater. Today. Energy. 2026, 57, 102232.

7. Van, M. H.; Verma, P.; Zhao, C.; Wu, X. A survey of AI for materials science: foundation models, LLM agents, datasets, and tools. ACM. Comput. Surv. 2026, 58, 1-37.

8. Wei, J.; Yang, Y.; Zhang, X.; et al. From AI for science to agentic science: a survey on autonomous scientific discovery. arXiv 2025, arXiv:2508.14111. Available online: https://doi.org/10.48550/arXiv.2508.14111 (accessed 1 September 2026).

9. Oliveira, O. N.; Christino, L.; Oliveira, M. C. F.; Paulovich, F. V. Artificial intelligence agents for materials sciences. J. Chem. Inf. Model. 2023, 63, 7605-9.

10. Calderon, C. E.; Plata, J. J.; Toher, C.; et al. The AFLOW standard for high-throughput materials science calculations. Comput. Mater. Sci. 2015, 108, 233-8.

11. Huber, S. P.; Bosoni, E.; Bercx, M.; et al. Common workflows for computing material properties using different quantum engines. npj. Comput. Mater. 2021, 7, 136.

12. Mathew, K.; Montoya, J. H.; Faghaninia, A.; et al. Atomate: a high-level interface to generate, execute, and analyze computational materials science workflows. Comput. Mater. Sci. 2017, 139, 140-52.

13. Yao, T.; Yang, Y.; Yan, Y.; et al. Knowledge-extractor: a self-evolving scientific framework for hydrogen energy research driven by AI agents. AI. Agent. 2025, 1, 7.

14. Jia, S.; Zhang, C.; Fung, V. LLMatDesign: autonomous materials discovery with large language models. arXiv 2024, arXiv:2406.13163. Available online: https://doi.org/10.48550/arXiv.2406.13163 (accessed 1 September 2026).

15. Ghafarollahi, A.; Buehler, M. J. Automating alloy design and discovery with physics-aware multimodal multiagent AI. Proc. Natl. Acad. Sci. U. S. A. 2025, 122, e2414074122.

16. Bran, A. M.; Cox, S.; Schilter, O.; Baldassari, C.; White, A. D.; Schwaller, P. Augmenting large language models with chemistry tools. Nat. Mach. Intell. 2024, 6, 525-35.

17. Zhang, D.; Jia, X.; Liu, H.; et al. Cloud synthesis: a global closed-loop feedback powered by autonomous AI-driven catalyst design agent. AI. Agent. 2025, 1, 2.

18. Zhou, L.; Ling, H.; Yan, K.; et al. Toward greater autonomy in materials discovery agents: unifying planning, physics, and scientists. In Transactions on Machine Learning Research, 2026. https://openreview.net/forum?id=Cwq1U8tbWW. (accessed 2026-09-01).

19. Boiko, D. A.; Macknight, R.; Kline, B.; Gomes, G. Autonomous chemical research with large language models. Nature 2023, 624, 570-8.

20. Hu, Z.; Talit, K.; Wang, Z.; et al. TritonDFT: automating DFT with a multi-agent framework. arXiv 2026, arXiv:2603.03372. Available online: https://doi.org/10.48550/arXiv.2603.03372 (accessed 1 September 2026).

21. Shi, Z.; A, H.; Shao, Y.; et al. MDAgent2: large language model for code generation and knowledge Q&A in molecular dynamics. arXiv 2026, arXiv:2601.02075. Available online: https://doi.org/10.48550/arXiv.2601.02075 (accessed 1 September 2026).

22. Yang, F.; Evans, J. D. QUASAR: a universal autonomous system for atomistic simulation and a benchmark of its capabilities. J. Chem. Inf. Model. 2026, 66, 5911-8.

23. Soleymanibrojeni, M.; Aydin, R.; Guedes-Sobrinho, D.; et al. GENIUS: an agentic AI framework for autonomous design and execution of simulation protocols. Commun. Mater. 2026, 7, 115.

24. Ding, M.; Huang, C.; Hu, Y.; et al. Automating computational chemistry workflows via OpenClaw and domain-specific skills. J. Chem. Theory. Comput. 2026, 22, 5919-29.

25. Zhang, C.; Yakobson, B. I. MatClaw: an autonomous code-first LLM agent for End-to-End materials exploration. arXiv 2026, arXiv:2604.02688. Available online: https://doi.org/10.48550/arXiv.2604.02688 (accessed 1 September 2026).

26. Gupta, T.; Zaki, M.; Krishnan, N. M. A. Mausam. MatSciBERT: a materials domain language model for text mining and information extraction. npj. Comput. Mater. 2022, 8, 102.

27. Ahlawat, D.; Mishra, V.; Singh, S.; et al. A family of large language models for materials research with insights into model adaptability in continued pretraining. Nat. Mach. Intell. 2026, 8, 435-48.

28. Tang, Y.; Xu, W.; Cao, J.; et al. A multimodal large language model for materials science. Nat. Mach. Intell. 2026, 8, 588-601.

29. Wu, Q.; Bansal, G.; Zhang, J.; et al. AutoGen: enabling next-gen LLM applications via multi-agent conversations. In First Conference on Language Modeling, 2024. https://openreview.net/forum?id=BAakY1hNKS. (accessed 2026-09-01).

30. Vriza, A.; Kornu, U.; Koneru, A.; Chan, H.; Sankaranarayanan, S. K. R. S. Multi-agentic AI framework for end-to-end atomistic simulations. Digit. Discov. 2026, 5, 440-52.

31. Ferrag, M. A.; Tihanyi, N.; Debbah, M. From LLM reasoning to autonomous AI agents: a comprehensive review. arXiv 2026, arXiv:2504.19678. Available online: https://doi.org/10.48550/arXiv.2504.19678 (accessed 1 September 2026).

32. Ramos, M. C.; Collison, C. J.; White, A. D. A review of large language models and autonomous agents in chemistry. Chem. Sci. 2025, 16, 2514-72.

33. Wang, Z.; Huang, H.; Zhao, H.; et al. DREAMS: density functional theory based research engine for agentic materials simulation. arXiv 2025, arXiv:2507.14267. Available online: https://doi.org/10.48550/arXiv.2507.14267 (accessed 1 September 2026).

34. Yang, P.; Zhang, Z.; Li, Y.; et al. AutoDFT: A closed-loop multi-agent framework for autonomous DFT calculations. arXiv 2026, arXiv:2605.26179. Available online: https://doi.org/10.48550/arXiv.2605.26179 (accessed 1 September 2026).

35. Liu, G.; Yang, S.; Zhong, Y. Masgent: an AI-assisted materials simulation agent. Digit. Discov. 2026, 5, 2151-71.

36. Yu, B.; Baker, F. N.; Chen, Z.; et al. Tooling or not tooling? The impact of tools on language agents for chemistry problem solving. In Findings of the Association for Computational Linguistics: NAACL 2025, Albuquerque, New Mexico, March 2025; Association for Computational Linguistics: Stroudsburg, PA, USA, 2025; pp 7635-55.

37. Zaki, M.; Jayadeva; Mausam; Krishnan, N. M. A. MaScQA: investigating materials science knowledge of large language models. Digit. Discov. 2024, 3, 313-27.

38. Cheung, J. J.; Shen, S.; Zhuang, Y.; Li, Y.; Ramprasad, R.; Zhang, C. MSQA: benchmarking LLMs on graduate-level materials science reasoning and knowledge. arXiv 2025, arXiv:2505.23982. Available online: https://doi.org/10.48550/arXiv.2505.23982 (accessed 1 September 2026).

39. Bajan, C.; Lambard, G. Exploring the expertise of large language models in materials science and metallurgical engineering. Digit. Discov. 2025, 4, 500-12.

40. Liu, S.; Hu, B.; Ye, B.; et al. MatTools: benchmarking large language models for materials science tools. arXiv 2025, arXiv:2505.10852. Available online: https://doi.org/10.48550/arXiv.2505.10852 (accessed 1 September 2026).

41. Shi, Z.; Xin, C.; Huo, T.; et al. A fine-tuned large language model based molecular dynamics agent for code generation to obtain material thermodynamic parameters. Sci. Rep. 2025, 15, 10295.

42. Kumar, S. G. H.; Zou, Y.; Wang, A.; et al. El Agente Sólido: a new age(nt) for solid state simulations. arXiv 2026, arXiv:2602.17886. Available online: https://doi.org/10.48550/arXiv.2602.17886 (accessed 1 September 2026).

43. Huang, Z.; Cao, Y.; Shargh, A. K.; et al. Can coding agents reproduce findings in computational materials science? arXiv 2026, arXiv:2605.00803. Available online: https://doi.org/10.48550/arXiv.2605.00803 (accessed 1 September 2026).

44. Holbrook, E.; Verduzco, J. C.; Strachan, A. Evaluating LLM-generated code for domain-specific languages: molecular dynamics with LAMMPS. Comput. Mater. Sci. 2026, 272, 114839.

45. Zhang, Z.; Yin, A.; Baweja, A.; et al. El agente forjador: task-driven agent generation for quantum simulation. In AI4X - Accelerate Conference 2026, 2026. https://openreview.net/forum?id=7aXeu0hHo5. (accessed 2026-09-01).

46. Liang, P.; Bommasani, R.; Lee, T.; et al. Holistic evaluation of language models. In Transactions on Machine Learning Research, 2023. https://openreview.net/forum?id=iO4LZibEqW. (accessed 2026-09-01).

47. Pham, T. D.; Tanikanti, A.; Keçeli, M. ChemGraph as an agentic framework for computational chemistry workflows. Commun. Chem. 2026, 9, 33.