[1]
Jati, H. et al. 2026. Closed-Source vs. Open-Weights Language Models: Accuracy, Stability, and Deployment Trade-Offs in Educational Assessment (GPT-4o vs. LLaMA 3.1). Journal of Soft Computing and Data Mining. 7, 2 (Jun. 2026), 58–74.