The Potential and Limitations of Large Language Models in Logical Problem Proving and Prompt Construction within Intelligent Tutoring Systems
原文英文,约100词,阅读约需1分钟。
📝
内容提要
本研究分析了智能辅导系统在个性化反馈中的不足,并提出了评估方法。结果显示,DeepSeek-V3在逻辑证明构建中的准确率为84.4%。尽管LLM生成的提示在一致性和清晰度上表现良好,但在解释背景时存在不足,需要改进以提高准确性和教育适宜性。
🏷️