Deprecated: Function curl_close() is deprecated since 8.5, as it has no effect since PHP 8.0 in /home/u483256323/domains/poorvam.com/public_html/subdomains/pore/includes/api.php on line 184
Back to Search View Original Cite This Article

Abstract

<title>Abstract</title> <p>We study fine-tuning large language models for automatic hint generation in reasoning tasks. Starting from five diverse question–answer datasets, we generate hints using four prompt-ing strategies: answer-aware, answer-agnostic, decomposed answer-aware, and decomposed answer-agnostic. This process is applied to three models: LLaMA-3.1-70B, Gemma-3n-E4B, and DeepSeek-V3.1. We evaluate hint quality using HintEval metrics capturing con-vergence, answer leakage, familiarity, and rel-evance, and propose a metric-guided dataset merging strategy for fine-tuning. Experimental results show that fine-tuning yields modest but consistent improvements in convergence and familiarity while maintaining high relevance and low answer leakage.</p>

Show More

Keywords

finetuning models hint using answeraware

Related Articles


Deprecated: Function curl_close() is deprecated since 8.5, as it has no effect since PHP 8.0 in /home/u483256323/domains/poorvam.com/public_html/subdomains/pore/includes/api.php on line 76
PORE

About

Connect