TAPR: Enhancing LLM Performance with a Task-Aware Prompt Rewriter

작성자

카테고리:

← 피드로
arXiv cs.AI · Oliver Savolainen, Emanuele Bastianelli, Hosein Azarbonyad · 2026-08-03 AI

[Submitted on 17 Jul 2026]

View PDF HTML (experimental)

Abstract:Large Language Models (LLMs) often require carefully crafted prompts to unlock their full potential, which can be a barrier for non-expert users. This work addresses the challenge by introducing a Task-Aware Prompt Rewriter (TAPR), a model that reformulates user prompts into task-optimized prompts with the explicit goal of improving downstream LLM performance. We train TAPR using reinforcement learning with Group Relative Policy Optimization (GRPO), where rewards are derived from LLM-as-judge evaluations of both the reformulated prompt and the corresponding task output. Experimental results on diverse tasks, such as question answering, summarization, and arithmetic reasoning, show that our method yields consistent gains over base models in prompt rewriting ability. Fine-tuning Phi-4-mini-instruct (as the base model for TAPR) produces prompts that contain clearer and more instructive language, leading to higher accuracy on established benchmarks such as Natural Questions and GSM8K. Our code is available at: this https URL

Submission history

From: Hosein Azarbonyad [view email]
[v1] Fri, 17 Jul 2026 15:37:07 UTC (4,618 KB)

원문에서 계속 ↗

추출 본문 · 출처: arxiv.org · https://arxiv.org/abs/2607.28657

코멘트

답글 남기기

이메일 주소는 공개되지 않습니다. 필수 필드는 *로 표시됩니다