AIToday
Large Language ModelsApple Machine LearningPublished: May 5, 2026, 10:00 JST1 min read

Apple researchers propose PORTool, a policy-optimization algorithm that improves LLM tool-use agents by assigning step-level rewards from outcome-only supervision.

Apple researchers propose PORTool, a policy-optimization algorithm that improves LLM tool-use agents by assigning step-level rewards from outcome-only supervision.

3 Key Points

  1. PORTool, developed by researchers at Purdue University and Apple, generates a rewarded rollout tree where trajectories share prefixes before branching, enabling direct comparisons among alternative tool-use decisions within the same context.

  2. The algorithm estimates each step's importance using a correctness-dominant signal—whether descendants of that step can ultimately produce a correct final answer—plus an auxiliary term indicating whether the step's tool calls execute successfully.

  3. Experiments show PORTool improves final-answer accuracy while reducing tool-call steps compared with state-of-the-art baselines, with ablation studies confirming the robustness of the proposed step-wise importance estimates.

  4. The paper was accepted at the Fifth Workshop on Natural Language Generation, Evaluation, and Metrics at ACL 2026.

Ask the AI about this article →

Apple Machine LearningRead Original Article

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • DeepMind chief: frontier AI leadership is all that mattersTHE DECODER · 23m ago
  • John Deere launches AI chatbot for farmersThe Verge AI · 24m ago
  • Google Pics launches with AI image editing for WorkspaceThe Verge AI · 24m ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · takes 30 seconds · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleTrump administration considers executive order creating AI working group and requiring government review of new models