---
title: Can We Rely on LLM Agents to Draft Long-Horizon Plans? Let's Take TravelPlanner as an Example
url: https://www.ml-quant.com/papers/arxiv/2408.06318/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2408.06318
source_url: https://arxiv.org/pdf/2408.06318v1
featured: 2024-08-15
citations: 30
topic: LLMs & Text
---


# Can We Rely on LLM Agents to Draft Long-Horizon Plans? Let's Take TravelPlanner as an Example

The study suggests a Feedback-Aware Fine-Tuning (FAFT) method to improve Large Language Models' (LLMs) performance in real-world planning tasks by using both positive and negative feedback.

- Source: https://arxiv.org/pdf/2408.06318v1
- Identifier: arXiv:2408.06318
- Released: 2024-08-12
- First featured: Quant Letter No. 61 (2024-08-15): https://www.ml-quant.com/issues/2024-08-15/
- Citations (Semantic Scholar): 30
- Published in: not yet
- Topic: LLMs & Text

## Related

- [Designing Heterogeneous LLM Agents for Financial Sentiment Analysis](https://www.ml-quant.com/papers/arxiv/2401.05799/): A study suggests using large language models without fine-tuning for financial sentiment analysis, offering a design framework that enhances accuracy.
- [Does Fine-Tuning LLMs on New Knowledge Encourage Hallucinations?](https://www.ml-quant.com/papers/arxiv/2405.05904/): Research shows that large language models have difficulty acquiring new factual knowledge through fine-tuning, learning new information slower than consistent knowledge, and are more likely to hallucinate, indicating the risks of introducing new facts through fine-tuning.
- [FinMem: A Performance-Enhanced LLM Trading Agent With Layered Memory and Character Design](https://www.ml-quant.com/papers/arxiv/2311.13743/): Performance-Enhanced Large Language Model Trading Agent: The research presents FinMe, a Large Language Model-based agent for financial decision-making, demonstrating its superior trading performance in stocks and funds.
- [Trial and Error: Exploration-Based Trajectory Optimization for LLM Agents](https://www.ml-quant.com/papers/arxiv/2403.02502/): The study explores an exploration-based trajectory optimization (ETO) method to enhance the performance of Large Language Models (LLMs) by learning from their exploration mistakes.
- [TableLlama: Towards Open Large Generalist Models for Tables](https://www.ml-quant.com/papers/arxiv/2311.09206/): The paper presents TableLlama, an open-source large language model fine-tuned for table-based tasks, and introduces a new dataset, TableInstruct, which improves the performance and generalizability of these models.
- [Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models](https://www.ml-quant.com/papers/arxiv/2411.14432/): The paper introduces Insight-V, a system that improves the reasoning abilities of large language models by generating extensive reasoning paths and integrating a multi-agent system for better visual reasoning performance.
