A neuro-symbolic path planning framework that integrates LLM-based planners with probabilistic model checking to provide guarantees for temporal and probabilistic requirements is presented.
Abstract
Recent advances in large language models (LLMs) have extended their capabilities beyond text generation to structured reasoning and planning, motivating their deployment in physical systems. A representative safety-critical setting is autonomous navigation, commonly abstracted as stochastic grid-world path planning, where an agent makes sequential decisions under uncertainty while satisfying temporal, probabilistic requirements. Classical deterministic planners offer formal guarantees but cannot adequately model stochastic dynamics, while reinforcement learning (RL) approaches are often problem-specific and may not ensure constraint satisfaction. Existing LLM-based planners largely focus on deterministic task planning and not stochastic grid-world planning with formal guarantees. In this paper, we present a neuro-symbolic path planning framework that integrates LLM-based planners with probabilistic model checking to provide guarantees for temporal and probabilistic requirements. We evaluate three state-of-the-art LLMs, assessing success rate, convergence behavior, and performance against a reinforcement learning baseline. Results show that our approach achieves high success rates across diverse grid-world configurations and outperforms the RL baseline.
Temporal logic is a formal language for reasoning about system behaviors over time. Signal temporal logic (STL), in particular, has been used to encode spatio-temporal requirements for control synthesis in multi-agent systems, often under the assumption that agents are cooperative and their dynamics are known. However,...
This work introduces a unified RL formulation that jointly optimizes agent and environment policies, where the environment policy learns graph edge costs to provide global movement guidance via backward Dijkstra search and achieves significant improvements over the strong search-based planner, Causal-PIBT, across multi...
He Jiang, Jingtian Yan, Yulun Zhang et al.· 0 citations
Improvements show that an explicit graph world model harness can substantially improve the reliability and efficiency of long-horizon embodied planning across compact and frontier hosted LLM capabilities.
Rui-Yang Wang, Hao-Lun Hsu, S. Mehta et al.· 0 citations
This work reduces target variance and yields well-calibrated predictive uncertainty under stochastic observations in continuous control, providing a principled framework for planning with learned distribution models.
Autonomous parking in nonconvex and narrow environments remains challenging. Although optimal-control methods can explicitly enforce vehicle dynamics and collision constraints, nonconvexity compromises solver robustness and can cause failures. Large language models (LLMs) exhibit strong semantic reasoning capabilities,...
This paper introduces an LLM-Augmented Reinforcement Learning Agent that integrates LLM-driven planning with RL-based action optimization, and highlights a promising direction for building more capable autonomous systems.
Christophe D. Hounwanou, John Emeka Eze, Yaé Ulrich Gaba· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.