Skip to content
Book Open access

LLM-Based Grid-World Path Planning With Probabilistic Model Checking

Jul 2026 · SIGSOFT FSE Companion · pp. 1684-1691 · 0 citations · 26 references
Computer Science

TL;DR

A neuro-symbolic path planning framework that integrates LLM-based planners with probabilistic model checking to provide guarantees for temporal and probabilistic requirements is presented.

Abstract

Recent advances in large language models (LLMs) have extended their capabilities beyond text generation to structured reasoning and planning, motivating their deployment in physical systems. A representative safety-critical setting is autonomous navigation, commonly abstracted as stochastic grid-world path planning, where an agent makes sequential decisions under uncertainty while satisfying temporal, probabilistic requirements. Classical deterministic planners offer formal guarantees but cannot adequately model stochastic dynamics, while reinforcement learning (RL) approaches are often problem-specific and may not ensure constraint satisfaction. Existing LLM-based planners largely focus on deterministic task planning and not stochastic grid-world planning with formal guarantees. In this paper, we present a neuro-symbolic path planning framework that integrates LLM-based planners with probabilistic model checking to provide guarantees for temporal and probabilistic requirements. We evaluate three state-of-the-art LLMs, assessing success rate, convergence behavior, and performance against a reinforcement learning baseline. Results show that our approach achieves high success rates across diverse grid-world configurations and outperforms the RL baseline.

Read PDF

Similar papers

Preprint Sep 2026

Scenario MPC with STL Specifications and Pareto-Based Feasibility Repair

Temporal logic is a formal language for reasoning about system behaviors over time. Signal temporal logic (STL), in particular, has been used to encode spatio-temporal requirements for control synthesis in multi-agent systems, often under the assumption that agents are cooperative and their dynamics are known. However,...

Tian-Hao Wu, Yiwei Lyu · 0 citations
Preprint Aug 2026

Search-Aided Joint Agent-Environment Reinforcement Learning for Robust Lifelong Multi-Agent Path Finding with Rotations

This work introduces a unified RL formulation that jointly optimizes agent and environment policies, where the environment policy learns graph edge costs to provide global movement guidance via backward Dijkstra search and achieves significant improvements over the strong search-based planner, Causal-PIBT, across multi...

He Jiang, Jingtian Yan, Yulun Zhang et al. · 0 citations
Preprint Aug 2026

Analytic Planning under Uncertainty with Moment Closure

This work reduces target variance and yields well-calibrated predictive uncertainty under stochastic observations in continuous control, providing a principled framework for planning with learned distribution models.

Shishir Sharma, D. Precup · 0 citations
#artificial intelligence Preprint Sep 2026

From Semantic Decisions to Feasible Trajectories: Self-Evolving LLM-Guided Optimal Control for Narrow-Space Parking

Autonomous parking in nonconvex and narrow environments remains challenging. Although optimal-control methods can explicitly enforce vehicle dynamics and collision constraints, nonconvexity compromises solver robustness and can cause failures. Large language models (LLMs) exhibit strong semantic reasoning capabilities,...

Zheng-Bao Yao, Yuan-Fu Luo, Ke-Han Xue · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.