Planning and acting in partially observable stochastic domains
From MaRDI portal
(Redirected from Publication:72343)
Recommendations
- Planning in partially-observable switching-mode continuous domains
- scientific article; zbMATH DE number 5547912
- Strong planning under partial observability
- Learning and planning in partially observable environments without prior domain knowledge
- Replanning in domains with partial information and sensing actions
- Randomized belief-space replanning in partially-observable continuous spaces
Cites work
- A survey of algorithmic methods for partially observed Markov decision processes
- A survey of solution techniques for the partially observed Markov decision process
- Application of Jensen's inequality to adaptive suboptimal design
- Fast planning through planning graph analysis
- scientific article; zbMATH DE number 3148886 (Why is no real title available?)
- scientific article; zbMATH DE number 4089320 (Why is no real title available?)
- scientific article; zbMATH DE number 700091 (Why is no real title available?)
- scientific article; zbMATH DE number 1095138 (Why is no real title available?)
- scientific article; zbMATH DE number 795587 (Why is no real title available?)
- OPTIMAL CONTROL FOR PARTIALLY OBSERVABLE MARKOV DECISION PROCESSES OVER AN INFINITE HORIZON
- Optimal control of Markov processes with incomplete state information
- Solution Procedures for Partially Observed Markov Decision Processes
- Solving H-horizon, stationary Markov decision problems in time proportional to log (H)
- State of the Art—A Survey of Partially Observable Markov Decision Processes: Theory, Models, and Algorithms
- The complexity of mean payoff games on graphs
- The complexity of stochastic games
- The Optimal Control of Partially Observable Markov Processes over a Finite Horizon
- The Optimal Control of Partially Observable Markov Processes over the Infinite Horizon: Discounted Costs
- The Optimal Search for a Moving Target When the Search Path Is Constrained
Cited in
(only showing first 100 items - show all)- Transfer in variable-reward hierarchical reinforcement learning
- Partially observable Markov decision processes with imprecise parameters
- A tutorial on partially observable Markov decision processes
- Permissive planning: Extending classical planning to uncertain task domains.
- Abstraction and approximate decision-theoretic planning.
- Stochastic dynamic programming with factored representations
- Finite-horizon LQR controller for partially-observed Boolean dynamical systems
- Autonomous agents modelling other agents: a comprehensive survey and open problems
- Open problems in universal induction \& intelligence
- Planning in hybrid relational MDPs
- Reasoning and predicting POMDP planning complexity via covering numbers
- Computation of weighted sums of rewards for concurrent MDPs
- Markov decision processes with sequential sensor measurements
- Task-structured probabilistic I/O automata
- Reinforcement learning with limited reinforcement: using Bayes risk for active learning in POMDPs
- A semi-Markov decision model for recognizing the destination of a maneuvering agent in real time strategy games
- Probabilistic may/must testing: retaining probabilities by restricted schedulers
- Policy iteration for bounded-parameter POMDPs
- An affective mobile robot educator with a full-time job
- Counterexample-guided inductive synthesis for probabilistic systems
- Knowledge-based programs as succinct policies for partially observable domains
- Partially observable environment estimation with uplift inference for reinforcement learning based recommendation
- Learning and planning in partially observable environments without prior domain knowledge
- Optimizing active surveillance for prostate cancer using partially observable Markov decision processes
- Simplified risk-aware decision making with belief-dependent rewards in partially observable domains
- Gradient-based mixed planning with symbolic and numeric action parameters
- Soft rumor control in mobile instant messengers
- Analyzing generalized planning under nondeterminism
- Gradient-descent for randomized controllers under partial observability
- Simultaneous learning and planning in a hierarchical control system for a cognitive agent
- Learning to steer nonlinear interior-point methods
- Dynamic optimization over infinite-time horizon: web-building strategy in an orb-weaving spider as a case study
- Multi-goal motion planning using traveling salesman problem in belief space
- Privacy stochastic games in distributed constraint reasoning
- Probabilistic reasoning about epistemic action narratives
- Algorithms and conditional lower bounds for planning problems
- A survey of inverse reinforcement learning: challenges, methods and progress
- An integrated approach to solving influence diagrams and finite-horizon partially observable decision processes
- Deliberative acting, planning and learning with hierarchical operational models
- Computer science and decision theory
- Recursively modeling other agents for decision making: a research perspective
- Partially observable game-theoretic agent programming in Golog
- Reasoning about uncertain parameters and agent behaviors through encoded experiences and belief planning
- Regression and progression in stochastic domains
- A Fenchel-Moreau-Rockafellar type theorem on the Kantorovich-Wasserstein space with applications in partially observable Markov decision processes
- A dynamic epistemic framework for reasoning about conformant probabilistic plans
- POMDPs under probabilistic semantics
- Meeting a deadline: shortest paths on stochastic directed acyclic graphs with information gathering
- Strong planning under uncertainty in domains with numerous but identical elements (a generic approach)
- Dynamic multiagent probabilistic inference
- Cross-entropic learning of a machine for the decision in a partially observable universe
- Recognizing and learning models of social exchange strategies for the regulation of social interactions in open agent societies
- Representations for robot knowledge in the \textsc{KnowRob} framework
- Robotic manipulation of multiple objects as a POMDP
- Geometric backtracking for combined task and motion planning in robotic systems
- State observation accuracy and finite-memory policy performance
- Strong planning under partial observability
- Performance prediction of an unmanned airborne vehicle multi-agent system
- Heuristic anytime approaches to stochastic decision processes
- The complexity of agent design problems: Determinism and history dependence
- Optimal cost almost-sure reachability in POMDPs
- Partially observable multistage stochastic programming
- Supervisor synthesis of POMDP via automata learning
- Optimal management of stochastic invasion in a metapopulation with Allee effects
- Quantitative controller synthesis for consumption Markov decision processes
- Rationalizing predictions by adversarial information calibration
- Partial-order planning with concurrent interacting actions
- An evidential approach to SLAM, path planning, and active exploration
- scientific article; zbMATH DE number 1728768 (Why is no real title available?)
- Goal-directed learning of features and forward models
- Multi-tasking arbitration and behaviour design for human-interactive robots
- Learning where to attend with deep architectures for image tracking
- Bounded-parameter partially observable Markov decision processes: framework and algorithm
- DESPOT: online POMDP planning with regularization
- Evidential Markov decision processes
- Integration of reinforcement learning and optimal decision-making theories of the basal ganglia
- Randomized belief-space replanning in partially-observable continuous spaces
- An educational management problem with continuous signal space
- Efficient planning under uncertainty with macro-actions
- Optimal speech motor control and token-to-token variability: a Bayesian modeling approach
- Integrated common sense learning and planning in POMDPs
- scientific article; zbMATH DE number 4174347 (Why is no real title available?)
- A two-state partially observable Markov decision process with three actions
- Solving for Best Responses and Equilibria in Extensive-Form Games with Reinforcement Learning Methods
- scientific article; zbMATH DE number 3871010 (Why is no real title available?)
- A synthesis of automated planning and reinforcement learning for efficient, robust decision-making
- Systems of Bounded Rational Agents with Information-Theoretic Constraints
- Locally-connected interrelated network: a forward propagation primitive
- Myopic bounds for optimal policy of POMDPs: an extension of lovejoy's structural results
- Exploiting expert knowledge in factored POMDPs
- scientific article; zbMATH DE number 4166888 (Why is no real title available?)
- Active inference and agency: optimal control without cost functions
- An online multi-agent co-operative learning algorithm in POMDPs
- POMDP planning for robust robot control
- Posterior weighted reinforcement learning with state uncertainty
- Planning for multiple measurement channels in a continuous-state POMDP
- A performance gradient perspective on gradient‐based policy iteration and a modified value iteration
- scientific article; zbMATH DE number 5547912 (Why is no real title available?)
- scientific article; zbMATH DE number 5547961 (Why is no real title available?)
- Probabilistic Reasoning by SAT Solvers
This page was built for publication: Planning and acting in partially observable stochastic domains
Report a bug (only for logged in users!)Click here to report a bug for this page (MaRDI item Q72343)