Online Replanning in Belief Space for Partially Observable Task and Motion Problems
-
Updated
Oct 18, 2022 - Python
Online Replanning in Belief Space for Partially Observable Task and Motion Problems
SymDer: Symbolic Derivative Approach to Discovering Sparse Interpretable Dynamics from Partial Observations
Solving pursuit-evasion problems on graphs using Reinfocement Learning and GNNs
Official PyTorch implementation of POEM (Partial Observation Experts Modelling) as introduced in the paper Contrastive Meta-Learning for Partially Observable Few-Shot Learning
Code for a multi-agent particle environment used in the paper "Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments"
Partially Observable Monte Carlo Planning algorithm (POMCP)
Comparing a full-map A* oracle with partially observable knowledge-based and genetic-policy agents in Wumpus World.
Hybrid planning agent for decision-making under partial observability (PG-WMA)
Decentralized traffic signal control under camera-limited observability (IEEE Access submission)
End-to-end empirical implementation of model-based off-policy evaluation and pessimistic policy selection for confounded POMDPs (Hong, Qi & Xu, ICML 2024)
Defender-side cyber incident simulator. An abstract state machine, not real infrastructure, and reinforcement-learning policies, not LLM agents. Host status is hidden behind noisy alerts, the attacker improvises, and a web console lets you play the same scenario by hand and compare your score to a trained agent's.
Minimal control-theoretic toy models for studying stability and failure modes under partial observability.
Priced information-gathering and abstention evaluation in a synthetic partially observed city.
Code and records for "Observational Aliasing Bounds Generalization in Constrained Active-Sensing MARL" (ICET 2026).
Controlled evaluation of persistent object memory for partially observable embodied agents in AI2-THOR.
Operator-geometric framework for metabolic identifiability and metabolite panel design under partial observability using Human-GEM, proteomics-informed reaction weighting, and spectral geometry.
Autonomous crewmate AI in a partially observable social-deduction simulator with strict information boundaries, belief modeling, and strategic RL.
Bachelor's thesis: reproducible multi-seed MiniGrid benchmark of PPO, A2C, DQN and RecurrentPPO under partial observability — memory, intrinsic motivation (RND, ICM, RIDE, NovelD) and curriculum learning.
Python library for studying elimination-based learning in deterministic MDPs under state aliasing
To associate your repository with the partial-observability topic, visit your repo's landing page and select "manage topics."