Policy-Invariant Reward Shaping from LLM Feedback: A Framework for Hybrid RL Agents
Researchers have proposed a framework for combining large language models (LLMs) with reinforcement learning agents. They formalized the hybrid LLM-pl...
Researchers have proposed a framework for combining large language models (LLMs) with reinforcement learning agents. They formalized the hybrid LLM-pl...
Researchers have found that GPT-style models, which excel at processing language, don't automatically work well for symbolic music. This is because to...
Researchers propose a new method for improving dynamic MRI reconstruction using both magnitude-only and complex measurements. They show that k-space m...
Researchers propose a new data infrastructure for generalist image generation that focuses on organizing heterogeneous supervision according to the de...
Researchers have investigated whether large language models (LLMs) can apply rules in a way that's similar to humans. They ran five experiments with 1...
Researchers have developed a new benchmark for navigation in environments with multiple humans. The HA-VLN 2.0 benchmark includes a standardized task ...
Researchers have developed dynamic shielding technology that can adapt to changing safety requirements in real-time. This is particularly useful for a...
Researchers have proposed a new framework called MoRA for learning geospatial representations at scale. The approach focuses on human activity pattern...
Researchers have proposed a new framework to reduce hallucinations in large language models (LLMs). Hallucinations occur when LLMs make incorrect pred...
Researchers have proposed a new approach to visual navigation called UniWM, which integrates planning and world modeling into a single framework. This...
Researchers propose a framework for planning in environments where conditions change over time. They use Partially Observable Markov Decision Processe...
Researchers have created a safety benchmark called Uni-SafeBench to evaluate the performance of Unified Multimodal Large Models (UMLMs). These models ...