PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 29, 2026Optical Memory and Neural Networks0 citations

Caged CrafText: Language-Grounded Safety Constraints for Multimodal Reinforcement Learning in CrafText

View Full Paper
GGG. GorbovDLD. LukashevskyASA. Skrynnik

Key Points

  • The study aims to improve safe reinforcement learning by utilizing natural language for both task objectives and safety constraints.
  • Introduced the Caged CrafText benchmark for multimodal reinforcement learning.
  • Designed two dataset levels: Main for aggregated constraint evaluation, and Debug for diagnosing individual failure modes.
  • Evaluated existing safety approaches against the benchmark to assess safe exploration capabilities.
  • Safe exploration capabilities of existing approaches were found inadequate.
  • Large Language Model agents demonstrated promising reasoning but frequently violated constraints due to grounding issues.

Abstract

Safe reinforcement learning under natural language instructions remains challenging. Current benchmarks are limited by providing only safety constraints in natural language while using structured goal representations, and by focusing on simple navigation tasks in static environments. We introduce Caged CrafText, the first benchmark where both task objectives and safety constraints are specified entirely through natural language. Our benchmark features complex, long-horizon tasks requiring reasoning in dynamic environments, and is structured around two dataset levels: the Main dataset, designed to evaluate language understanding and safe exploration capabilities across aggregated constraint types, and the Debug dataset, which enables detailed diagnosis of specific failure modes in individual constraint scenarios. Evaluation of existing safety approaches reveals their insufficient safe exploration capability, highlighting the need for improved methods capable of handling complex language-guided constraints. Our experiments also demonstrate that while Large Language Model(LLM) agents show promising reasoning capabilities, they suffer from grounding issues that lead to frequent constraint violations.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Gorbov et al. (2026) studied this question.

synapsesocial.com/papers/69c8c15ade0f0f753b39bc1ahttps://doi.org/10.3103/s1060992x25602799
Ask AI
Helpful
Bookmark
Share
View Full Paper