concept
active
concept:policy-recallPolicy Recall
Reasoning pattern where model explicitly invokes rules or refusal obligations during CoT, associated with lower ASR in QwQ-32B
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Sequence of actions considered by the agent; basis for planning.
- Choosing sequences of actions based on expected free energy; prior probability of policy is softmax of expected free energy
- In active inference, a policy is a sequence of actions through time, as opposed to state-action mappings in RL.
- In reinforcement learning, a policy maps states to actions, specifying behavior at each state.
- The preservation of unrelated model capabilities after a targeted intervention, operationalized via KL divergence on Alpaca
- Machine learning problem, avoided in biology via polycomputing adding new interpretations.
- Data structure generalizing Pair and Stream via indexable containers; demonstrates denotational design on memo structures
- The mechanism by which each step's effect is evaluated against the life of the whole, guiding the unfolding.