finding
active
finding:mismatch-negativityMismatch Negativity
ERP component reproduced by active inference: neural response to prediction violations.
Source paper
extracted_from(2017) · Karl Friston · Thomas FitzGerald · Francesco Rigoli · Philipp Schwartenbeck +1
Neighborhood — ranked by edge-count
Methods (1)
method
- Process by which neuronal dynamics minimize free energy; produces empirically observable neural phenomena.
Frameworks (1)
framework
- Active Inferenceassociated_withFoundational framework by Karl Friston; the paper extends it to three hierarchical levels for modeling meta-awareness.
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Distance between prior knowledge centroid and target pattern centroid, e.g., 1 - cos(eprior, eT).
- Distance between prior and target representations.
- Measures how far the target PT is from the prior P_prior; increases anchoring difficulty
- The broader phenomenon of misaligned behaviors generalizing beyond the fine-tuning distribution
- Attribute: an extreme attempt at undermining, actively contradicting or nullifying a text.
- Rubric-based thresholded GPT-4o grader scoring responses 1-5 on evil intent; scores 4-5 counted as misaligned
- A multi-dimensional characterization of a model's misaligned behaviors across different behavioral categories
- The phenomenon where finetuning on narrow-domain tasks produces broad misalignment extending far beyond the training domain