Language models struggle with compartmentalization - eval data
收藏官方服务:
资源简介:
Precomputed evaluation outputs for the paper Language models struggle with compartmentalization. Includes per-checkpoint validation losses on fineweb, finetune trajectories, cross-compartment cosine similarity sweeps, per-language multilingual curves, and raw training-time val-loss logs for the InfoNCE runs. Drop these into experiment/ of the companion code release (https://github.com/vinhowe/compartmentalization) and generate paper figures via provided plot scripts. See README.txt in this bundle for the file-by-file figure map.
提供机构:
Zenodo创建时间:
2026-05-15



