遇见数据集

Language models struggle with compartmentalization - eval data

收藏
Zenodo2026-05-15 更新2026-05-26 收录
官方服务:

资源简介:

Precomputed evaluation outputs for the paper Language models struggle with compartmentalization. Includes per-checkpoint validation losses on fineweb, finetune trajectories, cross-compartment cosine similarity sweeps, per-language multilingual curves, and raw training-time val-loss logs for the InfoNCE runs. Drop these into experiment/ of the companion code release (https://github.com/vinhowe/compartmentalization) and generate paper figures via provided plot scripts. See README.txt in this bundle for the file-by-file figure map.

提供机构:
Zenodo
创建时间:
2026-05-15
二维码
社区交流群
二维码
科研交流群
商业服务