Inner_Language_Model_Level_Data
Description
This repository contains two datasets supporting the systematic review “Emergent Inner Languages in AI: A Systematic Review of Cross-Lingual Representation Structures in Large Language Models.” The datasets contain compiled secondary-source evidence used for the model-level and dimension-level analyses reported in the manuscript.This repository contains two datasets supporting the systematic review “Emergent Inner Languages in AI.” The datasets include compiled secondary-source evidence on LLM characteristics, cross-lingual representations, probing results, CKA similarity, Inner-Language Scores, and related analyses. All data were derived from published and publicly available sources cited in the manuscript and are provided to support transparency and reproducibility.