Federated concept-based models aggregate distributed concept annotations across institutions, adapt architectures to evolving supervision, and enable interpretable inference for locally unavailable concepts while preserving privacy.
SG-Ego extends Ego4D with time-evolving scene graphs and GLEN reasons over them to model activity-driven scene dynamics, outperforming raw video and MLLM baselines on retrieval and long-horizon reasoning.