CASE contextualizes tabular embeddings via a dataset-anchored Gemma 3 language model to resolve feature semantics, substantially improving tabular learner accuracy especially with scarce data.
FlexTab uses a shared encoder and task-specific decoders for in-context tabular learning, achieving state-of-the-art results on classification, regression, anomaly detection, and entity matching.
TabFORGE introduces a tabular generative foundation model using causality-aware representations and two-stage diffusion-decoder training to generate high-fidelity synthetic data.
This paper benchmarks 2D tabular attention across GPU backends, finding optimal choices vary by row versus column attention, hardware, and sequence length.