Alessio Brini, Ekaterina Seregina
arXiv 22 Jan 2026 · Econometrics
arXiv:2601.16274 · PDF · DOI · OpenAlex · Extracted main text
We propose Mixed-Panels-Transformer Encoder (MPTE), a novel framework for estimating factor models in panel datasets with mixed frequencies and nonlinear signals. Traditional factor models rely on linear signal extraction and require homogeneous sampling frequencies, limiting their applicability to modern high-dimensional datasets where variables are observed at different temporal resolutions. Our approach leverages Transformer-style attention mechanisms to enable context-aware signal construction through flexible, data-dependent weighting schemes that replace fixed linear combinations with adaptive reweighting based on similarity and relevance. We extend classical principal component analysis (PCA) to accommodate general temporal and cross-sectional attention matrices, allowing the model to learn how to aggregate information across frequencies without manual alignment or pre-specified weights. For linear activation functions, we establish consistency and asymptotic normality of factor and loading estimators, showing that our framework nests Target PCA as a special case while providing efficiency gains through transfer learning across auxiliary datasets. The nonlinear extension uses a Transformer architecture to capture complex hierarchical interactions while preserving the theoretical foundations. In simulations, MPTE demonstrates superior performance in nonlinear environments, and in an empirical application to 13 macroeconomic forecasting targets using a selected set of 48 monthly and quarterly series from the FRED-MD and FRED-QD databases, our method achieves competitive performance against established benchmarks. We further analyze attention patterns and systematically ablate model components to assess variable importance and temporal dependence. The resulting patterns highlight which indicators and horizons are most influential for forecasting.
appendix boundary found by appendix_command · 73% of the source is main text. Read the extracted text to check this.
The works this paper leans on most, across its whole bibliography — not restricted to papers in our corpus. Ranked by composite intensity, which combines how often a work is mentioned, how many sections mention it, and how much of that falls in the main text rather than the appendix.
| Reference | Intensity | Mentions | Sections | Main text | |
|---|---|---|---|---|---|
| 1 | Duan, Junting and Pelger, Markus and Xiong, Ruoxuan (2024) Target PCA: Transfer learning large dimensional panel data | 0.985 | 23 | 5 | 96% |
| 2 | Gu, Shihao and Kelly, Bryan and Xiu, Dacheng (2021) Autoencoder asset pricing models | 0.874 | 5 | 2 | 100% |
| 3 | Vaswani, Ashish and Shazeer, Noam and Parmar, Niki and Uszkoreit, Ja… (2017) Attention is all you need | 0.874 | 5 | 2 | 100% |
| 4 | Lin, Jiahe and Michailidis, George (2024) A multi-task encoder-dual-decoder framework for mixed frequency data prediction | 0.737 | 3 | 2 | 100% |
| 5 | Bai, Jushan (2003) Inferential theory for factor models of large dimensions | 0.511 | 2 | 2 | 50% |
| 6 | Fan, Jianqing and Liao, Yuan and Wang, Weichen (2016) Projected principal component analysis in factor models | 0.511 | 2 | 1 | 100% |
| 7 | Ghysels, Eric and Sinko, Arthur and Valkanov, Rossen (2007) MIDAS regressions: Further results and new directions | 0.511 | 2 | 1 | 100% |
| 8 | Nadaraya, Elizbar A (1964) On estimating regression | 0.511 | 2 | 1 | 100% |
| 9 | Michael W. McCracken and Serena Ng (2016) FRED-MD: A Monthly Database for Macroeconomic Research | 0.405 | 1 | 1 | 100% |
| 10 | Bahdanau, Dzmitry and Cho, Kyunghyun and Bengio, Yoshua (2014) Neural machine translation by jointly learning to align and translate | 0.405 | 1 | 1 | 100% |
Showing the top 10 of 42 scored citations.