The Informational Content in Lepto-Variance and Its Relation to Higher Moments *
- 1 University of Limassol, Limassol, Cyprus
Abstract
Lepto-regression is defined as the machine learning process of constructing a Regression Tree of a target feature on itself. It is a novel, model-free method potentially revealing information on important sample structure properties. But it is yet not clear what the informational content of lepto-variance is and how it is related to other well-known statistics of a sample. One significant finding is that 58% of the historical US stock return variability is 1-bit lepto-variance that can not be explained by any financial factor. The central question investigated in this paper is to use small normal N(0, 1) drawn samples to explore how the 1-bit sample lepto-variance and lepto-ratio relate to sample variance, skewness and excess kurtosis. Using a large sample simulation, the lepto ratio of a normal is found to converge to 36.3%. For smaller normally distributed simulated N(0, 1) samples, while lepto-variance itself is highly correlated to sample variance, lepto-variance as a fraction of total variance is highly correlated to excess kurtosis. Both lepto-variance and lepto-ratio are orthogonal to sample skew. Another finding is that while lepto-ratio is strongly correlated to lepto-variance it remains orthogonal to sample variance.
- Bennett, W. R. (1948). Spectra of Quantized Signals. Bell System Technical Journal, 27, 446-472. https://doi.org/10.1002/j.1538-7305.1948.tb01340.x
- Breiman, L., Friedman, J., Olshen, R., & Stone, C. (1984). Classification and Regression Trees . Thomson Wadsworth.
- Demeterfi, K., Derman, E., Kamal, M., & Zou, J. (1999). More Than You Ever Wanted to Know about Volatility Swaps. Goldman Sachs Quantitative Strategies Research Notes , 41, 1-56.
- Fama, E. F., & French, K. R. (1993). Common Risk Factors in the Returns on Stocks and Bonds. Journal of Financial Economics, 33, 3-56. https://doi.org/10.1016/0304-405x(93)90023-5
- Fisher, W. D. (1958). On Grouping for Maximum Homogeneity. Journal of the American Statistical Association, 53, 789-798. https://doi.org/10.1080/01621459.1958.10501479
- Gray, R. M., & Neuhoff, D. L. (1998). Quantization. IEEE Transactions on Information Theory, 44, 2325-2383. https://doi.org/10.1109/18.720541
- Hastie, T., Tibshirani, R., & Friedman, J. (2009). Elements of Statistical Learning . Springer.
- Jenks, G. F., & Caspall, F. C. (1971). Error on Choroplethic Maps: Definition, Measurement, Reduction. Annals of the Association of American Geographers, 61, 217-244. https://doi.org/10.1111/j.1467-8306.1971.tb00779.x
- Oliver, B. M., Pierce, J. R., & Shannon, C. E. (1948). The Philosophy of PCM. Proceedings of the IRE, 36, 1324-1331. https://doi.org/10.1109/jrproc.1948.231941
- Polimenis, V. (2022). The Lepto-Variance of Stock Returns. In Proceedings of the 34th Panhellenic Statistics Conference (pp. 167-182). Greek Statistical Institute.
- Polimenis, V. (2024). The Historical Lepto-Variance of the US Stock Returns. Data Science in Finance and Economics, 4, 270-284. https://doi.org/10.3934/dsfe.2024011
- Ripley, B. D. (1996). Pattern Recognition and Neural Networks . Cambridge University Press. https://doi.org/10.1017/cbo9780511812651
- Whaley, R. E. (1993). Derivatives on Market Volatility. The Journal of Derivatives, 1, 71-84. https://doi.org/10.3905/jod.1993.407868