Desigualdades Geográficas e Impacto Ambiental no Desenvolvimento de Modelos de Inteligência Artificial
Resumo
O avanço da Inteligência Artificial (IA) levanta questões sobre concentração de recursos, desigualdades geográficas e impacto ambiental. Este estudo analisa dados do instituto de pesquisa EpochAI para investigar padrões no desenvolvimento de modelos de fronteira. Por meio de uma análise de clusters, distinguiu-se um grupo acadêmico de um industrial entre os principais desenvolvedores de modelos. Uma análise de regressão linear confirmou que o tamanho do modelo e o hardware utilizado são fatores dominantes para o custo financeiro (R2 = 0,871) e para o impacto ambiental (R2 = 0,762). Concluise que o desenvolvimento de IA é dominado por grandes empresas norteamericanas, mas com contribuições significativas de universidades, evidenciando uma forte concentração de recursos com focos de inovação acadêmica.
Referências
Cottier, B., Rahman, R., Fattorini, L., Maslej, N., and Owen, D. (2024). The rising costs of training frontier ai models. arXiv preprint arXiv:2405.21015.
Epoch AI (2025a). Data on large-scale ai models. [link]. Acessado em: 2025-06-21.
Epoch AI (2025b). Data on machine learning hardware. [link]. Acessado em: 2025-06-21.
Epoch AI (2025c). Data on notable ai models. [link]. Acessado em: 2025-06-21.
Epoch AI (2025d). Epoch AI. [link]. Acessado em: 2026-03-05.
Fox, J. (2016). Applied Regression Analysis and Generalized Linear Models. Sage Publications, Thousand Oaks, CA, 3rd edition.
Frymire, L. and Rahman, R. (2024). The power required to train frontier ai models is doubling annually. [link]. Acessado em: 2026-03-05.
Hair, J. F., Black, W. C., Babin, B. J., and Anderson, R. E. (2018). Multivariate Data Analysis. Cengage, Boston, MA, 8th edition.
Hotelling, H. (1933). Analysis of a complex of statistical variables into principal components. Journal of Educational Psychology, 24(6):417–441.
Liu, F. T., Ting, K. M., and Zhou, Z. H. (2008). Isolation forest. In 2008 Eighth IEEE International Conference on Data Mining, pages 413–422, Pisa, Italy.
MacQueen, J. B. (1967). Some methods for classification and analysis of multivariate observations. In Proceedings of the 5th Berkeley Symposium on Mathematical Statistics and Probability, volume 1, pages 281–297.
Montgomery, D. C. and Runger, G. C. (2018). Applied Statistics and Probability for Engineers. Wiley, Hoboken, NJ, 7th edition.
Rousseeuw, P. J. (1987). Silhouettes: a graphical aid to the interpretation and validation of cluster analysis. Journal of Computational and Applied Mathematics, 20:53–65.
Samborska, V. (2025). Scaling up: How increasing inputs has made artificial intelligence more capable. [link]. Acessado em: 2026-03-05.
Tsoumakas, G. and Katakis, I. (2007). Multi-label classification: An overview. International Journal of Data Warehousing and Mining, 3(3):1–13.
Vake, D., Šinik, B., Šinik, J., and Tošić, A. (2025). Is open source the future of ai? a data-driven approach. Applied Sciences, 15(5):2790.
Wu, C.-J. et al. (2022). Sustainable ai: Environmental implications, challenges and opportunities. In Proceedings of Machine Learning and Systems, volume 4, pages 795–813.
