Published July 17, 2025 | Version 1

A Holistic Approach to Transformer-Based Models: Architecture, Applications, and Ethical Layers

  • 1.  Mugla Sıtkı Kocman University, Institute of Science, Department of Artificial Intelligence, Mugla, Turkiye
  • 2.  Mugla Sıtkı Kocman University, Faculty of Science, Department of Statistics, Mugla, Turkiye

Description

This section presents a detailed description of transformer-based models, which have revolutionized the field of artificial intelligence, particularly in natural language processing, vision, speech, and multimodal applications, among others. This chapter begins with a description of the limitations of older architectures, such as Recurrent Neural Networks (RNNs) and Long Short-Term Memory (LSTM) networks, which were widely used before the advent of transformers, and how transformers addressed these deficiencies. The Vaswani et al. transformer model, introduced in 2017, offers a successful solution to such concerns, particularly through its parallel processing and self-attention mechanisms. The chapter covers the essential components of transformers, such as the multi-headed attention mechanism, positional encoding, residual connections, and the encoder-decoder model. Then it focuses on current leading models, such as BERT, GPT (through GPT-4), DALL·E, Claude 3, Gemini, LLaMA 3, Mistral, Whisper, and Sora. Concerns such as how these systems are powered, where they excel, and the type of output they produce are presented with examples. In addition to the technical data, the social, environmental, and ethical problems for which transformers provide solutions also come into focus. Serious issues, such as artificial intelligence sometimes producing content contrary to reality and not being grounded in reality, i.e., situations referred to as hallucinations; producing skewed results; inducing copyright infringement; posing a threat to personal data; and the detrimental effects of massive models on the environment, are being addressed. For example, a machine learning algorithm would generate a news title reporting accurately on a non-existent incident, demonstrating just how dangerous the threat of hallucinations can be when it comes to critical applications. Here, what is emphasized is that technology professionals must not only be capable of innovating but also behave by values such as transparency, fairness, ethics, and sustainability. Furthermore, the section offers practical advice on which models might be more suitable for other applications, such as creative text writing, classification, multi-modal tasks, or speech content generation. Specific key points that need to be remembered to minimize risks are also conveyed to the readers. Finally, the importance of collaboration among developers, decision-makers, and society to facilitate this growth in a manner that benefits society is highlighted.

Files

cp1.pdf

Files (693.9 kB)

Name Size Download all
md5:1442816d3800ebe6882d77bc402f0ab5
693.9 kB Preview Download

Additional details

Dates

Available
2025-07-17

References

  • Demir, S.T., & Gökçe Narin, N. (2025). A Holistic Approach to Transformer-Based Models: Architecture, Applications, and Ethical Layers. In O. Aydin & E. Karaarslan (Eds.), The Age of Generative Artificial Intelligence (pp. 1-27). Izmir Academy Association. Doi: 10.5281/zenodo.16008405