The Transformer is the neural network architecture behind today's LLMs, introduced in 2017. Its attention mechanism processes all the words of a text in parallel while weighing their relationships, where previous architectures read sequentially. The “T” in GPT is precisely that.
In practice at Gensai
All the models Gensai integrates (Mistral, Claude, GPT, Gemini) are built on variants of this architecture.
Same theme
LLM (Large Language Model)SLM (Small Language Model)DLLM (Diffusion Large Language Model)Hallucination (AI)Prompt chainingPre-trained modelFoundation modelGenerative AITokenContext windowPromptPrompt engineeringContext (AI)Memory (AI)Fine-tuningOpen-weightQuantizationDistillationLoRA (Low-Rank Adaptation)Mixture of Experts (MoE)Reasoning modelTemperature (AI)
Going beyond the definition?
From concept to project: Gensai builds custom, sovereign, GDPR-compliant AI solutions.
55 boulevard de Strasbourg, 75010 Paris