
Transformers Architecture (From Scratch)
Built a complete Transformer neural network with Multi-head attention and Positional encoding. Implemented Encoder-Decoder architecture for sequence-to-sequence NLP tasks, demonstrating deep understanding of LLM foundations.




