Transformers for Natural Language Processing and Computer Vision, Third Edition, explores Large Language Model (LLM) architectures, applications, and various platforms (Hugging Face, OpenAI, and Google Vertex AI) used for Natural Language Processing (NLP) and Computer Vision (CV).
The book guides you through different transformer architectures to the latest Foundation Models and Generative AI. You’ll pretrain and fine-tune LLMs and work through different use cases, from summarization to implementing question-answering systems with embedding-based search techniques. You will also learn the risks of LLMs, from hallucinations and memorization to privacy, and how to mitigate such risks using moderation models with rule and knowledge bases. You’ll implement Retrieval Augmented Generation (RAG) with LLMs to improve the accuracy of your models and gain greater control over LLM outputs.
Dive into generative vision transformers and multimodal model architectures and build applications, such as image and video-to-text classifiers. Go further by combining different models and platforms and learning about AI agent replication.
This book provides you with an understanding of transformer architectures, pretraining, fine-tuning, LLM use cases, and best practices.
Betala smidigt med kort, Klarna, Apple Pay eller Google Pay. Är du inte nöjd har du alltid 14 dagars ångerrätt. Läs mer i våra villkor. Har du några frågor, mejla oss på hello@memmo.org.
Memmo gör det enklare att plugga – var du än är i världen. Hos oss samlar du kursböcker och smarta studieverktyg på ett och samma ställe: sammanfattningar, quiz, poddar och flashcards. Och så Ted, din studiekompis som svarar på allt du undrar. Över 50 000 studenter pluggar redan här – byggt för att du ska lära dig snabbare och stressa mindre.