Transformers for Natural Language Processing and Computer Vision, Third Edition, explores Large Language Model (LLM) architectures, applications, and various platforms (Hugging Face, OpenAI, and Google Vertex AI) used for Natural Language Processing (NLP) and Computer Vision (CV).
The book guides you through different transformer architectures to the latest Foundation Models and Generative AI. You’ll pretrain and fine-tune LLMs and work through different use cases, from summarization to implementing question-answering systems with embedding-based search techniques. You will also learn the risks of LLMs, from hallucinations and memorization to privacy, and how to mitigate such risks using moderation models with rule and knowledge bases. You’ll implement Retrieval Augmented Generation (RAG) with LLMs to improve the accuracy of your models and gain greater control over LLM outputs.
Dive into generative vision transformers and multimodal model architectures and build applications, such as image and video-to-text classifiers. Go further by combining different models and platforms and learning about AI agent replication.
This book provides you with an understanding of transformer architectures, pretraining, fine-tuning, LLM use cases, and best practices.
Paga fácilmente con tarjeta, Klarna, Apple Pay o Google Pay. ¿No estás contento? Siempre tienes 14 días de garantía de devolución. Lee más en nuestros términos. Si tienes preguntas, escríbenos a hello@memmo.org.
Memmo hace que estudiar sea más fácil, estés donde estés. Aquí tienes tus libros de texto y herramientas de estudio inteligentes en un solo lugar: resúmenes, quizzes, podcasts y flashcards. Y también a Ted, tu compañero de estudio que responde a todo lo que te preguntes. Más de 50 000 estudiantes ya estudian aquí. Está hecho para que aprendas más rápido y te estreses menos.