Resource of free step by step video how to guides to get you started with machine learning.
Wednesday, October 7, 2020
Retrieval-Augmented Generation (RAG)
This video explains the Retrieval-Augmented Generation (RAG) model! This approach combines Dense Passage Retrieval with a Seq2Seq BART generator. This is tested out on knowledge intensive tasks like open-domain QA, jeopardy question generation, and FEVER fact verification. This looks like a really interesting paradigm for building language models that produce factually accurate generations! Thanks for watching! Please Subscribe! Paper Links: Original Paper: https://ift.tt/33bHTyO FB Blog Post (Animation used in Intro): https://ift.tt/3kVI4og HuggingFace RAG description: https://ift.tt/3iIO3eh Billion-scale similarity search with GPUs: https://ift.tt/2m8YPRc Language Models as Knowledge Bases? https://ift.tt/2MTGupS REALM: Retrieval-Augmented Language Models: https://ift.tt/2PeBxr4 Dense Passage Retrieval: https://ift.tt/3npz7FC FEVER: https://ift.tt/2F9QadF Natural Questions: https://ift.tt/3izTD2B TriviaQA: https://ift.tt/2SAI2WE MS MARCO: https://ift.tt/2GrcNuR Thanks for watching! Time Stamps 0:00 Introduction 2:05 Limitations of Language Models 4:10 Algorithm Walkthrough 5:48 Dense Passage Retrieval 7:44 RAG-Token vs. RAG-Sequence 10:47 Off-the-Shelf Models 11:54 Experiment Datasets 15:03 Results vs. T5 16:16 BART vs. RAG - Jeopardy Questions 17:20 Impact of Retrieved Documents zi 18:53 Ablation Study 20:25 Retrieval Collapse 21:10 Knowledge Graphs as Non-Parametric Memory 21:45 Can we learn better representations for the Document Index? 22:12 How will Efficient Transformers impact this?
Subscribe to:
Post Comments (Atom)
-
Using GPUs in TensorFlow, TensorBoard in notebooks, finding new datasets, & more! (#AskTensorFlow) [Collection] In a special live ep...
-
Quick answer: Qwen 3 32B needs about 22.7 GB of VRAM at Q4_K_M with a 4k-token context — a used 24 GB card like the RTX 3090 is the minimu...
-
Inside TensorFlow: Summaries and TensorBoard [Collection] Take an inside look into the TensorFlow team’s own internal training sessions-...
-
Hey everyone This is Ujjwal kapil B.tech 2nd year Ai and Ml In this video we covered one of the most important question "COLLEGE KAB O...
-
Quick answer: DeepSeek V4 Flash is a Mixture-of-Experts model: all 284 billion parameters must be stored, so combined RAM plus VRAM — not ...
-
This video is a crash course on understanding how finetuning on LLM models can be performed uing QLORA,LORA, Quantization using LLama2, Grad...
-
Quick answer: For an 8B model at Q4, local electricity costs roughly $0.05 per million output tokens — against about $0.53 per million ble...
-
Quick answer: For most learners, the best single machine learning book is Hands-On Machine Learning with Scikit-Learn, Keras & Tensor...
-
#minecraft #neuralnetwork #backpropagation I built an analog neural network in vanilla Minecraft without any mods or command blocks. The n...
-
In this video, you'll learn how to use machine learning, computer vision and deep learning to create a football analysis system. This pr...
No comments:
Post a Comment