RAG isn't dead, you just need a systematic workflow
Retrieval module, Augmentation module, and Generation module.
LATEST INVESTIGATIONS
Retrieval module, Augmentation module, and Generation module.
Curiosity: What makes Elasticsearch so versatile? How do organizations leverage its powerful search and analytics capabilities across different domain…
Curiosity: What insights can we retrieve from this? How does this connect to innovation in the field?
![["Workflow"]](/assets/img/llm/Llamma-finetune.jpeg)
Curiosity: How can we fine-tune Llama 3 8B to beat GPT-3.5? What makes Salesforce’s approach with Online Iterative RLHF so effective?
Curiosity: Can fine-tuned smaller models outperform large general-purpose models on specific tasks?
A summary of the explanation of the Transformer model is as follows:.
Self-Attention is a method where each word in the input sequence evaluates its relationships with all other words to assign weights.
Deep Dive into LlaMA 3 by Hand ✍️ by Srijanie Dey, PhD May, 2024 Towards Data Science.
Deep Dive into Sora’s Diffusion Transformer (DiT) by Hand ✍︎ by Srijanie Dey, PhD Apr, 2024 Towards Data Science.
Deep Dive into Vector Databases by Hand ✍︎ by Srijanie Dey, PhD Towards Data Science.
Search article titles, categories, and tags. Full text is searched when needed.
Type to search articles.