Uploaded September 2025 | Updated September 2026, 3 hours ago
Why does every RAG pipeline start with chunking? Because chunking defines what your vectors mean.
At its core, ๐ฐ๐ต๐๐ป๐ธ๐ถ๐ป๐ด is the preprocessing step of splitting texts into smaller pieces - and each chunk becomes the unit of information that gets vectorized and stored in your vector database.
In this short video, Femke breaks down simple chunking methods โ token, sentence, and document-based.
๐ย Get your copy of the free advanced RAG ebook: weaviate.io/ebooks/advanced-rag-techniques?utm_source=youtube&utm_campaign=rag&utm_content=680991368
Chapters:
00:00:00 - Why Large Docs Challenge AI Models
00:00:17 - Token-Chunking
00:00:29 - Sentence-Chunking for Better Context
00:00:45 - Document-Based Chunking Benefits & Limits
00:01:03 - Combining Chunking Methods
00:01:09 - Smarter Chunking Approaches
00:01:18 - Next Steps & Additional Resources
Paper review video: Late chunking improves context recall in RAG pipelines
youtube.com/watch?v=buzWGXOydD8
โฌโฌโฌโฌโฌโฌโฌโฌโฌโฌโฌโฌ CONNECT WITH US โฌโฌโฌโฌโฌโฌโฌโฌโฌโฌโฌโฌ
- Visit weaviate.io
- Star us on GitHub github.com/weaviate/weaviate
- Stay updated and subscribe to our newsletter: newsletter.weaviate.io
- Try out Weaviate Cloud Services for free here: https://console.weaviate.cloud/
Got a question?
- Forum: forum.weaviate.io
- Slack: weaviate.io/slack
Connect with us on
- Twitter: twitter.com/weaviate_io
- LinkedIn: linkedin.com/company/weaviate-io
Why does every RAG pipeline start with chunking? Because chunking defines what your vectors mean.
At its core, ๐ฐ๐ต๐๐ป๐ธ๐ถ๐ป๐ด is the preprocessing step of splitting texts into smaller pieces - and each chunk becomes the unit of information that gets vectorized and stored in your vector database.
In this short video, Femke breaks down simple chunking methods โ token, sentence, and document-based.
๐ย Get your copy of the free advanced RAG ebook: weaviate.io/ebooks/advanced-rag-techniques?utm_source=youtube&utm_campaign=rag&utm_content=680991368
Chapters:
00:00:00 - Why Large Docs Challenge AI Models
00:00:17 - Token-Chunking
00:00:29 - Sentence-Chunking for Better Context
00:00:45 - Document-Based Chunking Benefits & Limits
00:01:03 - Combining Chunking Methods
00:01:09 - Smarter Chunking Approaches
00:01:18 - Next Steps & Additional Resources
Paper review video: Late chunking improves context recall in RAG pipelines
youtube.com/watch?v=buzWGXOydD8
โฌโฌโฌโฌโฌโฌโฌโฌโฌโฌโฌโฌ CONNECT WITH US โฌโฌโฌโฌโฌโฌโฌโฌโฌโฌโฌโฌ
- Visit weaviate.io
- Star us on GitHub github.com/weaviate/weaviate
- Stay updated and subscribe to our newsletter: newsletter.weaviate.io
- Try out Weaviate Cloud Services for free here: https://console.weaviate.cloud/
Got a question?
- Forum: forum.weaviate.io
- Slack: weaviate.io/slack
Connect with us on
- Twitter: twitter.com/weaviate_io
- LinkedIn: linkedin.com/company/weaviate-io
![Reimagine Data Workflows with Weaviate Agents
Calling all AI devs, tech leads, and novices experimenting with agentic AI!
Join us for a hands-on walkthrough of Weaviate Agents โ a powerful suite of services designed to simplify and automate your AI data workflows.
In this live session, well showcase how the ๐๐ฎ๐๐ซ๐ฒ ๐๐ ๐๐ง๐ญ, ๐๐ซ๐๐ง๐ฌ๐๐จ๐ซ๐ฆ๐๐ญ๐ข๐จ๐ง ๐๐ ๐๐ง๐ญ, ๐๐ง๐ ๐๐๐ซ๐ฌ๐จ๐ง๐๐ฅ๐ข๐ณ๐๐ญ๐ข๐จ๐ง ๐๐ ๐๐ง๐ญ work under the hood to enable natural language querying, real-time data transformation, and context-aware personalization without heavy lifting.
This session will include live demos, example use cases, and implementation tips.
[๐๐ฆ๐ฉ๐จ๐ซ๐ญ๐๐ง๐ญ] ๐๐ ๐ฒ๐จ๐ฎ ๐ฅ๐ข๐ค๐ ๐ญ๐จ ๐ซ๐๐๐๐ข๐ฏ๐ ๐ฆ๐๐ญ๐๐ซ๐ข๐๐ฅ๐ฌ ๐ฎ๐ฌ๐๐ ๐๐ฎ๐ซ๐ข๐ง๐ ๐ญ๐ก๐ ๐ฌ๐๐ฌ๐ฌ๐ข๐จ๐ง ๐๐ง๐ ๐๐จ๐ฅ๐ฅ๐จ๐ฐ-๐ฎ๐ฉ ๐ข๐ง๐๐จ๐ซ๐ฆ๐๐ญ๐ข๐จ๐ง ๐ฆ๐๐ค๐ ๐ฌ๐ฎ๐ซ๐ ๐ญ๐จ ๐๐ ๐ซ๐๐ ๐ญ๐จ ๐๐๐๐ฏ๐ข๐๐ญ๐๐ฌ ๐๐ซ๐ข๐ฏ๐๐๐ฒ ๐๐จ๐ฅ๐ข๐๐ฒ Reimagine Data Workflows with Weaviate Agents](https://i.ytimg.com/vi/HMxhS8jyDYA/mqdefault.jpg)









