Getting your data ready for RAG systems can be tricky, but it doesn’t have to be a headache. In this session, I’ll share some practical tips from my projects, on how to make this process smoother. We’ll cover various methods of breaking your text into chunks for better LLM context. I’ll show you how I format documents for better results and how to handle references between paragraphs, documents, and images. We’ll also tackle how I process more complex documents, like legal texts or websites to make those more machine-friendly, without losing important details. Lastly, I’ll give you insights on tweaking human-created content to make it work better for machines. There will be memes
feeding rag systems with data without a headachee
Future Conf 2024
Wydarzenie i data
Future Conf 2024 · · Lubicz Park, Kraków, PL