Experiments in Organizing Information for Optimal Retrieval
Originally Published on my Substack
So recently, I have been experimenting with knowledge bases, and how information is structured is vital.
In some sense, it directly relates to the context and overall meaning. It also makes it easier for LLMs to answer questions accurately and reduces hallucinations.
And so we created a few experiments that you can play with and test yourself.
But first, why does information structure matter?
Imagine a library. One room has books strewn everywhere, titles mixed up, and no discernible order. In another room, the books are organized by subject, then by author, then by publication date. Which room would help you find the book you want more efficiently?
LLMs are similar. The way knowledge is presented and organized dictates not only their understanding but also their output.
Thank you for reading AI for Humanity with Stefan Speaks. This post is public so feel free to share it.
Our recent experiments highlighted two crucial aspects:
1. Organization of Informational Taxonomies and Hierarchies: This considers elements like URL structures, folders, and how information is interrelated. By defining the proper context, you can highlight what’s critical and organize it appropriately.
2. Organization Within a Document: Delving deeper, this looks at the composition of individual pieces of information, from structure and semantics to formatting and summaries.
Let’s dive into each experiment and the results.
To make this more fun and interactive, I am providing the documents and URLs used to train the AI. Additionally, you talk to the AI Agents, Good Bot and Bad Bot, and experience the difference for yourself.



