- Text content — the extracted or generated portion of the source document.
- Metadata — contextual information such as source, timestamp, or author.
- Optional vector embeddings — numerical representations used for similarity search and reasoning.
- Semantic search and retrieval
- Knowledge graph relationship extraction
- Vector similarity comparison
- Metadata-based filtering and organization