Quick Start
When to Use
Good For
- Research papers
- Topic-dense content
- Multi-subject documents
- Quality over speed
Consider Alternatives
- Speed-critical pipelines
- Uniform chunk sizes needed
- Simple structured content
- Very short documents
Parameters
Examples
Research Analysis
Knowledge Base
How It Works
Semantic chunking:- Splits document into sentences
- Generates embeddings for each sentence
- Groups consecutive similar sentences
- Creates new chunk when topic changes
Performance Note
Semantic chunking requires computing embeddings and is slower than token/sentence chunking. Use for quality-sensitive applications where retrieval accuracy matters more than speed.
Embedding Models
The default embedding model isall-MiniLM-L6-v2. You can use any model supported by the chonkie library:
Loading any of these models requires the
[knowledge] extra. Without it, the first search raises ImportError: chonkie package not found. Please install it using: pip install 'praisonaiagents[knowledge]'. If chonkie is installed but the import still fails, the underlying error (usually a broken sentence-transformers or torch install) is preserved as-is. Since PraisonAI PR #5105.
