Do you plan to offer content aware chunking?
Do you have any preferred frameworks?
They compute embeddings using a window of three sentences and then compute distance to find the largest deltas to break up the text into "topics". It is computationally expensive.