Chunk overlap is the number of tokens or characters duplicated between consecutive chunks when splitting a document. Without overlap, sentences or ideas that span a chunk boundary may be split across two chunks, with neither containing the complete information. Typical overlaps range from 10–20% of chunk size. Too much overlap wastes storage and embedding compute; too little risks losing boundary information.