Question 1 · choose 1
An aircraft maintenance team searches 3,000-page technical manuals. Technicians type precise questions such as "torque value for the left actuator bolt", and evaluations show that small chunks find the right passage but the model then lacks the surrounding procedure, while large chunks return the procedure but rank poorly. The manuals have a clear section structure. Which chunking strategy should the developer use for the knowledge base?
- ANo chunking, after splitting each manual into one separate file per chapter beforehand
- BSemantic chunking with a high breakpoint percentile threshold and a larger buffer size
- CFixed-size chunking with 100-token chunks and a 50% overlap between consecutive chunks
- DHierarchical chunking with small child chunks and larger parent chunks
Show the answer and why
ANo chunking, after splitting each manual into one separate file per chapter beforehand
Incorrect
With no chunking each file becomes one chunk, so whole chapters are embedded as single vectors, which hurts precision and inflates the context sent to the model.
BSemantic chunking with a high breakpoint percentile threshold and a larger buffer size
Incorrect
Semantic chunking groups sentences by meaning, which helps unstructured text, but a high threshold just makes chunks larger and reintroduces the ranking problem. It has no parent expansion.
CFixed-size chunking with 100-token chunks and a 50% overlap between consecutive chunks
Incorrect
Very small fixed chunks with heavy overlap improve matching but still hand the model only a fragment of the procedure.
DHierarchical chunking with small child chunks and larger parent chunks
Correct
Hierarchical chunking searches on small child chunks for precise matching and then replaces them with their parent chunks, so the model receives the broader procedure around the match.
The symptoms describe the precision versus context trade-off that hierarchical chunking was built for: match on small children, answer with the larger parent. Expect fewer results than numberOfResults, because children that share a parent collapse into one.
AWS documentation