Hacker News new | past | comments | ask | show | jobs | submit
Are you using blind chunking or section aware chunking?
for the example here the chunking is section aware -> but the general training data synthesis pipeline is agnostic to type of chunking
if chunking is section aware, how do you manage large section embeddings?