Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

The information you're seeking appears to be left out of the post. My best guess is that a separate embedding model, specifically tuned for document similarly, is used to generate the vectors and then a clustering algorithm is chosen to create the clusters. They may also use PCA to reduce the embedded vector dimensions before clustering.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: