Skip to content

Faster knn chunk merge strategy #4637

Description

@glookka

Proposal:

Currently knn index is fully rebuilt when disk chunks are merged. A faster approach would be to start with the index from the biggest chunk and add vectors from the remaining chunks being merged. That will improve cpu, i/o and/or memory usage (no need to fetch vectors for the bigger chunk) and chunk merge performance. That, in turn, will improve knn search performance until auto-optimize reached the set chunk count target.

Checklist:

To be completed by the assignee. Check off tasks that have been completed or are not applicable.

Details
  • Implementation completed
  • Tests developed
  • Documentation updated
  • Documentation reviewed
  • OpenAPI YAML updated and issue created to rebuild clients
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions