mirror of
https://github.com/qdrant/qdrant-client.git
synced 2026-07-30 06:31:00 -05:00
* wip: add draft implementation of batch processing * fix: embed dict and list of docs, remove redundant code * new: regen async, small refactor * refactor: add docstrings, rename methods * Upload points local inference (#881) * new: separate single and plural model embeddings * fix: fix lazy embed models * new: add inference object inspections to upload methods * wip: local inference upload parallel * new: add local inference to upload points and upload collection, refactor mixin * fix: remove redundant code * redundant import * tests: check is query for query points batch * refactor: refactor semi ordered map * tests: add test for local inference with batches with docs and vectors * tests: check the order of dict processing * new: distinguish models by options * fix: fix typing * fix: fix types * new: embed batches with different options * tests: add tests for batch with different options * fix: ignore ide incorrect type inspection * tests: wait for points to be inserted * fix: set threads to 1 in parallel inference * new: adjust max internal batch size * fix: fix type hints * function to get embeddings size (#892) * function to get embeddings size * async client * keep sync * new: extend embedding size to support image and late interaction models --------- Co-authored-by: George Panchuk <george.panchuk@qdrant.tech> * new: add local inference batch size (#894) --------- Co-authored-by: Andrey Vasnetsov <andrey@vasnetsov.com>
8.5 KiB
8.5 KiB