Commit Graph
23 Commits
Author SHA1 Message Date
Yuanqing ZHAOandYuanqing Zhao 2d84a6f13c perf(cuvs): improve recovery scalability and filtered-search efficiency (#3311)
* perf(index): avoid materializing descendant path strings

* perf(cuvs): compact host vector shadow

* perf(storage): prune redundant tenant path scopes

* fix(cuvs): synchronize worker-thread searches

* perf(cuvs): stream dense shadow recovery

* test(storage): cover cross-user path scopes

* perf(storage): page candidate recovery scans

* perf(cuvs): accelerate adaptive filter routing

* perf(cuvs): reduce filtered-search host overhead

---------

Co-authored-by: Yuanqing Zhao <2604121+yuanqingz@users.noreply.github.com>
2026-07-21 15:59:58 +08:00
Yuanqing ZHAOandYuanqing Zhao 7e6a0515f9 perf(cuvs): optimize filters, rebuilds, concurrency, and memory (#3092)
* perf(cuvs): fast-path cached native filter routes

* perf(cuvs): parallelize auto filter preflight

* perf(cuvs): add search route telemetry

* test(cuvs): use a valid telemetry vector dimension

* perf(cuvs): reuse native filter preflight results

* perf(cuvs): allow concurrent snapshot searches

* perf(cuvs): coalesce optional background rebuilds

* perf(cuvs): coordinate per-GPU build admission

* perf(cuvs): add opt-in float16 search

* build(cuvs): support vector benchmark harnesses

* perf(cuvs): bound concurrent GPU searches

* perf(cuvs): avoid partial background rebuilds

* fix(cuvs): address rebuild and telemetry review feedback

* fix(cuvs): defer rebuild until index initialization

---------

Co-authored-by: Yuanqing Zhao <2604121+yuanqingz@users.noreply.github.com>
2026-07-10 17:22:34 +08:00
Qin Haojie 2eb61fbbad feat(benchmark): 增加 OV 目录向量检索性能 benchmark (#3076)
补充基于 VikingVectorIndexBackend 的 dir-vector/synthetic benchmark,并修复本地 bitmap 读路径隐藏写入导致的并发检索崩溃。
2026-07-08 14:49:57 +08:00
Yuanqing ZHAOandYuanqing Zhao 39c778c953 feat: add cuVS vector search backend (#2974)
* feat: add cuVS vector search backend

* docs: add agent memory benchmark strategy

* bench: add cuVS index performance harness

* bench: add public ANN dataset tuning

* docs: record preliminary cuVS index results

* docs: clarify warm index latency

* docs: order cuVS before qdrant

* bench: aggregate independent index runs

* bench: order aggregate variants consistently

* docs: add repeatable index scaling results

* bench: add collection lifecycle benchmark

* docs: add collection lifecycle results

* perf: cache prepared cuvs filters

* docs: report prepared filter cache results

* bench: add async vector concurrency benchmark

* bench: aggregate service concurrency runs

* docs: add async concurrency results

* docs: clarify cuVS dtype behavior

* feat: add memory-aware cuVS auto mode

* feat: reuse native filters for cuVS search

* docs: publish cuVS integration plan as Markdown

* fix: route selective filters before cuVS rebuild

* docs: record selective-first routing results

---------

Co-authored-by: Yuanqing Zhao <2604121+yuanqingz@users.noreply.github.com>
2026-07-07 12:21:10 +08:00
Jiahui Zhou 21c58bd8f6 Set tags (#2706)
* feat(vectordb): add partial update api

fix(vectordb): support partial updates across adapters

fix(vectordb): return structured update results

test(vectordb): cover update result behavior

* feat(vectordb): support partial upsert semantics

* feat(content): support explicit search tags

* refactor unify search tag update api

* refactor(content): drop unrelated peer_id from semantic refresh

* fix(content): normalize set-tags targets and cli output
2026-06-18 17:48:49 +08:00
Qin Haojie f1afd40a66 fix(vectordb): reject oversized bytes row strings (#2171) 2026-05-21 19:50:58 +08:00
yepper ddcd3fb9c8 chore(format): align python and c++ file formatting (#2001)
* chore(format): align python and c++ file formatting

* chore: update urllib3 to 2.7.0 and clean test imports

1. bump urllib3 dependency from 2.6.3 to 2.7.0
2. remove unused pytest import and RoleScope import from test file

* style: format list comprehensions and lambda function for readability

Adjust the line breaks in the list comprehension in the VikingSearchTool class to follow standard Python formatting conventions, and rewrap the lambda assignment in the test case to improve code readability without changing functionality.

* style: fix line wrapping and remove extra blank line

- remove stray blank line in ov_server.py
- wrap long logger.info line in memory.py for better readability

* style: fix targeted ruff lint violations

* chore: clean up unused imports and reorder code

This commit removes unused imports, reorders import statements for better consistency,
and simplifies some test file imports. Changes include:
- Remove redundant blank lines and unused imports across multiple test files and core modules
- Reorder imports in openviking hooks module to follow standard layout
- Fix import ordering in memory isolation handler
- Simplify php parser type imports
- Move volcengine mock import to correct position in test file

* refactor(uri utils): remove extra blank lines in uri.py

clean up redundant whitespace to improve code readability
2026-05-13 17:53:09 +08:00
MaojiaShengandopenviking ce998873f9 lisence: change the main lisence to AGPL-3.0 (#1085)
* lisence: change the main lisence from Apache-2.0 to AGPL-v3

* lisence: change the main lisence from Apache-2.0 to AGPL-v3

* lisence: change the main lisence from Apache-2.0 to AGPL-v3

---------

Co-authored-by: openviking <openviking@example.com>
2026-03-30 14:37:42 +08:00
Jiahui Zhou 84d2cf8900 Fix Windows engine wheel runtime packaging (#993) 2026-03-26 13:17:44 +08:00
Jiahui Zhou 033826f028 Migrate vectordb engine to abi3 packaging (#897)
chore: remove pybind11 remnants after abi3 migration

fix: scope linux abi3 wheel smoke test to engine loader
2026-03-23 19:11:07 +08:00
zhoujiahui 4984ba8567 fix(vectordb): fix croaring avx (#771) 2026-03-19 17:29:55 +08:00
zhoujiahui 758ccd8986 feat: split vectordb engine by cpu variant (#656) 2026-03-16 17:49:59 +08:00
zhoujiahui ae4c64175b refactor: quit setup when fail (#504) 2026-03-10 09:30:30 +08:00
kkkwjx 5cf378dc00 fix(vectordb): default x86 build to AVX2 and disable AVX512 (#291) 2026-02-25 22:21:19 +08:00
Mingjian Queandmijamind719 498d560fd1 feat(vectordb): integrate KRL for ARM Kunpeng vector search optimization (#256)
- Add third_party/krl: Kunpeng Retrieval Library (KRL) source with ARM NEON/SIMD-optimized L2 and inner-product distance routines
- vector_base.h: add ARM platform macros OV_PLATFORM_ARM, OV_SIMD_NEON, OV_SIMD_SVE
- space_l2.h: on ARM use krl_L2sqr in l2_sqr_neon instead of scalar path
- space_ip.h: on ARM use krl_ipdis in inner_product_neon instead of scalar path
- CMakeLists.txt: enable OV_PLATFORM_ARM on aarch64, build and link KRL static library

On ARM, vectordb uses KRL-optimized paths; on x86 the existing AVX/SSE implementations are unchanged.

Co-authored-by: mijamind719 <mijamind@163.com>
2026-02-24 17:41:28 +08:00
kkkwjx c61aff3f0a fix: fix windows (#126) 2026-02-11 00:33:53 +08:00
kkkwjx c763ab125a refactor: cpp bytes rows (#105)
* refactor: cpp bytes rows

* refactor: cpp bytes row v2
2026-02-09 19:55:26 +08:00
kkkwjx de0f213a74 fix: fix sparse (#100) 2026-02-09 14:26:31 +08:00
kkkwjx 393a4c5878 Path filter (#92)
* feat: use path field

* refactor: bitmap filter
2026-02-07 17:27:42 +08:00
kkkwjx 9368d838f7 feat: add search_with_sparse_logit_alpha (#71) 2026-02-05 21:28:12 +08:00
kkkwjx 88d2d9f5d7 Fix compile (#62)
* fix: fix engine compile

* fix: fix linux release
2026-02-05 17:25:30 +08:00
kkkwjx 93dee92567 refactor: optimized pyproject.toml with optional groups (#26)
* refactor: optimized pyproject.toml with optional groups

* fix: fix observer
2026-02-02 21:54:01 +08:00
qin-ctx f98dc0ed1c first commit 2026-01-29 20:29:19 +08:00