Skip to content

Topic

Vector Similarity Search

Approximate nearest-neighbour retrieval over embedding vectors, including HNSW indexing, distance metrics and recall/speed tuning.

Current stories

build1 publisher

Plain-language questions push SEPA rulebook answers out of a top-5 vector search

One developer's test on 484 SEPA rulebook passages found that plain-English questions push several answers out of a top-5 vector search. Because the test measures each answer's rank directly, the failure shows up in retrieval, before the language model writes anything.

Publishers:dev.to

Reality

Evidence55
Adoption
Insufficient
Hype gap0
Incentives
Insufficient
Confidence45
build1 publisher

Full-precision embeddings push a 100-million-vector OpenSearch index to 1.3 TB of RAM

Amazon OpenSearch Service needs about 1.3 TB of resident RAM for 100 million 1,536-dimension FP32 vectors with one replica, a dev.to sizing post calculates. Raw vector values are about 98% of each entry, so the encoding picked before ingestion decides most of that memory and the node count behind it.

Publishers:dev.to

Reality

Evidence58
Adoption
Insufficient
Hype gap+5
Incentives
Insufficient
Confidence62
build1 publisher

One error code explains why RAG pipelines keep a keyword index

A dev.to post by Rijul argues that semantic similarity is the wrong tool for error codes, part numbers and filenames, and sketches a retrieval pipeline that runs keyword and vector search side by side. It reports no measurements.

Publishers:dev.to

Reality

Evidence28
Adoption
Insufficient
Hype gap+12
Incentives55
Confidence58

Earlier coverage

  1. JSONB and pgvector cover two of the five roles this Postgres consolidation absorbs

    Build · September 17, 2026 · 1 publisher

  2. Co-locating embeddings with permissions collapses the RAG fetch into one SQL statement

    Build · September 17, 2026 · 1 publisher

  3. Hashing chunk IDs into five shard keys widens a DynamoDB vector search to 500 candidates

    Build · September 16, 2026 · 1 publisher

  4. One Lambda fans a query into a prefix match and a 512-dimension cosine search

    Build · September 15, 2026 · 1 publisher

  5. Valkey 9.1 cuts per-key string overhead by 17 to 44 percent with no config change

    Build · September 12, 2026 · 1 publisher

  6. A boolean stream flag leaves callers guessing whether they get a string or a generator

    Build · September 12, 2026 · 1 publisher

  7. The server injects up to five project notes before the agent takes its first turn

    Build · September 11, 2026 · 1 publisher

  8. Cosine similarity ranks the menu above the peanut allergy note

    Build · September 11, 2026 · 1 publisher

  9. A Postgres default of 40 caps how many neighbors your vector search can return

    Build · September 10, 2026 · 1 publisher

  10. Adjudicating every extracted fact against its five nearest memories costs one LLM call apiece

    Build · September 6, 2026 · 1 publisher

  11. Only the corpus shrinks a billion-vector face index without costing recall

    Build · August 27, 2026 · 1 publisher

  12. Four control planes, one Postgres: a team's case against polyglot persistence

    Build · August 21, 2026 · 1 publisher

  13. DuckDB's vss extension removes a database from your RAG stack, then names the price

    Build · August 15, 2026 · 1 publisher