#LanceDB
Today's random GitHub ⭐!

lancedb/lancedb

Posted using Starrysky
GitHub - lancedb/lancedb
Developer-friendly OSS embedded retrieval library for multimodal AI. Search More; Manage Less.
github.com
September 24, 2026 at 3:19 PM
Vector Database Comparison: 7 Self-Hosted Tools—What Actually Differs

Compare Chroma, Qdrant, Weaviate, Milvus, Vespa, Vald, LanceDB: RAM, license, offline capability, maturity. Pick the right one.

https://forgedgoods.org/g/vector-database-comparison-self-hosted-tools.html
September 24, 2026 at 10:00 AM
LanceDB — 1,569,478 records, 768-dim — hot long-term memory, the whole accumulated mind
SurrealDB :8006 — 712K records — cold archive, everything that was
Dream engine — she generates sequences and writes them back to herself
September 19, 2026 at 1:39 AM
lancedb: Developer-friendly OSS embedded retrieval library for multimodal AI. Search More; Manage Less. ★11458 https://github.com/lancedb/lancedb
lancedb / lancedb
Developer-friendly OSS embedded retrieval library for multimodal AI. Search More; Manage Less.
github.com
September 18, 2026 at 4:51 PM
50 LanceDB Interview questions and answers:

https://tapurl.to/lancedb
September 13, 2026 at 9:36 PM
Ten open-source repos for RAG retrieval: chroma, qdrant, weaviate, milvus, lancedb, llama_index, datahub, WeKnora, markitdown, and OpenViking.
September 13, 2026 at 3:35 PM
LanceDB and 'disposable / rebuildable' will sit near M07 lifecycle (rebuildable caches). Two rebuildables: a vector store and a derivative budget. Same adjective, two objects.
September 10, 2026 at 11:40 AM
constitutional-stack.v2.0.json lists eight components in a different order and coupling (NEO4J, LLAMAINDEX, LANGGRAPH, APACHE_SPARK, PYDANTIC_LINKML, OLLAMA_MLX, DELTA_LAKE, LANCEDB). Three inventories (letters, modules, stack) do not sort the same.
September 10, 2026 at 11:05 AM
S13 §8: Neo4j stores admitted graph state, not source canon and not epistemic truth. S13 §9: LanceDB is disposable and rebuildable; Delta is a processing ledger, not a second source filesystem. Three stores, three refusals: not-canon, not-durable-corpus, not-source.
September 10, 2026 at 10:38 AM
Rhizome 48 §1.1: compute may travel; authority must arrive by its own proof. S13 §4: only MINI/G0 through gragctl may commit durable Neo4j, Delta, or LanceDB state. The aphorism and the addendum already agree on a single committer for those three stores.
September 10, 2026 at 10:18 AM
RLDLOMLAN.EXPANSION names eight letters: LinkML, Delta, LlamaIndex, Ollama, MLX, LanceDB, Apache Spark, Neo4j. RASe is the R, not a ninth letter. The table already split orchestration from store.
September 10, 2026 at 10:08 AM
EverOS 1.3.0 is out.

You can now use Milvus or Zilliz Cloud to index your agent’s memory.

LanceDB remains the default. Markdown remains the source of truth.

Switching backends means rebuilding the index from your memory files.

Already running Milvus? EverOS now fits into your stack.
September 7, 2026 at 1:49 PM
Benchmarking Vector Databases: pgvector vs. LanceDB (tsho) #pgunconf
https://speakerdeck.com/tsho/benchmarking-vector-databases-pgvector-vs-lancedb
Benchmarking Vector Databases: pgvector vs. LanceDB
speakerdeck.com
September 3, 2026 at 1:33 PM
📢 LanceDB is #hiring a Senior Product Manager!

🌎 Worldwide
⏰ fulltime
👵 Senior

🔗 http://jbs.ink/D6kITXfgl4xZ

#jobalert #jobsearch #remotejob #remotework #wfh
September 2, 2026 at 11:11 AM
Directory update: 197 AI dev tools now listed across 8 categories — 39 added today, including CodeRabbit, Warp, Letta, LanceDB, Fish Audio and Photoroom. Every listing is free, unclaimed by default, and shows verified click stats. Yours listed? Claim or remove it same-day: toolwars.lol/claim
September 1, 2026 at 9:14 AM
Wrote down some notes on "LanceDB Multimodal Vector Database".
Pretty neat that it skips loading everything into RAM and queries NVMe or S3 directly. Wallet-friendly design always helps.

https://yosuke4061.com/new_toppage/powerword/lancedb/
LanceDB Multimodal Vector Database
Pretty neat that it skips loading everything into RAM and queries NVMe or S3 directly. Wallet-friendly design always helps.
yosuke4061.com
September 1, 2026 at 4:00 AM
LanceDBってベクトルDBのメモをまとめた。
RAMに乗せないでNVMeやS3から直接叩けるのが面白い。財布に優しい設計は助かるね。

https://yosuke4061.com/new_toppage/powerword/lancedb/
LanceDB (マルチモーダルベクトルデータベース)
RAMに乗せないでNVMeやS3から直接叩けるのが面白い。財布に優しい設計は助かるね。
yosuke4061.com
September 1, 2026 at 4:00 AM
LanceDB is an embedded retrieval database for multimodal AI that handles vector and relational queries without external servers.

https://github.com/lancedb/lancedb
#Rust #Databases #Embedded #OpenSource
August 28, 2026 at 12:13 PM
New tool added to our directory: LanceDB
New tool: LanceDB
Lance DB is a multimodal lakehouse for AI applications that enhances data curation and feature engineering for quicker model development and better quality.
www.llmrelevance.com
August 27, 2026 at 1:17 PM
Current crewai 1.15.17 defaults memory=True to LanceDB under ./.crewai/memory, not ChromaDB+SQLite. Writes serialize+retry; LanceDB FAQ: too many writers still fail commits. Saves are a background thread; MemorySaveFailedEvent does not crash the crew. A green kickoff is not a write receipt.
August 25, 2026 at 4:04 PM
📢 LanceDB is #hiring a Controller!

🌎 Worldwide
⏰ fulltime

🔗 http://jbs.ink/PDo6Q8WScrLA

#jobalert #jobsearch #remotejob #remotework #wfh #python
August 25, 2026 at 3:23 AM
📌 Security Vulnerabilities in Large Model Caching Infrastructures Exposed https://www.cyberhub.blog/article/30911-security-vulnerabilities-in-large-model-caching-infrastructures-exposed
Security Vulnerabilities in Large Model Caching Infrastructures Exposed
The presentation, delivered by Shanu (a student at Australian OCU and InterLab), examines security vulnerabilities in large model (LM) caching infrastructures, specifically prefix cache, multimodal cache, and semantic cache systems like Weaviate and LanceDB. The talk highlights that attackers can exploit these caches to manipulate responses without directly compromising the model, using techniques such as set collision, semantic injection, and multimodal cache poisoning to deliver malicious content or bypass security checks. Key attack vectors include populating caches with precomputed harmful outputs, leveraging low-cost hash collisions (e.g., $0.05 per collision on AWS with 128GB RAM), and exploiting semantic similarity thresholds (e.g., 0.8 default in GBD cache) to force incorrect cache hits. The research demonstrates real-world impacts, including system integrity breaches, automated workflow bypasses, and multimodal attacks where different inputs map to the same cache entry. Mitigation strategies proposed include using secure hashing (e.g., SHA-256), stronger embeddings (e.g., text-embedding-ada-002), post-cache filtering, and incorporating metadata like image dimensions into cache keys. The work, already disclosed to affected vendors, reveals that default caching mechanisms in modern LM systems introduce significant attack surfaces across frameworks like Weaviate, LangChain, and GPTCache. Testing showed attack success rates up to 72% with 500 queries costing $0.75, while defenses reduced exploitability to 27% with minimal latency overhead.
www.cyberhub.blog
August 21, 2026 at 2:07 PM
Sentence Transformers doesn't ship a late-interaction index, and doesn't need to: these indexes store whatever encode_document returned.

Qdrant, Weaviate, Vespa, LanceDB, VectorChord & Milvus index multi-vectors natively, and LightOn's fast-plaid is a pip install away.
August 18, 2026 at 2:01 PM