Weaviate Podcast
Weaviate Podcast
Weaviate
Kevin Cohen on Neum AI - Weaviate Podcast #70!
55 minutes Posted Oct 18, 2023 at 2:00 pm.
Check this out!
Welcome Kevin!
Founding Neum AI
Data Ingestion, End-to-End Overview
Chunking and Metadata Extraction
Embedding Cache
Distributed Messaging Queues
Embeddings Cache ELI5
Customizing Weaviate Kubernetes
Multi-Tenancy and Resource Allocation
Billion-Scale Vector Search
Knowledge Graphs
Y Combinator Experience
0:00
55:02
Download MP3
Show notes
Hey everyone! Thank you so much for watching the 70th episode of the Weaviate podcast with Neum AI CTO and Co-Founder Kevin Cohen! I first met Kevin when he was debugging an issue with his distributed node utilization and have since learned so much from him about how he sees the space of Data Ingestion, also commonly referenced as ETL for LLMs! There are so many interesting parts to this from the general flow of data connectors, chunkers and metadata extractors, embedding inference, and the last leg of the mile of importing the vectors to a Vector DB such as Weaviate! I really loved how Kevin broke down the distributed messaging queue and system design for orchestrating data ingestion at massive scale such as dealing with failures and optimizing the infrastructure as code setup. We also discussed things like new use cases with quadrillion scale vector indexes and the role of knowledge graphs in all this! I really hope you enjoy the podcast, please check out this amazing article below from Neum AI!
https://medium.com/@neum_ai/retrieval-augmented-generation-at-scale-building-a-distributed-system-for-synchronizing-and-eaa29162521
Chapters