New AI Project I'm working on
Building a multi-agent RAG system spread across two servers.
Three collaborative agents (two on one machine, one on the other) powered by Llama-3.1-8B + Qwen2.5-7B. Using Qdrant vector DB, BGE-M3 embeddings, semantic chunking, rich metadata extraction, and a critic step for higher accuracy.
Goal: a reliable private knowledge base from all my internal docs.
Ingestion is currently running and will probably take 3 days. The waiting part is painful 😅
Suggestions?