Building a local RAG for complex semantic texts: when cloud AI fails and MiniLM breaks on philosophy
A developer recounts building a local RAG pipeline for philosophical texts by Jane Roberts, using multilingual-e5-large embeddings and Qwen2.5 via LM Studio. The project includes data preparation, mark-up, chunking, and retrieval. The author seeks community feedback on Chroma index architecture, dialogue state management, and reranker usage.
Sentence-Transformers (UKPLab)
LM Studio
Habr — хаб ИИ30.07 · 15:01
