Combining knowledge graphs with embeddings to enable multi-hop reasoning and contextual understanding in LLMs, while supporting natural language querying.
The ever-growing volume of research publications necessitates efficient methods for structuring this body of knowledge. This solution uses Machine Learning (UMAP, HDBSCAN), Embedding Quantization, and an LLM pipeline to classify 25,000 arXiv publications under a novel taxonomy.