Skip to main content
buildradar
Sign in
Topic · retrieval

retrieval

Tracked open-source repos tagged retrieval, sorted by stars.

Repos
30
Total stars
85,163
Avg. stars
2,839
Share
0.00%

Topics that frequently appear alongside retrieval on the same repo.

Recent risers

Repos created in the last 90 days, tagged retrieval.

No new repos tagged with this topic in the last 90 days.

  • PageIndex@VectifyAI

    📑 PageIndex: Document Index for Vectorless, Reasoning-based RAG

    35,495+89Star change over the last 7 days
  • semble@MinishLab

    Fast and Accurate Code Search for Agents. Uses 99% fewer tokens than grep+read

    5,985+16Star change over the last 7 days
  • OpenKB@VectifyAI

    OpenKB: Open LLM Knowledge Base

    4,411+356Star change over the last 7 days
  • SimpleMem@aiming-lab

    SimpleMem: Efficient Lifelong Memory for LLM Agents — Text & Multimodal

    3,735+9Star change over the last 7 days
  • mteb@embeddings-benchmark

    MTEB: State-of-the-art evaluation of embeddings across languages and modalities

    3,414+5Star change over the last 7 days
  • fastembed@qdrant

    Fast, Accurate, Lightweight Python library to make State of the Art Embedding

    3,179+6Star change over the last 7 days
  • sie@superlinked

    Open-source inference server and production cluster for all the models your agent needs.

    3,030+173Star change over the last 7 days
  • memobase@memodb-io

    User Profile-Based Long-Term Memory for AI Chatbot Applications.

    2,877+8Star change over the last 7 days
  • lucenenet@apache

    Apache Lucene.NET is an open-source full-text search library written in C#, ported from the Apache Lucene project.

    2,410+3Star change over the last 7 days
  • beir@beir-cellar

    A Heterogeneous Benchmark for Information Retrieval. Easy to use, evaluate your models across 15+ diverse IR datasets.

    2,283+5Star change over the last 7 days
  • bm25s@xhluca

    Fast BM25 search in Python, powered by Numpy and Numba

    1,779+3Star change over the last 7 days
  • raptor@parthsarthi03

    The official implementation of RAPTOR: Recursive Abstractive Processing for Tree-Organized Retrieval

    1,752+1Star change over the last 7 days
  • OpenResearcher@TIGER-AI-Lab

    OpenResearcher: A Fully Open Pipeline for Long-Horizon Deep Research Trajectory Synthesis

    1,214+3Star change over the last 7 days
  • SearchCLI@volcengine

    Open CLI for integrating AI search, recommendation, and conversational retrieval into agent systems and business systems

    1,176+1Star change over the last 7 days
  • CLIP4Clip@ArrowLuo

    An official implementation for "CLIP4Clip: An Empirical Study of CLIP for End to End Video Clip Retrieval"

    1,032+1Star change over the last 7 days
  • fastembed-rs@Anush008

    Rust library for generating vector embeddings and reranking locally!

    1,004+5Star change over the last 7 days
  • VisRAG@OpenBMB

    Parsing-free RAG supported by VLMs

    979-1Star change over the last 7 days
  • vectordb@epsilla-cloud

    Epsilla is a high performance Vector Database Management System

    875+0Star change over the last 7 days
  • sgpt@Muennighoff

    SGPT: GPT Sentence Embeddings for Semantic Search

    872+0Star change over the last 7 days
  • NeumAI@NeumTry

    Neum AI is a best-in-class framework to manage the creation and synchronization of vector embeddings at large scale.

    867+0Star change over the last 7 days
  • byaldi@AnswerDotAI

    Use late-interaction multi-modal models such as ColPali in just a few lines of code.

    852+0Star change over the last 7 days
  • RQ-VAE-Recommender@EdoardoBotta

    [Pytorch] Generative retrieval model using semantic IDs from "Recommender Systems with Generative Retrieval"

    845+4Star change over the last 7 days
  • searchGPT@michaelthwan

    Grounded search engine (i.e. with source reference) based on LLM / ChatGPT / OpenAI API. It supports web search, file content search etc.

    711+0Star change over the last 7 days
  • gritlm@ContextualAI

    Generative Representational Instruction Tuning

    700+1Star change over the last 7 days
  • Rankify@DataScienceUIBK

    🔥 Rankify: A Comprehensive Python Toolkit for Retrieval, Re-Ranking, and Retrieval-Augmented Generation 🔥. Our toolkit integrates 40 pre-retrieved benchmark datasets and supports 7+ retrieval techniques, 24+ state-of-the-art Reranking models, and multiple RAG methods.

    682-1Star change over the last 7 days
  • moss@usemoss

    The retrieval layer for production AI systems. Lightning-fast (<10ms) search without vector databases. Built for browser, edge, on-device, and cloud.

    671+4Star change over the last 7 days
  • EasyRAG@BUAADreamer

    Easy-to-Use RAG Framework; CCF AIOps International Challenge 2024 Top3 Solution; CCF AIOps 国际挑战赛 2024 季军方案

    639+1Star change over the last 7 days
  • sigmap@manojmallick

    ~97% token reduction for AI coding sessions — zero deps, 33 languages, MCP server

    626+5Star change over the last 7 days
  • ArXivChatGuru@redis-developer

    Use ArXiv ChatGuru to talk to research papers. This app uses LangChain, OpenAI, Streamlit, and Redis as a vector database/semantic cache.

    561+0Star change over the last 7 days
  • relik@SapienzaNLP

    Retrieve, Read and LinK: Fast and Accurate Entity Linking and Relation Extraction on an Academic Budget (ACL 2024)

    518+0Star change over the last 7 days
← Back to topics