nlp
Tracked open-source repos tagged nlp, sorted by stars.
- #331
💥 Use the latest Stanza (StanfordNLP) research models directly in spaCy
★ 746+0Star change over the last 7 days - #332★ 744+0Star change over the last 7 days
- #333★ 740+2Star change over the last 7 days
- #334
The prime repository for state-of-the-art Multilingual Question Answering research and development.
★ 739-1Star change over the last 7 days - #335
:clipboard: A Python Parser for PubMed Open-Access XML Subset and MEDLINE XML Dataset
★ 738+2Star change over the last 7 days - #336
A list of selected resources, methods, and tools dedicated to Legal Text Analytics.
★ 735+0Star change over the last 7 days - #337
Extend existing LLMs way beyond the original training length with constant memory usage, without retraining
★ 734-1Star change over the last 7 days - #338
PromptKG Family: a Gallery of Prompt Learning & KG-related research works, toolkits, and paper-list.
★ 734+0Star change over the last 7 days - #339★ 729+0Star change over the last 7 days
- #340★ 726+0Star change over the last 7 days
- #341
Simple implementation of OpenAI CLIP model in PyTorch.
★ 726+0Star change over the last 7 days - #342★ 721+2Star change over the last 7 days
- #343★ 719+6Star change over the last 7 days
- #344
Grounded search engine (i.e. with source reference) based on LLM / ChatGPT / OpenAI API. It supports web search, file content search etc.
★ 711+0Star change over the last 7 days - #345
VectorFlow is a high volume vector embedding pipeline that ingests raw data, transforms it into vectors and writes it to a vector DB of your choice.
★ 703+0Star change over the last 7 days - #346★ 702-1Star change over the last 7 days
- #347
A Full Stack ML (Machine Learning) Roadmap involves learning the necessary skills and technologies to become proficient in all aspects of machine learning, including data collection and preprocessing, model development, deployment, and maintenance.
★ 701+0Star change over the last 7 days - #348
:black_circle: A spaCy pipeline and model for NLP on unstructured legal text.
★ 696+0Star change over the last 7 days - #349★ 691+0Star change over the last 7 days
- #350
Build LLM-powered Dart/Flutter applications.
★ 686+0Star change over the last 7 days - #351
Stanford Open Information Extraction made simple!
★ 682+0Star change over the last 7 days - #352
🔥 Rankify: A Comprehensive Python Toolkit for Retrieval, Re-Ranking, and Retrieval-Augmented Generation 🔥. Our toolkit integrates 40 pre-retrieved benchmark datasets and supports 7+ retrieval techniques, 24+ state-of-the-art Reranking models, and multiple RAG methods.
★ 682-1Star change over the last 7 days - #353★ 674+0Star change over the last 7 days
- #354
SpaCy 中文模型 | Models for SpaCy that support Chinese
★ 673+0Star change over the last 7 days - #355
Ekphrasis is a text processing tool, geared towards text from social networks, such as Twitter or Facebook. Ekphrasis performs tokenization, word normalization, word segmentation (for splitting hashtags) and spell correction, using word statistics from 2 big corpora (english Wikipedia, twitter - 330mil english tweets).
★ 673+0Star change over the last 7 days - #356
Find dates inside text using Python and get back datetime objects
★ 665+2Star change over the last 7 days - #357
A fast, lightweight and easy-to-use Python library for splitting text into semantically meaningful chunks.
★ 664+0Star change over the last 7 days - #358★ 663+0Star change over the last 7 days
- #359
A Python multilingual toolkit for Sentiment Analysis and Social NLP tasks
★ 660+0Star change over the last 7 days - #360
Quickly format your notes with ChatGPT in Obsidian
★ 658-2Star change over the last 7 days