Natural language processing (NLP) is a field of computer science that studies how computers and humans interact. In the 1950s, Alan Turing published an article that proposed a measure of intelligence, now called the Turing test. More modern techniques, such as deep learning, have produced results in the fields of language modeling, parsing, and natural-language tasks.
⛔ [NOT MAINTAINED] An End-To-End Closed Domain Question Answering System.
Build chatbots and conversational experiences using React
A list of tools for annotating data, managing annotations, etc.
中文Mixtral混合专家大模型(Chinese Mixtral MoE LLMs)
基于法律裁判文书的事件抽取及其应用,包括数据的分词、词性标注、命名实体识别、事件要素抽取和判决结果预测等内容
[ICLR 2020] Lite Transformer with Long-Short Range Attention
A content-based recommender system that recommends movies similar to the movie the user likes and analyses the sentiments of the reviews given by the...
Tock, the open source conversational AI toolkit.
[ACL 2021] Learning Dense Representations of Phrases at Scale; EMNLP'2021: Phrase Retrieval Learns Passage Retrieval, Too https://arxiv.org/abs/2012.1...
Hierarchical Attention Networks for Document Classification in PyTorch
NLP in Python with Deep Learning
An implementation of the BERT model and its related downstream tasks based on the PyTorch framework. @跟我学机器学习
🍁 Sycamore is an LLM-powered search and analytics platform for unstructured data.
Building a Simple Chatbot from Scratch in Python (using NLTK)
PyTorch implementation for "Matching the Blanks: Distributional Similarity for Relation Learning" paper
Simple chatbot UI for the Web with JSON scripting 👋🤖🤙
MAD: The first work to explore Multi-Agent Debate with Large Language Models :D
Resources for learning about Text Mining and Natural Language Processing
Python package of Tomoto, the Topic Modeling Tool
NLP Zero to Hero in just 10 Kernels
A python toolkit for parsing captions (in natural language) into scene graphs (as symbolic representations).
微信公众号语料库
NLP Paper
A platform for building conversational interfaces with intelligent agents (chatbots)
PyTorch Re-Implementation of "Generating Sentences from a Continuous Space" by Bowman et al 2015 https://arxiv.org/abs/1511.06349
Offline semantic Text-to-Image and Image-to-Image search on Android powered by quantized state-of-the-art vision-language pretrained CLIP model and ON...
Span-level grounding verification for RAG, code, and tool-grounded AI outputs.
OpenICL is an open-source framework to facilitate research, development, and prototyping of in-context learning.
中文情感分析库(Chinese Sentiment))可对文本进行情绪分析、正负情感分析。Text analysis, supporting multiple methods including word count, readability, do...
⚡ boost inference speed of T5 models by 5x & reduce the model size by 3x.
A collection of datasets that pair questions with SQL queries.
[NeurIPS 2022] 🛒WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents
tensorflow를 사용하여 텍스트 전처리부터, Topic Models, BERT, GPT, LLM과 같은 최신 모델의 다운스트림 태스크들을 정리한 Deep Learning NLP 저장소입니다.
Simple XLNet implementation with Pytorch Wrapper
A list of NLP resources focused on event extraction task
Paper list of simultaneous translation / streaming translation, including text-to-text machine translation and speech-to-text translation.
🇨🇳Open Chinese Convert is an opensource project for conversion between Traditional Chinese and Simplified Chinese.(java 中文繁简体转换,支持台湾、香港...
Tensorflow Implementation of R-Net
REBEL is a seq2seq model that simplifies Relation Extraction (EMNLP 2021).
A collection of tools, datasets and resources on Bangla computing
Organize your experiments into discrete steps that can be cached and reused throughout the lifetime of your research project.
A Curated List of Dataset and Usable Library Resources for NLP in Bahasa Indonesia
The hands-on NLTK tutorial for NLP in Python
Medical Q&A with Deep Language Models
Julia Implementation of Transformer models
A suite of Arabic natural language processing tools developed by the CAMeL Lab at New York University Abu Dhabi.
Firefox Translations is a webextension that enables client side translations for web browsers.
An open collection of methodologies to help with successful training of large language models.
[EMNLP 2022] Unifying and multi-tasking structured knowledge grounding with language models
Fast Inference Solutions for BLOOM