Natural language processing (NLP) is a field of computer science that studies how computers and humans interact. In the 1950s, Alan Turing published an article that proposed a measure of intelligence, now called the Turing test. More modern techniques, such as deep learning, have produced results in the fields of language modeling, parsing, and natural-language tasks.
Oxford Deep NLP 2017 course - Practical 1: word2vec
Convolutional Neural Networks for Sentence Classification(TextCNN) implements by TensorFlow
Easy to use NLP library built on PyTorch and TorchText
Multilingual Rapid Automatic Keyword Extraction (RAKE) for Python
Deep-Learning Model Exploration and Development for NLP
Explains nlp building blocks in a simple manner.
A simple NLP library allows profiling datasets with one or more text columns. When given a dataset and a column name containing text data, NLP Profile...
A frame-semantic parsing system based on a softmax-margin SegRNN.
This repository contains an easy and intuitive approach to few-shot NER using most similar expansion over spaCy embeddings. Now with entity scoring.
中文ULMFiT 情感分析 文本分类
Toolkit to obtain and preprocess German text corpora, train models and evaluate them with generated testsets. Built with Gensim and Tensorflow.
OpenAI GPT2 pre-training and sequence prediction implementation in Tensorflow 2.0
Coding exercises for the Natural Language Processing concentration, part of Udacity's AIND program.
Backprop makes it simple to use, finetune, and deploy state-of-the-art ML models.
CNN for Chinese Text Classification in Tensorflow
Dynamic Memory Networks (https://arxiv.org/abs/1603.01417) in Tensorflow
💫 REST microservices for various spaCy-related tasks
MONPA 罔拍是一個提供正體中文斷詞、詞性標註以及命名實體辨識的多任務模型
Named Entity Recognition based on dictionaries
A Python wrapper for the ROUGE summarization evaluation package
Visualization Module for Natural Language Processing
Source code for paper: Improving Grammatical Error Correction via Pre-Training a Copy-Augmented Architecture with Unlabeled Data
An open-source text summarization toolkit for non-experts. EMNLP'2021 Demo
Deep Learning / NLP tutorial for Chatbot Developers
data resource untuk NLP bahasa indonesia
Fast + Non-Autoregressive Grammatical Error Correction using BERT. Code and Pre-trained models for paper "Parallel Iterative Edit Models for Local Seq...
:snake: Turkish Language Stemmer for Python
🧠 code-awareness
Implementing nlp papers relevant to classification with PyTorch, gluonnlp
A list of recent papers about Meta / few-shot learning methods applied in NLP areas.
Summarization, translation, sentiment-analysis, text-generation and more at blazing speed using a T5 version implemented in ONNX.
从零基础开始机器学习之旅
结合python一起学习自然语言处理 (nlp): 语言模型、HMM、PCFG、Word2vec、完形填空式阅读理解任务、朴素贝叶斯分类器、TFIDF、PCA、SVD
This repository contains code and datasets related to entity/knowledge papers from the VERT (Versatile Entity Recognition & disambiguation Toolkit) pr...
🏖 Easy training and deployment of seq2seq models.
Yet another Python binding for fastText
Sohu's 2018 content recognition competition 1st solution(搜狐内容识别大赛第一名解决方案)
Word Embeddings for Information Retrieval
All lecture notes, slides and assignments from CS224n: Natural Language Processing with Deep Learning class by Stanford
Punctuation restoration and spell correction experiments.
OCR, Archive, Index and Search: Implementation agnostic OCR framework.
R package for Tokenization, Parts of Speech Tagging, Lemmatization and Dependency Parsing Based on the UDPipe Natural Language Processing Toolkit
短文本聚类预处理模块 Short text cluster
Fast, DB Backed pretrained word embeddings for natural language processing.
FedNLP: An Industry and Research Integrated Platform for Federated Learning in Natural Language Processing, Backed by FedML, Inc. The Previous Researc...
A python module for English lemmatization and inflection.
ROUGE automatic summarization evaluation toolkit. Support for ROUGE-[N, L, S, SU], stemming and stopwords in different languages, unicode text evaluat...
Rule-based token, sentence segmentation for Russian language
INDRA (Integrated Network and Dynamical Reasoning Assembler) is an automated model assembly system interfacing with NLP systems and databases to colle...
Official source for spanish Language Models and resources made @ BSC-TEMU within the "Plan de las Tecnologías del Lenguaje" (Plan-TL).