Natural language processing (NLP) is a field of computer science that studies how computers and humans interact. In the 1950s, Alan Turing published an article that proposed a measure of intelligence, now called the Turing test. More modern techniques, such as deep learning, have produced results in the fields of language modeling, parsing, and natural-language tasks.
WRENCH: Weak supeRvision bENCHmark
An Efficient Chinese Text Classifier
Yet another Japanese IME for IBus/Linux
Codes for the IJCAI2022 paper: Inheriting the Wisdom of Predecessors: A Multiplex Cascade Framework for Unified Aspect-based Sentiment Analysis
Code for Java Deep Learning Cookbook
Abydos NLP/IR library for Python
✒️ Cedille is a large French language model (6B), released under an open-source license
Real-Time Open-Domain Question Answering with Dense-Sparse Phrase Index (DenSPI)
比做算法的懂工程落地,比做工程的懂算法模型。
IssuesにNLP(自然言語処理)に関連するの論文を読んだまとめを書いています.雑です.🚧 マークは編集中の論文です(事実上放置のものも多いです).🍡 マーク...
Reuters and Bloomberg
Build and train state-of-the-art natural language processing models using BERT
dna2vec: Consistent vector representations of variable-length k-mers
🏖TagEditor - Annotation tool for spaCy
🦅 Pretrained BigBird Model for Korean (up to 4096 tokens)
This repository contains all codes and materials of the current session. It contains the required code on Natural Language Processing, Artificial int...
Linguistic Annotation and Visualization Tool for PDF Documents
Text Mining and Topic Modeling Toolkit for Python with parallel processing power
Calculating ROUGE score between two files (line-by-line)
🧬 A JupyterLab extension for annotating data with Prodigy
Python library for Natural Language Preprocessing (NLPre)
Curso práctico: NLP de cero a cien 🤗
Community Curated NLP List
Материалы курса по компьютерной лингвистике Школы Лингвистики НИУ ВШЭ
A set of tutorials for torchtext
https://gurivr.com
Fast, Consistent Tokenization of Natural Language Text
One-Stop Solution to encode sentence to fixed length vectors from various embedding techniques
Efficient and clean PyTorch reimplementation of "End-to-end Neural Coreference Resolution" (Lee et al., EMNLP 2017).
Pre-trained Word2Vec Model for Turkish
Deep Learning based NLP modeling for Russian language
paper summary of Association for Computational Linguistics
A two-level morphological analyzer for Turkish.
Text tokenization and sentence segmentation (segtok v2)
A Natural Language Date Time Parser that Extract date and time from text with context and parse to the required format
适合中文程序员的变量命名助手,NLP+翻译,规范变量命名,定制化变量命名规则
Topic modeling helpers using managed language models from Cohere. Name text clusters using large GPT models.
Implementation of Siamese Neural Networks built upon multihead attention mechanism for text semantic similarity task.
:boom: :chart_with_upwards_trend: A curated list of data science, analysis and visualization tools
My (slightly modified) Keras implementation of the Recurrent Convolutional Neural Network (RCNN) described here: http://www.aaai.org/ocs/index.php/AAA...
Chinese GPT2: pre-training and fine-tuning framework for text generation
💙 Emoji handling and meta data for spaCy with custom extension attributes
Google USE (Universal Sentence Encoder) for spaCy
BERT for Finance : UC Berkeley MIDS w266 Final Project
Test your HN title against a neural network
2017-CCF-BDCI-让AI当法官(初赛):7th/415 (Top 1.68%)
Magento Chatbot Integration with Telegram, Messenger, Whatsapp, WeChat, Skype and wit.ai.
Text to sentence splitter using heuristic algorithm by Philipp Koehn and Josh Schroeder.
对收集的法律文档进行一系列分析,包括根据规范自动切分、案件相似度计算、案件聚类、法律条文推荐等(试验目前基于婚姻类案件,可扩展至其它领域)。
FairyTailor: Multimodal Generative Framework for Storytelling