Natural language processing (NLP) is a field of computer science that studies how computers and humans interact. In the 1950s, Alan Turing published an article that proposed a measure of intelligence, now called the Turing test. More modern techniques, such as deep learning, have produced results in the fields of language modeling, parsing, and natural-language tasks.
🎤 vibrato: Viterbi-based accelerated tokenizer
✔️Contextual word checker for better suggestions (not actively maintained)
A curated archive of breakthroughs in Agents, Architecture, Training, RAG, and On-Device AI.
A Japanese tokenizer based on recurrent neural networks
The CMU Link Grammar natural language parser
Cantonese Linguistics and NLP
An efficient framework for collaborative training of large language models
An open-source tool for sequence learning in NLP built on TensorFlow.
Automatic Web Article Summarizer
Natural Language Processing Papers
a Deep Learning Framework for Text https://delft.readthedocs.io/
TopicGPT: A Prompt-Based Framework for Topic Modeling [NAACL'24]
potato: the portable annotation tool
超轻量级bert的pytorch版本,大量中文注释,容易修改结构,持续更新
:page_facing_up: A PyTorch implementation of Paragraph Vectors (doc2vec).
API of Articut 中文斷詞 (兼具語意詞性標記):「斷詞」又稱「分詞」,是中文資訊處理的基礎。Articut 不用機器學習,不需資料模型,只用現代白話中文語法規則,...
Learn how to use PyTorch to solve some common NLP problems with deep learning.
A Great Collection of Deep Learning Tutorials and Repositories
Juman++ (a Morphological Analyzer Toolkit)
A dataset of millions of news articles scraped from a curated list of data sources.
JavaScript Web SDK for Dialogflow
Abstractive summarisation using Bert as encoder and Transformer Decoder
An easy to use Natural Language Processing library and framework for predicting, training, fine-tuning, and serving up state-of-the-art NLP models.
Applying Data Science and Machine Learning to Solve Real World Business Problems
:horse_racing: 聊天机器人,自然语言理解,语义理解
Hierarchical Attention Networks for document classification
BERT-NER (nert-bert) with google bert https://github.com/google-research.
A neural network architecture for NLP tasks, using cython for fast performance. Currently, it can perform POS tagging, SRL and dependency parsing.
End-to-End recipes for pre-training and fine-tuning BERT using Azure Machine Learning Service
A deep NLP library, based on Keras / tf, focused on question answering (but useful for other NLP too)
SpikeX - SpaCy Pipes for Knowledge Extraction
LiBai(李白): A Toolbox for Large-Scale Distributed Parallel Training
Veldra — talk an agent into existence, then watch it grow. A self-hostable, local-first agent platform: describe what you need in plain language and...
🧠 AI-powered Personalized Exam System: Integrating OpenPangu LLM, Knowledge Graph RAG, and BKT algorithm for adaptive question generation and recomme...
💬 Open Source App Framework to build streaming apps with real-time data - 💎 Build real-time data pipelines and make real-time data universally acc...
Zero and Few shot named entity & relationships recognition
中文文本分类任务,基于PyTorch实现(TextCNN,TextRNN,FastText,TextRCNN,BiLSTM_Attention, DPCNN, Transformer,Bert,ERNIE),开箱即用!
Python API for Kiwi
Code and data of ACL 2021 paper "Few-NERD: A Few-shot Named Entity Recognition Dataset"
AstrBot 自主学习插件 — 让 AI 聊天机器人自主学习对话风格、理解群组黑话、管理社交关系与好感度、自适应人格演化,像真人一样自然对话。
A Abstractive Summarization Implementation with Transformer and Pointer-generator
Language model fine-tuning on NER with an easy interface and cross-domain evaluation. "T-NER: An All-Round Python Library for Transformer-based Named...
Massive open Japanese speech corpus
Information extraction from English and German texts based on predicate logic
The Official Repo for "Quick Start Guide to Large Language Models"
More than 50+ collections of Thai Natural Language Processing libraries. Update daily.
基于BERT的中文命名实体识别
My NLP datasets for Russian language
活字通用大模型
Large Language Models: In this repository Language models are introduced covering both theoretical and practical aspects.