Natural language processing (NLP) is a field of computer science that studies how computers and humans interact. In the 1950s, Alan Turing published an article that proposed a measure of intelligence, now called the Turing test. More modern techniques, such as deep learning, have produced results in the fields of language modeling, parsing, and natural-language tasks.
Codes for ICML 2022 paper: Matching Structure for Dual Learning
啊哈自然语言处理包,提供包括分词、依存句法分析、语义角色标注、自动摘要、语义相似度计算、LDA 主题预测、词云等服务。
BNLP is a natural language processing toolkit for Bengali Language.
multilabel classification of EHR notes
A collection of datasets and other resources for legal text processing.
✨ 🤖 🤖 Build your own conversational bot on our Collaborative Bot Platform! 🤖🤖 ✨
A curated list of resources dedicated to Natural Language Processing (NLP) in polish. Models, tools, datasets.
AubAI brings you on-device gen-AI capabilities, including offline text generation and more, directly within your app.
A joint community effort to create one central leaderboard for LLMs.
JAX library for training sub-4B foundation models for edge
A pytorch implementation of the ACL2019 paper "Simple and Effective Text Matching with Richer Alignment Features".
turn natural language into structured data(支持中文,自定义了N种模型,支持不同的场景和任务)
A chatbot framework for Rasa NLU
Text2Text Language Modeling Toolkit
ACL 2018: Hybrid semi-Markov CRF for Neural Sequence Labeling (http://aclweb.org/anthology/P18-2038)
LSTM+CRF NER
A sentence segmenter that actually works!
This repo consists of multiple machine learning based projects with frontend
Textpipe: clean and extract metadata from text
a sklearn wrapper for Google's BERT model
Prosodic: a metrical-phonological parser, written in Python. For English and Finnish, with flexible language support.
A private nlp coding package, which quickly implements the SOTA solutions.
The high performance pinyin tool for java.(java 高性能中文转拼音工具。支持同音字。)
A modern, interlingual wordnet interface for Python
Open-sourced course notes for Artificial Intelligence and Data Science related topics, prepared in LaTeX
A large-scale (194k), Multiple-Choice Question Answering (MCQA) dataset designed to address realworld medical entrance exam questions.
📰Natural language processing (NLP) newsletter
Simple Solution for Multi-Criteria Chinese Word Segmentation
Human-AI Collaborative Data Science Using Visual Workflows
🛥 Vaporetto: Very accelerated pointwise prediction based tokenizer
gpttools extends gptstudio for package development to help you document code, write tests, or even explain code
ML based projects such as Spam Classification, Time Series Analysis, Text Classification using Random Forest, Deep Learning, Bayesian, Xgboost in Pyth...
An NLP system for generating reading comprehension questions
The hanzi similar tool.(汉字相似度计算工具,中文形近字算法。可用于手写汉字识别纠正,文本混淆等。)
AI for Ethical Hacking - Workshop
HanLP中文分词Lucene插件,支持包括Solr在内的基于Lucene的系统
Indic-BERT-v1: BERT-based Multilingual Model for 11 Indic Languages and Indian-English. For latest Indic-BERT v2, check: https://github.com/AI4Bharat/...
AI 小说分析可视化工具 — 角色关系图谱 · 地理地图 · 时间线 · 百科全书 | 支持 Ollama 本地 + 10 大云端 LLM | React + FastAPI + SQLite
Komputation is a neural network framework for the Java Virtual Machine written in Kotlin and CUDA C.
NLP library designed for reproducible experimentation management
LDA topic modeling for node.js
Data Augmentation for NLP. NLP数据增强
Parse SEC EDGAR HTML documents into a tree of elements that correspond to the visual (semantic) structure of the document.
Machine Learning and NLP: Text Classification using python, scikit-learn and NLTK
Text analysis with networks.
All NLP you Need Here. 目前包含15个NLP demo的pytorch实现(大量代码借鉴于其他开源项目,原先是自己玩的,后来干脆也开源出来)
RETVec is an efficient, multilingual, and adversarially-robust text vectorizer.
CNN/Daily Mail Reading Comprehension Task
最强接口测试平台
:owl: Snow Owl Terminology Server - a production-ready, fast, scalable, FHIR Terminology Service compliant server that supports SNOMED CT Internationa...