Natural language processing (NLP) is a field of computer science that studies how computers and humans interact. In the 1950s, Alan Turing published an article that proposed a measure of intelligence, now called the Turing test. More modern techniques, such as deep learning, have produced results in the fields of language modeling, parsing, and natural-language tasks.
An Open-Source Package for Textual Adversarial Attack.
DNABERT: pre-trained Bidirectional Encoder Representations from Transformers model for DNA-language in genome
A Neural Framework for MT Evaluation
PatrickStar enables Larger, Faster, Greener Pretrained Models for NLP and democratizes AI for everyone.
😎 A curated list of the Question Answering (QA)
Framework for enhancing LLMs for RAG tasks using fine-tuning.
深度学习近年来关于神经网络模型解释性的相关高引用/顶会论文(附带代码)
Deep Reinforcement Learning For Sequence to Sequence Models
Full text geoparsing as a Python library
BabyAI platform. A testbed for training agents to understand and execute language commands.
All-in-one text de-duplication
Kiwi(지능형 한국어 형태소 분석기)
DBpedia Spotlight is a tool for automatically annotating mentions of DBpedia resources in text.
This repository contains my full work and notes on Coursera's NLP Specialization (Natural Language Processing) taught by the instructor Younes Bensoud...
Python Implementations of Word Sense Disambiguation (WSD) Technologies.
💥 Use the latest Stanza (StanfordNLP) research models directly in spaCy
An AI-powered Personal Identifiable Information (PII) scanner.
The prime repository for state-of-the-art Multilingual Question Answering research and development.
Curated list of open source tooling for data-centric AI on unstructured data.
:clipboard: A Python Parser for PubMed Open-Access XML Subset and MEDLINE XML Dataset
Extend existing LLMs way beyond the original training length with constant memory usage, without retraining
地球上最全的华语现代诗歌语料库,3k+诗人,80K+诗歌,15M+字
PromptKG Family: a Gallery of Prompt Learning & KG-related research works, toolkits, and paper-list.
A list of selected resources, methods, and tools dedicated to Legal Text Analytics.
TensorFlow and Deep Learning Tutorials
Salesforce open-source LLMs with 8k sequence length.
Simple implementation of OpenAI CLIP model in PyTorch.
Repository of code for the tutorial on Transfer Learning in NLP held at NAACL 2019 in Minneapolis, MN, USA
🌟 A curated collection of free, high quality AI tools 🤖, APIs 🔗, datasets 📊, and learning resources 📚 covering machine learning 🧠, deep learning...
Revisiting Pre-trained Models for Chinese Natural Language Processing (MacBERT)
A Modern C++ Data Sciences Toolkit
A Lite Bert For Self-Supervised Learning Language Representations
NeuSpell: A Neural Spelling Correction Toolkit
《Natural Language Processing with PyTorch》中文翻译
Grounded search engine (i.e. with source reference) based on LLM / ChatGPT / OpenAI API. It supports web search, file content search etc.
A collections of public and free annotated datasets of relationships between entities/nominals (Portuguese and English)
VectorFlow is a high volume vector embedding pipeline that ingests raw data, transforms it into vectors and writes it to a vector DB of your choice.
Japanese Input Method System for Linux, macOS, Neural Kana-Kanji Conversion Engine
基于自然语言理解与机器学习的聊天机器人,支持多用户并发及自定义多轮对话
Original Implementation of Prompt Tuning from Lester, et al, 2021
The BiLSTM-CRF model implementation in Tensorflow, for sequence labeling tasks.
A Full Stack ML (Machine Learning) Roadmap involves learning the necessary skills and technologies to become proficient in all aspects of machine lear...
机器学习、深度学习、自然语言处理、计算机视觉、各种算法等AI领域相关技术的路线、教程、干货分享。笔记有:机器学习实战、剑指Offer、cs231n、cs131、吴恩达机...
:book: 收集NLP领域相关的数据集、论文、开源实现,尤其是情感分析、情绪原因识别、评价对象和评价词抽取方面。
Natural language detection library for Go
:black_circle: A spaCy pipeline and model for NLP on unstructured legal text.
:heavy_check_mark: 微信上的定时提醒 - Cron on WeChat
人工智能学习资料超全整理,包含机器学习基础ML、深度学习基础DL、计算机视觉CV、自然语言处理NLP、推荐系统、语音识别、图神经网路、算法工程师面试题
An opensource text-to-speech (TTS) voice building tool