Natural language processing (NLP) is a field of computer science that studies how computers and humans interact. In the 1950s, Alan Turing published an article that proposed a measure of intelligence, now called the Turing test. More modern techniques, such as deep learning, have produced results in the fields of language modeling, parsing, and natural-language tasks.
An opensource text-to-speech (TTS) voice building tool
Deep neural network framework for multi-label text classification
Build LLM-powered Dart/Flutter applications.
Stanford Open Information Extraction made simple!
Library for faster pinned CPU <-> GPU transfer in Pytorch
🔥 Rankify: A Comprehensive Python Toolkit for Retrieval, Re-Ranking, and Retrieval-Augmented Generation 🔥. Our toolkit integrates 40 pre-retrieved b...
甲言,专注于古代汉语(古汉语/古文/文言文/文言)处理的NLP工具包,支持文言词库构建、分词、词性标注、断句和标点。Jiayan, the 1st NLP toolkit designed for C...
A Vietnamese natural language processing toolkit (NAACL 2018)
Ekphrasis is a text processing tool, geared towards text from social networks, such as Twitter or Facebook. Ekphrasis performs tokenization, word norm...
Library for clinical NLP with spaCy.
SpaCy 中文模型 | Models for SpaCy that support Chinese
NBoost is a scalable, search-api-boosting platform for deploying transformer models to improve the relevance of search results on different platforms...
A fast, low-resource Natural Language Processing and Text Correction library written in Rust.
A list of open-source AI projects you can use to generate income easily.
Find dates inside text using Python and get back datetime objects
Modern spell checking library - accurate, fast, multi-language
A curated list of resources for NLP (Natural Language Processing) for Korean
A Pretrained BERT Model for Financial Communications. https://arxiv.org/abs/2006.08097
A fast, lightweight and easy-to-use Python library for splitting text into semantically meaningful chunks.
自然语言处理工具Macropodus,基于Albert+BiLSTM+CRF深度学习网络架构,中文分词,词性标注,命名实体识别,新词发现,关键词,文本摘要,文本相似度,科学计算...
Quickly format your notes with ChatGPT in Obsidian
A Python multilingual toolkit for Sentiment Analysis and Social NLP tasks
聚宝盆(Cornucopia): 中文金融系列开源可商用大模型,并提供一套高效轻量化的垂直领域LLM训练框架(Pretraining、SFT、RLHF、Quantize等)
The first-ever vast natural language processing benchmark for Indonesian Language. We provide multiple downstream tasks, pre-trained IndoBERT models,...
Automatically split your PyTorch models on multiple GPUs for training & inference
Awesome-LLM-Eval: a curated list of tools, datasets/benchmark, demos, leaderboard, papers, docs and models, mainly for Evaluation on LLMs. 一个由工具...
中文Mixtral-8x7B(Chinese-Mixtral-8x7B)
Python framework for AI workflows and pipelines with chain of thought reasoning, external tools, and memory.
Active Learning for Text Classification in Python
A simplified PyTorch implementation of "SeqGAN: Sequence Generative Adversarial Nets with Policy Gradient." (Yu, Lantao, et al.)
This repository collects an extensive list of awesome papers about Story Generation / Storytelling, exclusively focusing on the era of Large Language...
ConvoKit is a toolkit for extracting conversational features and analyzing social phenomena in conversations. It includes several large conversational...
[ICLR 2024] Sheared LLaMA: Accelerating Language Model Pre-training via Structured Pruning
A foundational library for Semantic Hypergraphs
TextCNN Pytorch实现 中文文本分类 情感分析
AutoPrompt: Automatic Prompt Construction for Masked Language Models.
Core Data of HowNet and OpenHowNet Python API
Pretrained ELECTRA Model for Korean
An open platform for artificial intelligence, chat bots, virtual agents, social media automation, and live chat automation.
👁️ + 💬 + 🎧 = 🤖 Curated list of top foundation and multimodal models! [Paper + Code + Examples + Tutorials]
Homer, a text analyser in Python, can help make your text more clear, simple and useful for your readers.
Deep NLP Course
Transformers for Longer Sequences
Examples and libraries for "Natural Language Processing in Action" book
Accurately generate all possible forms of an English word e.g "election" --> "elect", "electoral", "electorate" etc.
[ICML 2024] TrustLLM: Trustworthiness in Large Language Models
HTML to Markdown converter and crawler.
专注于可解释的NLP技术 An NLP Toolset With A Focus on Explainable Inference
OpenChatBI is an intelligent chat-based BI tool powered by large language models, designed to help users query, analyze, and visualize data through na...
A resource repository for machine unlearning in large language models