FreedomIntelligence/TextClassificationBenchmark
quality grade D, 37 out of 100A Benchmark of Text Classification in PyTorch
- stars
- 608
- stars gained this week
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Tokenization, parsing, classical NLP pipelines, translation and information extraction.
Signals: nlp, natural-language-processing, tokenizer, named-entity-recognition, text-classification, machine-translation, sentiment-analysis, spacy
637 results
A Benchmark of Text Classification in PyTorch
BERTweet: A pre-trained language model for English Tweets (EMNLP-2020)
DoTAT 是一款基于web、面向领域的通用文本标注工具,支持大规模实体标注、关系标注、事件标注、文本分类、基于字典匹配和正则匹配的自动标注以及用于实现归一化的标准名标注,同时也支持迭代标注、嵌套实体标注和嵌套事件标注。标注规范可自定义且同类型任务中可“一次创建多次复用”。通过分级实体集合扩大了实体类型的规模,并设计了全新高效的标注方式,提升了用户体验和标注效率。此外,本工具增加了审核环节,可对多人的标注结果进行一致性检验、自动合并和手动调整,提高了标注结果的准确率。
Fake News Detection in Python
Code and source for paper ``How to Fine-Tune BERT for Text Classification?``
TextCNN Pytorch实现 中文文本分类 情感分析
Active Learning for Text Classification in Python
A Modern C++ Data Sciences Toolkit
中文文本分析工具包(包括- 文本分类 - 文本聚类 - 文本相似性 - 关键词抽取 - 关键短语抽取 - 情感分析 - 文本纠错 - 文本摘要 - 主题关键词-同义词、近义词-事件三元组抽取)
Implementation of papers for text classification task on DBpedia
Tensorflow implementation of attention mechanism for text classification tasks.
多标签文本分类,多标签分类,文本分类, multi-label, classifier, text classification, BERT, seq2seq,attention, multi-label-classification
Text classification models implemented in Keras, including: FastText, TextCNN, TextRNN, TextBiRNN, TextAttBiRNN, HAN, RCNN, RCNNVariant, etc.
Text classification using deep learning models in Pytorch
Collection of papers and resources for data augmentation for NLP.
👑 spaCy building blocks and visualizers for Streamlit apps
This repository introduces PIXIU, an open-source resource featuring the first financial large language models (LLMs), instruction tuning data, and evaluation benchmarks to holistically assess financial LLMs. Our goal is to continually push forward the open-source development of financial artificial intelligence (AI).
🎯🗯 Dataset generation for AI chatbots, NLP tasks, named entity recognition or text classification models using a simple DSL!
This repo contains a PyTorch implementation of a pretrained BERT model for multi-label text classification.
Pre-training of Deep Bidirectional Transformers for Language Understanding: pre-train TextCNN
A tool for learning vector representations of words and entities from Wikipedia
1 line for thousands of State of The Art NLP models in hundreds of languages The fastest and most accurate way to solve text problems.
Text classifier for Hierarchical Attention Networks for Document Classification
Natural language detection library for Rust. Try demo online: https://whatlang.org/
24,523 repositories in the index in total.