29 repos
hppRC/defsent
DefSent: Sentence Embeddings using Definition Sentences
22
6 commits
AnswerDotAI/ModernBERT
Bringing BERT into modernity via both architecture changes and scaling
1,720
231 commits
izzetemredemir/turkish-sentiment-analysis-with-bert
Simple NLP project. Turkish sentiment analysis with bert.
4
4 commits
asahi417/relbert
The official implementation of "Distilling Relation Embeddings from Pre-trained Language Models,…
48
688 commits
SKTBrain/KoBERT
Korean BERT pre-trained cased (KoBERT)
1,420
90 commits
Beomi/KcBERT
🤗 Pretrained BERT model & WordPiece tokenizer trained on Korean Comments 한국어 댓글로 프리트레이닝한 BERT 모델과…
497
34 commits
mbahadirk/Offensive-Text-Classification
Turkish Toxic Comment Classification
3
60 commits
sagorbrur/bangla-bert
Bangla-Bert is a pretrained bert model for Bengali language
84
54 commits
huggingface/tokenizers
💥 Fast State-of-the-Art Tokenizers optimized for Research and Production
11,018
2,000 commits
MuhammetCanGumussu/Bert-Implementation-From-Scratch
No description
2
69 commits
a-v-ershov/russian_sentiment_analysis
Sequence classification model for sentiment analysis of Russian conversational text
1
7 commits
matheuscamposmt/food-hazard-bert
A Natural Language Processing (NLP) project for SemEval 2025 Task 9, using DistilBERT with a custom…
20 commits
mlwithme/BertWithPretrained
An implementation of the BERT model and its related downstream tasks based on the PyTorch…
603
240 commits
naver/splade
SPLADE: sparse neural search (SIGIR21, SIGIR22)
1,011
113 commits
Hugging-Face-Supporter/tftokenizers
Use Huggingface Transformer and Tokenizers as Tensorflow Reusable SavedModels
10
l3cube-pune/MarathiNLP
Marathi NLP - is a repository dedicated to development of tools and resources for Marathi language.
163
160 commits
Turkish-Word-Embeddings/Word-Embeddings-Repository-for-Turkish
Code for "A Comprehensive Analysis of Static Word Embeddings for Turkish". Expert Systems with…
29
114 commits
KRLabsOrg/LettuceDetect
Span-level grounding verification for RAG, code, and tool-grounded AI outputs.
606
218 commits
voorhs/dialogue-augmentation
Research project on Contrastive Learning with Augmentations for Training Dialogue Embeddings
0
147 commits
MaartenGr/BERTopic
Leveraging BERT and c-TF-IDF to create easily interpretable topics.
7,831
188 commits
wisdomify/wisdomify
A BERT-based reverse dictionary of Korean proverbs
96
170 commits
gerzin/irony-and-sarcasm-detector-italian
NLP project that analyses Italian tweets and finds out if they are ironic or not, and if they're…
132 commits
legacyai/tf-transformers
State of the art faster Transformer with Tensorflow 2.0 ( NLP, Computer Vision, Audio ).
85
383 commits
umarbutler/emubert-creator
The training code behind EmuBert, the largest open-source masked language model for Australian law.
8 commits
isaacus-dev/emubert-creator
metehan777/google-rerank-tool
A Python cli-command tool for creating reports for any Google query.
8
9 commits
ukairia777/tensorflow-nlp-tutorial
tensorflow를 사용하여 텍스트 전처리부터, Topic Models, BERT, GPT, LLM과 같은 최신 모델의 다운스트림 태스크들을 정리한 Deep Learning…
581
265 commits
yachiashen/DeFake-ZH
Chinese Fake News Detection based on MacBERT
38 commits
indobenchmark/indonlu
The first-ever vast natural language processing benchmark for Indonesian Language. We provide…
655
104 commits