Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Arailym-tleubayeva
/
kazakh-hybrid-duplicate-detector
like
0
Sentence Similarity
Joblib
Arailym-tleubayeva/KazakhTextDuplicates
Kazakh
kazakh
text-similarity
duplicate-detection
plagiarism-detection
hybrid-model
License:
apache-2.0
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
8536af5
kazakh-hybrid-duplicate-detector
283 MB
Ctrl+K
Ctrl+K
1 contributor
History:
6 commits
This model has 2 files scanned as suspicious.
Show
files
Arailym-tleubayeva
Add hybrid Kazakh duplicate detection model
8536af5
verified
6 months ago
.gitattributes
Safe
1.52 kB
initial commit
6 months ago
README.md
2.88 kB
Update README.md
6 months ago
count_vectorizer.joblib
Suspicious
pickle
Detected Pickle imports (2)
"sklearn.feature_extraction.text.CountVectorizer"
,
"numpy.int64"
How to fix it?
7.94 MB
xet
Add hybrid Kazakh duplicate detection model
6 months ago
lda_model.joblib
pickle
Detected Pickle imports (7)
"joblib.numpy_pickle.NumpyArrayWrapper"
,
"numpy.random._pickle.__randomstate_ctor"
,
"numpy.ndarray"
,
"sklearn.decomposition._lda.LatentDirichletAllocation"
,
"numpy.random._pickle.__bit_generator_ctor"
,
"numpy.dtype"
,
"numpy.random._mt19937.MT19937"
How to fix it?
19.1 MB
xet
Add hybrid Kazakh duplicate detection model
6 months ago
lsa_svd.joblib
pickle
Detected Pickle imports (4)
"sklearn.decomposition._truncated_svd.TruncatedSVD"
,
"joblib.numpy_pickle.NumpyArrayWrapper"
,
"numpy.dtype"
,
"numpy.ndarray"
How to fix it?
246 MB
xet
Add hybrid Kazakh duplicate detection model
6 months ago
meta.json
371 Bytes
Add hybrid Kazakh duplicate detection model
6 months ago
tfidf_vectorizer.joblib
Suspicious
pickle
Detected Pickle imports (6)
"joblib.numpy_pickle.NumpyArrayWrapper"
,
"sklearn.feature_extraction.text.TfidfTransformer"
,
"numpy.float64"
,
"numpy.dtype"
,
"numpy.ndarray"
,
"sklearn.feature_extraction.text.TfidfVectorizer"
How to fix it?
10.4 MB
xet
Add hybrid Kazakh duplicate detection model
6 months ago