Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Minh Le Duc
PRO
minhleduc
9
15
Follow
minhanhlt's profile picture
1 follower
·
17 following
MinLee0210
minh-le-duc-a62863172
AI & ML interests
Edge Device-oriented Models (language, audio, ...)
Recent Activity
updated
a dataset
1 day ago
LakoreAI/vivid-vietnamese-idioms
published
a dataset
1 day ago
LakoreAI/vivid-vietnamese-idioms
reacted
to
Felladrin
's
post
with 👍
16 days ago
I've open-sourced the trainer I've been using to build tiny language models from scratch, together with the 95M base model I trained with it. The trainer runs on Deno (https://deno.com, cross-platform), trains on WebGPU, and it writes GGUF directly. No Python/PyTorch. The weights live in a GGUF file from the first step to the last, so every checkpoint is already something llama.cpp can load. The model is https://huggingface.co/Felladrin/Minueza-3-95M-Base: 94.7M parameters, 1.95B tokens seen, 8192 context. And here’s the repository on GitHub: https://github.com/felladrin/gguf-trainer Here on Hugging Face, I published the optimizer state next to the weights, so you can continue the pretraining instead of starting over. Or start your own from nothing: `deno run -A cli.ts demo` trains a tiny one end to end in under a minute. And the docs are written for coding agents, so you can point your agent of choice at the GitHub repo and have it drive the whole pipeline.
View all activity
Organizations
Collections
1
LLM
TinyLlama: An Open-Source Small Language Model
Paper
•
2401.02385
•
Published
Jan 4, 2024
•
96
LLM
TinyLlama: An Open-Source Small Language Model
Paper
•
2401.02385
•
Published
Jan 4, 2024
•
96
models
2
Sort: Recently updated
minhleduc/xlm-roberta-multilang-finetuned-00
Text Classification
•
0.3B
•
Updated
Jul 26, 2025
•
4
minhleduc/toxigen-albert-binary-clsf
11.7M
•
Updated
Aug 16, 2024
•
10
datasets
4
Sort: Recently updated
minhleduc/wiki_famous_person_00
Viewer
•
Updated
Aug 3, 2025
•
1.3k
•
6
minhleduc/sentiment-classification-vietnamese-v1
Viewer
•
Updated
Jul 18, 2025
•
6.93k
•
12
minhleduc/multilang-classify-dataset-02
Viewer
•
Updated
Jul 13, 2025
•
119k
•
31
minhleduc/multilang-classify-dataset-01
Viewer
•
Updated
Jul 5, 2025
•
32.1k
•
10