Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Ta Xuan Duy
moiduy04
2
Follow
0 followers
·
5 following
AI & ML interests
None yet
Recent Activity
reacted
to
sergiopaniego
's
post
with 👍
16 days ago
Frontier models use distillation as a step of their post-training pipelines. In 2026 it has three jobs: compress a big model into a small one, merge RL experts into a single model, and let a model teach itself. I wrote up which frontier models use each one and how: https://huggingface.co/blog/sergiopaniego/distillation-2026 It pairs with Class 2 of the Training an Agent series Ben and I are doing, where we teach these techniques hands-on with TRL!
updated
a model
over 1 year ago
moiduy04/Llama-3-6.6B-LLM-Pruned
updated
a model
over 1 year ago
moiduy04/Llama-3-6.6B-R-Pruned
View all activity
Organizations
models
5
Sort: Recently updated
moiduy04/Llama-3-6.6B-LLM-Pruned
Text Generation
•
Updated
Feb 28, 2025
•
1
moiduy04/Llama-3-6.6B-R-Pruned
Text Generation
•
Updated
Feb 28, 2025
•
1
moiduy04/homemade_lo_vi
Text Generation
•
Updated
Jan 4, 2024
•
1
moiduy04/q-Taxi-v3
Reinforcement Learning
•
Updated
Jun 15, 2023
moiduy04/q-FrozenLake-v1-4x4-noSlippery
Reinforcement Learning
•
Updated
Jun 15, 2023
datasets
0
None public yet