Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
๐
In a Training Loop
42.5
TFLOPS
Milton Montiel
miltmont
77
4
Follow
0 followers
ยท
2 following
https://discretized.dev
MiltMont
AI & ML interests
Reinforcement learning
Recent Activity
upvoted
a
paper
about 11 hours ago
Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement
upvoted
a
paper
4 days ago
Understanding Evolution Strategies for LLM Reasoning: Broader Reasoning Coverage than GRPO
upvoted
a
paper
7 days ago
On-Policy Self-Distillation in Diffusion Models
View all activity
Organizations
miltmont
's buckets
1
Sort:ย Recently updated
miltmont/traces
407 kB