Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Milton Montiel's picture
๐Ÿ”„ In a Training Loop

Milton Montiel

miltmont
93 4
ยท
https://discretized.dev
  • MiltMont

AI & ML interests

Reinforcement learning

Recent Activity

upvoted a paper 5 days ago
The information geometry of large language models is shared, learned, and controllable
upvoted a paper 6 days ago
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses
upvoted a paper 6 days ago
OmniEdu: Open Foundation Models for Learning and Teaching
View all activity

Organizations

Prometeo's profile picture

liked a model about 1 month ago

Qwen/Qwen3.8-27B

Image-Text-to-Text โ€ข 28B โ€ข Updated Aug 14 โ€ข 7.02M โ€ข โ€ข 16.6k
liked a model about 2 months ago

LiquidAI/LFM2.5-2.6B-MLX

Text Generation โ€ข Updated Aug 6 โ€ข 27
liked a Space 2 months ago
Running
Featured
84

QED-Nano: Teaching a Tiny Model to Prove Hard Theorems

๐Ÿ“
84

Who needs 1T parameters? Olympiad proofs with a 4B model

liked a dataset 2 months ago

allenai/tulu-3-sft-mixture

Viewer โ€ข Updated Dec 2, 2024 โ€ข 939k โ€ข 54.8k โ€ข 265
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs