Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Nicole Jones's picture

Nicole Jones

nicoleju
5 6
·
  • zofiasilva
  • milesrivera
  • ethanwilliams
  • ruthgupta

AI & ML interests

Pizza is my favorite algorithm

Recent Activity

upvoted a paper 1 day ago
LoGRA: Scaling LLM Reinforcement Learning with Low-Rank Gradient Sketches
upvoted a paper 1 day ago
TRACE: Rollout-Guided Quantization-Aware Training for FP4 Reinforcement Learning of MoE Language Models
upvoted a paper 1 day ago
False Frontiers: Diagnosing and Mitigating Co-Cheating in Self-Evolving Search Agents
View all activity

Organizations

None yet

upvoted 4 papers 1 day ago

LoGRA: Scaling LLM Reinforcement Learning with Low-Rank Gradient Sketches

Paper • 2610.06647 • Published 7 days ago • 148

TRACE: Rollout-Guided Quantization-Aware Training for FP4 Reinforcement Learning of MoE Language Models

Paper • 2610.07767 • Published 6 days ago • 92

False Frontiers: Diagnosing and Mitigating Co-Cheating in Self-Evolving Search Agents

Paper • 2609.39102 • Published 12 days ago • 603

The Teacher Is a Direction, Not a Destination: Extrapolating RL-Induced Representation Residuals in On-Policy Distillation

Paper • 2609.36484 • Published 13 days ago • 553
upvoted a paper about 1 month ago

FlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience

Paper • 2609.03241 • Published Sep 3 • 51
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs