Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Pranav Upadhyaya
pranavupadhyaya52
6
Follow
webxos's profile picture
cascade-legend's profile picture
Quazim0t0's profile picture
17 followers
ยท
5 following
AI & ML interests
None yet
Recent Activity
reacted
to
LH-Tech-AI
's
post
with ๐ฅ
about 22 hours ago
Supra2-100M is out! Go check it out: - https://www.reddit.com/r/LocalLLaMA/comments/1velyl9/new_models_supra2100m_base_and_instruct_go_check/ - https://huggingface.co/SupraLabs/Supra2-100M - https://huggingface.co/SupraLabs/Supra2-100M-Instruct Give us a like and a follow!! HAVE FUN ๐ค๐ฅ๐ more coming soon...
reacted
to
ucr-max
's
post
with ๐ฅ
13 days ago
Introducing Limen0.2B We are releasing Limen0.2B, a 222.5M-parameter base language model developed as a research platform for efficient pretraining and superword tokenization at smaller scales. Limen0.2B was trained from scratch on 50B tokens and uses a compact 16K BoundlessBPE vocabulary. The project explores whether SuperBPE-style tokenization can remain effective in a substantially smaller model and vocabulary regime than those examined in earlier large-scale experiments. The model also combines a deep-and-narrow transformer design with Exclusive Self-Attention, grouped-query attention, and tied embeddings. Its compact vocabulary reduces the embedding footprint and leaves a larger share of the parameter budget available to the transformer layers. Despite its relatively modest training budget, Limen0.2B achieves competitive results for its scale across the reported language understanding, commonsense reasoning, and grammatical evaluation tasks. Comparisons with other compact models are provided as context rather than strict rankings, as their training data, token budgets, architectures, and evaluation settings differ. The release includes the model weights, implementation, training configuration, checkpoint progression, and evaluation results, all under Apache 2.0. https://huggingface.co/UniversalComputingResearch/Limen0.2B Technical feedback, independent evaluations, and further experiments with the model and tokenizer are welcome.
reacted
to
ucr-max
's
post
with ๐ฅ
13 days ago
Introducing Limen0.2B We are releasing Limen0.2B, a 222.5M-parameter base language model developed as a research platform for efficient pretraining and superword tokenization at smaller scales. Limen0.2B was trained from scratch on 50B tokens and uses a compact 16K BoundlessBPE vocabulary. The project explores whether SuperBPE-style tokenization can remain effective in a substantially smaller model and vocabulary regime than those examined in earlier large-scale experiments. The model also combines a deep-and-narrow transformer design with Exclusive Self-Attention, grouped-query attention, and tied embeddings. Its compact vocabulary reduces the embedding footprint and leaves a larger share of the parameter budget available to the transformer layers. Despite its relatively modest training budget, Limen0.2B achieves competitive results for its scale across the reported language understanding, commonsense reasoning, and grammatical evaluation tasks. Comparisons with other compact models are provided as context rather than strict rankings, as their training data, token budgets, architectures, and evaluation settings differ. The release includes the model weights, implementation, training configuration, checkpoint progression, and evaluation results, all under Apache 2.0. https://huggingface.co/UniversalComputingResearch/Limen0.2B Technical feedback, independent evaluations, and further experiments with the model and tokenizer are welcome.
View all activity
Organizations
pranavupadhyaya52
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
liked
4 datasets
3 months ago
lee-chang-93/qm9_magic
Viewer
โข
Updated
Jul 14, 2025
โข
2.47M
โข
11
โข
1
jonghyunlee/ZINC15
Viewer
โข
Updated
Nov 27, 2024
โข
36M
โข
50
โข
1
sagawa/pubchem-10m-canonicalized
Viewer
โข
Updated
Sep 4, 2022
โข
10M
โข
1.02k
โข
7
lukaskim/ChEMBL-36
Viewer
โข
Updated
24 days ago
โข
6.38M
โข
541
โข
1
liked
a dataset
4 months ago
meryyllebr543/lunaris-ultrafineweb-20b-tokenized
Updated
Jul 27, 2025
โข
165
โข
3
liked
a model
about 1 year ago
unsloth/Qwen2.5-Omni-7B-GGUF
Any-to-Any
โข
8B
โข
Updated
3 days ago
โข
7.41k
โข
76