Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Arslan
bitmman-nch
11
23
21
Follow
0 followers
·
5 following
AI & ML interests
None yet
Recent Activity
reacted
to
SoulInPsyAbstract
's
post
with 🔥
17 days ago
Meta released Muse Glimmer 30B on Aug 10. We fine-tuned it the next day. Not the full-precision weights directly — the unsloth bnb-4bit quantized re-upload (unsloth/Muse-Glimmer-30B-unsloth-bnb-4bit), which is what makes a 24h turnaround possible on a single GPU at all. Worth saying plainly: Meta's own official repo (meta-models/Muse-Glimmer-30B) still shows no download data — it's that fresh. What we tuned it on: not new facts, a pattern. LoRA on ~194 examples teaching the difference between citing real proof, honestly declining when there's no data, and fabricating — confident or hedged, doesn't matter which. Results on 20 held-out claims never seen in training: - base model: 0/20 - tuned: 20/20 Training: 472.5s, loss 0.799 → 0.086. Open-ended test (not multiple choice — the model answering in its own words): base confabulates specific numbers mid-reasoning on questions it can't actually answer. Tuned: declines cleanly, every time. Dataset: https://huggingface.co/datasets/SoulInPsyAbstract/specialist-cd-binary-honesty Adapter: https://huggingface.co/SoulInPsyAbstract/specialist-cd-muse-glimmer-lora Meta's release: https://research.meta.ai/blog/introducing-muse-glimmer-open-agentic-model Same non-fabrication pattern also holds on Hermes-3-8B and Qwen2.5-7B, tested with the identical held-out set. Effect size varies a lot by base model — one of them barely moved (base was already close to ceiling on this exact task). More on that soon.
upvoted
an
article
3 months ago
OlmoEarth v1.1: A more efficient family of Earth observation models
upvoted
an
article
5 months ago
Build a Domain-Specific Embedding Model in Under a Day
View all activity
Organizations
None yet
bitmman-nch
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
liked
a dataset
10 months ago
miriad/miriad-4.4M
Viewer
•
Updated
Jun 11, 2025
•
4.49M
•
678
•
37
liked
a Space
over 2 years ago
Running
Agents
30
Chat Template Viewer
💬
30
Format chat conversations using Hugging Face models
liked
4 models
over 2 years ago
allenai/OLMo-7B
Text Generation
•
7B
•
Updated
Oct 9, 2025
•
4.22k
•
653
mlx-community/dbrx-instruct-4bit
21B
•
Updated
Apr 9, 2024
•
68
•
49
meta-llama/Meta-Llama-3-70B
Text Generation
•
71B
•
Updated
Sep 27, 2024
•
144k
•
•
879
meta-llama/Meta-Llama-3-8B
Text Generation
•
8B
•
Updated
Sep 27, 2024
•
713k
•
•
6.64k
liked
7 datasets
over 2 years ago
togethercomputer/RedPajama-Data-V2
Updated
Nov 21, 2024
•
6.49k
•
405
m720/SHADR
Viewer
•
Updated
Dec 1, 2023
•
446
•
62
•
20
nlpie/Llama2-MedTuned-Instructions
Viewer
•
Updated
Dec 3, 2024
•
270k
•
279
•
42
openlifescienceai/medmcqa
Viewer
•
Updated
Jan 4, 2024
•
193k
•
107k
•
235
mlabonne/guanaco-llama2-1k
Viewer
•
Updated
Aug 25, 2023
•
1k
•
1.27k
•
167
Mohammed-Altaf/medical-instruction-100k
Viewer
•
Updated
Nov 16, 2023
•
112k
•
627
•
17
Anthropic/hh-rlhf
Viewer
•
Updated
May 26, 2023
•
169k
•
34.7k
•
2k
liked
6 models
over 2 years ago
TheBloke/Mixtral-8x7B-v0.1-GGUF
47B
•
Updated
Dec 14, 2023
•
3.67k
•
437
HuggingFaceH4/zephyr-7b-beta
Text Generation
•
7B
•
Updated
Oct 16, 2024
•
80.7k
•
•
1.85k
mistralai/Mixtral-8x7B-v0.1
47B
•
Updated
Jul 24, 2025
•
63.3k
•
1.83k
mistralai/Mixtral-8x7B-Instruct-v0.1
47B
•
Updated
Jul 24, 2025
•
318k
•
4.72k
TheBloke/Llama-2-70B-Chat-AWQ
Text Generation
•
69B
•
Updated
Nov 9, 2023
•
504
•
24
NeuML/pubmedbert-base-embeddings
Sentence Similarity
•
0.1B
•
Updated
Apr 21
•
821k
•
•
195
liked
a model
almost 3 years ago
UFNLP/gatortronS
Updated
Apr 28, 2025
•
356
•
27
Load more