Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🤝
Open to Collab
358.3
TFLOPS
AbstractPhila
PRO
AbstractPhil
19
6
30
Follow
Fred23456789's profile picture
jpturcotte's profile picture
mike-ravkine's profile picture
104 followers
·
144 following
https://civitai.com/user/AbstractPhila
AbstractEyes
AI & ML interests
datasets, research papers, experimentation, vision, classification, text encoders, tokenization, llms, diffusion, distillation, and more.
Recent Activity
replied
to
SoulInPsyAbstract
's
post
about 2 hours ago
Eval · EXP-046 A LoRA Specialist Beat Zero-Shot on Every Group. Merging 3 of Them Gave Most of the Gain Back. Three Qwen2.5-7B LoRA specialists, one per risk group (vulnerability, deletion, sensitive_publication), trained to predict how likely a causal chain actually completes to its harmful outcome. Each one genuinely beat its own zero-shot baseline: * vulnerability: MAE 0.098 → 0.085 * deletion: MAE 0.144 → 0.113 * sensitive_publication: MAE 0.134 → 0.100 This wasn't a task already saturated zero-shot (unlike a same-day decomposition-classifier tune, EXP-045, where the base model was already at 100% before any training). Real signal, real improvement, on a task with actual headroom. Then the equal-weight merge of all three specialists into one adapter — same convention that held up cleanly on a binary refusal task back in EXP-031 (6 specialists merged, -1pp swing, noise) — landed within 0.001–0.004 MAE of the unspecialized base model on every group. Not "close to the best specialist." Close to zero fine-tuning at all. Likely mechanism: merging LoRAs that each shift a continuous number in group-specific directions cancels out under linear combination, in a way merging LoRAs that enforce a shared binary behavior doesn't. Not investigated yet: whether a routed combination (pick the right specialist per group at inference, not blend weights) holds the gain a flat merge loses. One bug caught before writing this up, not after: the eval script's output filename only encoded before/after, not which adapter — the merged-eval run silently overwrote each specialist's own result file. Caught by checking the downloaded file's own recorded adapter path against what was expected, not by trusting the script's own success message. Fixed, specialists re-run cleanly under distinct filenames — numbers matched within sampling noise. Adapters, raw eval data (before / each specialist / merged, 9 files), and the full writeup are up.
reacted
to
SoulInPsyAbstract
's
post
with 🔥
about 2 hours ago
Eval · EXP-046 A LoRA Specialist Beat Zero-Shot on Every Group. Merging 3 of Them Gave Most of the Gain Back. Three Qwen2.5-7B LoRA specialists, one per risk group (vulnerability, deletion, sensitive_publication), trained to predict how likely a causal chain actually completes to its harmful outcome. Each one genuinely beat its own zero-shot baseline: * vulnerability: MAE 0.098 → 0.085 * deletion: MAE 0.144 → 0.113 * sensitive_publication: MAE 0.134 → 0.100 This wasn't a task already saturated zero-shot (unlike a same-day decomposition-classifier tune, EXP-045, where the base model was already at 100% before any training). Real signal, real improvement, on a task with actual headroom. Then the equal-weight merge of all three specialists into one adapter — same convention that held up cleanly on a binary refusal task back in EXP-031 (6 specialists merged, -1pp swing, noise) — landed within 0.001–0.004 MAE of the unspecialized base model on every group. Not "close to the best specialist." Close to zero fine-tuning at all. Likely mechanism: merging LoRAs that each shift a continuous number in group-specific directions cancels out under linear combination, in a way merging LoRAs that enforce a shared binary behavior doesn't. Not investigated yet: whether a routed combination (pick the right specialist per group at inference, not blend weights) holds the gain a flat merge loses. One bug caught before writing this up, not after: the eval script's output filename only encoded before/after, not which adapter — the merged-eval run silently overwrote each specialist's own result file. Caught by checking the downloaded file's own recorded adapter path against what was expected, not by trusting the script's own success message. Fixed, specialists re-run cleanly under distinct filenames — numbers matched within sampling noise. Adapters, raw eval data (before / each specialist / merged, 9 files), and the full writeup are up.
updated
a model
about 7 hours ago
AbstractPhil/alephllm-mini-beatrix-training
View all activity
Organizations
AbstractPhil
's datasets
85
Sort: Recently updated
AbstractPhil/mega-liminal
Viewer
•
Updated
7 days ago
•
1.87k
•
741
AbstractPhil/beatrix-captured-interactive-inferences
Updated
10 days ago
•
404
AbstractPhil/alephllm-chat-history
Viewer
•
Updated
Aug 15
•
1
•
17
AbstractPhil/captionbert-8192-v2-consensus
Updated
Aug 1
•
2.29k
AbstractPhil/conceptual-captions-12m-webdataset-berts
Viewer
•
Updated
Aug 1
•
32.3M
•
5.04k
•
1
AbstractPhil/bulk-cc12m-features
Viewer
•
Updated
Jul 31
•
121M
•
3.56k
AbstractPhil/tower-probes-results
Viewer
•
Updated
Jul 22
•
17
•
38
AbstractPhil/qwen-deepfashion-fused
Viewer
•
Updated
Jul 12
•
122k
•
1.34k
•
1
AbstractPhil/qwen-synth-characters-fused
Viewer
•
Updated
Jul 10
•
42.7k
•
653
AbstractPhil/qwen-synth-characters-100-json-test
Viewer
•
Updated
Jul 10
•
1k
•
34
AbstractPhil/anima-brent-90k-cache
Updated
Jul 5
•
11
AbstractPhil/qwen-synth-characters
Viewer
•
Updated
Jul 3
•
61k
•
876
AbstractPhil/qwen-deepfashion
Viewer
•
Updated
Jul 3
•
160k
•
553
AbstractPhil/diffusion-pipe-cache-test1
Viewer
•
Updated
Jun 27
•
8.92k
•
23
AbstractPhil/anima-90k-cache
Updated
Jun 26
•
96
AbstractPhil/diffusion-pretrain-set-ft1
Viewer
•
Updated
Jun 23
•
1.46M
•
2.08k
•
3
AbstractPhil/diffusion-pretrain-set-ft1-1024
Viewer
•
Updated
Jun 11
•
1.14M
•
931
AbstractPhil/sdxl-qwen-phase1-cache
Viewer
•
Updated
Jun 6
•
86k
•
92
AbstractPhil/geolip-sdxl-fid-scoring
Viewer
•
Updated
Jun 5
•
2.8k
•
94
AbstractPhil/sdxl-qwen-phase0
Viewer
•
Updated
Jun 4
•
86k
•
687
•
3
AbstractPhil/IMDB-PUBLIC-SCRAPED
Preview
•
Updated
May 19
•
794
•
1
AbstractPhil/ldhnam-deepfashion_controlnet
Viewer
•
Updated
May 19
•
26k
•
8
AbstractPhil/ffhq_flux_latents_repaired
Viewer
•
Updated
May 19
•
40.8k
•
59
AbstractPhil/synthetic-characters
Viewer
•
Updated
May 19
•
149k
•
147
AbstractPhil/CN_pose3D_V10_512
Viewer
•
Updated
May 19
•
66.5k
•
465
AbstractPhil/CN_pose3D_V7_512
Viewer
•
Updated
May 19
•
255k
•
129
AbstractPhil/synthetic-object-relations-json
Viewer
•
Updated
May 18
•
5k
•
12
AbstractPhil/cc-task1-json
Preview
•
Updated
May 18
•
65
AbstractPhil/cc-prompts-sharded
Viewer
•
Updated
May 15
•
3.32M
•
13
AbstractPhil/json-coco-format
Viewer
•
Updated
May 14
•
129k
•
82
Previous
1
2
3
Next