Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🏗️
Building on HF
John Locke
johnlockejrr
85
26
183
Follow
yuxinlu1's profile picture
erikbuck's profile picture
MehreenSaeed's profile picture
42 followers
·
115 following
johnlockejrr
AI & ML interests
OCR, HTR, ATR, NLP, AI
Recent Activity
liked
a model
about 6 hours ago
johnlockejrr/LightOnOCR-2-1B-base-nena-ussr
liked
a dataset
about 16 hours ago
oddadmix/dialectal-arabic-lahgtna-v2
reacted
to
AbstractPhil
's
post
with 👀
about 16 hours ago
The AlephLM results are rolling in and I'm very excited for the possibilities. I am very much looking forward to the coming weeks as I train the first AlephLM distillations from MANY teachers into AMOE arms. The AMOE arms hook cleanly to AlephLM structures and provide pos/neg learning elements. Hard positive and hard negatives coalesce to extend the capacity. https://huggingface.co/AbstractPhil/alephlm-0 https://huggingface.co/AbstractPhil/alephlm-adopt-0 As it stands they are structurally sound enough to fully pretrain. As or more stable than a standard Bert experimentally to distill using InfoNCE. AMOE legs improve these structures substantially. Structural behavior can be expanded in many ways on distilled and pretrained models alike. Attaching the AMOE to any model I've tried has created expanded or improved behavioral accumulations. They do have downsides but their upsides are very experimentally exciting. I've distilled multiple vits, multiple berts, and have begun distilling berts into AlephLM structures successfully. This is overall very exciting for me. I've begun formatting larger variants such as including GPT-2 and Qwen 3.5 4b as a paired combinator utilizing pathological T5 learned distilled encodings. It sounds odd, but the results show everything can be expanded and even be taught to cooperate. The CaptionBert-8192-v2 and v2-b are both structurally collapsing after token 480 or so, which is expected due to the small train. By distilling an AMOE arm to V2 by training with a longformer expert, the results are cutting through like butter. V2 has begun stabilizing rapidly for considerably longer token chains and sequences, the structure is repairing and building reusable capacity. I have discovered an improved methodology for sampling the AlephLM for text encoder benchmarks, which is predominantly L2 normalized outputs. Upcoming large paper for the distillation experiments and results within the next week or two. It's going to be a big one.
View all activity
Organizations
johnlockejrr
's models
52
Sort: Recently updated
johnlockejrr/hebrew_sefaria_preprocessor
Updated
Apr 22, 2025
johnlockejrr/yolo11_syr_synth
Updated
Apr 21, 2025
johnlockejrr/pylaya_syr_synth
Updated
Apr 20, 2025
johnlockejrr/norhand-regions-lines
Updated
Apr 16, 2025
johnlockejrr/yolo_samaritan
Updated
Apr 10, 2025
johnlockejrr/medieval-manuscript-yolov11-seg
Object Detection
•
Updated
Apr 5, 2025
johnlockejrr/pylaia_catmus_medieval
Image-to-Text
•
Updated
Apr 3, 2025
•
1
johnlockejrr/pylaia-heb_sam_v1
Image-to-Text
•
Updated
Apr 1, 2025
johnlockejrr/pylaia-mcdonald_v2
Image-to-Text
•
Updated
Apr 1, 2025
johnlockejrr/pylaia-ar-hand_v1
Image-to-Text
•
Updated
Apr 1, 2025
johnlockejrr/pylaia-samaritan_v1
Image-to-Text
•
Updated
Apr 1, 2025
johnlockejrr/trocr_syr_v1
61.6M
•
Updated
Nov 2, 2024
•
4
johnlockejrr/syrnt_v2_20-epoch
Updated
Oct 23, 2024
•
6
johnlockejrr/syrnt_v2_13-epoch
Updated
Oct 21, 2024
•
4
johnlockejrr/syrnt_v2
Updated
Oct 21, 2024
•
15
johnlockejrr/sbb_pixelwise_v1
Updated
Oct 18, 2024
•
4
johnlockejrr/yolov8-samaritan-segmentation
Updated
Sep 8, 2024
johnlockejrr/doc_ufcn_samaritan_v2
Image Segmentation
•
Updated
Sep 8, 2024
johnlockejrr/doc_ufcn_samaritan_v1
Image Segmentation
•
Updated
Aug 30, 2024
johnlockejrr/doc-ufcn-samaritan
Image Segmentation
•
Updated
Jun 28, 2024
johnlockejrr/heBERT-finetuned-samaritan
Fill-Mask
•
0.1B
•
Updated
May 3, 2024
•
5
johnlockejrr/BEREL_2.0-sam-v3
Fill-Mask
•
0.2B
•
Updated
May 3, 2024
•
4
Previous
1
2
Next