Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🏗️
Building on HF
John Locke
johnlockejrr
85
26
182
Follow
darkc0de's profile picture
cantillation's profile picture
MikeHarris's profile picture
42 followers
·
115 following
johnlockejrr
AI & ML interests
OCR, HTR, ATR, NLP, AI
Recent Activity
liked
a dataset
about 10 hours ago
oddadmix/dialectal-arabic-lahgtna-v2
reacted
to
AbstractPhil
's
post
with 👀
about 10 hours ago
The AlephLM results are rolling in and I'm very excited for the possibilities. I am very much looking forward to the coming weeks as I train the first AlephLM distillations from MANY teachers into AMOE arms. The AMOE arms hook cleanly to AlephLM structures and provide pos/neg learning elements. Hard positive and hard negatives coalesce to extend the capacity. https://huggingface.co/AbstractPhil/alephlm-0 https://huggingface.co/AbstractPhil/alephlm-adopt-0 As it stands they are structurally sound enough to fully pretrain. As or more stable than a standard Bert experimentally to distill using InfoNCE. AMOE legs improve these structures substantially. Structural behavior can be expanded in many ways on distilled and pretrained models alike. Attaching the AMOE to any model I've tried has created expanded or improved behavioral accumulations. They do have downsides but their upsides are very experimentally exciting. I've distilled multiple vits, multiple berts, and have begun distilling berts into AlephLM structures successfully. This is overall very exciting for me. I've begun formatting larger variants such as including GPT-2 and Qwen 3.5 4b as a paired combinator utilizing pathological T5 learned distilled encodings. It sounds odd, but the results show everything can be expanded and even be taught to cooperate. The CaptionBert-8192-v2 and v2-b are both structurally collapsing after token 480 or so, which is expected due to the small train. By distilling an AMOE arm to V2 by training with a longformer expert, the results are cutting through like butter. V2 has begun stabilizing rapidly for considerably longer token chains and sequences, the structure is repairing and building reusable capacity. I have discovered an improved methodology for sampling the AlephLM for text encoder benchmarks, which is predominantly L2 normalized outputs. Upcoming large paper for the distillation experiments and results within the next week or two. It's going to be a big one.
upvoted
a
paper
about 24 hours ago
ChronoLens: Measuring Language Change Across Time, Languages, and Linguistic Levels
View all activity
Organizations
johnlockejrr
's models
52
Sort: Recently updated
johnlockejrr/ppocrv6-sam-heb
Image-to-Text
•
Updated
7 days ago
•
1
johnlockejrr/LightOnOCR-2-1B-base-nena-ussr
Image-Text-to-Text
•
1B
•
Updated
Mar 16
•
6
johnlockejrr/lagfart_yolo
Updated
Feb 25
•
4
johnlockejrr/Qwen2.5-14B-Instruct-mxfp4
Text Generation
•
15B
•
Updated
Feb 19
•
22
johnlockejrr/Qwen2.5-7B-Instruct-mxfp4
Text Generation
•
1B
•
Updated
Feb 19
•
20
johnlockejrr/Qwen2.5-Coder-14b-mxfp4
Text Generation
•
15B
•
Updated
Feb 19
•
50
•
1
johnlockejrr/GLM-OCR-samaritan
Image-to-Text
•
1B
•
Updated
Feb 15
•
25
•
1
johnlockejrr/LightOnOCR-2-1B-base-samaritan
Image-Text-to-Text
•
1B
•
Updated
Feb 14
•
5
johnlockejrr/LightOnOCR-2-1B-base-samaritan-GGUF
Image-Text-to-Text
•
0.6B
•
Updated
Feb 13
•
19
johnlockejrr/Nordic_Multicentury
Object Detection
•
Updated
Jan 4
•
6
•
1
johnlockejrr/marianmt_syr_voc_eastern
Translation
•
0.2B
•
Updated
Nov 15, 2025
•
18
johnlockejrr/marianmt_heb_voc
Translation
•
61.4M
•
Updated
Nov 10, 2025
•
6
johnlockejrr/marianmt_syr_voc_western
Translation
•
0.2B
•
Updated
Nov 10, 2025
•
10
johnlockejrr/marianmt-smp-sam-onnx
Translation
•
Updated
Nov 2, 2025
•
7
johnlockejrr/marianmt-smp-sam
Translation
•
61.4M
•
Updated
Nov 2, 2025
•
6
johnlockejrr/aramaic-diacritization-model
61.4M
•
Updated
Jul 30, 2025
•
5
johnlockejrr/opus-arc-targum-vocalization
61.4M
•
Updated
Jul 25, 2025
•
3
johnlockejrr/marianmt-he2arc-targum-voc-shva
Translation
•
61.4M
•
Updated
Jul 22, 2025
•
5
johnlockejrr/marianmt-he2arc-targum-voc
Translation
•
61.4M
•
Updated
Jul 22, 2025
•
5
•
1
johnlockejrr/marianmt-he2yid-tanakh
Translation
•
77.1M
•
Updated
Jul 21, 2025
•
5
•
1
johnlockejrr/marianmt-he2arc-targum
Translation
•
61.4M
•
Updated
Jul 15, 2025
•
3
•
1
johnlockejrr/marianmt-he2arc-sam
Translation
•
61.4M
•
Updated
Jul 15, 2025
•
3
•
1
johnlockejrr/marianmt-en2he-nwt
77.9M
•
Updated
Jul 14, 2025
•
4
•
1
johnlockejrr/marianmt-he2en-nwt
77.1M
•
Updated
Jul 13, 2025
•
3
•
1
johnlockejrr/opus-mt-arc-heb
77.1M
•
Updated
Jun 25, 2025
•
6
johnlockejrr/eynollah-sam_40_mss-patches
Updated
May 10, 2025
•
10
johnlockejrr/eynollah-sam_40_mss-no_patches
Updated
May 10, 2025
•
13
johnlockejrr/pylaia-heb_synth_lines_pytorch_2
Text Generation
•
Updated
Apr 30, 2025
johnlockejrr/pylaia-yiddish_synth_pytorch_2
Image-to-Text
•
Updated
Apr 26, 2025
•
1
johnlockejrr/hebrew_sefaria_tokenizer
Updated
Apr 22, 2025
Previous
1
2
Next