Ā·
AI & ML interests
Retrieval & Small Models & PTBR Resources
Recent Activity
reacted to SeaWolf-AI's post with ā¤ļø 7 days ago š» Data-center AI, now on a laptop: POCKET-Darwin-180B
We're releasing a 4-bit GGUF build of Darwin-180B-RSI, #1 on seven official Hugging Face leaderboards (self-reported), that runs without a GPU.
š¦ 360 GB ā 111 GB (4-bit GGUF, 4 files)
š„ļø No GPU: one server CPU (16 threads) at 18.4ā21.0 tokens/s
š» RTX 5060 laptop (8 GB VRAM) + 32 GB RAM: 4.17 tokens/s
š§ 128 GB mini PC: whole model in memory, no GPU needed
šÆ MMLU-Pro, 2,000 questions, paired: original 87.65% = 4-bit 87.65%
How?
Ā· Only ~3B of 180B parameters are active per token (10 of 512 experts)
Ā· llama.cpp streams just the needed experts from SSD, so 32 GB RAM is enough
Ā· Graft quantization: we took the proven Unsloth UD-Q4_K_XL base build and swapped in only the 300 tensors our RSI training changed (300/300 verified)
Under the hood is Model-level Recursive Self-Improvement. The model solves verifiable problems, keeps only its own solutions that check out as correct, and trains on them. No human-written solutions or reasoning traces.
Built for teams that can't send data to an external cloud (defense, finance, public sector) to run a top-tier model fully offline.
š Article: https://huggingface.co/blog/FINAL-Bench/data-center-ai-now-on-a-laptop-pocket-darwin-180b
š¤ Model: https://huggingface.co/FINAL-Bench/POCKET-Darwin-180B-GGUF
𧬠Original: https://huggingface.co/FINAL-Bench/Darwin-180B-RSI
#Darwin #RSI #GGUF #llamacpp #OnDevice #MoE View all activity Organizations
cnmoro/Qwen2.5-0.5B-Chunk-Compressor-Q8_0-GGUF
Text Generation
⢠0.5B ⢠Updated ⢠23
cnmoro/Qwen2.5-0.5B-Rag-Thinking-Q8_0-GGUF
Text Generation
⢠0.5B ⢠Updated ⢠20
cnmoro/Qwen2.5-0.5B-Rag-Thinking
Text Generation
⢠0.5B ⢠Updated ⢠29
⢠7
cnmoro/Qwen2.5-0.5B-Chunk-Compressor
Text Generation
⢠0.5B ⢠Updated ⢠19
⢠5
cnmoro/QwenSlerp2-14B-Q3_K_M-GGUF
15B ⢠Updated ⢠11
⢠1
cnmoro/Qwen0.5b-RagSemanticChunker
Text Generation
⢠0.5B ⢠Updated ⢠29
⢠4
cnmoro/tangled-llama-33m-32k-instruct-v0.1-fix
33.3M ⢠Updated ⢠19
⢠1
cnmoro/granite-question-classifier
Text Classification
⢠30.3M ⢠Updated ⢠19
⢠2
cnmoro/nano-image-captioning
Image-to-Text
⢠10.1M ⢠Updated ⢠62
⢠3
cnmoro/tiny-image-captioning
Image-to-Text
⢠26.4M ⢠Updated ⢠237
⢠5
cnmoro/mini-image-captioning
Image-to-Text
⢠34.2M ⢠Updated ⢠225
⢠5
cnmoro/snowflake-arctic-embed-m-v2.0-cpu
Sentence Similarity
⢠0.3B ⢠Updated ⢠245
⢠4
cnmoro/Qwen3b-RagSemanticChunker
Text Generation
⢠3B ⢠Updated ⢠24
⢠2
cnmoro/micro-bertim-embeddings
Sentence Similarity
⢠4.43M ⢠Updated ⢠58
⢠1
cnmoro/static-retrieval-distilbert-ptbr
Sentence Similarity
⢠Updated ⢠4
4.43M ⢠Updated ⢠11
⢠1
cnmoro/bert-tiny-embeddings-english-portuguese
Sentence Similarity
⢠4.39M ⢠Updated ⢠110
⢠2
cnmoro/multilingual-e5-small-distilled-16m
16M ⢠Updated ⢠110
⢠1
cnmoro/TeenyTinyLlama-460m-Summarizer-PTBR
Text Generation
⢠0.5B ⢠Updated ⢠23
⢠1
cnmoro/Mistral-7B-Portuguese
Text Generation
⢠7B ⢠Updated ⢠27
⢠13
cnmoro/ptt5-small-named-entity-recognition
60.5M ⢠Updated ⢠7
⢠2
cnmoro/t5-small-named-entity-recognition
60.5M ⢠Updated ⢠12
⢠2
cnmoro/Mistral-7B-Portuguese-AWQ
Text Generation
⢠7B ⢠Updated ⢠24
⢠1
cnmoro/teenytinyllama-460m-text-simplification-ptbr
Text Generation
⢠0.5B ⢠Updated ⢠17
⢠2
cnmoro/teenytinyllama-160m-text-simplification-ptbr
Text Generation
⢠0.2B ⢠Updated ⢠41
⢠2
cnmoro/ahxt_llama2_xs_460M_experimental_ptbr_instruct
Text Generation
⢠Updated ⢠20
⢠4
cnmoro/ArgosTranslate-EN-PT
Updated ⢠2
cnmoro/ArgosTranslate-PT-EN
Updated ⢠1
cnmoro/TinyLlama-1.1B-intermediate-1.5T-PTBR-Instruct-v3-8k
Text Generation
⢠Updated ⢠29
⢠8
cnmoro/ptt5-base-ptbr-summarization
Summarization
⢠Updated ⢠16
⢠3