Ā·
AI & ML interests
Retrieval & Small Models & PTBR Resources
Recent Activity
reacted to SeaWolf-AI's post with ā¤ļø about 21 hours ago š» Data-center AI, now on a laptop: POCKET-Darwin-180B
We're releasing a 4-bit GGUF build of Darwin-180B-RSI, #1 on seven official Hugging Face leaderboards (self-reported), that runs without a GPU.
š¦ 360 GB ā 111 GB (4-bit GGUF, 4 files)
š„ļø No GPU: one server CPU (16 threads) at 18.4ā21.0 tokens/s
š» RTX 5060 laptop (8 GB VRAM) + 32 GB RAM: 4.17 tokens/s
š§ 128 GB mini PC: whole model in memory, no GPU needed
šÆ MMLU-Pro, 2,000 questions, paired: original 87.65% = 4-bit 87.65%
How?
Ā· Only ~3B of 180B parameters are active per token (10 of 512 experts)
Ā· llama.cpp streams just the needed experts from SSD, so 32 GB RAM is enough
Ā· Graft quantization: we took the proven Unsloth UD-Q4_K_XL base build and swapped in only the 300 tensors our RSI training changed (300/300 verified)
Under the hood is Model-level Recursive Self-Improvement. The model solves verifiable problems, keeps only its own solutions that check out as correct, and trains on them. No human-written solutions or reasoning traces.
Built for teams that can't send data to an external cloud (defense, finance, public sector) to run a top-tier model fully offline.
š Article: https://huggingface.co/blog/FINAL-Bench/data-center-ai-now-on-a-laptop-pocket-darwin-180b
š¤ Model: https://huggingface.co/FINAL-Bench/POCKET-Darwin-180B-GGUF
𧬠Original: https://huggingface.co/FINAL-Bench/Darwin-180B-RSI
#Darwin #RSI #GGUF #llamacpp #OnDevice #MoE View all activity Organizations
Other
⢠Updated ⢠2
cnmoro/static-nomic-384-pten-v2-st
Feature Extraction
⢠0.1B ⢠Updated ⢠376
⢠2
cnmoro/bert-base-portuguese-cased-safetensors
0.1B ⢠Updated ⢠10
cnmoro/domain-labeler-enpt
Text Classification
⢠Updated ⢠15
⢠2
cnmoro/Qwen3.5-2B-Pruned-English-Portuguese
Text Generation
⢠2B ⢠Updated ⢠184
⢠1
cnmoro/portuguese-multilingual-e5-small
Sentence Similarity
⢠39.8M ⢠Updated ⢠318
⢠1
cnmoro/inference-free-splade-co-condenser-en-ptbr-v2
Sentence Similarity
⢠Updated ⢠1
cnmoro/static-nomic-384-pten-v2
Feature Extraction
⢠12.8M ⢠Updated ⢠80
⢠2
cnmoro/nomic-embed-text-v2-moe-distilled-high-quality
Feature Extraction
⢠0.2B ⢠Updated ⢠39.2k
⢠⢠5
cnmoro/inference-free-splade-co-condenser-en-ptbr
Sentence Similarity
⢠Updated ⢠1
cnmoro/inference-free-splade-bert-tiny-en-ptbr
Sentence Similarity
⢠Updated 3B ⢠Updated ⢠39
cnmoro/static-nomic-384-pten
Feature Extraction
⢠12.8M ⢠Updated ⢠233
⢠1
cnmoro/RotorQuant-ModelWeights-Runtime
cnmoro/Qwen2.5-0.5B-Portuguese-v1
Text Generation
⢠0.5B ⢠Updated ⢠47
⢠5
cnmoro/LFM2-PTBR-imatrix-Quants
3B ⢠Updated ⢠36
Feature Extraction
⢠16.6M ⢠Updated ⢠39
⢠2
cnmoro/low-dimension-static-model
15k ⢠Updated ⢠8
cnmoro/custom-model2vec-tokenlearn-medium
64M ⢠Updated ⢠13
⢠1
cnmoro/custom-model2vec-tokenlearn-small
2.56M ⢠Updated ⢠18
⢠1
cnmoro/Qwen3-16B-A3B-REAP-PTBR
16B ⢠Updated ⢠18
⢠1
cnmoro/Qwen3-7B-A3B-REAP-PTBR
7B ⢠Updated ⢠20
⢠2
cnmoro/Qwen3-7B-A3B-REAP-PTBR-Q4_K_M-GGUF
7B ⢠Updated ⢠91
⢠1
cnmoro/Qwen3-16B-A3B-REAP-PTBR-SuperExp
16B ⢠Updated ⢠15
⢠1
cnmoro/distilbert-portuguese-tokenizer-lower-greedy
Updated
cnmoro/bert-hash-femto-mlm
Fill-Mask
⢠1.8M ⢠Updated ⢠17
cnmoro/static-nomic-distilled-vocab-quantized
1.78M ⢠Updated ⢠23
cnmoro/gpt-oss-20b-tokenizer-optional-reasoning
Updated ⢠3
cnmoro/gliclass-edge-v3.0-onnx
Text Classification
⢠Updated ⢠1
cnmoro/gliclass-large-v3.0-onnx
Text Classification
⢠Updated ⢠1