LM/hf: Difference between revisions

From Fundamental Ramen
< LM
Jump to navigation Jump to search
No edit summary
 
(One intermediate revision by the same user not shown)
Line 25: Line 25:
# search for latest tunning
# search for latest tunning
hf models ls --search "Muse-Glimmer" --apps llama.cpp --expand "downloads,likes,createdAt,lastModified" --sort last_modified --no-truncate --limit 25
hf models ls --search "Muse-Glimmer" --apps llama.cpp --expand "downloads,likes,createdAt,lastModified" --sort last_modified --no-truncate --limit 25
# search OpenVINO format
hf models ls --search "Qwen3.8" --filter openvino --expand "downloads,likes,createdAt,lastModified" --sort downloads --no-truncate --limit 25
</syntaxhighlight>
</syntaxhighlight>
|-
|-
Line 47: Line 49:


<syntaxhighlight lang="bash">
<syntaxhighlight lang="bash">
 
sudo apt update
sudo apt install pipx -y
pipx ensurepath
pipx install "huggingface_hub[cli]"
</syntaxhighlight>
</syntaxhighlight>



Latest revision as of 03:25, 27 August 2026

Quick references

Purpose Command
cache management
hf cache list
hf cache rm <model id>
hf cache prune
fix WiFi problem
HF_XET_FIXED_DOWNLOAD_CONCURRENCY=10 hf download "unsloth/Qwen3-Coder-30B-A3B-Instruct-GGUF" --include "*UD-Q4_K_XL*"
HF_XET_FIXED_DOWNLOAD_CONCURRENCY=10 hf download "unsloth/Devstral-Small-2-24B-Instruct-2512-GGUF" --include "*UD-Q4_K_XL*"
search
# search by population
hf models ls --search "Muse-Glimmer" --apps llama.cpp --expand "downloads,likes,createdAt,lastModified" --sort downloads --no-truncate --limit 25
# search for latest publish
hf models ls --search "Muse-Glimmer" --apps llama.cpp --expand "downloads,likes,createdAt,lastModified" --sort created_at --no-truncate --limit 25
# search for latest tunning
hf models ls --search "Muse-Glimmer" --apps llama.cpp --expand "downloads,likes,createdAt,lastModified" --sort last_modified --no-truncate --limit 25
# search OpenVINO format
hf models ls --search "Qwen3.8" --filter openvino --expand "downloads,likes,createdAt,lastModified" --sort downloads --no-truncate --limit 25
optimize Muse Glimmer
hf models ls -h unsloth/Muse-Glimmer-30B-GGUF
hf download unsloth/Muse-Glimmer-30B-GGUF --include "*UD-Q4_K_XL*"
hf download unsloth/Muse-Glimmer-30B-GGUF --include "*UD-Q5_K_XL*"
hf download unsloth/Muse-Glimmer-30B-GGUF --include "*UD-Q6_K_XL*"
hf download unsloth/Muse-Glimmer-30B-GGUF --include "dflash-kquant.gguf"
hf download unsloth/Muse-Glimmer-30B-GGUF --include "mmproj-kquant.gguf"
tree ~/.cache/huggingface/hub/models--unsloth--Devstral-Small-2-24B-Instruct-2512-GGUF/snapshots
tree ~/.cache/huggingface/hub/models--unsloth--Qwen3-Coder-30B-A3B-Instruct-GGUF/snapshots

Install

sudo apt update
sudo apt install pipx -y
pipx ensurepath
pipx install "huggingface_hub[cli]"

Environment Variables for hf