Production Artifact Mirror
/
downloads
/
tree
Expand All
Collapse All
models/
3 suites
llama-3.1-8b-instruct/
7 files
llama-3.1-8b-instruct.Q4_K_M.gguf
GGUF Quantized
4.92 GB
GET
tokenizer.json
Fast Tokenizer
2.10 MB
GET
tokenizer_config.json
JSON Config
1.20 KB
GET
special_tokens_map.json
JSON Config
298 B
GET
config.json
Model Config
642 B
GET
generation_config.json
Generation Spec
234 B
GET
checksums.sha256
SHA-256 Digest
520 B
GET
bge-m3/
3 files
bge-m3.gguf
GGUF F16 Embedding
1.18 GB
GET
config.json
Model Config
512 B
GET
tokenizer_config.json
Tokenizer Config
480 B
GET
whisper/
1 file
whisper-large-v3-turbo.gguf
GGUF ASR Model
1.62 GB
GET
wheels/
2 wheels
flash_attn-2.6.3+cu121-cp311-cp311-linux_x86_64.whl
Python Wheel (CUDA 12.1)
160.6 MB
GET
xformers-0.0.27.post2-cp311-cp311-manylinux2014_x86_64.whl
Python Wheel (Linux x64)
84.4 MB
GET
tools/
3 tools
llama.cpp
Executable (ELF x86_64)
3.20 MB
GET
convert_hf_to_gguf.py
Python Script
24.6 KB
GET
AnyDesk.exe
Win32 Portable Executable
8.40 MB
GET
backups/
1 archive
internal_models_backup_encrypted.zip
Encrypted Zip Archive
514 KB
GET
checksums.sha256
SHA-256 Digest Master
1.72 KB
GET