obelisk-49l-gguf

GGUF exports for the tanlaan/obelisk-49l CPT repair checkpoint.

Source model

  • Dense HF model: tanlaan/obelisk-49l
  • Source checkpoint line: repaired Granite 49L branch with staged Dolmino ingredient1 continual pretraining
  • Effective CPT depth for this export: ~64M tokens
  • Sequence length during CPT: 8192

Files

  • obelisk-49l-f16.gguf
    • unquantized reference GGUF
    • size: about 1.1 GB
  • obelisk-49l-Q4_K_M.gguf
    • intended for mobile / PocketPal / lighter local runtimes
    • size: about 332 MB
  • obelisk-49l-Q8_0.gguf
    • higher-quality reference quantization
    • size: about 568 MB

Evaluation snapshot

Checkpoint ARC HellaSwag MMLU Mini TruthfulQA MC2 Winogrande GSM8K Flex GSM8K Strict 5-task Mean
Base original 0.36 0.515 0.31 0.4200 0.60 0.18 n/a 0.3975
Recovered base 0.35 0.545 0.315 0.3955 0.55 0.20 0.20 0.4311
~64M effective @ 8192 continuation 0.38 0.52 0.255 0.4422 0.60 0.27 0.26 0.4394

Notes

  • This repo contains quantized GGUF exports derived from the dense Hugging Face checkpoint.
  • The HF source repo is the better choice if you want the original transformers model artifacts.
  • The Q4_K_M file is the first file to try for PocketPal and similar phone runtimes.
Downloads last month
15
GGUF
Model size
0.6B params
Architecture
granitehybrid
Hardware compatibility
Log In to add your hardware

4-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for tanlaan/obelisk-49l-gguf

Quantized
(1)
this model