Commit History
Add MLX int4 variants for 1B and 3B 4ba07c3 verified
Apply the model card standard ab3b8e2 verified
Apply the model card standard cca8aad verified
Remove the QLoRA variants, unsupported since v0.10 368750a verified
Stop restating quantized, and drop the unused default 800e60a
Correct the published metadata e351c65
Update chat template for strict monotonicity and tool support 424479f
Use min/max dynamic shape bounds for forward inputs/outputs 953c1e2
Add config.json for QLoRA variants b3368c8
Backfill constant metadata method values in config.json from .pte 82e2ba3
Add stub root config.json for HF download counter d5a8447 verified
Add spec-compliant config.json files f29b2f8 verified
Add spec-compliant config.json files 45b83f0 verified
Remove old-layout metadata orphaned by MODEL_SPEC.md restructure 7685cf5 verified
Restructure to MODEL_SPEC.md convention f50e172 verified
Fix invalid JSON in config.json eb03277
Update README.md 76ab87f verified
Update README.md f4b927e verified
update non-quantized llama models to include context len in metadata 58c482d
update README f8c68b3
update README 4f85a25
update models to work with v0.6.0 runtime fdba165
Update tokenizer_config 596e57a
pweglik commited on