Commit History

docs: remove Gemma-4-12B-Agentic configuration from llama-swap.yaml for cleanup
fb099e5
unverified

wrenchpilot commited on

docs: update hf reference in gemma-4-12b-agentic-fable5.ini for version consistency
130aa74
unverified

wrenchpilot commited on

docs: update spec-type and parameters in gemma-4-12b-agentic-fable5.ini for model configuration
67f1891
unverified

wrenchpilot commited on

docs: remove model-draft path from gemma-4-12b-agentic-fable5.ini for clarity
d2860ce
unverified

wrenchpilot commited on

docs: update model-draft path in gemma-4-12b-agentic-fable5.ini for accuracy
1b756cf
unverified

wrenchpilot commited on

docs: update model-draft path in gemma-4-12b-agentic-fable5.ini for version consistency
e0c0dab
unverified

wrenchpilot commited on

docs: update model-draft path in gemma-4-12b-agentic-fable5.ini for accuracy
1c84f9a
unverified

wrenchpilot commited on

docs: rename Gemma-4-12B-Agentic-Fable5 to Gemma-4-12B-Agentic for consistency
9649724
unverified

wrenchpilot commited on

docs: update model-draft path and simplify spec-type in gemma-4-12b-agentic-fable5.ini
0b33b3e
unverified

wrenchpilot commited on

docs: update device configuration to support multi-GPU setup in gemma-4-12b-agentic-fable5.ini
aa5c47b
unverified

wrenchpilot commited on

docs: update device configuration and enhance spec-type settings in gemma-4-12b-agentic-fable5.ini
9c8caa5
unverified

wrenchpilot commited on

docs: update hf reference in gemma-4-12b-agentic-fable5.ini for consistency
8717e9c
unverified

wrenchpilot commited on

docs: remove checkpoint settings and simplify spec-type in gemma-4-12b-agentic-fable5.ini
c58bb17
unverified

wrenchpilot commited on

docs: add model-draft path to gemma-4-12b-agentic-fable5.ini for improved model loading
18af853
unverified

wrenchpilot commited on

docs: update gemma-4-12b-it-qat.ini and gemma-4-12b-agentic-fable5.ini for enhanced specifications and add new agentic model configuration
64ac7bd
unverified

wrenchpilot commited on

docs: update gemma-4-12b-it-qat.ini to adjust np value and add chat-template-kwargs
dcae790
unverified

wrenchpilot commited on

docs: update gemma-4-12b-it-qat.ini to adjust np value and enable checkpoint settings
8d62b25
unverified

wrenchpilot commited on

docs: update README and .env.example for clarity on configuration and service lifecycle
73d54c5
unverified

wrenchpilot commited on

docs: enhance README with detailed preset file descriptions and usage examples
e5c3413
unverified

wrenchpilot commited on

docs: update README to clarify repository purpose and contents
b2e3640
unverified

wrenchpilot commited on

docs: add Hugging Face cache update script details to README
438d2db
unverified

wrenchpilot commited on

chore: remove unused PRESET variable and fix printf statement in stable-server script
ad11f60
unverified

wrenchpilot commited on

chore: replace server and stable scripts with llama-server and stable-server scripts
615457b
unverified

wrenchpilot commited on

Refactor and reorganize project structure
092192d
unverified

wrenchpilot commited on

chore: update qwopus3.6-35b-mtp.ini with ctx-checkpoints and checkpoint-min-step settings
0518ad9
unverified

wrenchpilot commited on

chore: comment out Qwen3.8-Flash-Next model configuration in llama-swap.yaml
7841864
unverified

wrenchpilot commited on

chore: update configuration settings in qwen3.8-flash-next.ini for improved performance
1baa120
unverified

wrenchpilot commited on

chore: add load-mode and lazy-mode settings in qwen3.8-flash-next configuration
1fe2bfe
unverified

wrenchpilot commited on

chore: enhance error handling and documentation in check-hf-updates script
a8d7ab4
unverified

wrenchpilot commited on

chore: refactor check-hf-updates script for improved clarity and functionality
48997bc
unverified

wrenchpilot commited on

chore: improve documentation and refactor check-hf-updates script for better clarity and functionality
4c999df
unverified

wrenchpilot commited on

chore: enhance virtual environment setup and improve cache handling in check-hf-updates script
e7281a6
unverified

wrenchpilot commited on

chore: refactor check-hf-updates script to add argument parsing and improve cache handling
f6a7453
unverified

wrenchpilot commited on

chore: enhance virtual environment setup and cache directory resolution in check-hf-updates script
b6b18f0
unverified

wrenchpilot commited on

chore: add check-hf-updates script for managing Hugging Face model cache and virtual environment setup
e6fa152
unverified

wrenchpilot commited on

chore: add Qwen3.8-Flash-Next model configuration and preset file
7761ed8
unverified

wrenchpilot commited on

chore: update device configuration and refine spec settings in qwen3.8-27b-pi
0870a0c
unverified

wrenchpilot commited on

chore: update device configuration and refine spec settings in qwen3.8-27b-pi
5970d59
unverified

wrenchpilot commited on

chore: update hf model reference in qwen3.8-27b-pi configuration
0cb29bb
unverified

wrenchpilot commited on

chore: simplify spec-type and comment out n-gram settings in qwen3.8-27b-pi configuration
15853bd
unverified

wrenchpilot commited on

chore: update spec-draft-n-max to 2 and add ctk, ctv settings in qwen3.8-27b-pi configuration
3b8d999
unverified

wrenchpilot commited on

chore: update spec settings and add load-on-startup in qwen3.8-27b-pi configuration
15835c6
unverified

wrenchpilot commited on

chore: update context size to 65536 in qwen3.8-27b-pi configuration
4eb8299
unverified

wrenchpilot commited on

chore: update context size for multiple models in llama-swap configuration
d0fa237
unverified

wrenchpilot commited on

chore: update spec-draft-n-max to 4 and add context setting in configuration files
3ead2e0
unverified

wrenchpilot commited on

chore: standardize model names by removing suffixes in llama-swap configuration
3cc768c
unverified

wrenchpilot commited on

chore: standardize section names by capitalizing preset headers in gemma-4-26b-a4b-it-qat-mtp and laya-bf16 configuration files
5e8872c
unverified

wrenchpilot commited on

chore: standardize preset section names by removing suffixes in multiple configuration files
a5a2975
unverified

wrenchpilot commited on

chore: increase context size from 65536 to 98304 in gemma-4-26b-a4b-it-qat-mtp preset
9f9973a
unverified

wrenchpilot commited on

chore: adjust spec-draft-n-max to 3 and restore device settings in Qwen3.6-35B-MTP preset
3e56f48
unverified

wrenchpilot commited on