Skip to content

Tags: LSDJesus/llama-cpp-python

Tags

v0.3.37-cu130-Basic-win-20260510

Toggle v0.3.37-cu130-Basic-win-20260510's commit message
Add build_output.txt and update llama_cpp integration

- Add build_output.txt
- Update llama_cpp/llama_chat_format.py and llama_cpp/llama_cpp.py
- Update pyproject.toml and vendor/llama.cpp

v0.3.37-cu130-Basic-linux-20260510

Toggle v0.3.37-cu130-Basic-linux-20260510's commit message
Add build_output.txt and update llama_cpp integration

- Add build_output.txt
- Update llama_cpp/llama_chat_format.py and llama_cpp/llama_cpp.py
- Update pyproject.toml and vendor/llama.cpp

v0.3.37-cu130-Basic-win-20260508

Toggle v0.3.37-cu130-Basic-win-20260508's commit message
fix image input

v0.3.27-cuda-6396cf3

Toggle v0.3.27-cuda-6396cf3's commit message
docs: add qwen3-vl hacking guide and penultimate layer extraction

- Add comprehensive hacking guide for llama.cpp qwen3-vl model modifications
- Add penultimate hidden states PR with patch for transformer layer capture
- Add activation steering script for semantic memory research
- Add KV fragment testing utilities
- Add GGUF dequantization test harness
- Update per-layer embeddings results
- Add steering delta coefficients for layer injection

v0.3.27-luna.1

Toggle v0.3.27-luna.1's commit message
ci: add CUDA wheel build workflow (abi3, sm_75-sm_120)

v0.3.27-cu130-Basic-win-20260223

Toggle v0.3.27-cu130-Basic-win-20260223's commit message
Bump version to 0.3.27

v0.3.27-cu130-Basic-linux-20260222

Toggle v0.3.27-cu130-Basic-linux-20260222's commit message
Bump version to 0.3.27

v0.3.27-cu128-Basic-win-20260223

Toggle v0.3.27-cu128-Basic-win-20260223's commit message
Bump version to 0.3.27

v0.3.27-cu128-Basic-linux-20260222

Toggle v0.3.27-cu128-Basic-linux-20260222's commit message
Bump version to 0.3.27

v0.3.27-cu126-Basic-win-20260223

Toggle v0.3.27-cu126-Basic-win-20260223's commit message
Bump version to 0.3.27