Tags: LSDJesus/llama-cpp-python
Tags
Add build_output.txt and update llama_cpp integration - Add build_output.txt - Update llama_cpp/llama_chat_format.py and llama_cpp/llama_cpp.py - Update pyproject.toml and vendor/llama.cpp
Add build_output.txt and update llama_cpp integration - Add build_output.txt - Update llama_cpp/llama_chat_format.py and llama_cpp/llama_cpp.py - Update pyproject.toml and vendor/llama.cpp
docs: add qwen3-vl hacking guide and penultimate layer extraction - Add comprehensive hacking guide for llama.cpp qwen3-vl model modifications - Add penultimate hidden states PR with patch for transformer layer capture - Add activation steering script for semantic memory research - Add KV fragment testing utilities - Add GGUF dequantization test harness - Update per-layer embeddings results - Add steering delta coefficients for layer injection
ci: add CUDA wheel build workflow (abi3, sm_75-sm_120)
PreviousNext