Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
Accelerator
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
LSDJesus
/
llama-cpp-python
Public
forked from
JamePeng/llama-cpp-python
Notifications
You must be signed in to change notification settings
Fork
0
Star
0
Code
Pull requests
0
Actions
Projects
Security and quality
0
Insights
Additional navigation options
Code
Pull requests
Actions
Projects
Security and quality
Insights
Commits
Breadcrumbs
History for
llama-cpp-python
llama_cpp
llama_cpp.py
on
main
User selector
All users
All time
Commit history
Commits on May 9, 2026
Add build_output.txt and update llama_cpp integration
Show description for f7314d2
LSDJesus
committed
f7314d2
View commit details
Copy full SHA for f7314d2
View code at this point
Browse repository at this point
Commits on May 3, 2026
Update llama.py, llama_cpp.py, and vendor/llama.cpp
LSDJesus
committed
60b68a9
View commit details
Copy full SHA for 60b68a9
View code at this point
Browse repository at this point
feat: add semantic memory injection research and layer capture/skip C API extensions
Show description for 1f20c99
LSDJesus
committed
1f20c99
View commit details
Copy full SHA for 1f20c99
View code at this point
Browse repository at this point
Update submodule to my fork + my changes
LSDJesus
committed
f79fa45
View commit details
Copy full SHA for f79fa45
View code at this point
Browse repository at this point
Commits on Apr 21, 2026
Update Submodule vendor/llama.cpp 81df3f7..82209ef
JamePeng
committed
8625836
View commit details
Copy full SHA for 8625836
View code at this point
Browse repository at this point
Commits on Apr 17, 2026
Update Submodule vendor/llama.cpp 9db77a0..45cac7c
JamePeng
committed
7a19575
View commit details
Copy full SHA for 7a19575
View code at this point
Browse repository at this point
Commits on Apr 9, 2026
Sync ggml: backend-agnostic tensor parallelism (experimental) (#19378)
JamePeng
committed
6866eba
View commit details
Copy full SHA for 6866eba
View code at this point
Browse repository at this point
Commits on Apr 8, 2026
Sync ggml: add Q1_0 1-bit quantization support (CPU) (#21273)
JamePeng
committed
9241b0f
View commit details
Copy full SHA for 9241b0f
View code at this point
Browse repository at this point
Commits on Apr 3, 2026
Update llama_vocab_pre_type varriable
JamePeng
committed
7ef09e9
View commit details
Copy full SHA for 7ef09e9
View code at this point
Browse repository at this point
Commits on Apr 1, 2026
Sync llama : refactor llama_model_quantize_params to expose a pure C interface (#20346)
Show description for 24f2562
JamePeng
committed
24f2562
View commit details
Copy full SHA for 24f2562
View code at this point
Browse repository at this point
refactor(logger): migrate from llama_log_callback to ggml_log_callback
Show description for 225c7ad
JamePeng
committed
225c7ad
View commit details
Copy full SHA for 225c7ad
View code at this point
Browse repository at this point
Commits on Mar 25, 2026
feat(internals): implement dynamic LoRA routing and Control Vector support
Show description for 0d379eb
JamePeng
committed
0d379eb
View commit details
Copy full SHA for 0d379eb
View code at this point
Browse repository at this point
Sync llama: fix llama-model-saver (#20503)
JamePeng
committed
6898d8e
View commit details
Copy full SHA for 6898d8e
View code at this point
Browse repository at this point
Commits on Mar 23, 2026
fix(types): correct llama_adapter_get_alora_invocation_tokens ctypes signature and use pointer for llama_token
Show description for ceb544c
JamePeng
committed
ceb544c
View commit details
Copy full SHA for ceb544c
View code at this point
Browse repository at this point
Commits on Mar 22, 2026
fix(types): correct `llama_set_adapters_lora` LoRA adapter ctypes signature and use pointer for scales
Show description for e42cf39
JamePeng
committed
e42cf39
View commit details
Copy full SHA for e42cf39
View code at this point
Browse repository at this point
Commits on Mar 18, 2026
Update Submodule vendor/llama.cpp d23355a..312cf03
Show description for ed5e212
JamePeng
committed
ed5e212
View commit details
Copy full SHA for ed5e212
View code at this point
Browse repository at this point
Commits on Mar 11, 2026
Update llama.cpp API 20260312
JamePeng
committed
c420a2a
View commit details
Copy full SHA for c420a2a
View code at this point
Browse repository at this point
Commits on Mar 9, 2026
Update llama.cpp API 20260310
JamePeng
committed
955ac33
View commit details
Copy full SHA for 955ac33
View code at this point
Browse repository at this point
Commits on Feb 20, 2026
Update llama_model_quantize params
JamePeng
committed
eedfa52
View commit details
Copy full SHA for eedfa52
View code at this point
Browse repository at this point
Commits on Feb 19, 2026
Update Submodule vendor/llama.cpp c78e682..abb9f3c
JamePeng
committed
6b38182
View commit details
Copy full SHA for 6b38182
View code at this point
Browse repository at this point
Commits on Feb 17, 2026
Update Submodule vendor/llama.cpp cceb1b4..afa6bfe
JamePeng
committed
16175e5
View commit details
Copy full SHA for 16175e5
View code at this point
Browse repository at this point
Commits on Feb 16, 2026
Update Submodule vendor/llama.cpp 079feab..cceb1b4
JamePeng
committed
2516555
View commit details
Copy full SHA for 2516555
View code at this point
Browse repository at this point
Commits on Feb 15, 2026
Update Submodule vendor/llama.cpp 1725e31..079feab
JamePeng
committed
dc5f7e5
View commit details
Copy full SHA for dc5f7e5
View code at this point
Browse repository at this point
Commits on Feb 14, 2026
Update llama.cpp API 20260214
JamePeng
committed
af9d925
View commit details
Copy full SHA for af9d925
View code at this point
Browse repository at this point
Commits on Feb 13, 2026
Update Submodule vendor/llama.cpp e463bbd..33a56f9
JamePeng
committed
f68fd9f
View commit details
Copy full SHA for f68fd9f
View code at this point
Browse repository at this point
Commits on Feb 11, 2026
Update Submodule vendor/llama.cpp 262364e..e463bbd
JamePeng
committed
974505e
View commit details
Copy full SHA for 974505e
View code at this point
Browse repository at this point
Commits on Feb 9, 2026
Separate the grammar sampler, improve the code stability of Sampler Chain processing, and fix some bugs.
JamePeng
committed
556976d
View commit details
Copy full SHA for 556976d
View code at this point
Browse repository at this point
Commits on Feb 7, 2026
Refactor sampling infrastructure to use llama.cpp sampler chain API
Show description for 1df39b4
JamePeng
committed
1df39b4
View commit details
Copy full SHA for 1df39b4
View code at this point
Browse repository at this point
Optimize the method definition of class llama_sampler_i and CtypesPointer definition
Show description for 672fa43
JamePeng
committed
672fa43
View commit details
Copy full SHA for 672fa43
View code at this point
Browse repository at this point
Commits on Jan 29, 2026
Sync llama : disable Direct IO by default
JamePeng
committed
5bdff0b
View commit details
Copy full SHA for 5bdff0b
View code at this point
Browse repository at this point
Commits on Jan 26, 2026
Update Submodule vendor/llama.cpp 8f91ca5..142cbe2
JamePeng
committed
8468ac2
View commit details
Copy full SHA for 8468ac2
View code at this point
Browse repository at this point
Commits on Jan 15, 2026
Update llama.cpp API 20260116
JamePeng
committed
736f2a1
View commit details
Copy full SHA for 736f2a1
View code at this point
Browse repository at this point
Commits on Jan 14, 2026
Update llama_vocab_pre_type enum variables
JamePeng
committed
fb06280
View commit details
Copy full SHA for fb06280
View code at this point
Browse repository at this point
Commits on Jan 9, 2026
Update llama_cpp.llama_detokenize API and remove unused comment
JamePeng
committed
457d711
View commit details
Copy full SHA for 457d711
View code at this point
Browse repository at this point
Update Submodule vendor/llama.cpp 55abc39..ec8fd78
JamePeng
committed
1fecb05
View commit details
Copy full SHA for 1fecb05
View code at this point
Browse repository at this point
Previous
Next
You can’t perform that action at this time.