Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
lgrammel
/
llama-cpp-provider
Public
Notifications
You must be signed in to change notification settings
Fork
3
Star
28
Code
Issues
0
Pull requests
0
Actions
Security and quality
0
Insights
Additional navigation options
Code
Issues
Pull requests
Actions
Security and quality
Insights
Actions: lgrammel/llama-cpp-provider
Actions
All workflows
Workflows
Format Check
Format Check
Lint
Lint
Type Check
Type Check
Unit Test
Unit Test
Show more workflows...
Management
Caches
Unit Test
Unit Test
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Show workflow options
Create status badge
Create status badge
Loading
Uh oh!
There was an error while loading.
Please reload this page
.
unit-test.yml
will be ignored since log searching is not yet available
101 workflow runs
101 workflow runs
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
Add experimental project note to README
Unit Test
#102:
Commit
5e900b0
pushed by
lgrammel
3m 53s
main
main
3m 53s
View workflow file
Harden Qwen cache token regression test
Unit Test
#101:
Commit
7d14aa9
pushed by
lgrammel
4m 55s
main
main
4m 55s
View workflow file
Release @lgrammel/llama-cpp-provider@0.4.3
Unit Test
#100:
Commit
ff93ea5
pushed by
lgrammel
6m 0s
main
main
6m 0s
View workflow file
Improve prompt cache and tool-call parsing
Unit Test
#99:
Commit
0518d0f
pushed by
lgrammel
6m 20s
main
main
6m 20s
View workflow file
Release @lgrammel/llama-cpp-provider@0.4.2
Unit Test
#98:
Commit
8af8561
pushed by
lgrammel
5m 47s
main
main
5m 47s
View workflow file
Handle Qwen XML tool calls and ChatML control tokens
Unit Test
#97:
Commit
d4c8034
pushed by
lgrammel
4m 50s
main
main
4m 50s
View workflow file
Release @lgrammel/llama-cpp-provider@0.4.1
Unit Test
#96:
Commit
a84e3c9
pushed by
lgrammel
3m 51s
main
main
3m 51s
View workflow file
Add release wrapper for package publishing
Unit Test
#95:
Commit
f4bdcee
pushed by
lgrammel
6m 35s
main
main
6m 35s
View workflow file
Support Qwen 3.6 XML tool calls and reasoning
Unit Test
#94:
Commit
9f617a8
pushed by
lgrammel
6m 12s
main
main
6m 12s
View workflow file
Remove duplicate sampler token acceptance
Unit Test
#93:
Commit
d0a6b5c
pushed by
lgrammel
6m 1s
main
main
6m 1s
View workflow file
Add agent-friendly E2E scripts and docs
Unit Test
#92:
Commit
d3cc278
pushed by
lgrammel
8m 48s
main
main
8m 48s
View workflow file
Release 0.4.0
Unit Test
#91:
Commit
3548780
pushed by
lgrammel
4m 19s
main
main
4m 19s
View workflow file
Update AGENTS guidance for changesets
Unit Test
#90:
Commit
d6043d0
pushed by
lgrammel
7m 54s
main
main
7m 54s
View workflow file
Pass enableThinking through reasoning and chat templates
Unit Test
#89:
Commit
5adf393
pushed by
lgrammel
6m 12s
main
main
6m 12s
View workflow file
Enforce formatting checks for native bindings
Unit Test
#88:
Commit
24b8d18
pushed by
lgrammel
6m 34s
main
main
6m 34s
View workflow file
Refactor streamed tool-call parsing for reasoning content
Unit Test
#87:
Commit
b568b14
pushed by
lgrammel
5m 29s
main
main
5m 29s
View workflow file
Support OpenAI-compatible structured output options
Unit Test
#86:
Commit
0cdd0c5
pushed by
lgrammel
6m 1s
main
main
6m 1s
View workflow file
Stream tool input parts during tool-call parsing
Unit Test
#85:
Commit
91bbcc4
pushed by
lgrammel
4m 27s
main
main
4m 27s
View workflow file
Add reasoning budget mapping to llama.cpp provider
Unit Test
#84:
Commit
1f225bd
pushed by
lgrammel
5m 3s
main
main
5m 3s
View workflow file
Constrain JSON responses without schemas to objects
Unit Test
#83:
Commit
d624b33
pushed by
lgrammel
5m 36s
main
main
5m 36s
View workflow file
Handle stream cancellation during active generation
Unit Test
#82:
Commit
9850e73
pushed by
lgrammel
5m 29s
main
main
5m 29s
View workflow file
Release 0.3.0
Unit Test
#81:
Commit
b5ac877
pushed by
lgrammel
5m 42s
main
main
5m 42s
View workflow file
Align defaults with llama.cpp server behavior
Unit Test
#80:
Commit
fd0e0d6
pushed by
lgrammel
6m 23s
main
main
6m 23s
View workflow file
Add native llama.cpp tool-calling support
Unit Test
#79:
Commit
dbbeb2d
pushed by
lgrammel
5m 51s
main
main
5m 51s
View workflow file
Update CI and docs for Node 22 and reasoning behavior
Unit Test
#78:
Commit
d3a4da0
pushed by
lgrammel
5m 22s
main
main
5m 22s
View workflow file
Previous
1
2
3
4
5
Next
You can’t perform that action at this time.