Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
ggml-org
llama.cpp
Repository navigation
Code
Issues
855
(855)
Pull requests
1.6k
(1.6k)
Discussions
Actions
Projects
Wiki
Security and quality
13
(13)
Insights
More
items
Actions: ggml-org/llama.cpp
Actions
All workflows
Workflows
Make Release
Make Release
Publish Docker image
Publish Docker image
Publish Release
Publish Release
Release
Release
Update Winget Package
Update Winget Package
EditorConfig Checker
EditorConfig Checker
.github/workflows/build-and-test-hexagon.yml
.github/workflows/build-and-test-hexagon.yml
Build Actions Cache
Build Actions Cache
Build and Test - Hexagon Android (QDC)
Build and Test - Hexagon Android (QDC)
Build on RISCV Linux Machine by Cloud-V
Build on RISCV Linux Machine by Cloud-V
Build relocatable cmake package
Build relocatable cmake package
Show more workflows...
Management
Caches
Deployments
EditorConfig Checker
EditorConfig Checker
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Show workflow options
Create status badge
Create status badge
Loading
Uh oh!
There was an error while loading.
Please reload this page
.
editorconfig.yml
will be ignored since log searching is not yet available
2,500+ workflow runs
2,500+ workflow runs
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
cuda: update uncoalesced memory reads in pool2d (#29425)
EditorConfig Checker
#61348:
Commit
42c787e
pushed by
JohannesGaessler
45s
master
master
45s
View workflow file
context : simplify setters
EditorConfig Checker
#61347:
Pull request
#29730
synchronize by
nikwen
54s
nikwen:nikwen/simplify-context
nikwen:nikwen/simplify-context
54s
View #29730
View workflow file
server : preserve context checkpoints across slot save/restore
EditorConfig Checker
#61346:
Pull request
#26004
synchronize by
Tough-Respawn
Action required
Tough-Respawn:fix/kv-restore-checkpoint
Tough-Respawn:fix/kv-restore-checkpoint
Action required
View #26004
View workflow file
sycl: use ExternalProject to let this backend be built with SYCL compiler while everything else could use another
EditorConfig Checker
#61345:
Pull request
#29506
synchronize by
a1batross
Action required
a1batross:mix-rocm-sycl-build
a1batross:mix-rocm-sycl-build
Action required
View #29506
View workflow file
server : accumulate generated text and tokens as parse input (#29876)
EditorConfig Checker
#61344:
Commit
18b5f8b
pushed by
aldehir
3m 56s
master
master
3m 56s
View workflow file
model : use exact GELU for ModernBERT encoders
EditorConfig Checker
#61343:
Pull request
#30108
opened by
boshjerns
Action required
boshjerns:modernbert-encoder-gelu-erf
boshjerns:modernbert-encoder-gelu-erf
Action required
View #30108
View workflow file
vulkan: use 4 rows for NVIDIA MUL_MAT_ID MMVQ except pre-Turing
EditorConfig Checker
#61342:
Pull request
#29274
synchronize by
SG-Amadeus
Action required
SG-Amadeus:sg-amadeus/vulkan-mmid-mmvq-rows4
SG-Amadeus:sg-amadeus/vulkan-mmid-mmvq-rows4
Action required
View #29274
View workflow file
chat : name tool and argument parser rules by index
EditorConfig Checker
#61341:
Pull request
#30088
synchronize by
aldehir
56s
Frost-54:master
Frost-54:master
56s
View #30088
View workflow file
llama : add a GPU cache for MoE experts kept in host memory
EditorConfig Checker
#61340:
Pull request
#29887
synchronize by
am17an
1m 27s
am17an:aman/moe-cache
am17an:aman/moe-cache
1m 27s
View #29887
View workflow file
vulkan: use 4 rows for NVIDIA MUL_MAT_ID MMVQ except pre-Turing
EditorConfig Checker
#61339:
Pull request
#29274
synchronize by
SG-Amadeus
Action required
SG-Amadeus:sg-amadeus/vulkan-mmid-mmvq-rows4
SG-Amadeus:sg-amadeus/vulkan-mmid-mmvq-rows4
Action required
View #29274
View workflow file
model-loader: add --reclaim-mmap-source to drop dormant mmap pages (Fixes #16761)
EditorConfig Checker
#61338:
Pull request
#24156
synchronize by
markkobo
Action required
markkobo:feature/cpu-repack-reclaim
markkobo:feature/cpu-repack-reclaim
Action required
View #24156
View workflow file
vulkan : fix TOP_K for +inf/NaN inputs and k = 1 on negative values
EditorConfig Checker
#61337:
Pull request
#30107
opened by
gianni-cor
Action required
gianni-cor:vulkan-topk-nonfinite
gianni-cor:vulkan-topk-nonfinite
Action required
View #30107
View workflow file
vulkan : one warp tile per subgroup in the small matmul tiles
EditorConfig Checker
#61336:
Pull request
#30106
opened by
ianloic
Action required
ianloic:vulkan-small-warptile-subgroups
ianloic:vulkan-small-warptile-subgroups
Action required
View #30106
View workflow file
Formalizing scheduler and async backend behavior through tests
EditorConfig Checker
#61335:
Pull request
#27258
synchronize by
aendk
48s
aendk:akieslinger/sched-backend-test-coverage-master
aendk:akieslinger/sched-backend-test-coverage-master
48s
View #27258
View workflow file
cpu: add F16 SWIGLU_OAI reference and activation-op test coverage
EditorConfig Checker
#61334:
Pull request
#29200
synchronize by
cqderek
Action required
qualcomm:upstream-f16-swiglu-reference
qualcomm:upstream-f16-swiglu-reference
Action required
View #29200
View workflow file
vulkan : honor FP32 source precision in MUL_MAT_ID
EditorConfig Checker
#61333:
Pull request
#30105
opened by
linuxid10t
Action required
linuxid10t:fix/mistral4-vulkan-offload
linuxid10t:fix/mistral4-vulkan-offload
Action required
View #30105
View workflow file
hexagon: improved GELU accuracy
EditorConfig Checker
#61332:
Pull request
#30104
opened by
kurquhar
Action required
qualcomm:kurquhar_htp_gelu_accuracy
qualcomm:kurquhar_htp_gelu_accuracy
Action required
View #30104
View workflow file
CUDA/HIP: fix race in flash_attn_ext_f16_process_tile when nbatch_combine != DKQ/2
EditorConfig Checker
#61331:
Pull request
#30103
opened by
IMbackK
56s
IMbackK:fattn_fix
IMbackK:fattn_fix
56s
View #30103
View workflow file
vulkan: use density gate for MUL_MAT_VEC_ID path
EditorConfig Checker
#61330:
Pull request
#27332
synchronize by
theycallmeloki
Action required
theycallmeloki:vulkan-moe-density-gate
theycallmeloki:vulkan-moe-density-gate
Action required
View #27332
View workflow file
CUDA : fix PAD for more than 65535 rows or slices
EditorConfig Checker
#61329:
Pull request
#30102
opened by
pskrunner14
Action required
pskrunner14:cuda-pad-large-grid
pskrunner14:cuda-pad-large-grid
Action required
View #30102
View workflow file
llama: share the nextn tensor flags between models (#30097)
EditorConfig Checker
#61328:
Commit
448147d
pushed by
ServeurpersoCom
8m 38s
master
master
8m 38s
View workflow file
WIP: Persist checkpoints when saving hybrid model slots to disk
EditorConfig Checker
#61327:
Pull request
#30101
opened by
rain-sk
-1s
matarbot:pr/ckpt-persist
matarbot:pr/ckpt-persist
-1s
View #30101
View workflow file
Add TML Inkling architecture
EditorConfig Checker
#61326:
Pull request
#25731
synchronize by
danielhanchen
1m 58s
danielhanchen:add-inkling
danielhanchen:add-inkling
1m 58s
View #25731
View workflow file
mtmd: add cohere2 vision support
EditorConfig Checker
#61325:
Pull request
#30062
synchronize by
Terrencezzj
1m 5s
Terrencezzj:cohere2-vision
Terrencezzj:cohere2-vision
1m 5s
View #30062
View workflow file
metal : few-row MMA mat-mul for the remaining src0 types (#30065)
EditorConfig Checker
#61324:
Commit
9881906
pushed by
ggerganov
3m 47s
master
master
3m 47s
View workflow file
Previous
1
2
3
4
5
…
99
100
101
Next
You can’t perform that action at this time.