Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
PrismML-Eng
llama.cpp
Repository navigation
Code
Issues
44
(44)
Pull requests
52
(52)
Discussions
Actions
Projects
Security and quality
Insights
More
items
Actions: PrismML-Eng/llama.cpp
Actions
All workflows
Workflows
Release (Prism)
Release (Prism)
Build relocatable cmake package
Build relocatable cmake package
Check Pre-Tokenizer Hashes
Check Pre-Tokenizer Hashes
Check vendor
Check vendor
Code Style Checker
Code Style Checker
CodeQL
CodeQL
Convert PR to draft
Convert PR to draft
Copilot
Copilot
Copilot cloud agent
Copilot cloud agent
Copilot code review
Copilot code review
Show more workflows...
Management
Caches
All workflows
All workflows
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Showing runs from all workflows
will be ignored since log searching is not yet available
2,500+ workflow runs
2,500+ workflow runs
Workflow
Filter by Workflow
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching workflows.
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
cuda : take the short-query MMA FA shortcut only for the f16 K/V kernel
Server
#455:
Pull request
#337
opened by
sb32445
Action required
sb32445:hotfix/fattn-189-native-kv
sb32445:hotfix/fattn-189-native-kv
Action required
View #337
View workflow file
cuda : take the short-query MMA FA shortcut only for the f16 K/V kernel
Pull Request Labeler
#487:
Pull request
#337
opened by
sb32445
19s
19s
View #337
View workflow file
Push on prism
CodeQL
#291:
by
bri-prism
15m 10s
prism
prism
15m 10s
Push on prism
CodeQL
#290:
by
bri-prism
20m 2s
prism
prism
20m 2s
Push on prism
CodeQL
#289:
by
bri-prism
25m 11s
prism
prism
25m 11s
Push on prism
CodeQL
#288:
by
bri-prism
21m 13s
prism
prism
21m 13s
Push on prism
CodeQL
#287:
by
bri-prism
26m 38s
prism
prism
26m 38s
Push on prism
CodeQL
#286:
by
bri-prism
15m 1s
prism
prism
15m 1s
Push on prism
CodeQL
#285:
by
bri-prism
16m 43s
prism
prism
16m 43s
grammar: clone finds each stack element's rule by binary search
Server
#454:
Pull request
#336
opened by
professorpalmer
Action required
professorpalmer:grammar-clone-bsearch
professorpalmer:grammar-clone-bsearch
Action required
View #336
View workflow file
grammar: clone finds each stack element's rule by binary search
Pull Request Labeler
#486:
Pull request
#336
opened by
professorpalmer
17s
17s
View #336
View workflow file
speculative: --spec-draft-window (MTP drafting at every depth)
Server
#453:
Pull request
#320
synchronize by
professorpalmer
Action required
professorpalmer:claude/spec-draft-window
professorpalmer:claude/spec-draft-window
Action required
View #320
View workflow file
speculative: --spec-draft-window (MTP drafting at every depth)
Pull Request Labeler
#485:
Pull request
#320
synchronize by
professorpalmer
1m 9s
1m 9s
View #320
View workflow file
Musa FWHT fix
Python Type-Check
#225:
Commit
b6f1bb5
pushed by
bri-prism
3m 8s
fwht-musa-smem
fwht-musa-smem
3m 8s
View workflow file
Runtime support for Prism Bonsai 2 27B
Update Operations Documentation
#43:
Commit
3bfabd2
pushed by
bri-prism
38s
hadamard-folded-runtime
hadamard-folded-runtime
38s
View workflow file
Runtime support for Prism Bonsai 2 27B
Python Type-Check
#224:
Commit
3bfabd2
pushed by
bri-prism
3m 51s
hadamard-folded-runtime
hadamard-folded-runtime
3m 51s
View workflow file
Runtime support for Prism Bonsai 2 27B
Check Pre-Tokenizer Hashes
#67:
Commit
3bfabd2
pushed by
bri-prism
1m 3s
hadamard-folded-runtime
hadamard-folded-runtime
1m 3s
View workflow file
Runtime support for Prism Bonsai 2 27B
Python check requirements.txt
#41:
Commit
3bfabd2
pushed by
bri-prism
3m 5s
hadamard-folded-runtime
hadamard-folded-runtime
3m 5s
View workflow file
cuda: use one-row scheduling for PTQ1_0 planar GEMV
Pull Request Labeler
#484:
Pull request
#335
opened by
Max-sm-yc
33s
33s
View #335
View workflow file
cuda: gate (SwiGLU) fused PTQ1_0 mat-vec for 2-4 columns
Server
#452:
Pull request
#310
synchronize by
sb32445
Action required
sb32445:pr/ptq1-gate-up-fuse-mc
sb32445:pr/ptq1-gate-up-fuse-mc
Action required
View #310
View workflow file
cuda: gate (SwiGLU) fused PTQ1_0 mat-vec for 2-4 columns
Pull Request Labeler
#483:
Pull request
#310
synchronize by
sb32445
21s
21s
View #310
View workflow file
cuda: prefetch the next PTQ1_0 mat-vec's weights into L2 from the last CTAs
Server
#451:
Pull request
#314
synchronize by
sb32445
19m 35s
sb32445:pr/ptq1-l2-prefetch
sb32445:pr/ptq1-l2-prefetch
19m 35s
View #314
View workflow file
cuda: prefetch the next PTQ1_0 mat-vec's weights into L2 from the last CTAs
Pull Request Labeler
#482:
Pull request
#314
synchronize by
sb32445
14s
14s
View #314
View workflow file
cuda: use the MMA flash attention kernel for GQA above 4 with quantized K/V on Ada
Server
#450:
Pull request
#307
synchronize by
sb32445
Action required
sb32445:pr/fattn-gqa-mma
sb32445:pr/fattn-gqa-mma
Action required
View #307
View workflow file
cuda: use the MMA flash attention kernel for GQA above 4 with quantized K/V on Ada
Pull Request Labeler
#481:
Pull request
#307
synchronize by
sb32445
24s
24s
View #307
View workflow file
Previous
1
2
3
4
5
…
99
100
101
Next
You can’t perform that action at this time.