Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
Accelerator
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
Uh oh!
There was an error while loading.
Please reload this page
.
AMD-Ecosystem
/
llama.cpp
Public
forked from
ggml-org/llama.cpp
Notifications
You must be signed in to change notification settings
Fork
8
Star
18
Code
Issues
1
Pull requests
17
Discussions
Actions
Projects
Security and quality
0
Insights
Additional navigation options
Code
Issues
Pull requests
Discussions
Actions
Projects
Security and quality
Insights
Actions: AMD-Ecosystem/llama.cpp
Actions
All workflows
Workflows
Code Style Checker
Code Style Checker
.github/workflows/build-cann.yml
.github/workflows/build-cann.yml
AI review (issues)
AI review (issues)
Build Actions Cache
Build Actions Cache
Build gfx11 + ROCm
Build gfx11 + ROCm
CI
CI
CI (3rd-party)
CI (3rd-party)
CI (android)
CI (android)
CI (cross)
CI (cross)
CI (CUDA, windows)
CI (CUDA, windows)
CI (msys)
CI (msys)
Show more workflows...
Management
Caches
Code Style Checker
Code Style Checker
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Show workflow options
Create status badge
Create status badge
Loading
Uh oh!
There was an error while loading.
Please reload this page
.
code-style.yml
will be ignored since log searching is not yet available
194 workflow runs
194 workflow runs
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
ggml-cuda: GEMM weight row padding + one-time K-padded f16 dequant for prefill
Code Style Checker
#96:
Pull request
#57
opened by
roberteg16
32s
rogarcia.cuda-gemm-pad-prefill-dequant
rogarcia.cuda-gemm-pad-prefill-dequant
32s
View #57
View workflow file
ggml-cuda: enable GGML_CUDA_GRAPH_OPT by default on RDNA3.5
Code Style Checker
#95:
Pull request
#56
opened by
roberteg16
46s
rogarcia.enable-by-default-ggml_cuda_graph_opt-on
rogarcia.enable-by-default-ggml_cuda_graph_opt-on
46s
View #56
View workflow file
Merge pull request #53 from ROCm/rogarcia.roofline-moe-tokens-per-expert
Code Style Checker
#94:
Commit
0941c4c
pushed by
mgehre-amd
42s
gfx11
gfx11
42s
View #60
View workflow file
CUDA: fused chunked gated_delta_net kernel (RDNA3.5)
Code Style Checker
#91:
Pull request
#54
synchronize by
roberteg16
22s
rogarcia.gdn-shared-kq
rogarcia.gdn-shared-kq
22s
View #54
View workflow file
CUDA: fused chunked gated_delta_net kernel (RDNA3.5)
Code Style Checker
#90:
Pull request
#54
opened by
roberteg16
29s
rogarcia.gdn-shared-kq
rogarcia.gdn-shared-kq
29s
View #54
View workflow file
roofline: emit per-expert routing histogram (tokens_per_expert) for MoE launches
Code Style Checker
#89:
Pull request
#53
opened by
roberteg16
1m 12s
rogarcia.roofline-moe-tokens-per-expert
rogarcia.roofline-moe-tokens-per-expert
1m 12s
View #53
View workflow file
Merge pull request #52 from ROCm/rogarcia.qwen3.6-moe-optimization
Code Style Checker
#88:
Commit
f6ea7bc
pushed by
roberteg16
20s
gfx11
gfx11
20s
View workflow file
ssm_conv: channels-major input mode to drop the delta-net transpose
Code Style Checker
#87:
Pull request
#52
synchronize by
roberteg16
1m 22s
rogarcia.qwen3.6-moe-optimization
rogarcia.qwen3.6-moe-optimization
1m 22s
View #52
View workflow file
Merge pull request #36 from ROCm/rogarcia.cuda-multistream-shexp
Code Style Checker
#86:
Commit
367c4d0
pushed by
roberteg16
17s
gfx11
gfx11
17s
View workflow file
ssm_conv: channels-major input mode to drop the delta-net transpose
Code Style Checker
#85:
Pull request
#52
synchronize by
roberteg16
53s
rogarcia.qwen3.6-moe-optimization
rogarcia.qwen3.6-moe-optimization
53s
View #52
View workflow file
ggml-cuda: overlap the MoE shared expert on a separate stream
Code Style Checker
#84:
Pull request
#36
synchronize by
roberteg16
23s
rogarcia.cuda-multistream-shexp
rogarcia.cuda-multistream-shexp
23s
View #36
View workflow file
Merge pull request #50 from ROCm/rogarcia.roofline-moe-experts-used
Code Style Checker
#83:
Commit
442b17e
pushed by
roberteg16
1m 5s
gfx11
gfx11
1m 5s
View workflow file
ssm_conv: channels-major input mode to drop the delta-net transpose
Code Style Checker
#82:
Pull request
#52
synchronize by
roberteg16
25s
rogarcia.qwen3.6-moe-optimization
rogarcia.qwen3.6-moe-optimization
25s
View #52
View workflow file
ssm_conv: channels-major input mode to drop the delta-net transpose
Code Style Checker
#81:
Pull request
#52
synchronize by
roberteg16
19s
rogarcia.qwen3.6-moe-optimization
rogarcia.qwen3.6-moe-optimization
19s
View #52
View workflow file
roofline: capture per-launch MoE active-expert count (experts_used)
Code Style Checker
#80:
Pull request
#50
synchronize by
roberteg16
1m 6s
rogarcia.roofline-moe-experts-used
rogarcia.roofline-moe-experts-used
1m 6s
View #50
View workflow file
Merge pull request #48 from ROCm/rogarcia.roofline-fused-raw-facts
Code Style Checker
#79:
Commit
ffe2ad4
pushed by
mgehre-amd
44s
gfx11
gfx11
44s
View workflow file
Merge pull request #51 from ROCm/rogarcia.fix-ci-llamacpp
Code Style Checker
#78:
Commit
77939b0
pushed by
mgehre-amd
27s
gfx11
gfx11
27s
View workflow file
Merge pull request #49 from ROCm/rogarcia.cuda-mmvf-single-col-swap
Code Style Checker
#77:
Commit
a4498c2
pushed by
mgehre-amd
44s
gfx11
gfx11
44s
View workflow file
ssm_conv: channels-major input mode to drop the delta-net transpose
Code Style Checker
#76:
Pull request
#52
synchronize by
roberteg16
1m 8s
rogarcia.qwen3.6-moe-optimization
rogarcia.qwen3.6-moe-optimization
1m 8s
View #52
View workflow file
ssm_conv: channels-major input mode to drop the delta-net transpose
Code Style Checker
#75:
Pull request
#52
opened by
roberteg16
20s
rogarcia.qwen3.6-moe-optimization
rogarcia.qwen3.6-moe-optimization
20s
View #52
View workflow file
fix: ci(hip): drop web UI provisioning from hip-quality-check
Code Style Checker
#74:
Pull request
#51
opened by
roberteg16
36s
rogarcia.fix-ci-llamacpp
rogarcia.fix-ci-llamacpp
36s
View #51
View workflow file
RDNA3.5 (gfx11 / gfx1151) MMQ prefill optimizations
Code Style Checker
#73:
Pull request
#32
synchronize by
liangliangchang
1m 15s
liangliangchang:lichang.mmq_opt
liangliangchang:lichang.mmq_opt
1m 15s
View #32
View workflow file
roofline: emit raw per-tensor data for DDR byte accounting (single + fused)
Code Style Checker
#72:
Pull request
#48
synchronize by
roberteg16
53s
rogarcia.roofline-fused-raw-facts
rogarcia.roofline-fused-raw-facts
53s
View #48
View workflow file
CUDA: route single-column f32 mul_mat through mmvf (transpose-free operand swap)
Code Style Checker
#71:
Pull request
#49
synchronize by
roberteg16
30s
rogarcia.cuda-mmvf-single-col-swap
rogarcia.cuda-mmvf-single-col-swap
30s
View #49
View workflow file
CUDA: route single-column f32 mul_mat through mmvf (transpose-free operand swap)
Code Style Checker
#70:
Pull request
#49
opened by
roberteg16
50s
rogarcia.cuda-mmvf-single-col-swap
rogarcia.cuda-mmvf-single-col-swap
50s
View #49
View workflow file
Previous
1
2
3
4
5
6
7
8
Next
You can’t perform that action at this time.