Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
Accelerator
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
Uh oh!
There was an error while loading.
Please reload this page
.
Tencent
/
ncnn
Public
Notifications
You must be signed in to change notification settings
Fork
4.5k
Star
23.7k
Code
Issues
1.1k
Pull requests
138
Discussions
Actions
Projects
Wiki
Security and quality
2
Insights
Additional navigation options
Code
Issues
Pull requests
Discussions
Actions
Projects
Wiki
Security and quality
Insights
Actions: Tencent/ncnn
Actions
All workflows
Workflows
android
android
code-format
code-format
code-format-msg
code-format-msg
CodeQL
CodeQL
compare-binary-size
compare-binary-size
compare-binary-size-pr-comment
compare-binary-size-pr-comment
Copilot
Copilot
Copilot cloud agent
Copilot cloud agent
Copilot code review
Copilot code review
CPU Support Test
CPU Support Test
Show more workflows...
Management
Caches
Deployments
linux-x64-gpu-gcc
All workflows
Actions
Loading...
Loading
Sorry, something went wrong.
Uh oh!
There was an error while loading.
Please reload this page
.
will be ignored since log searching is not yet available
Showing runs from all workflows named linux-x64-gpu-gcc
will be ignored since log searching is not yet available
1,718 workflow run results
1,718 workflow run results
Workflow
Filter by Workflow
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching workflows.
Event
Filter by Event
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching events.
Status
Filter by Status
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching statuses.
Branch
Filter by Branch
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching branches.
Actor
Filter by Actor
Sorry, something went wrong.
Filter
Loading
Sorry, something went wrong.
No matching users.
vulkan: fix fp16 arithmetic types in erf and celu shaders
linux-x64-gpu-gcc
#9339:
Pull request
#6927
opened by
futz12
Action required
futz12:fix/erf-celu-fp16-arithmetic
futz12:fix/erf-celu-fp16-arithmetic
Action required
View #6927
View workflow file
let float2bfloat round to nearest, even or ties away from zero
linux-x64-gpu-gcc
#9338:
Pull request
#6925
synchronize by
nihui
1h 7m 34s
nihui:bf16-rne
nihui:bf16-rne
1h 7m 34s
View #6925
View workflow file
let float2bfloat round to nearest, even or ties away from zero
linux-x64-gpu-gcc
#9337:
Pull request
#6925
opened by
nihui
7m 15s
nihui:bf16-rne
nihui:bf16-rne
7m 15s
View #6925
View workflow file
sdpa flash attention x86 optimization, llm models enable gqa, dispatch for avx2 bfloat2float
linux-x64-gpu-gcc
#9336:
Pull request
#6923
synchronize by
nihui
1h 1m 14s
nihui:sdpa-fa2
nihui:sdpa-fa2
1h 1m 14s
View #6923
View workflow file
sdpa flash attention x86 optimization, llm models enable gqa, dispatch for avx2 bfloat2float
linux-x64-gpu-gcc
#9335:
Pull request
#6923
synchronize by
nihui
13m 43s
nihui:sdpa-fa2
nihui:sdpa-fa2
13m 43s
View #6923
View workflow file
sdpa flash attention x86 optimization, llm models enable gqa, dispatch for avx2 bfloat2float
linux-x64-gpu-gcc
#9334:
Pull request
#6923
synchronize by
nihui
41m 15s
nihui:sdpa-fa2
nihui:sdpa-fa2
41m 15s
View #6923
View workflow file
sdpa flash attention x86 optimization, llm models enable gqa, dispatch for avx2 bfloat2float
linux-x64-gpu-gcc
#9333:
Pull request
#6923
synchronize by
nihui
13s
nihui:sdpa-fa2
nihui:sdpa-fa2
13s
View #6923
View workflow file
sdpa flash attention x86 optimization, llm models enable gqa, dispatch for avx2 bfloat2float
linux-x64-gpu-gcc
#9332:
Pull request
#6923
synchronize by
nihui
4m 17s
nihui:sdpa-fa2
nihui:sdpa-fa2
4m 17s
View #6923
View workflow file
sdpa flash attention x86 optimization, llm models enable gqa, dispatch for avx2 bfloat2float
linux-x64-gpu-gcc
#9331:
Pull request
#6923
synchronize by
nihui
51m 41s
nihui:sdpa-fa2
nihui:sdpa-fa2
51m 41s
View #6923
View workflow file
sdpa flash attention x86 optimization, llm models enable gqa, dispatch for avx2 bfloat2float
linux-x64-gpu-gcc
#9330:
Pull request
#6923
synchronize by
nihui
1h 4m 32s
nihui:sdpa-fa2
nihui:sdpa-fa2
1h 4m 32s
View #6923
View workflow file
sdpa flash attention x86 optimization, llm models enable gqa, dispatch for avx2 bfloat2float
linux-x64-gpu-gcc
#9329:
Pull request
#6923
synchronize by
github-actions
Bot
Action required
nihui:sdpa-fa2
nihui:sdpa-fa2
Action required
View #6923
View workflow file
sdpa flash attention x86 optimization, llm models enable gqa, dispatch for avx2 bfloat2float
linux-x64-gpu-gcc
#9328:
Pull request
#6923
synchronize by
nihui
36m 45s
nihui:sdpa-fa2
nihui:sdpa-fa2
36m 45s
View #6923
View workflow file
fix: stabilize logaddexp expression evaluation
linux-x64-gpu-gcc
#9327:
Pull request
#6924
opened by
primorLee
Action required
primorLee:codex/fix-expression-logaddexp-overflow
primorLee:codex/fix-expression-logaddexp-overflow
Action required
View #6924
View workflow file
sdpa flash attention x86 optimization, llm models enable gqa, dispatch for avx2 bfloat2float
linux-x64-gpu-gcc
#9326:
Pull request
#6923
synchronize by
nihui
1h 0m 26s
nihui:sdpa-fa2
nihui:sdpa-fa2
1h 0m 26s
View #6923
View workflow file
vulkan 1D convolution family: conv1d winograd/subgroup, convolutiondepthwise1d, deconvolution1d
linux-x64-gpu-gcc
#9325:
Pull request
#6909
synchronize by
futz12
Action required
futz12:1d-conv-vulkan
futz12:1d-conv-vulkan
Action required
View #6909
View workflow file
sdpa flash attention x86 optimization, llm models enable gqa, dispatch for avx2 bfloat2float
linux-x64-gpu-gcc
#9324:
Pull request
#6923
synchronize by
github-actions
Bot
Action required
nihui:sdpa-fa2
nihui:sdpa-fa2
Action required
View #6923
View workflow file
sdpa flash attention x86 optimization, llm models enable gqa, dispatch for avx2 bfloat2float
linux-x64-gpu-gcc
#9323:
Pull request
#6923
opened by
nihui
59m 41s
nihui:sdpa-fa2
nihui:sdpa-fa2
59m 41s
View #6923
View workflow file
fix unvalidated layer counts, out-of-bounds blob index and division by zero in load_param and load_param_bin
linux-x64-gpu-gcc
#9322:
Pull request
#6922
synchronize by
kaltsit-10
Action required
kaltsit-10:fix/text-layer-count
kaltsit-10:fix/text-layer-count
Action required
View #6922
View workflow file
fix unvalidated layer counts, out-of-bounds blob index and division by zero in load_param and load_param_bin
linux-x64-gpu-gcc
#9321:
Pull request
#6922
synchronize by
kaltsit-10
Action required
kaltsit-10:fix/text-layer-count
kaltsit-10:fix/text-layer-count
Action required
View #6922
View workflow file
fix unvalidated layer counts, out-of-bounds blob index and division by zero in load_param and load_param_bin
linux-x64-gpu-gcc
#9320:
Pull request
#6922
opened by
kaltsit-10
Action required
kaltsit-10:fix/text-layer-count
kaltsit-10:fix/text-layer-count
Action required
View #6922
View workflow file
initialize softmax layer pointer in YoloDetectionOutput
linux-x64-gpu-gcc
#9319:
Pull request
#6921
opened by
kaltsit-10
Action required
kaltsit-10:fix/ydo-softmax-init
kaltsit-10:fix/ydo-softmax-init
Action required
View #6921
View workflow file
x86/loongarch/mips: refine rsqrt with newton-raphson iterations && vu…
linux-x64-gpu-gcc
#9316:
Commit
946fe3f
pushed by
nihui
1h 56m 48s
master
master
1h 56m 48s
View workflow file
x86/loongarch/mips: refine rsqrt with newton-raphson iterations && vulkan: use fp32 for coefficient
linux-x64-gpu-gcc
#9315:
Pull request
#6903
synchronize by
nihui
1h 5m 56s
futz12:rsqrt-precise-fix
futz12:rsqrt-precise-fix
1h 5m 56s
View #6903
View workflow file
x86/loongarch/mips: refine rsqrt with newton-raphson iterations && vulkan: use fp32 for coefficient
linux-x64-gpu-gcc
#9314:
Pull request
#6903
synchronize by
futz12
1h 2m 12s
futz12:rsqrt-precise-fix
futz12:rsqrt-precise-fix
1h 2m 12s
View #6903
View workflow file
x86/loongarch/mips: refine rsqrt with newton-raphson iterations && vulkan: use fp32 for coefficient
linux-x64-gpu-gcc
#9311:
Pull request
#6903
synchronize by
futz12
19m 3s
futz12:rsqrt-precise-fix
futz12:rsqrt-precise-fix
19m 3s
View #6903
View workflow file
Previous
1
2
3
4
5
…
68
69
Next
You can’t perform that action at this time.