llama: refactor llama_decode_impl (#11381) #3882
build.yml
on: push
Matrix: windows-2019-cmake-cuda
Matrix: windows-latest-cmake-hip-release
Matrix: windows-latest-cmake
macOS-latest-cmake-arm64
12m 12s
macOS-latest-cmake-x64
4m 58s
ubuntu-cpu-cmake
3m 0s
ubuntu-latest-cmake-rpc
2m 46s
ubuntu-22-cmake-vulkan
19m 18s
ubuntu-22-cmake-hip
20m 51s
ubuntu-22-cmake-musa
12m 34s
ubuntu-22-cmake-sycl
5m 2s
ubuntu-22-cmake-sycl-fp16
5m 19s
macOS-latest-cmake-ios
1m 47s
macOS-latest-cmake-tvos
2m 4s
ubuntu-latest-cmake-cuda
12m 11s
windows-latest-cmake-sycl
10m 14s
windows-latest-cmake-hip
27m 31s
ios-xcode-build
1m 42s
android-build
7m 9s
Matrix: macOS-latest-swift
Matrix: openEuler-latest-cmake-cann
Matrix: ubuntu-latest-cmake-sanitizer
Matrix: windows-msys2
release
1m 55s
Annotations
1 error and 7 warnings
Artifacts
Produced during runtime
Name | Size | |
---|---|---|
cudart-llama-bin-win-cu11.7-x64.zip
|
303 MB |
|
cudart-llama-bin-win-cu12.4-x64.zip
|
372 MB |
|
llama-bin-macos-arm64.zip
|
20.9 MB |
|
llama-bin-macos-x64.zip
|
22.4 MB |
|
llama-bin-ubuntu-x64.zip
|
24.2 MB |
|
llama-bin-win-avx-x64.zip
|
13.8 MB |
|
llama-bin-win-avx2-x64.zip
|
13.8 MB |
|
llama-bin-win-avx512-x64.zip
|
13.8 MB |
|
llama-bin-win-cu11.7-x64.zip
|
150 MB |
|
llama-bin-win-cu12.4-x64.zip
|
150 MB |
|
llama-bin-win-hip-x64-gfx1030.zip
|
236 MB |
|
llama-bin-win-hip-x64-gfx1100.zip
|
238 MB |
|
llama-bin-win-hip-x64-gfx1101.zip
|
238 MB |
|
llama-bin-win-kompute-x64.zip
|
14.1 MB |
|
llama-bin-win-llvm-arm64-opencl-adreno.zip
|
17.5 MB |
|
llama-bin-win-llvm-arm64.zip
|
17.5 MB |
|
llama-bin-win-msvc-arm64.zip
|
56.5 MB |
|
llama-bin-win-noavx-x64.zip
|
13.8 MB |
|
llama-bin-win-openblas-x64.zip
|
24.8 MB |
|
llama-bin-win-sycl-x64.zip
|
95.3 MB |
|
llama-bin-win-vulkan-x64.zip
|
15.9 MB |
|