Help is available by moving the cursor above any
symbol or by checking MAQAO website.
| Metric | r0 | r1 | r2 | r3 | r4 | r5 | r6 | r7 | r8 | r9 | r10 | r11 | r12 | r13 | r14 | |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Total Time (s) | 127.81 | 65.04 | 33.68 | 18.07 | 10.26 | 8.14 | 6.92 | 6.20 | 5.67 | 5.30 | 4.96 | 4.78 | 4.58 | 4.43 | 4.38 | |
| Max (Thread Active Time) (s) | 126.26 | 63.51 | 32.14 | 16.47 | 8.67 | 6.56 | 5.36 | 4.62 | 4.18 | 3.70 | 3.32 | 3.19 | 2.91 | 2.70 | 2.66 | |
| Average Active Time (s) | 126.26 | 63.22 | 31.70 | 16.00 | 8.16 | 6.02 | 4.79 | 4.05 | 3.53 | 3.13 | 2.77 | 2.60 | 2.38 | 2.23 | 2.12 | |
| Activity Ratio (%) | 98.8 | 98.6 | 98.1 | 97.2 | 95.4 | 94.2 | 92.9 | 91.8 | 90.7 | 89.5 | 87.7 | 87.0 | 85.6 | 84.8 | 83.5 | |
| Average number of active threads | 0.988 | 1.944 | 3.765 | 7.087 | 12.733 | 17.756 | 22.155 | 26.137 | 29.849 | 33.068 | 35.716 | 39.121 | 41.455 | 44.263 | 46.494 | |
| Affinity Stability (%) | 100.0 | 98.4 | 98.0 | 97.2 | 95.8 | 94.8 | 93.9 | 93.1 | 92.3 | 91.5 | 90.3 | 89.9 | 89.1 | 88.5 | 88.0 | |
| Time in analyzed loops (%) | 99.1 | 98.8 | 98.4 | 97.7 | 96.5 | 94.5 | 94.5 | 92.1 | 90.7 | 89.8 | 90.3 | 86.9 | 86.7 | 85.8 | 83.4 | |
| Time in analyzed innermost loops (%) | 96.8 | 96.5 | 96.1 | 95.5 | 94.1 | 92.4 | 92.4 | 90.2 | 88.6 | 88.0 | 88.2 | 85.0 | 84.9 | 83.9 | 81.6 | |
| Time in user code (%) | 99.3 | 99.0 | 98.7 | 97.9 | 96.7 | 94.7 | 94.7 | 92.3 | 90.9 | 90.0 | 90.5 | 87.2 | 86.9 | 86.0 | 83.7 | |
| Compilation Options Score (%) | 100.0 | 100.0 | 100.0 | 100.0 | 100.0 | 99.9 | 99.9 | 99.9 | 99.9 | 99.9 | 99.9 | 99.8 | 99.9 | 99.9 | 99.9 | |
| Array Access Efficiency (%) | 74.6 | 74.6 | 74.6 | 74.6 | 74.6 | 74.6 | 74.5 | 74.6 | 74.6 | 74.6 | 74.6 | 74.6 | 74.7 | 74.7 | 74.7 | |
| Potential Speedups | ||||||||||||||||
| Perfect Flow Complexity | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | |
| Perfect OpenMP/MPI/Pthread/TBB | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.01 | 1.01 | 1.01 | 1.01 | 1.02 | 1.02 | 1.03 | 1.04 | 1.05 | 1.04 | |
| Perfect OpenMP/MPI/Pthread/TBB + Perfect Load Distribution | 1.00 | 1.01 | 1.02 | 1.04 | 1.09 | 1.14 | 1.17 | 1.22 | 1.29 | 1.30 | 1.31 | 1.40 | 1.39 | 1.39 | 1.48 | |
| Scalability - Gap | 1.00 | 1.02 | 1.05 | 1.13 | 1.28 | 1.53 | 1.73 | 1.94 | 2.13 | 2.32 | 2.48 | 2.70 | 2.87 | 3.05 | 3.29 | |
| No Scalar Integer | Potential Speedup | 1.01 | 1.01 | 1.01 | 1.01 | 1.01 | 1.01 | 1.01 | 1.01 | 1.01 | 1.01 | 1.01 | 1.01 | 1.01 | 1.01 | 1.01 |
| Nb Loops to get 80% | 3 | 3 | 3 | 3 | 3 | 3 | 3 | 3 | 3 | 3 | 3 | 3 | 3 | 3 | 3 | |
| FP Vectorised | Potential Speedup | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.00 | 1.01 | 1.01 | 1.01 | 1.01 |
| Nb Loops to get 80% | 3 | 4 | 4 | 3 | 4 | 3 | 2 | 2 | 2 | 1 | 2 | 2 | 1 | 1 | 1 | |
| Fully Vectorised | Potential Speedup | 1.78 | 1.77 | 1.77 | 1.75 | 1.74 | 1.72 | 1.73 | 1.70 | 1.68 | 1.67 | 1.68 | 1.64 | 1.64 | 1.62 | 1.59 |
| Nb Loops to get 80% | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | |
| Only FP Arithmetic | Potential Speedup | 4.48 | 4.44 | 4.37 | 4.25 | 4.06 | 3.80 | 3.79 | 3.53 | 3.39 | 3.30 | 3.35 | 3.06 | 3.01 | 2.96 | 2.80 |
| Nb Loops to get 80% | 2 | 2 | 2 | 2 | 2 | 2 | 2 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | 1 | |
| Source Object | Issue |
|---|---|
| ▼libllama.so | |
| ▼hashtable.h | |
| ○ | -mcpu=native is missing. |
| ▼llama-vocab.cpp | |
| ○ | -mcpu=native is missing. |
| ▼hashtable_policy.h | |
| ○ | -mcpu=native is missing. |
| ▼libggml-cpu.so | |
| ▼vec.cpp | |
| ○ | |
| ▼traits.cpp | |
| ○ | |
| ▼kai_rhs_pack_nxk_qsi4c32pscalef16_qsu4c32s16s0.c | |
| ○ | |
| ▼kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c | |
| ○ | |
| ▼kleidiai.cpp | |
| ○ | |
| ▼binary-ops.cpp | |
| ○ | |
| ▼kai_lhs_quant_pack_qsi8d32p4x8sb_f32_neon.c | |
| ○ | |
| ▼ops.cpp | |
| ○ | |
| ▼ggml-cpu.c | |
| ○ | |
| ▼quants.c | |
| ○ | |
| ▼exec | |
| ▼ | |
| ○ | -g is missing for some functions (possibly ones added by the compiler), it is needed to have more accurate reports. Other recommended flags are: -O2/-O3, -march=(target) |
| ○ | -O2, -O3 or -Ofast is missing. |
| ○ | -mcpu=native is missing. |
| ▼libggml-base.so | |
| ▼ggml-quants.c | |
| ○ | -mcpu=native is missing. |
| ▼[vdso] | |
| ▼ | |
| ○ | -g is missing for some functions (possibly ones added by the compiler), it is needed to have more accurate reports. Other recommended flags are: -O2/-O3, -march=(target) |
| ○ | -O2, -O3 or -Ofast is missing. |
| ○ | -mcpu=native is missing. |
| r0 | r1 | r2 | r3 | r4 | r5 | r6 | r7 | r8 | r9 | r10 | r11 | r12 | r13 | r14 | |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Application | /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/run/oneview_runs/defaults/orig/exec | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Timestamp | 2025-10-24 14:39:41 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Experiment Type | MPI; | MPI; OpenMP; | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 | same as r1 |
| Machine | ip-172-31-47-249.ec2.internal | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Architecture | aarch64 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Micro Architecture | ARM_NEOVERSE_V2 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Model Name | |||||||||||||||
| Cache Size | |||||||||||||||
| Number of Cores | |||||||||||||||
| Maximal Frequency | 0 GHz | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| OS Version | Linux 6.1.155-176.282.amzn2023.aarch64 #1 SMP Tue Oct 7 15:53:00 UTC 2025 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Architecture used during static analysis | aarch64 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Micro Architecture used during static analysis | ARM_NEOVERSE_V2 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Compilation Options | exec: N/A libggml-cpu.so: Arm C/C++/Fortran Compiler version 24.10.1 (build number 4) (based on LLVM 19.1.0) /opt/arm/arm-linux-compiler-24.10.1_AmazonLinux-2023/llvm-bin/clang-19 -D GGML_BACKEND_BUILD -D GGML_BACKEND_SHARED -D GGML_SCHED_MAX_COPIES=4 -D GGML_SHARED -D GGML_USE_CPU_KLEIDIAI -D GGML_USE_CPU_REPACK -D GGML_USE_LLAMAFILE -D GGML_USE_OPENMP -D _GNU_SOURCE -D _XOPEN_SOURCE=600 -D ggml_cpu_EXPORTS -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/.. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/ggml-cpu -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_fp32_bf16p_bf16p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/pack -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/../include -O3 -g -fno-omit-frame-pointer -fcf-protection=none -no-pie -grecord-command-line -O3 -D NDEBUG -std=gnu11 -fPIC -Wshadow -Wstrict-prototypes -Wpointer-arith -Wmissing-prototypes -Werror=implicit-int -Werror=implicit-function-declaration -Wall -Wextra -Wpedantic -Wcast-qual -Wno-unused-function -Wunreachable-code-break -Wunreachable-code-return -Wdouble-promotion -mcpu=native+dotprod+i8mm+sve+nosme -fopenmp=libomp -mcpu=native+dotprod+i8mm+sve+nosme -MD -MT ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o -MF ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o.d -o ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o -c /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c libllama.so: Arm C/C++/Fortran Compiler version 24.10.1 (build number 4) (based on LLVM 19.1.0) /opt/arm/arm-linux-compiler-24.10.1_AmazonLinux-2023/llvm-bin/clang-19 --driver-mode=g++ -D GGML_BACKEND_SHARED -D GGML_SHARED -D GGML_USE_BLAS -D GGML_USE_CPU -D LLAMA_BUILD -D LLAMA_SHARED -D llama_EXPORTS -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/../include -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/../include -O3 -g -fno-omit-frame-pointer -fcf-protection=none -no-pie -grecord-command-line -O3 -D NDEBUG -fPIC -Wmissing-declarations -Wmissing-noreturn -Wall -Wextra -Wpedantic -Wcast-qual -Wno-unused-function -Wunreachable-code-break -Wunreachable-code-return -Wmissing-prototypes -Wextra-semi -MD -MT src/CMakeFiles/llama.dir/unicode.cpp.o -MF src/CMakeFiles/llama.dir/unicode.cpp.o.d -o src/CMakeFiles/llama.dir/unicode.cpp.o -c /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/unicode.cpp | libggml-base.so: Arm C/C++/Fortran Compiler version 24.10.1 (build number 4) (based on LLVM 19.1.0) /opt/arm/arm-linux-compiler-24.10.1_AmazonLinux-2023/llvm-bin/clang-19 -D GGML_BUILD -D GGML_COMMIT=\"unknown\" -D GGML_SCHED_MAX_COPIES=4 -D GGML_SHARED -D GGML_USE_CPU_KLEIDIAI -D GGML_VERSION=\"0.0.0\" -D _GNU_SOURCE -D _XOPEN_SOURCE=600 -D ggml_base_EXPORTS -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_fp32_bf16p_bf16p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/pack -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/../include -O3 -g -fno-omit-frame-pointer -fcf-protection=none -no-pie -grecord-command-line -O3 -D NDEBUG -std=gnu11 -fPIC -Wshadow -Wstrict-prototypes -Wpointer-arith -Wmissing-prototypes -Werror=implicit-int -Werror=implicit-function-declaration -Wall -Wextra -Wpedantic -Wcast-qual -Wno-unused-function -Wunreachable-code-break -Wunreachable-code-return -Wdouble-promotion -MD -MT ggml/src/CMakeFiles/ggml-base.dir/ggml-quants.c.o -MF ggml/src/CMakeFiles/ggml-base.dir/ggml-quants.c.o.d -o ggml/src/CMakeFiles/ggml-base.dir/ggml-quants.c.o -c /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/ggml-quants.c libggml-cpu.so: Arm C/C++/Fortran Compiler version 24.10.1 (build number 4) (based on LLVM 19.1.0) /opt/arm/arm-linux-compiler-24.10.1_AmazonLinux-2023/llvm-bin/clang-19 -D GGML_BACKEND_BUILD -D GGML_BACKEND_SHARED -D GGML_SCHED_MAX_COPIES=4 -D GGML_SHARED -D GGML_USE_CPU_KLEIDIAI -D GGML_USE_CPU_REPACK -D GGML_USE_LLAMAFILE -D GGML_USE_OPENMP -D _GNU_SOURCE -D _XOPEN_SOURCE=600 -D ggml_cpu_EXPORTS -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/.. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/ggml-cpu -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_fp32_bf16p_bf16p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/pack -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/../include -O3 -g -fno-omit-frame-pointer -fcf-protection=none -no-pie -grecord-command-line -O3 -D NDEBUG -std=gnu11 -fPIC -Wshadow -Wstrict-prototypes -Wpointer-arith -Wmissing-prototypes -Werror=implicit-int -Werror=implicit-function-declaration -Wall -Wextra -Wpedantic -Wcast-qual -Wno-unused-function -Wunreachable-code-break -Wunreachable-code-return -Wdouble-promotion -mcpu=native+dotprod+i8mm+sve+nosme -fopenmp=libomp -mcpu=native+dotprod+i8mm+sve+nosme -MD -MT ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o -MF ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o.d -o ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o -c /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c libllama.so: Arm C/C++/Fortran Compiler version 24.10.1 (build number 4) (based on LLVM 19.1.0) /opt/arm/arm-linux-compiler-24.10.1_AmazonLinux-2023/llvm-bin/clang-19 --driver-mode=g++ -D GGML_BACKEND_SHARED -D GGML_SHARED -D GGML_USE_BLAS -D GGML_USE_CPU -D LLAMA_BUILD -D LLAMA_SHARED -D llama_EXPORTS -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/../include -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/../include -O3 -g -fno-omit-frame-pointer -fcf-protection=none -no-pie -grecord-command-line -O3 -D NDEBUG -fPIC -Wmissing-declarations -Wmissing-noreturn -Wall -Wextra -Wpedantic -Wcast-qual -Wno-unused-function -Wunreachable-code-break -Wunreachable-code-return -Wmissing-prototypes -Wextra-semi -MD -MT src/CMakeFiles/llama.dir/unicode.cpp.o -MF src/CMakeFiles/llama.dir/unicode.cpp.o.d -o src/CMakeFiles/llama.dir/unicode.cpp.o -c /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/unicode.cpp | + [vdso]: N/A libggml-cpu.so: Arm C/C++/Fortran Compiler version 24.10.1 (build number 4) (based on LLVM 19.1.0) /opt/arm/arm-linux-compiler-24.10.1_AmazonLinux-2023/llvm-bin/clang-19 -D GGML_BACKEND_BUILD -D GGML_BACKEND_SHARED -D GGML_SCHED_MAX_COPIES=4 -D GGML_SHARED -D GGML_USE_CPU_KLEIDIAI -D GGML_USE_CPU_REPACK -D GGML_USE_LLAMAFILE -D GGML_USE_OPENMP -D _GNU_SOURCE -D _XOPEN_SOURCE=600 -D ggml_cpu_EXPORTS -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/.. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/ggml-cpu -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_fp32_bf16p_bf16p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/pack -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/../include -O3 -g -fno-omit-frame-pointer -fcf-protection=none -no-pie -grecord-command-line -O3 -D NDEBUG -std=gnu11 -fPIC -Wshadow -Wstrict-prototypes -Wpointer-arith -Wmissing-prototypes -Werror=implicit-int -Werror=implicit-function-declaration -Wall -Wextra -Wpedantic -Wcast-qual -Wno-unused-function -Wunreachable-code-break -Wunreachable-code-return -Wdouble-promotion -mcpu=native+dotprod+i8mm+sve+nosme -fopenmp=libomp -mcpu=native+dotprod+i8mm+sve+nosme -MD -MT ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o -MF ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o.d -o ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o -c /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c libllama.so: Arm C/C++/Fortran Compiler version 24.10.1 (build number 4) (based on LLVM 19.1.0) /opt/arm/arm-linux-compiler-24.10.1_AmazonLinux-2023/llvm-bin/clang-19 --driver-mode=g++ -D GGML_BACKEND_SHARED -D GGML_SHARED -D GGML_USE_BLAS -D GGML_USE_CPU -D LLAMA_BUILD -D LLAMA_SHARED -D llama_EXPORTS -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/../include -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/../include -O3 -g -fno-omit-frame-pointer -fcf-protection=none -no-pie -grecord-command-line -O3 -D NDEBUG -fPIC -Wmissing-declarations -Wmissing-noreturn -Wall -Wextra -Wpedantic -Wcast-qual -Wno-unused-function -Wunreachable-code-break -Wunreachable-code-return -Wmissing-prototypes -Wextra-semi -MD -MT src/CMakeFiles/llama.dir/unicode.cpp.o -MF src/CMakeFiles/llama.dir/unicode.cpp.o.d -o src/CMakeFiles/llama.dir/unicode.cpp.o -c /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/unicode.cpp | + [vdso]: N/A exec: N/A libggml-base.so: Arm C/C++/Fortran Compiler version 24.10.1 (build number 4) (based on LLVM 19.1.0) /opt/arm/arm-linux-compiler-24.10.1_AmazonLinux-2023/llvm-bin/clang-19 -D GGML_BUILD -D GGML_COMMIT=\"unknown\" -D GGML_SCHED_MAX_COPIES=4 -D GGML_SHARED -D GGML_USE_CPU_KLEIDIAI -D GGML_VERSION=\"0.0.0\" -D _GNU_SOURCE -D _XOPEN_SOURCE=600 -D ggml_base_EXPORTS -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_fp32_bf16p_bf16p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/pack -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/../include -O3 -g -fno-omit-frame-pointer -fcf-protection=none -no-pie -grecord-command-line -O3 -D NDEBUG -std=gnu11 -fPIC -Wshadow -Wstrict-prototypes -Wpointer-arith -Wmissing-prototypes -Werror=implicit-int -Werror=implicit-function-declaration -Wall -Wextra -Wpedantic -Wcast-qual -Wno-unused-function -Wunreachable-code-break -Wunreachable-code-return -Wdouble-promotion -MD -MT ggml/src/CMakeFiles/ggml-base.dir/ggml-quants.c.o -MF ggml/src/CMakeFiles/ggml-base.dir/ggml-quants.c.o.d -o ggml/src/CMakeFiles/ggml-base.dir/ggml-quants.c.o -c /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/ggml-quants.c libggml-cpu.so: Arm C/C++/Fortran Compiler version 24.10.1 (build number 4) (based on LLVM 19.1.0) /opt/arm/arm-linux-compiler-24.10.1_AmazonLinux-2023/llvm-bin/clang-19 -D GGML_BACKEND_BUILD -D GGML_BACKEND_SHARED -D GGML_SCHED_MAX_COPIES=4 -D GGML_SHARED -D GGML_USE_CPU_KLEIDIAI -D GGML_USE_CPU_REPACK -D GGML_USE_LLAMAFILE -D GGML_USE_OPENMP -D _GNU_SOURCE -D _XOPEN_SOURCE=600 -D ggml_cpu_EXPORTS -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/.. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/ggml-cpu -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_fp32_bf16p_bf16p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/pack -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/../include -O3 -g -fno-omit-frame-pointer -fcf-protection=none -no-pie -grecord-command-line -O3 -D NDEBUG -std=gnu11 -fPIC -Wshadow -Wstrict-prototypes -Wpointer-arith -Wmissing-prototypes -Werror=implicit-int -Werror=implicit-function-declaration -Wall -Wextra -Wpedantic -Wcast-qual -Wno-unused-function -Wunreachable-code-break -Wunreachable-code-return -Wdouble-promotion -mcpu=native+dotprod+i8mm+sve+nosme -fopenmp=libomp -mcpu=native+dotprod+i8mm+sve+nosme -MD -MT ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o -MF ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o.d -o ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o -c /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c libllama.so: Arm C/C++/Fortran Compiler version 24.10.1 (build number 4) (based on LLVM 19.1.0) /opt/arm/arm-linux-compiler-24.10.1_AmazonLinux-2023/llvm-bin/clang-19 --driver-mode=g++ -D GGML_BACKEND_SHARED -D GGML_SHARED -D GGML_USE_BLAS -D GGML_USE_CPU -D LLAMA_BUILD -D LLAMA_SHARED -D llama_EXPORTS -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/../include -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/../include -O3 -g -fno-omit-frame-pointer -fcf-protection=none -no-pie -grecord-command-line -O3 -D NDEBUG -fPIC -Wmissing-declarations -Wmissing-noreturn -Wall -Wextra -Wpedantic -Wcast-qual -Wno-unused-function -Wunreachable-code-break -Wunreachable-code-return -Wmissing-prototypes -Wextra-semi -MD -MT src/CMakeFiles/llama.dir/unicode.cpp.o -MF src/CMakeFiles/llama.dir/unicode.cpp.o.d -o src/CMakeFiles/llama.dir/unicode.cpp.o -c /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/unicode.cpp | same as r2 | same as r3 | + [vdso]: N/A exec: N/A libggml-cpu.so: Arm C/C++/Fortran Compiler version 24.10.1 (build number 4) (based on LLVM 19.1.0) /opt/arm/arm-linux-compiler-24.10.1_AmazonLinux-2023/llvm-bin/clang-19 -D GGML_BACKEND_BUILD -D GGML_BACKEND_SHARED -D GGML_SCHED_MAX_COPIES=4 -D GGML_SHARED -D GGML_USE_CPU_KLEIDIAI -D GGML_USE_CPU_REPACK -D GGML_USE_LLAMAFILE -D GGML_USE_OPENMP -D _GNU_SOURCE -D _XOPEN_SOURCE=600 -D ggml_cpu_EXPORTS -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/.. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/ggml-cpu -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_fp32_bf16p_bf16p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/pack -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/../include -O3 -g -fno-omit-frame-pointer -fcf-protection=none -no-pie -grecord-command-line -O3 -D NDEBUG -std=gnu11 -fPIC -Wshadow -Wstrict-prototypes -Wpointer-arith -Wmissing-prototypes -Werror=implicit-int -Werror=implicit-function-declaration -Wall -Wextra -Wpedantic -Wcast-qual -Wno-unused-function -Wunreachable-code-break -Wunreachable-code-return -Wdouble-promotion -mcpu=native+dotprod+i8mm+sve+nosme -fopenmp=libomp -mcpu=native+dotprod+i8mm+sve+nosme -MD -MT ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o -MF ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o.d -o ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o -c /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c libllama.so: Arm C/C++/Fortran Compiler version 24.10.1 (build number 4) (based on LLVM 19.1.0) /opt/arm/arm-linux-compiler-24.10.1_AmazonLinux-2023/llvm-bin/clang-19 --driver-mode=g++ -D GGML_BACKEND_SHARED -D GGML_SHARED -D GGML_USE_BLAS -D GGML_USE_CPU -D LLAMA_BUILD -D LLAMA_SHARED -D llama_EXPORTS -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/../include -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/../include -O3 -g -fno-omit-frame-pointer -fcf-protection=none -no-pie -grecord-command-line -O3 -D NDEBUG -fPIC -Wmissing-declarations -Wmissing-noreturn -Wall -Wextra -Wpedantic -Wcast-qual -Wno-unused-function -Wunreachable-code-break -Wunreachable-code-return -Wmissing-prototypes -Wextra-semi -MD -MT src/CMakeFiles/llama.dir/unicode.cpp.o -MF src/CMakeFiles/llama.dir/unicode.cpp.o.d -o src/CMakeFiles/llama.dir/unicode.cpp.o -c /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/unicode.cpp | same as r3 | same as r2 | same as r2 | same as r3 | same as r2 | same as r2 | same as r2 | + [vdso]: N/A libggml-base.so: Arm C/C++/Fortran Compiler version 24.10.1 (build number 4) (based on LLVM 19.1.0) /opt/arm/arm-linux-compiler-24.10.1_AmazonLinux-2023/llvm-bin/clang-19 -D GGML_BUILD -D GGML_COMMIT=\"unknown\" -D GGML_SCHED_MAX_COPIES=4 -D GGML_SHARED -D GGML_USE_CPU_KLEIDIAI -D GGML_VERSION=\"0.0.0\" -D _GNU_SOURCE -D _XOPEN_SOURCE=600 -D ggml_base_EXPORTS -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_fp32_bf16p_bf16p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/pack -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/../include -O3 -g -fno-omit-frame-pointer -fcf-protection=none -no-pie -grecord-command-line -O3 -D NDEBUG -std=gnu11 -fPIC -Wshadow -Wstrict-prototypes -Wpointer-arith -Wmissing-prototypes -Werror=implicit-int -Werror=implicit-function-declaration -Wall -Wextra -Wpedantic -Wcast-qual -Wno-unused-function -Wunreachable-code-break -Wunreachable-code-return -Wdouble-promotion -MD -MT ggml/src/CMakeFiles/ggml-base.dir/ggml-quants.c.o -MF ggml/src/CMakeFiles/ggml-base.dir/ggml-quants.c.o.d -o ggml/src/CMakeFiles/ggml-base.dir/ggml-quants.c.o -c /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/ggml-quants.c libggml-cpu.so: Arm C/C++/Fortran Compiler version 24.10.1 (build number 4) (based on LLVM 19.1.0) /opt/arm/arm-linux-compiler-24.10.1_AmazonLinux-2023/llvm-bin/clang-19 -D GGML_BACKEND_BUILD -D GGML_BACKEND_SHARED -D GGML_SCHED_MAX_COPIES=4 -D GGML_SHARED -D GGML_USE_CPU_KLEIDIAI -D GGML_USE_CPU_REPACK -D GGML_USE_LLAMAFILE -D GGML_USE_OPENMP -D _GNU_SOURCE -D _XOPEN_SOURCE=600 -D ggml_cpu_EXPORTS -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/.. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/ggml-cpu -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_fp32_bf16p_bf16p -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/pack -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/../include -O3 -g -fno-omit-frame-pointer -fcf-protection=none -no-pie -grecord-command-line -O3 -D NDEBUG -std=gnu11 -fPIC -Wshadow -Wstrict-prototypes -Wpointer-arith -Wmissing-prototypes -Werror=implicit-int -Werror=implicit-function-declaration -Wall -Wextra -Wpedantic -Wcast-qual -Wno-unused-function -Wunreachable-code-break -Wunreachable-code-return -Wdouble-promotion -mcpu=native+dotprod+i8mm+sve+nosme -fopenmp=libomp -mcpu=native+dotprod+i8mm+sve+nosme -MD -MT ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o -MF ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o.d -o ggml/src/CMakeFiles/ggml-cpu.dir/__/__/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c.o -c /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/build/_deps/kleidiai_download-src/kai/ukernels/matmul/matmul_clamp_f32_qsi8d32p_qsi4c32p/kai_matmul_clamp_f32_qsi8d32p4x8_qsi4c32p4x8_16x4_neon_i8mm.c libllama.so: Arm C/C++/Fortran Compiler version 24.10.1 (build number 4) (based on LLVM 19.1.0) /opt/arm/arm-linux-compiler-24.10.1_AmazonLinux-2023/llvm-bin/clang-19 --driver-mode=g++ -D GGML_BACKEND_SHARED -D GGML_SHARED -D GGML_USE_BLAS -D GGML_USE_CPU -D LLAMA_BUILD -D LLAMA_SHARED -D llama_EXPORTS -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/. -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/../include -I /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/ggml/src/../include -O3 -g -fno-omit-frame-pointer -fcf-protection=none -no-pie -grecord-command-line -O3 -D NDEBUG -fPIC -Wmissing-declarations -Wmissing-noreturn -Wall -Wextra -Wpedantic -Wcast-qual -Wno-unused-function -Wunreachable-code-break -Wunreachable-code-return -Wmissing-prototypes -Wextra-semi -MD -MT src/CMakeFiles/llama.dir/unicode.cpp.o -MF src/CMakeFiles/llama.dir/unicode.cpp.o.d -o src/CMakeFiles/llama.dir/unicode.cpp.o -c /home/eoseret/Tools/QaaS/qaas_runs/ip-172-31-47-249.ec2.internal/176-131-3962/llama.cpp/build/llama.cpp/src/unicode.cpp |
| Number of processes observed | 1 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Number of threads observed | 1 | 2 | 4 | 8 | 16 | 24 | 32 | 40 | 48 | 56 | 64 | 72 | 80 | 88 | 96 |
| Frequency Driver | NA | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Frequency Governor | NA | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Huge Pages | madvise | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Hyperthreading | off | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Number of sockets | 1 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Number of cores per socket | 96 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| MAQAO version | 2025.1.3 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| MAQAO build | b489783858807c9a72e0923fd7b399a22c81991c::20251024-122946 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |
| Comments | OV scalability run using armclang | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 | same as r0 |