Concedo
|
4e43c21e58
|
Merge commit '9d5d882d8c' into concedo_experimental
# Conflicts:
# .github/labeler.yml
# app/CMakeLists.txt
# app/llama.cpp
# build-xcframework.sh
# common/CMakeLists.txt
# common/download.h
# docs/backend/SYCL.md
# docs/backend/snapdragon/CMakeUserPresets.json
# docs/speculative.md
# ggml/CMakeLists.txt
# ggml/include/ggml-sycl.h
# ggml/src/ggml-hexagon/CMakeLists.txt
# ggml/src/ggml-hexagon/ggml-hexagon.cpp
# ggml/src/ggml-hexagon/htp/CMakeLists.txt
# ggml/src/ggml-hexagon/htp/cmake-toolchain.cmake
# ggml/src/ggml-hexagon/htp/flash-attn-ops.c
# ggml/src/ggml-hexagon/htp/hex-dma.h
# ggml/src/ggml-hexagon/htp/hex-utils.h
# ggml/src/ggml-hexagon/htp/hmx-flash-attn-ops.c
# ggml/src/ggml-hexagon/htp/htp-ctx.h
# ggml/src/ggml-hexagon/htp/htp-ops.h
# ggml/src/ggml-hexagon/htp/htp_iface.idl
# ggml/src/ggml-hexagon/htp/hvx-base.h
# ggml/src/ggml-hexagon/htp/main.c
# ggml/src/ggml-hexagon/htp/matmul-ops.c
# ggml/src/ggml-hexagon/libggml-htp.inf
# ggml/src/ggml-opencl/ggml-opencl.cpp
# ggml/src/ggml-opencl/kernels/norm.cl
# ggml/src/ggml-sycl/conv3d.cpp
# ggml/src/ggml-sycl/ggml-sycl.cpp
# scripts/snapdragon/adb/run-completion.sh
# scripts/snapdragon/adb/run-tool.sh
# scripts/snapdragon/ggml-hexagon-profile.py
# tests/CMakeLists.txt
# tests/test-backend-ops.cpp
# tests/test-thread-safety.cpp
# tools/llama-bench/llama-bench.cpp
# tools/mtmd/CMakeLists.txt
# tools/mtmd/tests/test-deepseek-ocr.py
|
2026-06-27 10:18:52 +08:00 |
|
Xuan-Son Nguyen
|
60bc8866b1
|
common: refactor model handling (#24980)
* common: refactor models handling
* remote preset
* cont
* rm skip_download option
* missing header
* fix plan.model_files
* fix --offline case
* move hf_plan to download
* refactor
* rm redundant curr_ex, add comments
* adapt
|
2026-06-25 15:17:51 +02:00 |
|
Adrien Gallouët
|
683b04cc4a
|
app : add the llama download subcommand (#24982)
* app : add the download command (with llama-download)
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
* Remove llama-download tool for now
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
---------
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
|
2026-06-25 13:36:36 +02:00 |
|
Adrien Gallouët
|
6ec59ddaea
|
app : enable self-update only when built with llama-install.sh (#24754)
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
|
2026-06-18 09:57:59 +02:00 |
|
Pascal
|
5a46b46acd
|
app: add llama update self updater (#23865)
* wip: llama update POC
* cleaning: llama update
* llama-gen-docs
* app: delegate llama update to the install script
* app: spawn the installer detached so llama update can replace a running binary
* cleaning: inline llama update into llama.cpp, drop app-update.{cpp,h}
* app: make llama_update static
Address review from @angt
|
2026-05-29 23:02:40 +02:00 |
|
Adrien Gallouët
|
98e480a32e
|
app : move licences to llama-app (#23824)
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
|
2026-05-29 07:46:11 +02:00 |
|
Adrien Gallouët
|
479a9a1b03
|
app : improve help output (#23805)
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
|
2026-05-28 16:45:06 +02:00 |
|
Concedo
|
632c41a72f
|
Merge branch 'upstream' into concedo_experimental
# Conflicts:
# .github/workflows/build-apple.yml
# .github/workflows/build-cmake-pkg.yml
# .github/workflows/release.yml
# .pi/gg/SYSTEM.md
# CMakeLists.txt
# CODEOWNERS
# README.md
# build-xcframework.sh
# ci/run.sh
# docs/build.md
# examples/CMakeLists.txt
# examples/llama.android/lib/build.gradle.kts
# ggml/src/ggml-webgpu/wgsl-shaders/flash_attn_tile.wgsl
# tests/CMakeLists.txt
# tests/test-backend-ops.cpp
# tests/test-save-load-state.cpp
# tools/batched-bench/CMakeLists.txt
# tools/cli/CMakeLists.txt
# tools/completion/CMakeLists.txt
# tools/llama-bench/CMakeLists.txt
# tools/perplexity/CMakeLists.txt
# tools/quantize/CMakeLists.txt
# tools/server/CMakeLists.txt
|
2026-05-22 20:42:51 +08:00 |
|
Pascal
|
c9021714e8
|
server: re-inject subcommand when router spawns children under unified binary (#23442)
|
2026-05-21 10:09:19 +02:00 |
|
Adrien Gallouët
|
1d7ab2b947
|
app : add batched-bench, fit-params, quantize & perplexity (#23459)
Python Type-Check / python type-check (push) Waiting to run
* app : add batched-bench, fit-params, quantize & perplexity
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
* Add missing main.cpp
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
* Add EOL
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
---------
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
|
2026-05-21 10:29:44 +03:00 |
|
Adrien Gallouët
|
ce02093fdd
|
app : show version (#23426)
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
|
2026-05-21 06:21:13 +02:00 |
|
Adrien Gallouët
|
29f1482221
|
app : introduce the llama unified executable (#23296)
* app : introduce the llama unified executable
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
* Use serve for server
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
* Hide completion and bench, add help command
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
* Remove STATIC
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
* Use -impl targets instead of -lib
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
* Revert "Remove STATIC"
This reverts commit cc44caccb9902b34a3531633edac911e5b3d65cd.
---------
Signed-off-by: Adrien Gallouët <angt@huggingface.co>
|
2026-05-20 13:22:22 +02:00 |
|