koboldcpp/src
Concedo 6d7ef10671 Merge branch 'upstream' into concedo_experimental
Renable qwen2vl GPU for vulkan https://github.com/ggml-org/llama.cpp/pull/11902

# Conflicts:
#	.github/workflows/build.yml
#	.github/workflows/docker.yml
#	.gitignore
#	CONTRIBUTING.md
#	Makefile
#	common/CMakeLists.txt
#	common/arg.cpp
#	common/common.cpp
#	examples/main/main.cpp
#	examples/run/run.cpp
#	examples/server/tests/README.md
#	ggml/src/ggml-cuda/mma.cuh
#	scripts/get_chat_template.py
#	tests/test-backend-ops.cpp
#	tests/test-chat-template.cpp
#	tests/test-chat.cpp
2025-02-20 23:17:20 +08:00
..
llama-adapter.cpp llama : add llama_vocab, functions -> methods, naming (#11110) 2025-01-12 11:32:42 +02:00
llama-adapter.h llama : add llama_vocab, functions -> methods, naming (#11110) 2025-01-12 11:32:42 +02:00
llama-arch.cpp llama : add support for GLM-Edge and GLM-Edge-V series models (#10573) 2025-02-02 09:48:46 +02:00
llama-arch.h Add Jinja template support (#11016) 2025-01-21 13:18:51 +00:00
llama-batch.cpp llama : refactor src/llama.cpp (#10902) 2025-01-03 10:18:53 +02:00
llama-batch.h llama : refactor src/llama.cpp (#10902) 2025-01-03 10:18:53 +02:00
llama-chat.cpp llama : add support for GLM-Edge and GLM-Edge-V series models (#10573) 2025-02-02 09:48:46 +02:00
llama-chat.h llama : add support for GLM-Edge and GLM-Edge-V series models (#10573) 2025-02-02 09:48:46 +02:00
llama-context.cpp broken commit 2025-01-16 21:41:18 +08:00
llama-context.h llama : add llama_vocab, functions -> methods, naming (#11110) 2025-01-12 11:32:42 +02:00
llama-cparams.cpp llama : refactor src/llama.cpp (#10902) 2025-01-03 10:18:53 +02:00
llama-cparams.h llama : refactor src/llama.cpp (#10902) 2025-01-03 10:18:53 +02:00
llama-grammar.cpp llama : fix indentation in llama-grammar [no ci] (#11943) 2025-02-19 06:16:23 +01:00
llama-grammar.h llama : fix typo in llama-grammar.h [no ci] (#11816) 2025-02-12 09:40:01 +02:00
llama-hparams.cpp llama: add support for QRWKV6 model architecture (#11001) 2025-01-10 09:58:08 +08:00
llama-hparams.h llama : add llama_vocab, functions -> methods, naming (#11110) 2025-01-12 11:32:42 +02:00
llama-impl.cpp GGUF: C++ refactor, backend support, misc fixes (#11030) 2025-01-07 18:01:58 +01:00
llama-impl.h cleanup: fix compile warnings associated with gnu_printf (#11811) 2025-02-12 10:06:53 -04:00
llama-kv-cache.cpp reduce some spamminess 2025-02-08 22:49:48 +08:00
llama-kv-cache.h llama : update llama_decode_internal ref [no ci] (#11840) 2025-02-13 08:07:51 +02:00
llama-mmap.cpp Merge branch 'upstream' into concedo_experimental 2025-01-21 00:25:07 +08:00
llama-mmap.h llama-mmap: fix missing include (#11796) 2025-02-10 20:58:18 +02:00
llama-model-loader.cpp tts can now set a length limit 2025-01-28 22:06:59 +08:00
llama-model-loader.h llama : add llama_model_load_from_splits (#11255) 2025-01-16 13:54:08 +01:00
llama-model.cpp Merge branch 'upstream' into concedo_experimental 2025-02-08 22:57:18 +08:00
llama-model.h rpc : early register backend devices (#11262) 2025-01-17 10:57:09 +02:00
llama-quant.cpp Merge branch 'upstream' into concedo_experimental 2025-01-17 23:13:50 +08:00
llama-quant.h llama : refactor src/llama.cpp (#10902) 2025-01-03 10:18:53 +02:00
llama-sampling.cpp sampling: add Top-nσ sampler (#11223) 2025-02-13 08:45:57 +02:00
llama-sampling.h llama : add llama_vocab, functions -> methods, naming (#11110) 2025-01-12 11:32:42 +02:00
llama-vocab.cpp Merge branch 'upstream' into concedo_experimental 2025-02-01 17:14:59 +08:00
llama-vocab.h survived the storm, again 2025-01-16 22:25:18 +08:00
llama.cpp Merge branch 'upstream' into concedo_experimental 2025-02-08 22:57:18 +08:00
unicode-data.cpp server : better security control for public deployments (#9776) 2024-10-08 13:27:04 +02:00
unicode-data.h llama : reduce compile time and binary size (#9712) 2024-10-02 15:49:55 +02:00
unicode.cpp Merge branch 'upstream' into concedo_experimental 2025-02-16 02:08:39 +08:00
unicode.h unicode : improve naming style (#10838) 2024-12-16 12:31:45 +02:00