llama.cpp

History

Aaron Teo 40be51152d ggml-zdnn: fix #15414 , activate FP16 and BF16 acceleration and incorrect zTensor free (#15839 )		2025-09-13 02:39:52 +08:00
..
backend	CANN: Disable acl_graph for prefill stage (#15933 )	2025-09-11 15:59:37 +08:00
development	docs : update HOWTO‑add‑model.md for ModelBase and new model classes (#14874 )	2025-07-25 16:25:05 +02:00
multimodal	model : support MiniCPM-V 4.5 (#15575 )	2025-08-26 10:05:55 +02:00
ops	ggml-zdnn: fix #15414 , activate FP16 and BF16 acceleration and incorrect zTensor free (#15839 )	2025-09-13 02:39:52 +08:00
android.md	repo : update links to new url (#11886 )	2025-02-15 16:40:57 +02:00
build-s390x.md	ggml-zdnn: fix #15414 , activate FP16 and BF16 acceleration and incorrect zTensor free (#15839 )	2025-09-13 02:39:52 +08:00
build.md	Update build.md to remove MSVC arm64 notes (#15684 )	2025-08-30 23:51:28 +08:00
docker.md	musa: upgrade musa sdk to rc4.2.0 (#14498 )	2025-07-24 20:05:37 +01:00
function-calling.md	server : add documentation for `parallel_tool_calls` param (#15647 )	2025-08-29 20:25:40 +03:00
install.md	docs : add "Quick start" section for new users (#13862 )	2025-06-03 13:09:36 +02:00
llguidance.md	llguidance build fixes for Windows (#11664 )	2025-02-14 12:46:08 -08:00
multimodal.md	mtmd : add support for Voxtral (#14862 )	2025-07-28 15:01:48 +02:00
ops.md	ggml-zdnn: fix #15414 , activate FP16 and BF16 acceleration and incorrect zTensor free (#15839 )	2025-09-13 02:39:52 +08:00