llama.cpp

History

Francis Couture-Harpin 9cd402cbe1 tests : add test-model-random This generates random models and then tests different concurrencies of batches to check if the output is consistent. This can detect when e.g. the recurrent cache has been broken, or anything else which would affect the consistency of the output when inferencing multiple distinct sequences. More architectures will be added, but for now this starts with Mamba. Eventually, consistency of pooled embeddings will also be tested. The goal is to reduce accidental regressions by making it easy to quickly test a lot of edge cases on the supported architectures, without having to download any model.		2025-06-12 01:00:57 -04:00
..
.gitignore	tests : gitignore ggml-common.h	2024-03-09 14:17:11 +02:00
CMakeLists.txt	tests : add test-model-random	2025-06-12 01:00:57 -04:00
get-model.cpp	ci : add model tests + script wrapper (#4586 )	2024-01-26 14:18:00 +02:00
get-model.h	ci : add model tests + script wrapper (#4586 )	2024-01-26 14:18:00 +02:00
run-json-schema-to-grammar.mjs	llama : move end-user examples to tools directory (#13249 )	2025-05-02 20:27:13 +02:00
test-arg-parser.cpp	tests : avoid github urls due to throttling (#13654 )	2025-05-20 12:03:17 +02:00
test-autorelease.cpp	llama : add `llama_vocab`, functions -> methods, naming (#11110 )	2025-01-12 11:32:42 +02:00
test-backend-ops.cpp	ggml-vulkan: adds support for op CONV_TRANSPOSE_1D (#13813 )	2025-06-04 22:02:00 +02:00
test-barrier.cpp	ggml : move CPU backend to a separate file (#10144 )	2024-11-03 19:34:08 +01:00
test-c.c	Nomic Vulkan backend (#4456 )	2024-01-29 15:50:50 -05:00
test-chat-parser.cpp	tests : remove json.hpp from a test (#13880 )	2025-05-29 12:17:16 +03:00
test-chat-template.cpp	llama-chat : reset glmedge chat template (#13253 )	2025-05-02 11:06:09 +02:00
test-chat.cpp	llama : allow building all tests on windows when not using shared libs (#13980 )	2025-06-09 20:03:09 +02:00
test-double-float.cpp	ggml : minor naming changes (#8433 )	2024-07-12 10:46:02 +03:00
test-gbnf-validator.cpp	cmake : do not include ./src as public for libllama (#13062 )	2025-04-24 16:00:10 +03:00
test-gguf.cpp	gguf: fix failure on version == 0 (#13956 )	2025-06-01 18:08:05 +02:00
test-grammar-integration.cpp	sync : vendor (#13901 )	2025-05-30 16:25:45 +03:00
test-grammar-llguidance.cpp	cmake : do not include ./src as public for libllama (#13062 )	2025-04-24 16:00:10 +03:00
test-grammar-parser.cpp	cmake : do not include ./src as public for libllama (#13062 )	2025-04-24 16:00:10 +03:00
test-json-partial.cpp	`server`: streaming of tool calls and thoughts when `--jinja` is on (#12379 )	2025-05-25 01:48:08 +01:00
test-json-schema-to-grammar.cpp	sync : vendor (#13901 )	2025-05-30 16:25:45 +03:00
test-llama-grammar.cpp	cmake : do not include ./src as public for libllama (#13062 )	2025-04-24 16:00:10 +03:00
test-log.cpp	common : use common_ prefix for common library functions (#9805 )	2024-10-10 22:57:42 +02:00
test-lora-conversion-inference.sh	ci : use -no-cnv in gguf-split tests (#11254 )	2025-01-15 18:28:35 +02:00
test-model-load-cancel.cpp	llama : update llama_model API names (#11063 )	2025-01-06 10:55:18 +02:00
test-model-random.cpp	tests : add test-model-random	2025-06-12 01:00:57 -04:00
test-mtmd-c-api.c	mtmd : add C public API (#13184 )	2025-05-04 23:43:42 +02:00
test-opt.cpp	llama/ggml: add LLM training support (#10544 )	2025-05-12 14:44:49 +02:00
test-quantize-fns.cpp	tests : fix test-quantize-fns to init the CPU backend (#12306 )	2025-03-10 14:07:15 +02:00
test-quantize-perf.cpp	ggml : inttypes.h -> cinttypes (#0 )	2024-11-17 08:30:29 +02:00
test-quantize-stats.cpp	docker : do not build tests (#13204 )	2025-04-30 10:44:07 +02:00
test-regex-partial.cpp	`common`: add partial regex support (#12808 )	2025-05-14 19:50:57 +01:00
test-rope.cpp	llama : add Qwen2VL support + multimodal RoPE (#10361 )	2024-12-14 14:43:46 +02:00
test-sampling.cpp	sampling : make sure samplers return at least 1 token (#13822 )	2025-05-27 12:07:52 +03:00
test-tokenizer-0.cpp	llama : add `llama_vocab`, functions -> methods, naming (#11110 )	2025-01-12 11:32:42 +02:00
test-tokenizer-0.py	py : logging and flake8 suppression refactoring (#7081 )	2024-05-05 08:07:48 +03:00
test-tokenizer-0.sh	tests : fix test-tokenizer-0.sh	2024-05-28 15:04:09 +03:00
test-tokenizer-1-bpe.cpp	cmake : do not include ./src as public for libllama (#13062 )	2025-04-24 16:00:10 +03:00
test-tokenizer-1-spm.cpp	cmake : do not include ./src as public for libllama (#13062 )	2025-04-24 16:00:10 +03:00
test-tokenizer-random.py	llama : add `llama_vocab`, functions -> methods, naming (#11110 )	2025-01-12 11:32:42 +02:00
test-tokenizers-repo.sh	tests : add test-tokenizers-repo (#14017 )	2025-06-11 17:16:32 +02:00