llama.cpp

History

Sirui He 073bb2c20b mtmd : add MERaLiON-2 multimodal audio support (#21756 ) * mtmd : add MERaLiON-2 multimodal audio support Adds support for ASTAR's MERaLiON-2 audio-language model (3B and 10B) to the multimodal framework. Architecture: - Whisper large-v2 encoder for audio feature extraction - Gated MLP adaptor: ln_speech -> frame stack (x15) -> Linear+SiLU -> GLU -> out_proj - Gemma2 3B / 27B decoder The mmproj GGUF is generated via convert_hf_to_gguf.py --mmproj on the full MERaLiON-2 model directory (architecture: MERaLiON2ForConditionalGeneration). The decoder is converted separately as a standard Gemma2 model after stripping the text_decoder. weight prefix. New projector type: PROJECTOR_TYPE_MERALION Supports tasks: speech transcription (EN/ZH/MS/TA), translation, spoken QA. Model: https://huggingface.co/MERaLiON/MERaLiON-2-3B https://huggingface.co/MERaLiON/MERaLiON-2-10B simplify comments in meralion adaptor * meralion: use format_tensor_name, ascii arrows in comments		2026-04-11 14:15:48 +02:00
..
scripts	ggml : add NVFP4 quantization type support (#19769 )	2026-03-11 21:02:54 +01:00
__init__.py	convert-*.py: GGUF Naming Convention Refactor and Metadata Override Refactor (#7499 )	2024-07-18 20:40:15 +10:00
constants.py	mtmd : add MERaLiON-2 multimodal audio support (#21756 )	2026-04-11 14:15:48 +02:00
gguf.py	gguf-py: Refactor and allow reading/modifying existing GGUF files (#3981 )	2023-11-11 08:04:50 +03:00
gguf_reader.py	ggml/gguf : prevent integer overflows (#19856 )	2026-02-24 20:17:11 +02:00
gguf_writer.py	model, mtmd: fix gguf conversion for audio/vision mmproj (#21309 )	2026-04-02 17:10:32 +02:00
lazy.py	ci : switch from pyright to ty (#20826 )	2026-03-21 08:54:34 +01:00
metadata.py	chore : correct typos [no ci] (#20041 )	2026-03-05 08:50:21 +01:00
py.typed	convert : various script cleanups/fixes + merges and special token handling (#2842 )	2023-08-30 11:25:50 +03:00
quants.py	ci : switch from pyright to ty (#20826 )	2026-03-21 08:54:34 +01:00
tensor_mapping.py	mtmd : add MERaLiON-2 multimodal audio support (#21756 )	2026-04-11 14:15:48 +02:00
utility.py	gguf-py : do not align the data start offset (#18291 )	2025-12-22 20:25:16 +01:00
vocab.py	requirements : update transformers to 5.5.1 (#21617 )	2026-04-09 12:36:29 +02:00