model: EmbeddingGemma Adding Support for SentenceTransformers Dense Modules (#16367)

mirror of https://github.com/ggml-org/llama.cpp.git synced 2025-11-15 11:17:31 +00:00

* model: EmbeddingGemma sentence-transformers dense linear projections support

* model: add support for EmbeddingGemma SentenceTransformers dense linear projections

Adding support for the Dense modules used in EmbeddingGemma models.
EmbeddingGemma is a SentenceTransformers model with additional modules beyond the base Transformer backbone.

See: https://developers.googleblog.com/en/gemma-explained-embeddinggemma-architecture-and-recipe/

* model: add support for EmbeddingGemma SentenceTransformers dense linear projections

- converting model with dense-layers is optional
- introduced dense config params

* Update convert_hf_to_gguf.py

Co-authored-by: Daniel Bevenius <daniel.bevenius@gmail.com>

* fixed formatting issues

* Update src/llama-graph.cpp

Co-authored-by: Georgi Gerganov <ggerganov@gmail.com>

* - removed pooling_type_opt, always allow overriding pooling_type
- asserts checking dense features dims

* fix python lint

* fix ubuntu gcc build warning

* - fixed thread-safety test
- moved asserts to load_hparams

* - tidying up code
- simplifying graph-context expecting both dense weights

* minor : add TODO

---------

Co-authored-by: Daniel Bevenius <daniel.bevenius@gmail.com>
Co-authored-by: Georgi Gerganov <ggerganov@gmail.com>

This commit is contained in:

Saba Fallah

2025-10-09 08:39:18 +02:00

committed by

GitHub

parent 12bbc3fa50

commit e08db42595

12 changed files with 170 additions and 7 deletions

									
										8

src/llama-graph.h
									
												View File
												
				@@ -814,6 +814,14 @@ struct llm_graph_context {

				            ggml_tensor * cls_b,

				            ggml_tensor * cls_out,

				            ggml_tensor * cls_out_b) const;

				    //

				    // dense (out)

				    //

				    void build_dense_out(

				            ggml_tensor * dense_2,

				            ggml_tensor * dense_3) const;

				};

				// TODO: better name

model: EmbeddingGemma Adding Support for SentenceTransformers Dense Modules (#16367)

8 src/llama-graph.h Unescape Escape View File

8

src/llama-graph.h

View File