EmbeddingGemma 300M

Google DeepMind

Tylko w katalogu Otwarte wagi

Suwerenność i hosting

Tylko w katalogu

Ten model jest skatalogowany jako odniesienie. Nie jest jeszcze hostowany w LLM EU i żaden własny endpoint go tu nie obsługuje. Żądania do niego trafiają bezpośrednio do dostawcy, na warunkach tego dostawcy.

Podsumowanie

EmbeddingGemma is a 308M-parameter text embedding model with a 768-dimensional output that can be truncated to 512, 256 or 128 dimensions through Matryoshka representation learning. The model card states a maximum input length of 2048 tokens and training data in over 100 spoken languages. It is released under Google’s Gemma terms rather than an OSI licence.

Specyfikacje

Dostawca
Google DeepMind (US)
Okno kontekstowe
2.0k
Parametry
0.3B
Modalności
embedding
Licencja
gemma · licence
Otwarte wagi
yes
Data wydania
17 lip 2025
Zadania
rag, multilingual-support, private, cheap

Języki

Mocny: en, de, fr, es, it, pt, nl, pl

Wystarczający: sv, da, fi, cs, ro, el, uk, ru, ar, hi, zh, ja, ko, tr

Języki

Mocny

  • en
  • de
  • fr
  • es
  • it
  • pt
  • nl
  • pl

Wystarczający

  • sv
  • da
  • fi
  • cs
  • ro
  • el
  • uk
  • ru
  • ar
  • hi
  • zh
  • ja
  • ko
  • tr

Niezweryfikowane: dostawca nie udokumentował tego twierdzenia. — a language is listed as strong only where a source claims it; the card links the evidence where one exists.

Kiedy używać

Mocne strony

  • flexible output dimensions for cheaper vector indexes
  • small enough to embed on CPU
  • training data in over 100 languages

Ograniczenia

  • 2048-token input limit
  • Gemma terms of use, not an OSI-approved licence
  • no retrieval benchmark has been run by LLM EU