← model directory

cloudflare

google/embeddinggemma-300m

cloudflare/google/embeddinggemma-300m

↓ runs free on your own hardware⚙ tool calling

EmbeddingGemma is a 300M parameter, state-of-the-art for its size, open embedding model from Google, built from Gemma 3 (with T5Gemma initialization) and the same research and technology used to create Gemini models. EmbeddingGemma produces vector representations of text, making it well-suited for search and retrieval tasks, including classification, clustering, and semantic similarity search. This model was trained with data in 100+ spoken languages.

specs & pricing
Type
embeddings
Provider
cloudflare
Model ID
cloudflare/google/embeddinggemma-300m
Capabilities
tools
Self-hostable
Yes — runs on your own GPU

Listed in the directory for discovery. embeddingsmodels aren't callable through the chat gateway yet — image and video models can run on your own nodes today.

/// initialize

Call google/embeddinggemma-300m through one OpenAI-compatible endpoint.

Serve it from hardware you own, with cloud failover when your nodes are busy or offline.

no credit card · 2 nodes free · openai-compatible

Frequently asked questions

Can I run google/embeddinggemma-300m on my own hardware?
Yes — google/embeddinggemma-300m is an open-weight model you can run on a workstation or on-prem server you own. Wide Area Intelligence routes requests to your own nodes first and fails over to the cloud only when they can't serve.

More from cloudflare

← browse all models