Sentence Similarity
ONNX
Safetensors
English
ogma
embeddings
dense-retrieval
matryoshka
rag
agents
mteb
semantic-search
text-embeddings
text-embedding
vector-search
document-retrieval
similarity-search
classification
clustering
edge-ai
on-device
local-inference
efficient-ai
rag-retrieval
custom_code
Eval Results (legacy)
Download tokenizer_config.json from axiotic/ogma-base: direct link, hf CLI and curl.
- Browser
- Download file 428 Bytes
-
https://huggingface.co/axiotic/ogma-base/resolve/main/tokenizer_config.json
- Command line
-
hf download hf://axiotic/ogma-base/tokenizer_config.json
-
curl -L -o tokenizer_config.json https://huggingface.co/axiotic/ogma-base/resolve/main/tokenizer_config.json
428 Bytes
| { | |
| "tokenizer_class": "OgmaTokenizerFast", | |
| "auto_map": { | |
| "AutoTokenizer": [ | |
| null, | |
| "tokenization_ogma.OgmaTokenizerFast" | |
| ] | |
| }, | |
| "model_max_length": 1024, | |
| "padding_side": "right", | |
| "pad_token": "<pad>", | |
| "unk_token": "<unk>", | |
| "cls_token": "[CLS]", | |
| "sep_token": "[SEP]", | |
| "bos_token": "[CLS]", | |
| "eos_token": "[SEP]", | |
| "mask_token": "[MASK]", | |
| "do_lower_case": true, | |
| "backend": "tokenizers" | |
| } | |