[Bug]: vLLM recognize the bge-m3-korean (embedding model) max length, 512 tokens.

### Your current environment

vLLM 0.10.2

### 🐛 Describe the bug

bge-m3-korean config.json
{
--
"architectures": [
"XLMRobertaModel"
],
"attention_probs_dropout_prob": 0.1,
"bos_token_id": 0,
"classifier_dropout": null,
"eos_token_id": 2,
"hidden_act": "gelu",
"hidden_dropout_prob": 0.1,
"hidden_size": 1024,
"initializer_range": 0.02,
"intermediate_size": 4096,
"layer_norm_eps": 1e-05,
"max_position_embeddings": 8194,
"model_type": "xlm-roberta",
"num_attention_heads": 16,
"num_hidden_layers": 24,
"output_past": true,
"pad_token_id": 1,
"position_embedding_type": "absolute",
"torch_dtype": "float32",
"transformers_version": "4.42.4",
"type_vocab_size": 1,
"use_cache": true,
"vocab_size": 250002
}


vLLM recognize the bge-m3-korean (embedding model) max length, 512 tokens despite of the config.json above.



### Before submitting a new issue...

- [x] Make sure you already searched for relevant issues, and asked the chatbot living at the bottom right corner of the [documentation page](https://docs.vllm.ai/en/latest/), which can answer lots of frequently asked questions.

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Uh oh!

Uh oh!

[Bug]: vLLM recognize the bge-m3-korean (embedding model) max length, 512 tokens. #25865

Your current environment

🐛 Describe the bug

bge-m3-korean config.json
{

Before submitting a new issue...

Metadata

Assignees

Labels

Type

Projects

Milestone

Relationships

Development

Uh oh!

[Bug]: vLLM recognize the bge-m3-korean (embedding model) max length, 512 tokens. #25865

Description

Your current environment

🐛 Describe the bug

bge-m3-korean config.json {

Before submitting a new issue...

Metadata

Metadata

Assignees

Labels

Type

Projects

Milestone

Relationships

Development

Issue actions

bge-m3-korean config.json
{