Qwen/Qwen3-Reranker-4BQwen3 Reranker 4B by Reranker.
Select an endpoint and copy a working example for this model.
from openai import OpenAI client = OpenAI( api_key="YOUR_API_KEY", base_url="https://api.apertis.ai/v1") response = client.chat.completions.create( model="Qwen/Qwen3-Reranker-4B", messages=[ {"role": "user", "content": "Hello!"} ], max_tokens=1024, temperature=0.7) print(response.choices[0].message.content) # Optional: Enable context compression to reduce token usage# response = client.chat.completions.create(# model="Qwen/Qwen3-Reranker-4B",# messages=[{"role": "user", "content": "Hello!"}],# extra_body={"compression": {"enabled": True, "model": "gpt-4.1-mini"}}# )modelquerydocumentstop_nreturn_documentsmax_chunks_per_docUse these namespaced identifiers in Cursor IDE to avoid conflicts with built-in models.
See how this model compares to others from the same provider.
See how this model compares to others from the same provider.
No observed failures in the current observation window