GLM Embedding 3 是智谱AI的嵌入模型,可生成针对中英双语检索优化的2048维向量。该模型广泛应用于中文RAG流水线,以极低的成本在C-MTEB基准上达到前沿英文嵌入模型的性能水平。
直接替换OpenAI。只需改base_url和模型。
# Python - pip install openai from openai import OpenAI client = OpenAI( base_url="https://meshtok.com/v1", api_key="sk-your-MeshTok-key", ) resp = client.chat.completions.create( model="bigmodel/embedding-3", messages=[{"role":"user","content":"Hello!"}], ) print(resp.choices[0].message.content)
# cURL curl https://meshtok.com/v1/chat/completions \ -H "Authorization: Bearer sk-your-MeshTok-key" \ -H "Content-Type: application/json" \ -d '{"model":"bigmodel/embedding-3","messages":[{"role":"user","content":"Hello!"}]}'
你可能考虑的其他模型。
Zhipu’s flagship with a 1M context window.
Balanced daily-driver from Zhipu.
Fast GLM variant for low-latency apps.
General-purpose chat model.