LLM প্রদানকারী
বিভাগ
টেবিলের মান registry থেকে আসে। URL scheme ও পুরো path রাখুন; placeholder বাস্তব deployment দিয়ে বদলান। অধিকার, মূল্য ও model অবসরের সময় provider-এর বিষয়।
ক্লাউড
| প্রদানকারী | প্রাথমিক মডেল | প্রাথমিক URL |
|---|---|---|
| DeepSeek | deepseek-v4-pro | https://api.deepseek.com |
| Qwen | qwen3-235b-a22b | https://dashscope.aliyuncs.com/compatible-mode/v1 |
| Qwen Code | qwen3-coder-plus | https://dashscope.aliyuncs.com/compatible-mode/v1 |
| Doubao | ep-xxxxxxxxxxxxxxxx | https://ark.cn-beijing.volces.com/api/v3 |
| Moonshot | kimi-k2-0905-preview | https://api.moonshot.cn/v1 |
| Xiaomi MiMo | mimo-v2.5-pro | https://api.xiaomimimo.com/v1 |
| GLM | glm-5 | https://open.bigmodel.cn/api/paas/v4 |
| Z AI | glm-5 | https://api.z.ai/api/paas/v4 |
| MiniMax | MiniMax-M2.7 | https://api.minimaxi.com/v1 |
| Baidu Qianfan | ernie-4.5-turbo-32k | https://qianfan.baidubce.com/v2 |
| SiliconFlow | Qwen/QwQ-32B | https://api.siliconflow.cn/v1 |
| Huawei Cloud MaaS | DeepSeek-V3 | https://api.modelarts-maas.com/v1 |
| OpenAI | gpt-4o | https://api.openai.com/v1 |
| Anthropic | claude-3-5-sonnet-20240620 | https://api.anthropic.com |
gemini-2.0-flash-exp | https://generativelanguage.googleapis.com/v1 | |
| Mistral | mistral-large-latest | https://api.mistral.ai/v1 |
| Azure OpenAI | gpt-4o | নিজের Azure resource endpoint সেট করুন |
| xAI | grok-4 | https://api.x.ai/v1 |
| Groq | moonshotai/kimi-k2-instruct-0905 | https://api.groq.com/openai/v1 |
| Together | meta-llama/Meta-Llama-3.1-70B-Instruct-Turbo | https://api.together.xyz/v1 |
| Fireworks | accounts/fireworks/models/kimi-k2p5 | https://api.fireworks.ai/inference/v1 |
| Nebius | openai/gpt-oss-120b | https://api.studio.nebius.com/v1 |
| Cerebras | gpt-oss-120b | https://api.cerebras.ai/v1 |
Gateway ও proxy
| প্রদানকারী | প্রাথমিক মডেল | প্রাথমিক URL |
|---|---|---|
| OpenRouter | anthropic/claude-3.7-sonnet | https://openrouter.ai/api/v1 |
| AIHubMix | gpt-4o-mini | https://aihubmix.com/v1 |
| GitHub Models | gpt-4o-mini | https://models.github.ai/inference |
| PPIO | qwen/qwen3-32b | https://api.ppinfra.com/v3/openai |
| New API | gpt-4.1 | http://localhost:3000/v1 |
| LiteLLM | your-proxy-model | http://localhost:4000/v1 |
| Hugging Face | openai/gpt-oss-120b | https://router.huggingface.co/v1 |
| Vercel AI Gateway | anthropic/claude-sonnet-4.5 | https://ai-gateway.vercel.sh/v1 |
| Requesty | anthropic/claude-3-7-sonnet-latest | https://router.requesty.ai/v1 |
| OpenAI Compatible | your-model-id | https://your-openai-compatible-endpoint/v1 |
স্থানীয়
| প্রদানকারী | প্রাথমিক মডেল | প্রাথমিক URL |
|---|---|---|
| OVMS | openvino-model | http://localhost:8000/v3 |
| LMStudio | local-model | http://localhost:1234/v1 |
| Ollama | llama3 | http://localhost:11434/api |
চীনা প্রদানকারী
আঞ্চলিক setup দেখুন। বিভাগটি geographic availability বা Chinese-language quality নিশ্চিত করে না। Account region, endpoint ও model entitlement মেলান।
কাজভিত্তিক নির্বাচন
useMultiModelSettings আলাদা provider/model override দেয়; না হলে activeProvider ব্যবহৃত হয়। ফাঁকা override সংশ্লিষ্ট provider model নেয়। Invalid saved provider settings load-এর সময় active provider-এর সঙ্গে সমন্বিত হয়।
| কাজ | Provider setting | Model override |
|---|---|---|
| লিংক | addLinksProvider | addLinksModel |
| Title generation | generateTitleProvider | generateTitleModel |
| Research summary | researchProvider | researchModel |
| Translation | translateProvider | translateModel |
| Concept extraction | extractConceptsProvider | extractConceptsModel |
| Mermaid summary | summarizeToMermaidProvider | summarizeToMermaidModel |
| Original-text extraction | extractOriginalTextProvider | extractOriginalTextModel |
Active ও saved task provider প্রথমে DeepSeek। শুধু active বদলালে enabled individual routing বদলায় না। Private content-এর আগে প্রতিটি route দেখুন।
API-এর গঠন
Transport
Obsidian requests ও protocol-aware fallback ব্যবহৃত হয়। Desktop-এ Node HTTP streams, অন্য environment-এ fetch হতে পারে। OpenAI-compatible gateway-এর Claude ওই gateway profile-এ চলে, native Anthropic-এ নয়।
Retry
Stable mode শুরুতে বন্ধ। প্রযোজ্য default তিন retries এবং পাঁচ সেকেন্ড interval। Transient transport failure shared retry/fallback নিতে পারে। Authentication বা ভুল model ঠিক করতে হয়। Long request timeout, model limit বা quota-তে ব্যর্থ হতে পারে।
Response cache
সীমিত memory cache সফল response পুনর্ব্যবহার করতে পারে। Endpoint, model, generation settings, prompt ও content request identity-তে থাকে। এটি persistence বা cost meter নয়, নতুন paid request না হওয়ার নিশ্চয়তাও নয়।
Reasoning model
Reasoning/thinking profile ও model rules অনুসরণ করে। অন্য protocol-এর parameter কপি করবেন না। নির্দিষ্ট task test করুন; chat চললেও সব token/sampling setting গ্রহণ নাও করতে পারে।
Token অনুমান
Text length ও output limits exact provider token count নয়। Billing-এর জন্য actual usage দেখুন। Universal cost estimate নেই।
Model discovery
Fetch models যেখানে supported সেখানে চলে। কিছু deployment manual নাম চায়। List ও generation rights আলাদা। Test connection-এর models-then-chat শুধু list দিয়েই সফল হতে পারে; chat-only ছোট generation করে। আসল task ও output form যাচাই করতে হবে।
- Anthropic ও Google native discovery ব্যবহার করে।
- Ollama pulled tags, LM Studio server-exposed models দেখায়।
- Azure deployment ও API version চায়।
- Gateway list সীমিত রেখেও explicit model chat দিতে পারে।
শুরু
- Settings → Notemd-এ সঠিক profile নিন।
- Endpoint, credentials ও উপলব্ধ model/deployment দিন।
- Connection ও disposable note test করুন।
- Output/report দেখে batch বা per-task routing চালু করুন।
Local model ও server আগে প্রস্তুত করুন। Ollama key বাধ্যতামূলক নয়; LM Studio optional key-তে EMPTY দিয়ে শুরু করে। সুরক্ষিত server-এ বাস্তব token দিন। Local LLM web search offline করে না।