মূল বিষয়বস্তুতে যান

LLM প্রদানকারী

💡TL;DR
36 profiles পাঁচ পরিবারে রয়েছে: OpenAI-compatible, Anthropic, Google, Azure OpenAI ও Ollama। Endpoint-এর protocol, credentials ও উপলব্ধ model সঠিকভাবে মিলান। Preset-এর প্রাথমিক মান পরিবর্তনযোগ্য, বর্তমান model availability-এর নিশ্চয়তা নয়।

বিভাগ​

টেবিলের মান registry থেকে আসে। URL scheme ও পুরো path রাখুন; placeholder বাস্তব deployment দিয়ে বদলান। অধিকার, মূল্য ও model অবসরের সময় provider-এর বিষয়।

ক্লাউড​

প্রদানকারীপ্রাথমিক মডেলপ্রাথমিক URL
DeepSeekdeepseek-v4-prohttps://api.deepseek.com
Qwenqwen3-235b-a22bhttps://dashscope.aliyuncs.com/compatible-mode/v1
Qwen Codeqwen3-coder-plushttps://dashscope.aliyuncs.com/compatible-mode/v1
Doubaoep-xxxxxxxxxxxxxxxxhttps://ark.cn-beijing.volces.com/api/v3
Moonshotkimi-k2-0905-previewhttps://api.moonshot.cn/v1
Xiaomi MiMomimo-v2.5-prohttps://api.xiaomimimo.com/v1
GLMglm-5https://open.bigmodel.cn/api/paas/v4
Z AIglm-5https://api.z.ai/api/paas/v4
MiniMaxMiniMax-M2.7https://api.minimaxi.com/v1
Baidu Qianfanernie-4.5-turbo-32khttps://qianfan.baidubce.com/v2
SiliconFlowQwen/QwQ-32Bhttps://api.siliconflow.cn/v1
Huawei Cloud MaaSDeepSeek-V3https://api.modelarts-maas.com/v1
OpenAIgpt-4ohttps://api.openai.com/v1
Anthropicclaude-3-5-sonnet-20240620https://api.anthropic.com
Googlegemini-2.0-flash-exphttps://generativelanguage.googleapis.com/v1
Mistralmistral-large-latesthttps://api.mistral.ai/v1
Azure OpenAIgpt-4oনিজের Azure resource endpoint সেট করুন
xAIgrok-4https://api.x.ai/v1
Groqmoonshotai/kimi-k2-instruct-0905https://api.groq.com/openai/v1
Togethermeta-llama/Meta-Llama-3.1-70B-Instruct-Turbohttps://api.together.xyz/v1
Fireworksaccounts/fireworks/models/kimi-k2p5https://api.fireworks.ai/inference/v1
Nebiusopenai/gpt-oss-120bhttps://api.studio.nebius.com/v1
Cerebrasgpt-oss-120bhttps://api.cerebras.ai/v1

Gateway ও proxy​

প্রদানকারীপ্রাথমিক মডেলপ্রাথমিক URL
OpenRouteranthropic/claude-3.7-sonnethttps://openrouter.ai/api/v1
AIHubMixgpt-4o-minihttps://aihubmix.com/v1
GitHub Modelsgpt-4o-minihttps://models.github.ai/inference
PPIOqwen/qwen3-32bhttps://api.ppinfra.com/v3/openai
New APIgpt-4.1http://localhost:3000/v1
LiteLLMyour-proxy-modelhttp://localhost:4000/v1
Hugging Faceopenai/gpt-oss-120bhttps://router.huggingface.co/v1
Vercel AI Gatewayanthropic/claude-sonnet-4.5https://ai-gateway.vercel.sh/v1
Requestyanthropic/claude-3-7-sonnet-latesthttps://router.requesty.ai/v1
OpenAI Compatibleyour-model-idhttps://your-openai-compatible-endpoint/v1

স্থানীয়​

প্রদানকারীপ্রাথমিক মডেলপ্রাথমিক URL
OVMSopenvino-modelhttp://localhost:8000/v3
LMStudiolocal-modelhttp://localhost:1234/v1
Ollamallama3http://localhost:11434/api

চীনা প্রদানকারী​

আঞ্চলিক setup দেখুন। বিভাগটি geographic availability বা Chinese-language quality নিশ্চিত করে না। Account region, endpoint ও model entitlement মেলান।

কাজভিত্তিক নির্বাচন​

useMultiModelSettings আলাদা provider/model override দেয়; না হলে activeProvider ব্যবহৃত হয়। ফাঁকা override সংশ্লিষ্ট provider model নেয়। Invalid saved provider settings load-এর সময় active provider-এর সঙ্গে সমন্বিত হয়।

কাজProvider settingModel override
লিংকaddLinksProvideraddLinksModel
Title generationgenerateTitleProvidergenerateTitleModel
Research summaryresearchProviderresearchModel
TranslationtranslateProvidertranslateModel
Concept extractionextractConceptsProviderextractConceptsModel
Mermaid summarysummarizeToMermaidProvidersummarizeToMermaidModel
Original-text extractionextractOriginalTextProviderextractOriginalTextModel

Active ও saved task provider প্রথমে DeepSeek। শুধু active বদলালে enabled individual routing বদলায় না। Private content-এর আগে প্রতিটি route দেখুন।

API-এর গঠন​

Transport​

Obsidian requests ও protocol-aware fallback ব্যবহৃত হয়। Desktop-এ Node HTTP streams, অন্য environment-এ fetch হতে পারে। OpenAI-compatible gateway-এর Claude ওই gateway profile-এ চলে, native Anthropic-এ নয়।

Retry​

Stable mode শুরুতে বন্ধ। প্রযোজ্য default তিন retries এবং পাঁচ সেকেন্ড interval। Transient transport failure shared retry/fallback নিতে পারে। Authentication বা ভুল model ঠিক করতে হয়। Long request timeout, model limit বা quota-তে ব্যর্থ হতে পারে।

Response cache​

সীমিত memory cache সফল response পুনর্ব্যবহার করতে পারে। Endpoint, model, generation settings, prompt ও content request identity-তে থাকে। এটি persistence বা cost meter নয়, নতুন paid request না হওয়ার নিশ্চয়তাও নয়।

Reasoning model​

Reasoning/thinking profile ও model rules অনুসরণ করে। অন্য protocol-এর parameter কপি করবেন না। নির্দিষ্ট task test করুন; chat চললেও সব token/sampling setting গ্রহণ নাও করতে পারে।

Token অনুমান​

Text length ও output limits exact provider token count নয়। Billing-এর জন্য actual usage দেখুন। Universal cost estimate নেই।

Model discovery​

Fetch models যেখানে supported সেখানে চলে। কিছু deployment manual নাম চায়। List ও generation rights আলাদা। Test connection-এর models-then-chat শুধু list দিয়েই সফল হতে পারে; chat-only ছোট generation করে। আসল task ও output form যাচাই করতে হবে।

  • Anthropic ও Google native discovery ব্যবহার করে।
  • Ollama pulled tags, LM Studio server-exposed models দেখায়।
  • Azure deployment ও API version চায়।
  • Gateway list সীমিত রেখেও explicit model chat দিতে পারে।

শুরু​

  1. Settings → Notemd-এ সঠিক profile নিন।
  2. Endpoint, credentials ও উপলব্ধ model/deployment দিন।
  3. Connection ও disposable note test করুন।
  4. Output/report দেখে batch বা per-task routing চালু করুন।

Local model ও server আগে প্রস্তুত করুন। Ollama key বাধ্যতামূলক নয়; LM Studio optional key-তে EMPTY দিয়ে শুরু করে। সুরক্ষিত server-এ বাস্তব token দিন। Local LLM web search offline করে না।

পরবর্তী ধাপ​