.env handling के ज़रिये model preference और
endpoint configuration लिखता है। यह weights डाउनलोड नहीं करता और न API secrets
स्टोर करता है; launch commands और provider-specific अगले steps print किए जाते हैं।
API keys के लिए arka integration setup <provider> का उपयोग करें, फिर runtime
readiness जाँचने के लिए arka model doctor। लोकल और hosted मॉडलों को
arka hybrid status और arka hybrid run --policy parallel के साथ जोड़ा जा सकता है।
LAN inference cluster के लिए Exo को एक लोकल OpenAI-compatible host के रूप में
configure करें:
local-only के तहत इसे कभी hosted fallbacks पर नहीं भेजता।
Model advisor Apple Silicon/MLX अवसरों, MoE model hints, और speculative/MTP
decoding hints को भी annotate करता है। ये evidence labels हैं, guarantees नहीं:
deployment से पहले runtime support और active-parameter metadata की पुष्टि करनी
चाहिए।