Models & inference
In-engine speculative decoding
Configure draft models for faster local generation.
Related topics
Easy AI model hosting setupVoice control with wake word, STT, and TTSEngineering loopsReverse-engineer a repositoryRunning the autonomous goal agentWas this page helpful?