Blog
RSSUpdates, tutorials, and best practices from the MindRouter team.
-
The Almost Free Lunch: How We Just Doubled MindRouter's AI Speed with Speculative Decoding
Your GPU spends most of its life waiting. Speculative decoding gives it something useful to do.
Read more → -
Conversations That Remember: MindRouter Now Supports OpenAI Responses and Conversations APIs
MindRouter now natively supports OpenAI's Responses and Conversations APIs, giving your programs server-side memory, live web search, automatic context managem…
Read more → -
From Audio to Text: Secure On-Campus Dictation and Transcription with MindRouter
Secure, on-premises speech-to-text and text-to-speech via MindRouter's Voice API, powered by OpenAI Whisper and Kokoro on dedicated GPUs, with a focus on resea…
Read more → -
Bring Your Own Search: Giving Your LLMs and Agents a Window to the Live Web
How to wire MindRouter's web search into chats, scripts, apps, and agents using skills and MCP.
Read more → -
Writing Robust, Resilient Python Clients for MindRouter
A comprehensive guide to building production-grade client code against MindRouter's OpenAI-compatible API
Read more → -
What Happens to Your Request Inside MindRouter's Scheduling and Routing
When you send a request to MindRouter, whether it's a simple chat completion, an image analysis, or an embedding query, it kicks off a carefully orchestrated s…
Read more → -
Agentic AI on Campus: Powering Claude Code with MindRouter
How we run autonomous coding agents on local GPU infrastructure: faster, cheaper, and without sending a single byte off campus.
Read more →