Original caption
Mira Murati's startup Thinking Machines Lab just dropped the technical details for an "interaction model", and it's a direct shot at the autonomy consensus the rest of the industry is racing toward. The thesis: AI labs are designing for a future where you're not in the loop. Every architectural choice in ChatGPT, Claude, Gemini assumes you delegated and walked away. Thinking Machine Lab is going the opposite direction: building AI trained natively to *collaborate* with you in real time, not replace you. The architecture is genuinely different. Every 200 milliseconds, the model processes whatever's happening — your voice, your typing, what's on your screen — and produces an output (speech, text, an action, or silence). No more stitching together speech-to-text + LLM + text-to-speech + tool calls. One model, trained on the interaction itself. And to keep both real-time speed AND deep intelligence, they split the system into two models: a small fast one that stays with you, a bigger slow one that handles heavy reasoning in the background. Research preview only, coming in the next few months. Whether the architecture holds up across longer real-world use is still an open question. But the philosophical break is real, and it's the first one any major lab has made. #airesearch #thinkingmachineslab