

Monarch
Explore MonarchDive deeper into our family of thought-grounded multimodal models. is our family of thought-grounded models. Most AI systems are built around tokens, breaking information into pieces and predicting what comes next. We believe intelligence begins earlier, with understanding.
Instead of reasoning through words alone, Monarch
Explore MonarchDive deeper into our family of thought-grounded multimodal models. forms and refines thought vectors that bring language, vision, audio, and memory into a shared understanding. It spends more time on difficult problems and less on simple ones, adapting its reasoning to the task at hand.
Our goal is to build intelligence that can connect ideas across domains, adapt to new situations, and reason from meaning rather than patterns.
Multimodal input

# 02

Multimodal input
Text, images, audio, video, code, and structured data are projected into one shared vector space. No input type is treated as the main one.



Text, images, audio, video, code, and structured data are projected into one shared vector space. No input type is treated as the main one.
This lets Monarch form one combined understanding from many signals, instead of stitching together separate text and image interpretations later.
Cloudy is Ctrl's calm, familiar voice making powerful models feel natural to talk to, and a little more human.
Research at atom.
Atom Ctrl is an AI research lab building Thinking Machines that can understand, reason, and learn from the world.