Cataloging the robot uprising, one model at a time 🍓
OpenAI
Cinematic text-to-video model producing long, coherent, physics-aware clips with synced audio.
Google's flagship natively-multimodal model with a massive 2M-token context window.
Ultra-fast, efficient multimodal model for scale.
DeepSeek
Open-weight MoE model with remarkable efficiency and low API cost.
Alibaba
Alibaba's flagship open-weight model with strong multilingual and coding skills.
OpenAI
OpenAI's flagship frontier model with unified reasoning, agentic tool use, and near-expert performance across domains.
OpenAI
Cost-efficient version of GPT-5 for high-volume production apps.
Anthropic
Anthropic's most capable model, excelling at long-horizon agentic coding and nuanced reasoning.
Anthropic
Balanced workhorse model with huge context and strong coding at a fraction of Opus cost.
Anthropic
Fast, cheap model for lightweight tasks and high throughput.
AKA 'Nano Banana' — best-in-class image generation and editing.
Mistral
Mistral's top-tier reasoning and coding model, strong on European languages.
xAI
xAI's frontier model with real-time X data access and strong reasoning.
xAI
Multi-agent version that runs parallel reasoning for hardest problems.
Alibaba
Agentic coding model rivaling closed frontier coders.
ElevenLabs
Most expressive text-to-speech with emotional control and 70+ languages.
OpenAI
Deep reasoning model that thinks step-by-step for math, science, and complex code.
State-of-the-art video generation with native audio.
High-fidelity image generation with excellent text rendering.
Mistral
Frontier-class performance at 8x lower cost.
Suno
Full-song music generation with vocals from a text prompt.
Meta
Meta's largest open-weight MoE teacher model with astonishing context length.
Meta
Open-weight multimodal MoE model competitive with closed frontier models.
Meta
Efficient open model that fits on a single GPU with enormous context.
Midjourney
Beloved artistic image generator with distinctive aesthetic quality.
Runway
Consistent character/world video generation for filmmakers.
DeepSeek
Open reasoning model that shook the industry with frontier-level math at tiny cost.
Perplexity
Search-grounded model powering Perplexity answers with citations.
xAI
xAI's photorealistic image generation model.
Microsoft
Small open model punching far above its weight on reasoning.
Amazon
Amazon's multimodal model available on Bedrock.
Mistral
Vision-language model with strong document understanding.
Cohere
Open multilingual model covering 23 languages.
Stability AI
Open-weight text-to-image model family for local generation.
Black Forest Labs
Fast, high-fidelity image model with excellent prompt adherence.
Cohere
Enterprise RAG-optimized model with strong tool use.
Anthropic
The model that pioneered Artifacts and computer use.
Nvidia
Open model optimized for synthetic data generation.
OpenAI
Omni model handling text, vision and audio in real time with voice mode.
Mistral
Specialized code generation and fill-in-the-middle model.
OpenAI
Text-to-image model integrated into ChatGPT for prompt-faithful generation.
OpenAI
Robust open-source multilingual speech-to-text model.