
Grok 4.7
Frontier reasoning model for coding, agentic tasks, and knowledge work
Grok 4.7
Frontier reasoning model for coding, agentic tasks, and knowledge work
Grok 4.7 Overview
Grok 4.7 is xAI's frontier model for coding, agentic tasks, and knowledge work, built to work longer on difficult tasks and check its own work more carefully. It reads text and images, calls tools, and offers reasoning effort from low to xhigh, with a 500,000-token context window and no fixed output limit. It suits software engineering, multi-step agents, and document-heavy analysis.
More models from xAI
Grok Imagine Video 1.5 Lite is the lower-cost tier of xAI's Grok Imagine Video 1.5. It generates video with native audio from a text prompt or from a single starting frame, at 480p, 720p, or 1080p and durations from 1 to 15 seconds. It suits high-volume work such as social clips, ad variations, and fast iteration on prompts and motion.
Grok Imagine Image 2.0 is xAI's next-generation image generation and editing model for both text-to-image and prompt-guided image transformation. It keeps the same core workflow as the current Grok Imagine image family, including aspect-ratio control and image-based editing, while adding a dedicated quality parameter so teams can tune output fidelity within the same API shape. It is a strong fit for creative production pipelines that want one Grok image endpoint for generation, editing, and quality-sensitive iteration without switching to a separate model family.
Grok Imagine Video 1.5 is xAI's newer image-to-video model. It is positioned above the earlier Grok Imagine Video release with higher per-second pricing, supports durations up to 15 seconds, and generates 480p or 720p video from a single still-image starting frame for cinematic clips, animated visuals, and prompt-guided short-form video creation.
Grok Imagine Image Quality is xAI's quality-focused image generation and editing model. It is designed for higher realism, stronger multilingual text rendering, tighter prompt following, deeper scene understanding, and more consistent brand-oriented output across both text-to-image and image editing workflows.
Grok 4.3 is xAI's flagship language model for agentic reasoning, strong instruction following, and minimal hallucinations. It supports text and image input, a 1 million token context window, configurable reasoning effort including non-reasoning mode, function calling, and structured outputs for production assistants, coding workflows, and long-context analysis.
xAI Text-to-Speech converts text into natural-sounding spoken audio with a single API call. It offers more than two dozen expressive voices, inline speech tags for fine-grained control over pauses, laughter, whispers, and emphasis, and supports over 20 auto-detected languages.
Grok Imagine Image is a multimodal generative image model that creates high-quality still images from text prompts or image inputs. It supports flexible visual synthesis across a range of styles, enabling developers to generate creative imagery directly from structured prompts or to expand on existing visuals with coherent, detailed outputs.
Grok Imagine Video is a multimodal generative video model that produces short video clips with native audio from text descriptions or static images. It supports text-to-video and image-to-video generation with synchronized sound effects and dialogue, enabling developers to animate scenes with motion, camera dynamics, and audio in a single API workflow.







