Grok 正式发布图像生成功能
Grok 推出图像生成功能,支持通过自然语言指令直接创建图片。该功能已向 X 平台用户开放,标志着这款 AI 助手从文本交互向多模态能力扩展,用户可在对话中直接生成并编辑视觉内容。
xAI发布Grok图像生成功能,拓展多模态应用场景
A newer model is available
Grok Imagine API is here
Our most powerful generative model yet — state-of-the-art video generation, video editing, and image creation, all in one API.
We've enhanced Grok's image generation abilities with a new model, code-named Aurora. Aurora is an autoregressive mixture-of-experts network trained to predict the next token from interleaved text and image data. We trained the model on billions of examples from the internet, giving it a deep understanding of the world. As a result, it excels at photorealistic rendering and precisely following text instructions. Beyond text, the model also has native support for multimodal input, allowing it to take inspiration from or directly edit user-provided images.
Grok's new capabilities are now available on the 𝕏 platform in select countries and will roll out to all users within a week.
Image Generation
Grok can now generate high-quality images across several domains where other image generation models often struggle. It can render precise visual details of real-world entities, text, logos, and can create realistic portraits of humans.
Prompt
Cybertruck under an aurora

Grok

Imagen 3

Flux.1 Pro

Ideogram 2.0

Dall-E 3
Image Editing
Our new image generation model can now take images as input, giving users greater creative control and flexibility. We will release this capability to users on the 𝕏 platform soon.
Prompt
Make the cat anime style

Input image

Output image
Looking Forward
At xAI, we are advancing the frontier of multimodal understanding and generation. If this goal inspires you, we invite you to join us on this journey — we are hiring!
来源:xAI:News(网页) · x.ai