我们正在更新 Grok 的功能,引入一款代号为 Aurora 的新型自回归图像生成模型,该模型已在 𝕏 平台上可用。
Grok 图像生成 发布 图像生成 图像编辑 展望未来
新模型已上线 Grok Imagine API 现已推出 这是我们迄今为止最强大的生成模型——集最先进的视频生成、视频编辑和图像创建于一体,全部整合在一个 API 中。查看新功能 我们通过一款代号为 Aurora 的新模型增强了 Grok 的图像生成能力。Aurora 是一个自回归混合专家网络,经过训练可从交错的文本和图像数据中预测下一个 token。我们在互联网上数十亿个样本上训练了该模型,使其对世界有了深刻的理解。因此,它在照片级逼真渲染和精确遵循文本指令方面表现出色。除了文本,该模型还原生支持多模态输入,使其能够从用户提供的图像中获取灵感或直接对其进行编辑。
Grok 的新功能现已在 𝕏 平台的部分国家/地区上线,并将在未来一周内向所有用户推出。
洛克希德 SR-71 黑鸟侦察机飞过一片抽象的天空。
一位宇航员站在外星行星表面,背景中有一艘飞船,天空中挂着多颗卫星。
一座被冰包围的火山。
一只猫在双曲时间舱中的叠加态,以梵高风格呈现。
一辆折纸风格的 Cybertruck。
夕阳天空下的樱花。
纸上画的一个多面三维几何形状的草图。
一只在图书馆里喝茶的狗。
一位吉他手握着拨片的手部特写。
一幅漫画:一个年轻人站在海边,回头凝视,表情坚定。对话框里印着文字:“Make it happen, yesterday.”
一个放在盘子里的双层肉饼汉堡。
一位手持剑的女战士,身着精致盔甲,表情自信。
一幅使用几何形状和鲜艳色彩构成的抽象作品,唤起能量与动感。
埃隆·马斯克作为动画剧集《瑞克和莫蒂》中的角色。
日落时分宁静的山间湖泊,水面升起薄雾,山峰完美倒映在平静的湖面上。
一座赛博朋克风格的城市夜景,霓虹灯闪烁,飞行汽车穿梭,摩天大楼高耸入云。
一位老人,捕捉到了每一道皱纹和每一个表情。
蜘蛛网上的露珠,蛛网的精巧纹理与光线的折射清晰可见。
图像生成
Grok 现在能够生成多个领域的高质量图像,而这些领域往往是其他图像生成模型难以应对的。它可以精准渲染真实世界实体、文字、Logo 的视觉细节,还能创作逼真的人物肖像。
实体生成 艺术文字 梗图生成 逼真肖像 名人
提示词
极光下的 Cybertruck
Grok
Imagen 3
Flux.1 Pro
Ideogram 2.0
Dall-E 3
图像编辑
我们全新的图像生成模型现在可以接收图像作为输入,为用户带来更强的创意控制力和灵活性。我们很快将在 𝕏 平台上向用户开放这一功能。
提示词
把猫变成动漫风格
输入图像
输出图像
展望未来
在 xAI,我们正在推进多模态理解与生成的前沿。如果你也为此目标感到振奋,欢迎加入我们的旅程——我们正在招聘!
We are updating Grok's capabilities with a new autoregressive image generation model, code-named Aurora, available on the 𝕏 platform.
Grok Image Generation Release Image Generation Image Editing Looking Forward
A newer model is available Grok Imagine API is here Our most powerful generative model yet — state-of-the-art video generation, video editing, and image creation, all in one API. See what's new We've enhanced Grok's image generation abilities with a new model, code-named Aurora. Aurora is an autoregressive mixture-of-experts network trained to predict the next token from interleaved text and image data. We trained the model on billions of examples from the internet, giving it a deep understanding of the world. As a result, it excels at photorealistic rendering and precisely following text instructions. Beyond text, the model also has native support for multimodal input, allowing it to take inspiration from or directly edit user-provided images.
Grok's new capabilities are now available on the 𝕏 platform in select countries and will roll out to all users within a week.
Lockheed SR-71 Blackbird flying through an abstract sky.
An astronaut standing on the surface of an alien planet, with a spaceship in the background and multiple moons in the sky.
A volcano surrounded by ice.
A superposition of a cat in a hyperbolic time chamber in the style of Van Gogh.
An origami Cybertruck.
Cherry blossoms beneath a sunset sky.
A sketch of a multi-sided 3d geometric shape on paper.
A dog drinking a cup of tea in a library.
A closeup of a guitar player's hand holding a pick.
A comic of a young man standing by the sea, gazing back over his shoulder with a determined expression. In a speech bubble, printing the text, 'Make it happen, yesterday.'
A burger with double meat patty placed on a plate.
A female warrior holding a sword, with intricate armor and a confident expression.
An abstract composition using geometric shapes and vibrant colors, evoking a sense of energy and movement.
Elon Musk as a character in the animated series Rick and Morty.
A serene mountain lake at sunset, with mist rising from the water and the peaks reflected perfectly in the still surface.
A cyberpunk-inspired city at night, with neon lights, flying cars, and towering skyscrapers.
An elderly person, capturing every wrinkle and expression.
A dewdrop on a spider web, with the intricate patterns of the web and the refraction of light.
Image Generation
Grok can now generate high-quality images across several domains where other image generation models often struggle. It can render precise visual details of real-world entities, text, logos, and can create realistic portraits of humans.
Entity generation Artistic text Meme generation Realistic portraits Celebrities
Prompt
Cybertruck under an aurora
Grok
Imagen 3
Flux.1 Pro
Ideogram 2.0
Dall-E 3
Image Editing
Our new image generation model can now take images as input, giving users greater creative control and flexibility. We will release this capability to users on the 𝕏 platform soon.
Prompt
Make the cat anime style
Input image
Output image
Looking Forward
At xAI, we are advancing the frontier of multimodal understanding and generation. If this goal inspires you, we invite you to join us on this journey — we are hiring!