Imagine a digital canvas where your ideas come to life without the need for words, where a simple image can trigger an explosion of visual creativity.
In the fast-paced world of artificial intelligence (AI), the Gemini digital universe has taken a bold step with Whisk, an experimental tool that redefines how we generate images.
Launched under the Google Labs umbrella, Whisk allows users to create art from other images, eliminating the need for thorough textual descriptions.
This approach, as intuitive as it is innovative, is capturing the imagination of artists, designers and the curious alike. We invite you to discover what makes Whisk special, how it works, and how it stacks up against other options in the creative AI landscape.
What is Whisk?
Whisk was launched in December 2024 as part of Google Labs, the platform where Google experiments with disruptive ideas. Initially restricted to the United States, its success led to a global expansion in February 2025, reaching more than 100 countries.
This deployment reflects Google’s commitment to bringing generative AI to a broader audience, democratizing tools that previously seemed reserved for experts.
How It Works
Unlike traditional image generators, which rely on textual prompts, Whisk relies on visual input.
It combines two cutting-edge AI models: Gemini and Image 3 (although the integration of Image 4 is expected before the end of May). Gemini analyzes the images uploaded by the user (which can define subject, scene and style) and generates automatic descriptions.
Image 3 then transforms those descriptions into final images. This process makes Whisk an ideal tool for those who think in images rather than words, offering a direct and fluid creative experience.
The Veo 2 video model has recently been incorporated, which allows generating animated shorts from images.
Features that define Whisk
Google Whisk shines for its simplicity. Its interface allows you to drag and drop images to configure three basic elements: the subject (what appears), the scene (where it appears) and the style (how it looks).
In addition, it includes presets such as “Ornament”, “Sticker”, “Enamel Pin” or “Plushie”, which apply predefined styles with a single click. These details make it accessible to both newbies and experienced creators looking for speed.
A creative playground
Whisk is not intended for precise editing, but rather for experimentation. It is perfect for sketching out ideas, testing concepts, or simply letting yourself be carried away by inspiration.
Users can combine images and see instant results, encouraging a dynamic creative flow. This spontaneity is one of its greatest attractions: it is not about perfection, but about discovery.
Sharing and archiving features
After generating an image, Whisk allows you to share it through direct links, ideal for collaborating or showing your work. It also has «My Library», a section where all your creations are saved, making it easy to review and reuse past ideas. This functionality adds a practical touch to your experimental approach.
Whisk in Practice
Let’s take a real case: a designer wants to create a stuffed animal character for a children’s project. Upload a photo of a bear as your subject, a snowy landscape as your scene, and select the “Plushie” preset.
In seconds, Whisk generates an adorable image of a teddy bear in the snow. What previously required hours of sketching is now accomplished in minutes, showing the tool’s potential to streamline creative processes.
User opinions
The community has received Whisk with enthusiasm. On social media, they describe it as “charming” and “addictive,” highlighting how “time flies” when using it.
However, it is not without criticism: some regret its limited availability by country or the lack of fine control over the results. Still, his fresh approach has made a notable impact.
Alternatives to Compare
Whisk is not alone in the field of AI imaging. There are other tools that also incorporate visual input, each with their own strengths.
Adobe Firefly
Adobe Firefly allows you to use images as guides to define structure and style, with a focus on accuracy and security for commercial use. It’s ideal for professionals, although it requires a Creative Cloud subscription, making it less accessible than Whisk.
Canva and its AI Image Generator
Canva offers a feature to upload reference images and create content inspired by them. It’s perfect for quick projects and non-technical users, such as social media posts. Although less versatile than Whisk, its integration into a well-known platform makes it very practical.
Runway.ml
Runway.ml is an advanced option that combines generative AI with personalization. Although it does not use multiple images as prompts, it allows you to transform visual content with text. It’s aimed at technical users looking for deep control, unlike Whisk’s lightweight approach.
Final Thoughts
Whisk is more than a tool; It is a window to a future where AI adapts to how we think and create. By prioritizing images over text, it breaks barriers for those who find words an obstacle in their creative process.
However, its experimental nature means that it does not replace more precise tools, but rather complements them as a space for play and exploration.
The most intriguing thing about Whisk is not just what it offers today, but what it suggests for tomorrow. As AI evolves, we could see an even greater fusion between our visual ideas and the machines that interpret them, blurring the lines between creator and technology.
For those willing to experiment, Whisk isn’t just an alternative: it’s a glimpse into the next frontier of digital creativity.
This post is also available in: