Amazon has introduced its new family of foundational models called Amazon Nova. Designed to deliver cutting-edge AI at a low cost, these models are available exclusively through Amazon Bedrock, the AI ​​services platform managed by AWS.

With multimodal capabilities spanning text, images and video, Amazon Nova not only seeks to compete with other solutions on the market, but also lead a new era in content creation, application personalization and advanced data analysis.

Let’s explore the key features of Amazon Nova, its main models and its applications in industries such as advertising, music, technology and media.

A family of models designed for every need

Amazon Nova includes a range of models to suit different use cases and levels of complexity:

Understanding models: Micro, Lite and Pro

  • Amazon Nova Micro: Specialized in text, it stands out for its speed and low cost. It is ideal for tasks such as language comprehension, translation, mathematical reasoning, and code generation. With a speed of more than 200 tokens per second, it is optimized for applications that require immediate responses (low latency).
  • Amazon Nova Lite: This multimodal model processes text, images and videos with impressive speed, being suitable for high-volume interactive applications where cost is a key factor.
  • Amazon Nova Pro: The most advanced model in this category, it combines precision, speed and cost efficiency. Its applications include video summarization, Q&A, software development, and complex workflows in AI agents. It also excels in reasoning and content generation benchmarks.

Creative models: Canvas and Reel

  • ·Amazon Nova Canvas: Specialized in image generation, it allows you to create professional quality visual content from text or images. It also offers editing tools and custom settings to suit users’ specific needs.
  • Amazon Nova Reel: This model focuses on creating high-quality videos from textual or visual input. Among its notable features is the possibility of controlling the visual style and camera movements using natural language commands.

Innovation and customization through Amazon Bedrock

One of the most notable aspects of Amazon Nova is its integration with Amazon Bedrock, a platform that allows customers to access and use AI models through a unified API.

With this integration, companies can experiment, adjust, and deploy foundational models quickly and easily.

Personalization is one of the pillars of this proposal. Amazon Nova models support both fine-tuning and distillation.

This means that companies can train models with first-party data to improve accuracy on specific tasks ortransfer knowledge from a large model to a more efficient one.

Additionally, the models are designed for RAG (Recovery Augmented Generation), which allows them to provide accurate and contextualized answers based on organizations’ own data.

This is especially useful in business applications that require a deep and contextualized understanding of information.

Use cases in various industries

For a few months now, Amazon Nova has been available to some companies, where it is showing its impact in a variety of sectors, from media to music and advertising.

Advertising and visual creativity

Companies such as Dentsu Digital and Shutterstock have highlighted the advantages of Canvas and Reel models in creating visual content. According to Dentsu, Nova Reel has transformed their creative processes by allowing them to generate high-quality videos in a matter of days, instead of weeks.

For its part, Shutterstock points out that Canvas significantly increases the quality of the images generated, facilitating a more intuitive experience for users.

Media and data processing

Hearst Corporation, a media giant, is using Nova Pro to summarize videos and analyze documents with astonishing accuracy.

These capabilities not only improve internal workflows, but also offer new opportunities to personalize experiences for subscribers.

Music and audiovisual content

In the music sector, Musixmatch is using Nova Canvas and Reel to democratize music video creation.

Emerging artists can now generate high-quality videos using their own songs as a base, something that previously required significant resources.

Technology and logistics

Companies like Palantir Technologies and Caylent are leveraging Amazon Nova models to optimize complex processes like supply chain management and video analytics.

Randall Hunt, CTO of Caylent, has praised the simplicity and effectiveness of Nova, describing its integration as “magical” for its ability to deliver cutting-edge results without the need for complex techniques.

A step towards the future: More models arrive in 2025

Amazon has already announced that in 2025 it will launch two new models that will further expand the capabilities of Nova:

A voice-to-speech model: Capable of interpreting natural spoken language, including nuances such as tone and cadence, to generate more human and natural interactions.

A multimodal-to-multimodal (“any-to-any”) model: This model will be able to process text, images, audio and video as inputs and outputs, simplifying complex applications that require translating or transforming content between different modalities.

Amazon seeks to improve the price per token

The presentation of Amazon Nova is a significant advance for the company, which seemed stagnant. Since Amazon introduced the Titan models a couple of years ago, we had had no news about its internal developments.

The strategy appears to be to offer solutions at a reduced price per token. This strategy positions Amazon as the option for companies looking for economical and scalable models without compromising advanced functionalities.

This development is crucial for a company like Amazon, which already masters the manufacturing of clusters with its own chips and is investing heavily in Anthropic to stay at the forefront of the development of generative AIs.

This post is also available in: Español Français Русский Italiano