-
1
Chirp 3
Google
Create unique voices effortlessly with advanced audio synthesis technology.
Google Cloud has introduced Chirp 3 within its Text-to-Speech API, enabling users to create personalized voice models using their own high-quality audio samples. This advancement simplifies the creation of distinctive voices for audio synthesis through the Cloud Text-to-Speech API, making it suitable for both streaming content and extensive text applications. However, due to security measures, this feature is currently available only to a limited group of users, who must contact the sales team to be considered for access. The Instant Custom Voice functionality accommodates various languages, including English (US), Spanish (US), and French (Canada), which broadens its usability. Additionally, this service functions across multiple Google Cloud regions and supports an array of output formats such as LINEAR16, OGG_OPUS, PCM, ALAW, MULAW, and MP3, depending on the selected API method. As advancements in voice technology progress, the potential for tailored audio experiences continues to grow, offering exciting opportunities for innovation in communication and entertainment. This evolution not only enhances creativity but also fosters deeper connections between content creators and their audiences.
-
2
Lyria
Google
Transform words into captivating soundtracks for every project.
Lyria is an advanced text-to-music model that transforms text descriptions into fully composed, high-quality music tracks. Whether you're crafting soundtracks for a marketing campaign, enhancing video content, or creating immersive brand experiences, Lyria delivers music that reflects your desired tone and energy. With its ability to generate diverse musical styles and compositions, Lyria offers businesses an efficient and creative solution to enhance their media production. By leveraging Lyria, companies can significantly reduce the time and costs associated with finding and licensing music.
-
3
Imagen 4
Google
Unleash creativity with stunning, rapid, photorealistic images!
Imagen 4 represents the cutting edge of image generation technology, combining photorealism with powerful creative features to produce high-quality images. This model allows users to generate realistic visuals with breathtaking detail, from the texture of surfaces to accurate lighting and typography. Whether you’re looking to create landscapes, portraits, or more abstract concepts, Imagen 4 offers the tools to render a wide variety of artistic styles with impressive precision. Notably, it enhances the sharpness of generated images, producing crisp and accurate results that surpass previous versions. Users can now benefit from an ultra-fast mode, enabling them to generate multiple images in a fraction of the time it took before—up to 10x faster. Imagen 4 supports 2K resolution, delivering exceptional clarity that’s perfect for both large-scale prints and digital media. It also features improvements in color rendering, with more vivid and accurate tones, making it ideal for artists, designers, and marketers. With the ability to generate complex compositions with minimal effort, Imagen 4 is a powerful tool for professionals across a wide range of industries.
-
4
Lyria 3
Google
Unleash your creativity with AI-driven music innovation.
Lyria 3 represents Google DeepMind’s most advanced step forward in AI-powered music generation, offering creators the ability to produce professional-quality audio using natural language prompts. Designed to understand musicality at a structural level, it captures rhythm, harmony, arrangement, and vocal nuance to create tracks that feel cohesive and intentional. Users can start with a simple idea, such as a mood or theme, and progressively refine technical elements like tempo, genre, instrumentation, and vocal style. The model supports multilingual vocals and spans a broad spectrum of global genres, enabling experimentation across cultural and stylistic boundaries. A unique feature allows users to upload images and transform them into custom musical compositions, blending visual inspiration with sonic creativity. Lyria 3 was developed with feedback from musicians and producers to ensure outputs reflect authentic musical flow rather than fragmented loops. Tracks can be exported in high-fidelity formats suitable for background scoring, digital content, or large-scale performance use. The model family also includes real-time and open creative variants, expanding options for interactive and experimental workflows. To promote responsible AI development, Lyria 3 incorporates robust content filtering and imperceptible SynthID watermarking to identify AI-generated audio. While powerful, the system acknowledges ongoing improvements and encourages creators to review outputs carefully. Integrated into Gemini and YouTube Shorts through Dream Track, Lyria 3 fits seamlessly into modern creative ecosystems. Overall, it functions as a collaborative creative partner, helping artists, creators, and storytellers explore new musical possibilities while maintaining control over their artistic vision.
-
5
GPT-5.4
OpenAI
Elevate productivity with advanced reasoning and seamless workflows.
GPT-5.4 is a frontier artificial intelligence model developed by OpenAI to perform complex reasoning, coding, and knowledge-based tasks. It is designed to support professionals across industries by helping them automate workflows, analyze information, and produce detailed work outputs. The model integrates advanced reasoning capabilities with powerful coding performance derived from earlier Codex systems. GPT-5.4 can generate and edit documents, spreadsheets, presentations, and structured data used in business operations. One of its major improvements is its ability to interact with tools and external systems to complete multi-step workflows across different applications. This capability allows AI agents built on GPT-5.4 to perform tasks such as data entry, research, and automated software interactions. The model also supports extremely large context windows, enabling it to process long documents and maintain awareness across extended tasks. Improved visual understanding allows GPT-5.4 to interpret images, screenshots, and complex documents more effectively. It also introduces better web browsing and research capabilities for locating and synthesizing information online. Compared with previous versions, GPT-5.4 reduces factual errors and produces more consistent responses. Developers can access the model through APIs and integrate it into software applications, automation systems, and enterprise workflows. Overall, GPT-5.4 represents a significant step forward in AI capabilities for knowledge work, software development, and intelligent automation.
-
6
Lyria 3 Pro
Google
Create dynamic, high-quality music effortlessly with advanced AI.
Lyria 3 Pro is a cutting-edge AI music generation model created by Google DeepMind, designed to produce longer, more structured, and highly customizable music tracks for a wide range of users. It allows users to generate compositions up to three minutes in length, offering detailed control over musical elements such as intros, verses, choruses, bridges, and transitions. The model’s enhanced understanding of musical structure ensures that outputs are cohesive, dynamic, and professionally arranged. Lyria 3 Pro is integrated into multiple Google platforms, including Gemini Enterprise Agent Platform for enterprise-scale applications, Google AI Studio for developers, and the Gemini app for creators. It is also available in tools like Google Vids and ProducerAI, enabling seamless integration into video production and collaborative music workflows. The platform supports diverse use cases, from creating soundtracks for games and videos to generating personalized music for content creators. Its scalability allows businesses to produce high-quality audio content efficiently and at scale. Lyria 3 Pro is built with a strong focus on responsible AI, ensuring that it does not replicate specific artists while still allowing stylistic inspiration. It includes built-in safeguards, such as content filters and SynthID watermarking, to protect intellectual property and identify AI-generated content. The model is designed to enhance creativity by allowing users to experiment with different musical styles and structures effortlessly. It also helps streamline production workflows by reducing the time and effort required to compose original music. Overall, Lyria 3 Pro represents a significant advancement in AI-driven music creation, enabling users to bring their creative ideas to life with greater flexibility and precision.