« A next-generation speech-to-text model that detects more than 85 languages, identifies different speakers, adds precise timestamps, and adapts to industry-specific vocabulary. Capable of converting raw audio into clean, formatted text, even in the presence of background noise »
« The first natively multimodal model in the GLM-5 family, capable of observing interfaces and generating code based on what it sees, with a context of 1 million tokens at a very low cost »
« Turn any PDF, Word document, or web link into a structured PowerPoint presentation. AI analyzes the content to extract key ideas. Features professional templates, automatically generated presenter notes, and PowerPoint-compatible export »
« Swap a face in a video or photo quickly and without a watermark. This AI tool also works on GIFs and group shots (up to four people tracked separately). Expressions, head movements, and lighting variations are preserved frame by frame »
« Turn your ideas into anime episodes with the help of AI. Manga, comics, and animation can be created from a storyboard. The characters, costumes, and styles remain consistent from one panel or episode to the next »
« A suite of AI music tools all in one place: generate, remix, extend tracks, add vocals or instruments, and even separate audio stems. Create complete songs from a prompt or your own lyrics »
« A single gateway that provides access to more than 200 text, image, video, and music models via a single API key. It features automatic failover in the event of an outage and does not store any requests »
« An AI-powered workspace designed for writing, presentations, image and video generation, and design. You can also take advantage of AI agents specifically designed to automate your work tasks »
« A free training platform created by Anthropic to help you learn how to collaborate with AI, write documents, analyze data, and automate tasks through conversation. No technical prerequisites required »
« A long-term coding and agent model based on GLM-5.2, featuring enhanced training, a 1-million-token context, and always-active reasoning that can be adjusted across three levels. This LLM stands out for its performance in logical reasoning and semantic understanding. »
« An all-in-one platform for generating AI-generated videos, images, and music. It also includes text-to-video and image-to-video modes, as well as a built-in editor »
« Build modular agents by freely choosing their tools, memory, interface, and runtime environment. Each component can be replaced or extended with plugins to adapt the system to different uses and workflows »