Video Generation 26

AI Minecraft is an AI that generates playable Minecraft versions from user input. AI Minecraft, also called Oasis, is a new AI trained on Minecraft gameplay. It allows users to play a completely AI-generated version of Minecraft. It utilizes a 3D spatiotemporal joint attention mechanism, better modeling complex spatiotemporal motions to generate video content with significant motion while adhering to motion rules. It can generate videos up to 2 minutes long with a frame rate of 30fps. Based on proprietary model architecture and the strong modeling capabilities inspired by Scaling Law, AI Minecraft can simulate real-world physical characteristics, generating videos that conform to physical laws. With a deep understanding of text-to-video semantics and the powerful Diffusion Transformer architecture, AI Minecraft can transform users' rich imagination into concrete images, creating scenarios that do not exist in the real world. Based on proprietary 3D VAE, AI Minecraft can generate cinema-grade videos with 1080p resolution, vividly presenting everything from vast and magnificent scenes to detailed close-ups. AI Minecraft adopts a variable resolution training strategy, allowing for the output of the same content in various video aspect ratios during inference, meeting the needs of more diverse video material usage scenarios.
All-in-One AI Video, Image Creation Platform Monet AI delivers a true All-in-One experience that makes everyone a visual art master. Say goodbye to platform switching, enjoy true All-in-One experience.
All-in-one AI platform for image to video generation. VidFlux.ai is an all-in-one AI platform designed to transform images into stunning, dynamic videos. It provides access to 6+ industry-leading AI models, including Veo 3, Sora 2, Kling AI, Runway, Seedance, Wan AI, and Grok Imagine, all within a single unified platform. This eliminates the need for users to switch between multiple tools, allowing for effortless creation of professional image-to-video content. Users can generate creative videos, optionally with audio, from their uploaded photos, producing both realistic scenes and animations quickly and precisely.
Single API access to 200+ AI models, 80% lower costs than OpenAI, seamless integration. AI/ML API revolutionizes tech by offering developers access to over 100 AI models via a single API, ensuring round-the-clock innovation. Offering GPT-4 level performance at 80% lower costs, and seamless OpenAI compatibility for easy transitions. AI/ML API Tokens offer the flexibility to precisely allocate resources, enhancing performance and cost efficiency.
AI Image/Video API platform for rapid generation. Muapi is an AI Image/Video API Platform designed to accelerate AI image and video generation. It offers a comprehensive suite of cutting-edge AI models to transform creative visions into stunning images and videos in seconds with professional quality results.
Runable is a general AI agent that can execute any task, from building web apps, slides, reports, and documents to generating images, videos, and podcasts, all in one place. It doesn’t just create it connects. Runable integrates with thousands of your favorite apps so you can simply ask it to do the work for you.
Platform with 30+ open-source LLMs for chat, image, and video generation with unlimited usage. TavonnAI is a platform that provides access to over 30 open-source Large Language Models (LLMs) for chat, image, and video generation. It offers unlimited usage for a monthly subscription fee, making it a playground for AI enthusiasts, creators, and innovators to explore open-source artificial intelligence.
API for automated image and video generation from templates. Bannerbear is an API-driven platform that helps individuals and teams automatically generate social media visuals, e-commerce banners, podcast videos, and more. It turns graphic templates into an API, allowing for the automated creation of variations. It offers APIs for image, multi-image, video, and PDF generation, along with integrations for popular tools like Zapier, Airtable, Make.com, Forms, and WordPress, enabling no-code automated workflows for repetitive marketing tasks.
AI Image to Video Generator powered by VEO3, SORA2. KOOX AI Image To Video AI Generator is an AI-powered platform designed to convert static images into dynamic videos. The tool leverages advanced AI model 'SORA 2' by OpenAI to generate professional videos from images with text prompts. The tool supports JPEG, PNG, GIF, and WebP file formats. Users input their image and add a text prompt describing the scene which the AI uses to generate a video. There are customization settings for video duration, quality, and aspect ratio. The AI ensures consistency, fidelity, and dynamic motion, maintaining elements true to the original image and produces convincing movement without sacrificing detail or context. Additionally, users can leverage this tool to generate matching audio, including custom music and sound effects for their videos, providing a complete and immersive storytelling experience. The generator can be useful for a broad audience including influencers, content creators, product marketers, and artists. Overall, KOOX AI Image To Video AI Generator offers a comprehensive solution for transforming images into captivating video stories.
A resource hub for generative AI models, tools, prompt engineering, guides, and tutorials. FraxAI is a website dedicated to generative AI, offering models, tools, prompt engineering resources, guides, and tutorials for Stable Diffusion, ChatGPT, and other AI technologies. It provides a comprehensive collection of resources for users interested in exploring and utilizing generative AI.
AI-powered lip sync video generator that turns voice or text into realistic talking videos in seconds. Lip Sync AI is an advanced AI video generation platform that transforms text, voice, or audio into highly realistic talking videos with accurate lip synchronization. It is designed for creators, marketers, educators, and developers who need fast and scalable video production without traditional filming or editing. The platform uses AI-driven facial animation and audio analysis to generate natural lip movements, facial expressions, and timing alignment. Users can create professional-quality talking videos in seconds, significantly reducing production time and cost. Lip Sync AI supports multiple use cases, including content creation for social media, marketing videos, educational explainers, product presentations, and AI-powered dubbing. It is built to deliver high efficiency, consistent output quality, and ease of use for both beginners and professionals.
Browser-based ComfyUI for easy workflow discovery, building, and running with zero setup. Floyo is a browser-based platform that brings the full power of ComfyUI to users without any setup or installation. It allows users to easily find, launch, and build open-source ComfyUI workflows, offering creative freedom and faster, easier sharing. It supports custom nodes and models and provides various pre-made workflows for different use cases like image and video generation, LoRA training, and more. Floyo aims to make ComfyUI accessible and efficient for creators and businesses.
Cloud-based ComfyUI platform for AI art creation with fast GPUs and easy workflows. RunComfy is a premier cloud-based ComfyUI platform designed for stable diffusion, empowering AI art creation with high-speed GPUs and efficient workflows without the need for technical setup. It offers a native ComfyUI experience, allowing users to seamlessly transition between local and cloud environments. RunComfy provides tools and resources to focus on art creation, including easy model downloads, node installation, and reproducible workflow environments.
MiniMax is an AI company offering text, speech, and video generation models via API. MiniMax is a leading global technology company and one of the pioneers of large language models (LLMs) in Asia. They offer a range of AI models and capabilities, including text, speech, and video generation, through their API platform. Their mission is to build a world where intelligence thrives with everyone.
AI engine for cinema-grade audiovisual production. MuseSteamer AI is a revolutionary multimedia intelligence engine that leverages groundbreaking computational creativity for professional-grade audiovisual production. It achieves 89.38% VBench performance metrics and transforms concepts and visuals into premium content, enabling worldwide creative collaboration. The platform utilizes sophisticated algorithms to produce cinema-grade output, converting textual stories or visual assets into premium audiovisual experiences with harmonized sound design. It revolutionizes digital storytelling by offering a pioneering multimedia intelligence framework that delivers professional audiovisual narratives with coordinated sound landscapes.
AI app that transforms kids' drawings into vibrant artworks and animated videos. Drawings Alive is an AI-powered application that transforms children's drawings into vibrant artworks and animated videos. It allows users to upload drawings, add descriptions, and watch as the AI brings them to life, creating fun and magical experiences for kids and families.
AI tool to animate static images into stunning videos. AI Animate Image is an advanced online tool that leverages cutting-edge AI video models like Veo 3, Kling, and Runway to transform static images into stunning, lifelike animations. It provides a browser-based platform for content creators, marketers, and hobbyists to generate professional-quality animated content with ease, offering features like intelligent analysis of photos, natural motion effects, and high-resolution output. The tool aims to revolutionize how animated content is created from static images, delivering professional results without requiring technical expertise.
Stability AI develops open-source AI models for image, video, 3D, and audio generation. Stability AI is a company that develops cutting-edge open models in image, video, 3D, and audio generation. Their flagship product, Stable Diffusion, is a deep learning, text-to-image model used to generate detailed images conditioned on text descriptions. It can also be applied to other tasks such as inpainting, outpainting, and generating image-to-image translations guided by a text prompt. Stability AI offers various tools and platforms for deploying and utilizing their models, including self-hosted licenses, a platform API, and cloud platform integrations.
AI-powered creative workspace for designing workflows across all mediums. Fuser is an AI-powered creative workspace designed for professionals to design and manage workflows across any medium (text, images, videos, audio, 3D) all on a single canvas. It integrates with a vast array of AI models and LLMs, allowing tools to work together and ideas to evolve. Fuser emphasizes exploration and iteration, providing tailored workflows and templates for various creative modalities, aiming to be the central hub where creative flow meets creative control.
Unified AI hub for text, image, video, and audio generation. GPTunneL is a unified AI hub providing official access to a wide range of leading neural networks, including ChatGPT, Claude, Gemini, MidJourney, Stable Diffusion, DALL-E, and more, in Russia and in Russian. It allows users to generate text, images, audio, and video content, as well as music, all within a single interface. The service operates on a pay-as-you-go model, meaning users only pay for their actual usage without subscriptions or auto-payments. It also offers various built-in AI tools like voice synthesis, image editing, face swapping, background removal, sticker generation, audio/video transcription, code generation, and AI assistants. API functionality and payment options via bank cards or cryptocurrency are available.
All-in-one AI platform for SEO content, image, video, and 500+ AI tools. i10X is an all-in-one AI platform that provides access to various advanced AI models such as ChatGPT, Claude, Gemini, and Perplexity, along with over 500 specialized AI tools. Its core offering, the AI Content Creator, leverages proprietary RankeSense Technology to generate SEO-ready, competitor-beating articles. This technology analyzes top-ranking Google results to understand ranking factors and then creates content designed to outperform them, optimized for GPT and all LLMs, and ready for immediate publication. Beyond content creation, i10X supports tasks like image and video generation, PDF and document management, marketing, business planning, and more.
Powerful, modular, open-source visual AI for generating video, images, 3D, audio. ComfyUI is the most powerful and modular visual AI application and engine, serving as an open-source node-based platform for generative AI. It enables users to generate video, images, 3D, and audio using AI. The platform offers full control over AI workflows through a visual node-based canvas, allowing for branching, remixing, and real-time adjustments. Workflows are reusable, with exported files carrying metadata for easy reconstruction. Comfy Cloud, a related product, provides instant, hardware-free creative tools and custom solutions for design studios and production houses, offering access to powerful creation on demand with ready-to-use models and high-performance server GPUs.
Unified generative AI platform for creative teams, featuring 50+ models and scalable workflows. FLORA is an Intelligent Canvas and generative media platform designed for creative professionals and teams. It unifies every creative AI tool—including access to 50+ state-of-the-art multimodal models (like GPT-5, Imagen 4, and Veo 3)—into one unified process. FLORA enables users to accelerate creation from ideation to production by providing scalable workflows, real-time collaboration features (with unlimited seats), advanced editing tools (Inpaint, outpaint, crop), and predictable credit-based pricing where unused credits roll over.
Unified AI gateway and platform providing free multimodal AI APIs and applications. Agnes AI by Sapiens AI is a comprehensive AI gateway, free AI API platform, and AI application ecosystem featuring flagship models like Agnes, Echo, and Pavo. It provides developers and users with access to full-stack, scalable generative AI capabilities including text, image, and video generation, chatbots, AI agents, and intelligent workflows through a single, unified platform.