text-to-video

Tag

Cards List
#text-to-video

MiniMax H3 (10 minute read)

TLDR AI ↗ · 2026-07-31 Cached

MiniMax launches H3, an open multimodal generation model that handles text, images, video, and audio, generating up to 15 seconds of 2K video with native stereo sound, and plans to open-source the weights.

0 favorites 0 likes
#text-to-video

Comfy-Org/MiniMax-H3

Hugging Face Models Trending ↗ · 2026-07-30 Cached

Comfy-Org repackaged MiniMax-H3 model files for ComfyUI, including diffusion models, text encoders, and VAEs, with workflow templates for text-to-video, image-to-video, and reference-to-video generation.

0 favorites 0 likes
#text-to-video

FilmBench: A Film-Grade Benchmark for Cinematic Video Generation

Hugging Face Daily Papers ↗ · 2026-07-27 Cached

FilmBench is a new benchmark for cinematic video generation, using prompts derived from award-winning films across 20 genres and a three-level cinematic taxonomy with 35+ sub-metrics. It includes an open-source automatic evaluation agent (FilmOps) that reproduces human model rankings with high correlation, revealing significant gaps in dynamic aesthetics and multi-shot performance compared to prior web-style benchmarks.

0 favorites 0 likes
#text-to-video

Flux 3

Hacker News Top ↗ · 2026-07-24 Cached

Black Forest Labs announces FLUX 3, a multimodal foundation model that jointly learns from images, videos, and audio, enabling unified generation and understanding across modalities with early access now available.

0 favorites 0 likes
#text-to-video

Lightricks/LTX-2.5

Hugging Face Models Trending ↗ · 2026-07-23 Cached

Lightricks releases LTX-2.5, an open-weights world model for generating synchronized video and audio from text, image, and video inputs, with features like native multishot generation and a new diffusion video decoder.

0 favorites 0 likes
#text-to-video

Neill Blomkamp’s new zombie AI ‘film’ is just slop warmed over

The Verge ↗ · 2026-07-21 Cached

Neill Blomkamp released a 13-minute sci-fi short 'Nightborne' made entirely with ByteDance's Seedance 2.0 text-to-video AI. The Verge criticizes it as slop, noting the AI's limitations and lack of creative voice.

0 favorites 0 likes
#text-to-video

Moving Alphabet: A Controlled Study of Training Data for Text-to-Video Generation

Hugging Face Daily Papers ↗ · 2026-07-21 Cached

This paper introduces Moving Alphabet, a procedural testbed for controlled experiments on how data distribution and caption quality affect text-to-video models, revealing key insights for data curation.

0 favorites 0 likes
#text-to-video

Japan accelerates video generation with new series of anime generation models.

Reddit r/singularity ↗ · 2026-07-20 Cached

Japan's AIdea Labs released AnimeGen, a free AI model for anime-style video generation, capable of text-to-video and image-to-video, with commercial use allowed.

0 favorites 0 likes
#text-to-video

ByteDance set to launch Seedance 2.5 with 3-minute AI video output (2 minute read)

TLDR AI ↗ · 2026-07-06 Cached

ByteDance is set to launch Seedance 2.5, an AI video model capable of generating up to 3-minute videos, with features like 30-second scenes, multimodal references, and integration into Dreamina and CapCut.

0 favorites 0 likes
#text-to-video

@Smartpigai: https://x.com/Smartpigai/status/2073753428212478384

X AI KOLs Timeline ↗ · 2026-07-05 Cached

This is a detailed tutorial on how to create Chinese self-media videos using Claude Code and the Remotion framework. Users only need to provide ideas and prompts, and the AI handles setting up the environment, writing code, generating voiceovers, and rendering the final MP4 video. The tutorial covers the complete process from pipeline setup and script writing to video generation and adjustments.

0 favorites 0 likes
#text-to-video

@seclink: Mimo Short Drama Creator Platform Backend Generated by One Sentence

X AI KOLs Following ↗ · 2026-07-05 Cached

Mimo is a creator platform that generates short dramas from a single sentence, and includes an operation backend.

0 favorites 0 likes
#text-to-video

@iluciddreaming: Build a 'Video Content Factory' on Codex – Install These Plugins to Turn Text into Videos: HyperFrames – 'Ideation', Creative Planning; ElevenLabs Skill – 'Voice', AI Dubbing; Remotion Skill – 'Motion', Procedural Animation…

X AI KOLs Timeline ↗ · 2026-07-04 Cached

On the Codex platform, using plugins like HyperFrames (ideation), ElevenLabs (AI dubbing), Remotion (procedural animation), and Captions (subtitles), you can build an automated pipeline that turns text into video.

0 favorites 0 likes
#text-to-video

Physics Question Scene Graph: Fine-grained Evaluation of Physical Plausibility in Text-to-Video Generation

Hugging Face Daily Papers ↗ · 2026-06-24 Cached

Physics Question Scene Graph (PQSG) is a hierarchical question-based pipeline using VLMs to evaluate video generation models' physical plausibility with fine-grained violation detection. It introduces the FinePhyEval dataset and shows higher correlation with human judgments than prior work.

0 favorites 0 likes
#text-to-video

DomainShuttle: Freeform Open Domain Subject-driven Text-to-video Generation

Hugging Face Daily Papers ↗ · 2026-06-24 Cached

DomainShuttle introduces a method for open domain subject-driven text-to-video generation, achieving high fidelity and flexibility across in-domain and cross-domain scenarios using domain-aware modeling and dual RoPE schemes.

0 favorites 0 likes
#text-to-video

@rohanpaul_ai: AI video is moving into its real-time reaction era, with MaineCoon now leading in low-latency AI video. @catnips_ai jus…

X AI KOLs Following ↗ · 2026-06-23 Cached

MaineCoon is a 22B real-time text-to-audio-video model that achieves up to 47.5 FPS on a single H100 GPU, enabling low-cost, long-duration streaming with synchronized speech and visuals for live AI characters.

0 favorites 0 likes
#text-to-video

TapVid

Product Hunt ↗ · 2026-06-22

TapVid is a tool that allows users to turn any idea into a motion video.

0 favorites 0 likes
#text-to-video

Pixlie

Product Hunt ↗ · 2026-06-19

Pixlie is an AI video studio that enables text and image to video conversion with real control.

0 favorites 0 likes
#text-to-video

LooseControlVideo: Directorial Video Control using Spatial Blocking

Hugging Face Daily Papers ↗ · 2026-06-17 Cached

LooseControlVideo introduces a framework for intuitive 3D spatial control in text-to-video generation using sparse oriented 3D boxes as proxies, achieving superior trajectory accuracy and occlusion handling. It fine-tunes a Wan 2.2 backbone and demonstrates significant improvements over existing methods on multiple benchmarks.

0 favorites 0 likes
#text-to-video

@laowangbabababa: Shocked! Dr. Qi on Douyin sells a 500k digital human agent per day, and I built it in 2 minutes. Using the Pixelle-Video project, which already has 22k stars. It supports digital human lip-syncing, motion transfer, and image-to-video. Supports ComfyUI, input a topic, from script writing to adding...

X AI KOLs Timeline ↗ · 2026-06-13 Cached

Introducing the open-source project Pixelle-Video: a fully automated AI short video engine. Input a topic and it automatically generates a video with script, images, voiceover, and background music. Supports local and cloud models, modular design allows flexible replacement of each component model.

0 favorites 0 likes
#text-to-video

Avatars in ElevenCreative

Product Hunt ↗ · 2026-06-12

ElevenLabs introduces Avatars in ElevenCreative, a dedicated entry point for generating talking-head videos, enabling users to create AI-driven avatar videos with realistic speech and lip-sync.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback