I’ve been experimenting with AI video generation recently, and I’m honestly surprised by how quickly the technology is improving. The videos are becoming more cinematic and realistic, and tools today can already create scenes that would have taken much more time and resources in the past. But after creating a few videos myself, I realized that making something look realistic is only one part of the challenge.
Character consistency is still difficult, especially across multiple scenes. Sometimes the movement or physics can feel unnatural, and getting exactly the camera angle or action you imagine isn’t always easy either. So I’m curious about what the community thinks. What do you think is the biggest challenge AI video generation needs to solve next? Is it longer video consistency, better motion and physics, more creative control, or something else? I’d love to hear your thoughts, especially from people who are actively experimenting with video models.