Skip to content
Home

Kling O1

Generates and edits videos from text, images, and reference clips, with features like start/end frame generation, restyling, and camera extension.

A unified multimodal video generation model that handles text-to-video, image-to-video, reference-to-video, and editing tasks in one workflow, solving the problem of fragmented video creation tools. It serves marketers, content creators, and film professionals in advertising, fashion, and post-production, targeting both B2B and B2C users with a free tier. The model differentiates by accepting multiple input types (images, clips, text) simultaneously and maintaining character and scene consistency across shots, with shot lengths from 3 to 10 seconds.

Key features

  • Text-to-video generation
  • Image-to-video generation
  • Reference-to-video generation
  • Start and end frame generation
  • Content editing and transformations
  • Restyling and camera extension
  • Multimodal input understanding
  • Character and scene consistency
  • 3-10 second shot length control
  • Element-based controls
Social posts
  • No social media activity detected
GTM channels
  • API
  • Docs
ICP
  • Content creators
  • Marketing teams
  • Agencies