Skip to content
Home

Edit Mind

Video knowledge base that transcribes audio, analyzes frames, and makes every scene, object, and word searchable to locate moments in footage

Launch screenshots from Product Hunt, May 2026

The product is a video knowledge base that makes footage searchable by transcribing spoken words, detecting objects and faces, reading on-screen text, and describing scenes to solve the problem of locating moments across large libraries of raw files. It is for video editors, designers, and individual creators as well as business users who manage extensive footage collections. It is delivered as a local desktop application and as a self-hosted Docker option that keeps all data on the user's machine.

Key features

  • Timecoded transcription of spoken words
  • Object detection in frames
  • Face recognition in frames
  • On-screen text reading (OCR)
  • Natural language scene description
  • Multi-modal vector search
  • Natural language footage search
  • Jump to exact frame
  • Download clip or ZIP
  • One-click send to timeline
  • Reddit<0.1/day
GTM channels
  • Blog
ICP
  • Content creators
  • Freelancers solopreneurs
  • Design teams