Skip to content

Latest commit

 

History

History
41 lines (27 loc) · 4.93 KB

File metadata and controls

41 lines (27 loc) · 4.93 KB

Get started with Cutawan

Cutawan is a desktop app. Download an installer, choose how to handle speech and AI on first launch, then import a video. You do not need to clone the repository or install Node.js.

Install the app

Download the current build from GitHub Releases. Expand Assets and choose the file for your computer:

Platform Choose First launch
Windows .exe installer Run the installer. The build is not code-signed yet, so Windows may show a SmartScreen warning; verify that you downloaded it from this repository before proceeding.
macOS, Apple Silicon (M1 or newer) arm64.dmg Move Cutawan to Applications. The build is not signed or notarised yet. After trying to open it, macOS may require System Settings → Privacy & Security → Open Anyway. Intel Macs do not currently have a published installer.
Linux .AppImage Make it executable, then run it. For example, in a terminal: chmod +x Cutawan*.AppImage and ./Cutawan*.AppImage.

Signing and notarisation are planned. Until then, only download installers from the official release page, not a reposted copy. Builds before v0.10.0 do not have the first-run wizard; use Settings → General → AI connection for setup instead.

Mac updates: The toolbar notifies you when a release is available. Click it to open Updates, then choose Download Mac installer. The app shows progress, verifies the download, and offers Open installer. Quit Cutawan, then replace the app in Applications. Your projects and settings are stored separately and are preserved. The current unsigned Mac updater requires this final replacement step. Update checks run on launch and every six hours; failed checks retry automatically.

Choose a connection

The first-run wizard offers four paths. You can switch later in Settings → General → AI connection. Explore without setup lets you look around, but processing a new video still needs one of the routes below.

Route What it does What you need
ChatGPT sign-in (beta) Uses Codex for clip finding and local Whisper for transcription. No OpenAI API key or automatic paid-API fallback. An installed Codex CLI signed in with ChatGPT, Python 3.10+, and a local Whisper model. Your plan's Codex limits apply. Detailed setup.
OpenAI-compatible API Uses your configured API endpoint for analysis. Speech can go to the transcription API or run locally with Whisper. An API key and separately billed provider usage. A local-speech choice also needs Python 3.10+ and a model.
OpenRouter One OpenRouter key for clip finding with any chat model OpenRouter lists (OpenAI, Anthropic, Google and others), picked from a searchable list with suggestions at the top. Speech goes to an OpenRouter-hosted Whisper model or runs locally. An OpenRouter key and OpenRouter credits. A local-speech choice also needs Python 3.10+ and a model.
Local captions only Transcribes and captions a whole video without sending it to an AI service. It does not find or score short clips. Python 3.10+ and a local Whisper model.

For either local-speech route, install Python separately first. The wizard can then create a private environment and download faster-whisper plus a Small or Large v3 model into Cutawan's app-data folder after you request it. Small is the lighter download; Large v3 needs several gigabytes. The setup check does not make an AI model request. CPU transcription works but can be slow.

Make your first video

  1. Drag a local video onto the home screen, click Drag a video here, or click to choose, or paste a supported video URL and click Import.
  2. Pick Find viral clips for AI-suggested short moments or Caption whole video for one continuous captioned edit. The local-only route supports the latter, not AI clip finding.
  3. For clips, click Get clips and review the suggestions. For a whole video, click Transcribe and caption. Processing time depends on video length and your computer or provider.
  4. Open the result, adjust the trim, framing and caption style, then export an MP4. Review the finished file before posting; automatic selections and captions can need correction.

The original video stays on your computer. Rendering, face tracking and export are local. On the API route, extracted audio may be sent for transcription, and transcript text and sampled frames are sent for analysis. With ChatGPT/Codex, speech stays local but transcript text and sampled frames are sent for analysis. The local-only route does not run AI clip analysis.

If you hit a problem, open a bug report with your Cutawan version, operating system and what happened. Please do not include API keys or private footage in an issue.