Compare

Kekoso vs Layercode CLI

CLI for building voice AI agents, does not provide local dictation or transcription UI

Side by side

What it is
Kekoso:Kekoso provides on-device dictation, audio/video transcription and call recording for macOS, running locally on the Neural Engine.
Layercode CLI:Layercode CLI lets you create, test, and deploy voice AI agents from the terminal with built-in tunneling and edge-native audio delivery.
Best for
Kekoso:Local, no-cloud dictation and transcription
Layercode CLI:CLI-driven voice AI agent development
Who it’s for
Kekoso:Mac users who need private, offline speech-to-text for apps, files, calls and AI agents.
Layercode CLI:Developers building voice AI agents who want fast local testing and zero infrastructure setup
Pricing
Kekoso:Paid
Layercode CLI:Freemium
Works with
Kekoso:Claude Code, Claude Desktop
Layercode CLI:—
DevHunt upvotes
Kekoso:11
Layercode CLI:13
Launched on DevHunt
Kekoso:Sep 2026
Layercode CLI:Oct 2025

Kekoso features

  • Hotkey dictation. Press a shortcut to dictate into any active app; text appears where the cursor is.
  • File & YouTube transcription. Drop audio/video files or paste YouTube links to get local transcripts with timestamps.
  • Call recording with speaker separation. Record calls locally, split your microphone and the other side, and export labeled transcripts.
  • Local MCP transcription server. Runs a socket server so Claude, Codex and other agents can transcribe without API keys.
  • Custom vocabulary. Add or import word lists to enforce your terminology across all transcriptions.
  • Multiple model families. Choose between Parakeet, Whisper and SenseVoice models for speed, language coverage or CJK support.

Layercode CLI features

  • One-command init. Create a voice AI agent with a single CLI command, including real-time speech-to-text and text-to-speech.
  • Built-in tunneling. Local testing via an integrated tunnel eliminates manual webhook URL copying.
  • Sample backend. Provides a starter backend that processes transcripts and generates responses based on prompts.
  • LLM-agnostic. Use any large language model while retaining full control over agent logic and tools.
  • Edge deployment. Deploys audio processing to 330+ global edge locations with ~50 ms latency.

Based on each tool's website and DevHunt data. Details may change; check the official sites.