Android · beta now iOS · in development Runs offline

SID3KICK

Your AI sidekick for everything you say out loud

Point it at a meeting, a conversation, or nothing in particular. SID3KICK records, transcribes, summarizes, pulls out the calendar events and to-dos you mentioned, and rolls the whole day into a digest you can actually search later. On your device by default — the offline engine never sends a byte anywhere.

$5/month at launch  ·  plus your own AI API usage  ·  $0 if you run it fully offline

Speech-to-text
~20offline languages, on-device via Vosk
Accounts required
Zerono login, no server we operate
AI providers
YoursOpenAI, Anthropic, Ollama, LM Studio
Cost tracking
Built inevery call logged to the cent
What it does

A dictation app and a life-logger, fused

Tap to record when you know you want a transcript. Or leave it running and let it work out for itself when someone's actually talking.

One-tap capture

Start recording and the pipeline takes over — transcribe, summarize, title, and scan for anything that sounds like a date or an appointment. An optional floating bubble sits over other apps so you never have to open SID3KICK to hit record.

Passive life-assist mode

A foreground service listens with the screen off and the phone in your pocket. Voice activity detection slices continuous audio into real speech, so you're not paying to transcribe silence — and each chunk runs the same pipeline.

Genuinely multi-language

It transcribes what's actually being spoken, not what your phone's locale expects. Then pin one output language and every title, summary, and digest comes back in it — which also inoculates you against cloud models hallucinating a wrong-language transcript on noisy audio.

Events that add themselves

Dates, times, and appointments get pulled out of what you said and surfaced as editable one-tap calendar suggestions. They write straight to your device's own calendar — no Google Cloud project, no OAuth screen, no API key.

A second brain you can ask

"What did I say I'd send Priya?" "When was that dentist thing?" The Ask tab answers across every transcript at once, and because each capture is tagged with when — and optionally where — it happened, "yesterday" resolves against your real history instead of being guessed at.

Action items, tracked

To-dos get pulled out of what you actually said, land on a tickable list, show up in the daily digest, and — if you paste a token — push one-way to your Google Tasks default list.

Where you were, if you want

Location tagging is off, country-only, or precise — your call. Country-only is the default: it reads the last-known fix with no GPS wake-up, keeps the place name, and blurs the coordinates to roughly a kilometre before storing anything.

Sync with no account

Keep transcripts, summaries, to-dos, and digests in step across your devices through your Dropbox, Google Drive, or FTP server — never through anything we run. Each device writes its own snapshot and merges the others.

Three themes, one of them silly

Dark and Light are ordinary Material 3. LCARS is not: black panels, asymmetric corners, condensed uppercase type, and synthesized interaction beeps generated at runtime — no ripped audio, no licensing headache.

How it works

Four steps, and you choose where each one runs

Every stage is either on your phone or on a provider you hold the key to. There is no middle tier operated by us.

Capture

16 kHz mono audio, manual or voice-activated. Snap photos or short video alongside it — with video, only the audio is ever extracted and sent; the file itself never leaves the device.

Transcribe

Vosk on-device across ~20 downloadable language models, with zero network calls and zero cost. Or route to OpenAI's cloud transcription when you want the accuracy and don't mind the upload.

Summarize

Title, summary, extracted events, action items, and optional speaker diarization — through OpenAI, Anthropic, a local Ollama or LM Studio box, or any endpoint that speaks the OpenAI chat-completions schema.

Recall

Day-by-day digests, a map of where you were, describe-and-recall search for half-remembered conversations, and a chat box that answers across the lot.

Privacy

The offline path is a real one

Plenty of apps say "privacy-first" and mean "we encrypt it on the way to our servers." SID3KICK has a configuration where there are no servers.

Fully offline is a supported mode

On-device Vosk for speech-to-text plus a local Ollama or LM Studio box for summaries, and nothing leaves your network. Not a degraded fallback — a first-class setup.

Your keys, your provider, your bill

When you do use the cloud, you paste your own API key and talk to that provider directly. We are not a proxy, we never see the traffic, and there is no account to create.

Serverless multi-device sync

Sync runs through storage you already own. Raw audio, photos, and video stay on the device that captured them; optional cloud archive and auto-prune only ever delete audio that's already been backed up.

Pricing

$5 a month, plus whatever your AI usage costs

Two separate bills, and we're upfront about both. The subscription is ours. The AI usage is billed by whichever provider you point the app at — or it's nothing at all, because you're running the models on your own hardware.

The app

$5 / month

One subscription, every feature, both platforms as they ship.

  • Manual capture, passive mode, and voice activity detection
  • Daily digests, action items, calendar extraction, second-brain search
  • Multi-device sync through your own storage
  • All three themes, ~20 offline language models
  • No usage caps on anything running on-device

Your AI usage

$0 + / month, at cost

Paid straight to your provider, never marked up by us. Run everything on-device and this line is genuinely $0.

  • $0 — Vosk on-device transcription and a local Ollama or LM Studio model
  • Cents per hour — cloud transcription plus a small hosted model
  • Your call — swap in a frontier model whenever accuracy matters more than cost
  • Every single call is logged with provider, model, tokens, and computed USD, against a price table you can edit
Estimated AI cost per hour of recorded speech
Setup Transcription Summaries Per hour
Fully offline Vosk, on-device Ollama / LM Studio, local $0.00
Hybrid — the sweet spot Vosk, on-device Claude Haiku 4.5 ~$0.02
Cheapest all-cloud gpt-4o-mini-transcribe GPT-5.4 nano ~$0.18
Accuracy-first whisper-1 Claude Sonnet 5 ~$0.41

How these are worked out: one hour of conversational speech is roughly 9,000 words, or about 12,000 tokens of transcript, plus around 600 tokens of generated title, summary, and extracted events. Transcription is billed per audio minute (gpt-4o-mini-transcribe $0.003/min, whisper-1 $0.006/min); summaries are billed per token at each provider's published rate. These are estimates, not quotes — your real numbers depend on how densely people actually talk and on provider pricing at the time, which is exactly why the in-app price table is editable.

What a realistic month looks like
How you'd use it Setup AI usage Total / month
Any amount, fully offline Vosk + local model $0.00 $5.00
~3 h/day of talk, hybrid Vosk + Claude Haiku 4.5 ~$1.50 ~$6.50
~1 h/day of meetings Cloud transcription + GPT-5.4 nano ~$5.50 ~$10.50
Passive all day, ~2 h of actual speech Cloud transcription + GPT-5.4 nano ~$11.00 ~$16.00
Heavy — ~3 h/day, all cloud Cloud transcription + GPT-5.4 nano ~$16.50 ~$21.50

Notice the pattern: transcription is the expensive part, and it's the part you can move on-device for free. That's why the hybrid row is a rounding error — Vosk does the heavy lifting locally and a small hosted model handles the writing. Note too that passive mode bills for detected speech, not for hours of runtime: voice activity detection means eight hours in your pocket is nothing like eight hours of billed audio.

Let's move faster

We're looking for partners, influencers & investors

The app works. The Android build is in testers' hands and the iOS port is underway. What gets this to a lot more people, a lot sooner, is the right people alongside it.

Partners

Integrations, distribution, and bundles. If you build for people who live in meetings, field work, journalism, accessibility, research, or healthcare admin — there's an obvious fit here, and an app deliberately built to talk to whatever endpoint you point it at.

Influencers & creators

Productivity, privacy, self-hosting, local-LLM, Android power-user, or second-brain audiences — this demos well and it holds up to scrutiny, because the offline mode is real and the cost tracker is honest. Early builds, direct access to the developer, and affiliate terms available.

Investors

A shipping product with a clear $5/month subscription, no inference costs on our balance sheet — users bring their own keys — and a second platform in development. Happy to walk through the roadmap, the architecture, and the numbers.

Why the economics are unusual: because SID3KICK is on-device-first and bring-your-own-key, we don't pay for inference and we don't operate a storage backend. There's no per-user cloud cost quietly scaling underneath the subscription — which means the $5/month is margin, not a loss-leader waiting to be repriced.

Contact

Start a conversation

Partnership, coverage, investment, or you just want on the early-access list — it all comes to the same inbox and it all gets a reply.

  • Early access to Android builds, and to iOS as soon as it's testable
  • Direct line to the developer — no support-ticket carousel
  • Affiliate and review terms for creators
  • Deck, roadmap, and architecture walkthrough for investors

Prefer plain email? [email protected]