Toolso.AI
Toolso.AI
All ToolsCategoriesTrendingLatest ToolsPricingBlog
Toolso.AI
Toolso.AI
Toolso.AI
Toolso.AI

Discover the best AI tools to boost your productivity

GitHubGitHubTwitterX (Twitter)YouTubeYouTubeTikTokEmail

Popular Categories

  • AI Writing
  • AI Image
  • AI Video
  • AI Coding
  • More Categories

Explore

  • Latest Tools
  • Popular Tools
  • More Tools
  • Submit Tool
  • Pricing

About

  • About Us
  • Contact
  • Blog
  • Changelog

Legal

  • Cookie Policy
  • Privacy Policy
  • Terms of Service
  • Refund Policy
© 2026 Toolso.AI All Rights Reserved
Limited timeLimited-time offerFeatured Listing24h priority review · No backlink · 30 days featured$29.90then $59.90Price rises to $59.90 after Oct 31Ends in--:--:--Submit now
  1. Home
  2. All Tools
  3. Subtitle Generation
  4. Typist
Typist interface previewVisit Website
Typist logo

Typist

Turn audio and video recordings into editable text, choose a transcription model, and export documents or timed subtitles. Typist also offers browser recording and read-only access to your transcript library from AI clients.

Subtitle GenerationAudio Tool#Transcription#Speech To Text#Subtitles
Start Free Trial
Saves
Visits
Views
Pricing
Free Trial
Published
Oct 8, 2026
Domain
iamtypist.dev
Community rating

Used this tool? Rate it

Rate this tool

Typist Product Information

Start Free Trial
Tool Information
Saves
Visits
Views
Pricing
Free Trial
Published
Oct 8, 2026
Domain
iamtypist.dev
Community rating

Used this tool? Rate it

Rate this tool

Featured Tools

Related Tools

Start Free Trial

What is Typist?

Typist is an AI audio and video transcription service operated by Kollektiv LLC. It turns a recording into editable text that can be searched, quoted and repurposed. The product combines file transcription with a browser recorder, a transcript library and a connection that lets compatible AI clients read completed transcripts.

Audio transcription is its central task, with document exports for interview and lecture work and timed text for video captions.

Core features

Model selection

Typist offers xAI Grok, ElevenLabs Scribe and OpenAI Whisper, with language-specific labels for Most accurate, Fastest and Lowest cost. Those labels describe the vendor's intended tradeoffs for a particular language, rather than an independently verified ranking of every recording or accent.

  • Whisper v3 Turbo: Described as OpenAI Whisper running on Groq, built for volume, with word timestamps and support for 100 languages. Its positioning centers on long recordings and backlogs.
  • Grok 2.0: Attributed to xAI, with speaker labels and formatting for numbers and dates. The vendor positions it around combining transcription speed with accuracy in common languages.
  • Scribe v2: Attributed to ElevenLabs, with support for 100 languages, speaker labels and roles, and audio-event tags. Typist positions this model around its highest-accuracy option.

Transcript search and timed subtitles

  • Search across interviews: Typist advertises full-transcript search across an interview corpus, including mentions of a phrase across multiple files. This library search is a distinct description from the metadata-based search exposed through MCP.
  • Subtitle generator: The subtitle generator produces SRT or VTT files with short, timed lines and no speaker labels burned into the captions.
  • Speech-based timing: For video, Typist transcribes the audio track and ignores the picture; it uses word-level timing to start and end subtitle cues. The vendor says video resolution and frame rate do not determine the captions, while the speech and the way lines are split affect the result.

Guide

File transcription

Typist accepts MP3, M4A, WAV, FLAC, OGG and video files through its upload control. The published audio-to-text flow is:

  1. Drop a recording into the upload area or click to select it.
  2. Select the spoken language and a transcription model, using the language-specific speed or accuracy labels.
  3. Read the resulting transcript, then copy it or export it as TXT, DOCX, PDF or SRT.

Microphone recording

Typist also provides a microphone recorder with a Start Recording control; transcription begins after the user stops recording.

Audio and video use cases

Interviews and qualitative work

Typist positions interview transcription for research workflows, qualitative coding and investigative journalism, with timestamped text that can be cited, coded or quoted.

Every interview sentence is described as carrying a timestamp for citations and quotes. This makes the time reference part of the published interview workflow.

Lectures and recorded teaching

Typist positions lecture transcription around recordings made in class, Zoom recordings and M4A files, producing searchable, timestamped text. The use case is access to what was said in a recording, with a written form that can be read again and located by time.

Public video transcripts and summaries

The free YouTube transcript utility retrieves a video's captions or closed captions and supports copying with or without timestamps, TXT/SRT/VTT downloads, text search, seeking to a moment and sharing a public page. That caption-retrieval route is different from uploading a private audio file for a selected speech-recognition model.

The YouTube summary utility produces a TL;DR, key topics, key moments with clickable timestamps, and structured markdown for AI clients.

Who is it for

  • People working from recordings: Typist markets to students, teachers, podcasters, researchers, consultants, therapists, coaches and film or video users. The common task is turning spoken material into text or captions that can be revisited and reused.
  • Work that does not require certified transcripts: The terms say transcriptions are unsuitable for legal proceedings, medical records or any situation requiring certified accuracy. This is a substantial boundary even though the marketing includes therapy, coaching and research audiences.

Platforms

Browser access and AI clients

The voice memo workflow uses a browser on a phone or laptop and starts by sharing the original recording as a file.

Typist MCP uses Streamable HTTP and browser-based OAuth, without requiring the user to create or copy an API key. A compatible AI client connects to the remote service, and the user signs in through the browser to authorize access. The public AI entry mentions clients such as ChatGPT and Claude; actual access depends on a client supporting the required remote transport and authorization flow.

Transcript access through MCP

  • Read-only scope: Typist MCP cannot create, rename, edit or delete transcripts; access resolves to one account and does not expose recordings or billing. Connecting an AI client therefore adds a way to consume transcript text, rather than a way for that client to manage the account's files or subscription.
  • Metadata search: MCP search covers completed transcript titles, exact topics, category and an inclusive UTC date range; omitting filters lists recent transcripts. It is not the same as the interview page's advertised full-text corpus search.
  • Paged reading: Transcript reads default to 2,000 characters, can request up to 90,000 characters per call, and return a nextOffset for continuation. A long recording can require multiple reads; the first returned page is not necessarily the whole transcript.
  • Locked previews: Locked transcripts remain visible in search, but reading them returns the preview available to the account without an error. Discovery in the library consequently does not guarantee access to the full text.
  • Temporary exports: MCP export links are signed for one hour and stop working afterward. That link lifetime is separate from the file-retention period included in a subscription.

Pricing

Monthly plans and retention

The public monthly cards showed the following prices and limits as captured on October 8, 2026. Paid cards explicitly inherit the preceding plan's features; the table preserves that inheritance where a card does not repeat a value.

PlanMonthly priceIncluded usageFile limitFile retentionConcurrent transcriptions
Free$060 free minutes advertised; every transcription modelUp to 500 MB7 daysNot stated in the card
Lite$6.99/month10 credits per month; inherits FreeUp to 5 GB30 days3
Premium$19.99/month75 credits per month; inherits LiteInherited from LiteUnlimited10
Max$49.99/month200 credits per month; inherits PremiumInherited from PremiumInherited from PremiumInherited from Premium

The Free card lists TXT, DOCX, PDF, SRT and VTT exports. Premium adds unlimited AI Insights, described as TL;DR, chapters, quotes and actions. Max adds priority support.

Model-dependent credits

One credit covers 1 hour on Whisper v3 Turbo, 30 minutes on Grok 2.0 or Whisper v3, or 15 minutes on Scribe v2; credits reset each billing cycle, and speaker labels and filler removal work with Grok 2.0 or Scribe v2. The purchased unit is consequently a credit with a model-dependent duration, rather than a uniform hour across all engines.

The free headline is specifically advertised as 60 minutes. Because the same pricing section explains different credit consumption for different models, that headline alone does not establish identical free recording time on every engine. Monthly credit resets also describe a billing-cycle allowance rather than a lifetime pool of purchased transcription time.

Renewal, cancellation and refunds

Monthly or yearly subscriptions automatically renew, can be cancelled in account settings, and remain active until the billing period ends after a mid-period cancellation. Cancelling therefore ends the next renewal rather than automatically ending current-period access or reversing a completed purchase.

The stated 7-day refund window covers transcription-preventing service or technical failures, billing errors or duplicate charges, and first-time subscribers unsatisfied with transcription quality. The window is conditional, rather than an unconditional refund promise for every subscription.

Refund exclusions include successfully completed transcriptions, usage beyond free limits, requests after 7 days, buyer's remorse and changed circumstances. The eligibility language and the exclusions appear together in the terms; the stated quality-related category does not erase the separate completed-transcription or usage exclusions.

A refund request requires contacting support within 7 days with order details and a reason; the terms state a response within 48 hours.

Fees exclude applicable taxes unless stated otherwise. Subscription price changes apply at renewal with reasonable notice of increases.

Limitations

Accuracy, access and page inconsistencies

  • Transcription errors: The terms warn that AI transcripts can contain mistakes, omissions and misinterpretations and are not 100% accurate. Timestamped words and speaker metadata still come from automated transcription; a precise time reference is not a guarantee of correct recognition.
  • Service continuity: The terms do not guarantee uninterrupted service and allow interruptions for maintenance, updates or unforeseen circumstances. Model-speed marketing does not establish a continuous-availability commitment.
  • Age requirement: The terms require users to be at least 18. The privacy policy separately says the service is not directed to children under 13 and does not knowingly collect their personal information; that privacy statement does not replace the higher contractual access threshold.
  • Export inconsistency: The voice memo page says Free exports TXT and DOCX, while paid plans add PDF and SRT. This differs from the Free pricing card, which lists TXT, DOCX, PDF, SRT and VTT. The two public descriptions do not provide a single consistent free-export boundary.

The same distinction between a capability and its conditions applies to model labels. The service provides choices with different feature sets and consumption rates, and a named model's language coverage is not an independently tested accuracy percentage. The reviewed public material does not supply a product-level benchmark that would turn the promotional accuracy language into a reliable result for every uploaded file.

Ownership and processing rights

Users retain ownership of uploaded content and own the transcriptions generated from it.

Uploading grants Typist a limited, non-exclusive license solely to provide transcription, ending when the content or account is deleted.

Storage, sharing and deletion

The audio-to-text entry claims files are uploaded only to transcribe, removed afterward, and never sold, shared or used to train models. This is vendor privacy marketing attached to that conversion entry. It is not an independently verified statement about every downstream provider or every route through the product.

The privacy policy separately states that uploaded audio and resulting transcripts are stored, associated with the account, and deletable by users. The account-storage statement and the plans' retention periods do not describe the same immediate-removal promise as the conversion entry. The public wording therefore leaves different storage descriptions in place rather than a single universal deletion timeline.

The privacy policy says database, authentication, transcription, hosting, monitoring, analytics and payment providers access information as needed for their functions. That service-provider disclosure also differs in scope from the conversion entry's unqualified no-sharing wording. The training statement is the vendor's declaration; the reviewed privacy text does not separately specify the training policies of each processing provider.

Account deletion leads to personal-data deletion within 30 days, except required legal or regulatory retention; aggregate anonymized analytics may remain. This account-deletion rule is separate from deletion of an individual file, the retention allowance of a plan and expiry of a temporary export link.

The privacy policy claims encryption in transit and at rest, secure authentication and regular security monitoring. It also says providers may store and process data in countries including the United States.

Free-tool outputs and public video pages

Most free-tool outputs are automatically deleted within 15 minutes to 1 hour, while YouTube transcripts may instead be stored and displayed publicly and are subject to removal requests. The short temporary-output range covers downloaded files, converted media and similar free-tool outputs as described in the terms; it does not replace the account library's file-retention rules.

FAQ

Q1. Does connecting an AI client let it upload or change recordings?

Typist's published MCP tools are read-only. They search, read and provide transcript exports; they cannot create, rename, edit or delete transcripts, and the scope excludes recordings and billing.

Q2. Are YouTube transcripts treated like private uploaded files?

The YouTube transcript utility supports public pages and links. Its terms allow transcripts from public videos to be stored and displayed for sharing and discovery, unlike the separate descriptions of account storage and plan retention.

Q3. Does video transcription analyze the picture?

The subtitle entry says only the video's audio track is transcribed and the picture is ignored. It describes word-level timing for subtitle cues, rather than visual analysis of scenes or video frames.

Know a Similar Tool?
If you know other great AI tools, feel free to submit them to us