Pick by the job: GoTranscript when a wrong word costs something, Vook.ai when volume matters more than features, Sonix for difficult audio, Otter for meetings, Descript for video, Good Tape for confidential material, and OpenAI Whisper when the budget is nothing at all. This page names 12 tools, links to none of them, and earns nothing if you buy any. Three of the nine pages ranking for this search put their own transcription product first and a fourth turns every row into an affiliate redirect, which matters here more than usual, because the number those pages disagree about is the only number most buyers look at.
The accuracy figure you are shown depends on who published it, and for the same tool the published range runs from 83% to over 99%. Sonix’s own comparison page rates Otter at 83 to 85%; Media Copilot’s hands-on test gives Otter four stars out of five for accuracy and calls it the best all-round choice for journalists; Wirecutter, after 35 hours with 15 services, reports the AI field averaging 96% and puts OpenAI’s free Whisper Large model at 98.7%. Nobody is lying. They are measuring different things, on different audio, with different reasons to publish. This page prints the spread first and the shortlist second, because reading the rows without the spread is how people buy the wrong tool.
Who publishes which accuracy number
A vendor publishes a ceiling, an independent tester publishes a mean, and a rival vendor publishes your floor. All three appear in this search result, often for the same product on the same day, and the pattern holds well enough to use as a rule. The table below puts four published figures for the same handful of tools next to the party who published each one and the conditions each figure was measured under, as recorded on those pages and checked on 12 August 2026.
| Tool | Figure published | Published by | Measured on |
|---|---|---|---|
| Otter | 83 to 85%, English only | Sonix, a direct competitor | Not stated on the page |
| Otter | Four stars out of five for accuracy | Media Copilot, independent | Three 10-minute files: a clean video call, a poor phone recording, a noisy press gaggle |
| Otter | Below the transcription-focused tools | Wirecutter, independent | 15 services, 35 hours, one shared batch |
| Rev | 90% | Sonix, a direct competitor | Not stated on the page |
| Rev AI | 95 to 98% on English interviews | CleverX, independent | Research interviews, banded rather than exact |
| Sonix | Consistently over 99% | Sonix, its own page | Clear audio |
| Sonix | Best of the field on difficult audio | Media Copilot, independent | The same three files, including the press gaggle |
| The AI field | 96% average, weakest 94% | Wirecutter, independent | One batch across 15 services |
| Whisper Large | 98.7%, and Small at 97.7% | Wirecutter, independent | The same batch, run locally |
Two things follow from that table. A vendor’s “up to 99%” is a real measurement of its best case and tells you nothing about your worst one, and a competitor’s number for your tool is measured with an interest in the result. The figures worth weighting are the ones where a publisher with nothing to sell in the category describes what it fed the tool, which in this search result means Wirecutter, Zapier, Media Copilot, PCMag and CleverX.
What word error rate measures and what it ignores
Word error rate counts substitutions, deletions and insertions against a human-made reference transcript, so 99% accuracy means one wrong word in every hundred under the conditions tested. It is a good measure and a narrow one. It treats every word as equally important, which means a misheard “the” and a misheard drug name cost the same, and it is measured on audio somebody chose, which is almost always cleaner than yours.
Two ordinary conditions move the number, and neither is exotic. Overlapping speakers break the segmentation the model depends on, so a four-person call degrades in a way a two-person interview does not. Unfamiliar proper nouns have no strong statistical prior, so names, products and places come back as the nearest common word. Wirecutter’s testing produced the clearest examples published anywhere in this search result: “as large as” was returned by multiple services as “F/f”, one service rendered the decimal expression “1.82%” as “1.8199999999999998%”, and Trint turned “steady” into “study”, “nebulously” into “nebulous Lee”, and “That’s always gonna” into “I was always going to”. Every one of those is a meaning change that a spellcheck pass would never flag.
The historical anchor is worth carrying too, because it explains why the category feels finished when it is not. Wirecutter found the best AI transcription in 2018 to be 73% accurate. Today the weakest service it tested reaches 94% and the best free model beats several paid ones. The remaining gap is not in the average; it is concentrated in exactly the material professionals record.
The 12 tools at a glance
The table below lists each tool with the job it is built for, the exact limit of its free tier, and its published starting price, so the row you need is the row for your workflow rather than the row somebody ranked first. Every figure comes from the publication named beside it in the entries that follow, checked on 12 August 2026, and prices in this category moved materially during the past year.
| Tool | Best for | Free tier | Published starting price |
|---|---|---|---|
| GoTranscript | Accuracy that has to hold up | None | $1.60 per minute human, $0.20 per minute AI |
| Vook.ai | Volume on a budget | First 30 minutes | About $0.05 per minute, or $10 a month |
| Sonix | Difficult audio and a full editor | 30 minutes | $22 a month plus $5 per transcribed hour |
| Otter | Meetings and live notes | 300 minutes a month | $8.33 per user a month billed annually |
| Rev | An AI pass with a human tier behind it | 45 AI minutes a month, English only | $25.49 per user a month billed yearly |
| Descript | Editing video by editing text | One hour a month | From $16 per user a month |
| Good Tape | Confidential and source-protected work | Limited free use | $16.85 a month |
| Google Pinpoint | Reporters with no budget | Free in full | Free |
| Trint | Newsroom volume and shared workflow | Trial | From $60 a month for unlimited AI transcription |
| Fireflies | Meeting bots feeding a CRM | Yes | $10 to $19 a month |
| AssemblyAI | Developers building transcription in | Credits | $0.30 an hour base |
| OpenAI Whisper | Zero budget, technical setup | Free and open source | $0 locally, $0.006 a minute via API |
What the table cannot show is the shape of each product, which is what actually decides the choice: whether a transcript is the deliverable or the raw material, whether anyone else has to read it, and whether the recording is something you are allowed to upload at all.
GoTranscript
GoTranscript is the answer when an error costs more than the transcript does. Wirecutter made it the top pick of 15 services tested over 35 hours, describing AI-assisted human transcripts at over 99% accuracy delivered in under a day, and PCMag independently named it an Editors’ Choice for accurate and affordable human transcription across the education, legal and medical fields.
Pricing is quoted per minute of audio rather than per seat, which makes it directly comparable with everything else on this page. Wirecutter records $1.60 per minute for one-day turnaround on human transcripts, PCMag records $1.20 per minute for human work with discounts on large jobs, and both put the AI-only tier at $0.20 per minute. The gap between those two figures is the gap between a machine pass and a human reading every line of it.
Speed is better than the price suggests. In Wirecutter’s test the AI service transcribed the entire batch of recordings in six minutes, under one minute per minute of audio, faster than Reduct and Rev AI though slower than Vook.ai, Whisper and Descript. Human transcripts came back the same day. Uploads are capped at 4 GB, the largest limit among the services Wirecutter compared.
The flaw is the interface. Wirecutter calls the web app fairly basic, with simple sharing, a text editor and audio playback that highlights words as they are spoken, and nothing like the collaboration in Descript or Trint. Choose it when the transcript has to be right and somebody else’s workflow does not depend on living inside the tool.
Otter
Otter is the meeting default, and it is also the clearest illustration of the accuracy spread described above. Sonix’s comparison table rates it at 83 to 85% and English only. Media Copilot, testing it against four rivals, gives it four stars out of five for accuracy, five for ease of use, and calls it the closest thing to a universal transcription tool for journalists. Wirecutter, testing it against transcription-focused services rather than meeting tools, found its accuracy lower than the field and its speaker separation weaker. All three are describing the same product used three different ways.
The free tier is specific and worth reading before signing up: 300 live transcription minutes a month with a maximum session length of 30 minutes, plus three pre-recorded file uploads per lifetime rather than per month, as PCMag documents it. Paid plans start at $8.33 per user a month billed annually according to both PCMag and Zapier, which is the same product Media Copilot lists at $16.99 on monthly billing.
What it does well is the meeting itself. A bot joins calendared calls, screenshots presentation slides into the transcript, and produces an outline with action items you can assign, and Zapier notes you can paste an ad-hoc meeting link and have the bot join that too. Media Copilot rates the automatic filler-word removal and the AI summaries linked to transcript timestamps as the best features-per-dollar in its test.
The flaw is data handling rather than accuracy. Media Copilot records that Otter uses customer data to train its AI with an opt-out available, and that its servers are not necessarily EU-based, which is enough to disqualify it for confidential source material even where the transcript quality would do.
Rev
Rev sells an AI transcript with a human tier behind it, and the second half is the reason to choose it. Zapier’s reviewer, who went hands-on with 68 transcription apps, describes the human review add-on as extending accuracy to 99% and frames the case precisely: in a contract review, a service that renders “defective” as “effective” leaves everyone who was not on the call acting on wrong information.
The free tier is 45 minutes of AI transcription a month, English only. Essentials is $25.49 per user a month billed yearly for 5,000 AI minutes. Human review is ordered per file and carries a two-hour minimum turnaround, while the AI pass is quick: Zapier timed a six-minute video at about 40 seconds.
Published accuracy for Rev spans the usual range. Sonix’s table gives it 90% and quotes $0.25 a minute or $15 an hour. CleverX puts Rev AI at 95 to 98% on English interviews, the highest band in its own comparison, and recommends it specifically when verbatim quoting matters. Neither figure includes the human tier, which is where the 99% claim lives.
The flaw Zapier reports is commercial rather than technical: billing complaints centred on refund windows for transcription orders. Worth knowing before buying minutes in bulk, and not a reason to avoid the human tier where the alternative is a correction.
Sonix
Sonix is the tool that wins on hard audio in the only hands-on test in this search result that used hard audio deliberately. Media Copilot fed five services a clean video-call interview, a degraded phone recording and a press gaggle on Air Force One full of overlapping voices and difficult proper nouns, and Sonix outperformed the field, taking five stars for accuracy and 4.5 out of 5 overall.
Its own pricing is unusual and needs stating in both forms, because the two published versions look like different products. Media Copilot records $22 a month plus $5 for every transcribed hour. Sonix’s own page quotes $10 an hour pay-as-you-go or $5 an hour with a subscription, alongside claims of consistently over 99% accuracy, support for 53 or more languages, SOC 2 Type 2 compliance, AES-256 encryption, a 30-minute free trial and a 10-minute file transcribed in under two minutes. The first set is what an independent observed; the second is what the vendor publishes.
The editor is the most complete in this group, with subtitle export, filler-word control that lets you keep or remove them, adjustable AI summaries, and XML export straight into Adobe Premiere and Final Cut Pro, which is why Media Copilot recommends it to video and podcast producers over print reporters.
The flaw is cost at volume. Media Copilot calls it country-club pricing and the highest of everything it tested, and the hourly component means the bill scales with the archive rather than the team. Summaries are also not linked back to the transcript, which Otter does and which matters when you are hunting a quote.
Vook.ai
Vook.ai is Wirecutter’s affordable pick, and it is the clearest case in this category of a small product built well on an open model. The team describes it as added layers on OpenAI’s Whisper algorithm run on serverless GPUs, and in Wirecutter’s testing the results came back slightly more accurate than default Whisper models, at over 98%, delivered in minutes.
Pricing is where it separates from everything above it. The first 30 minutes are free, after which credit is bought in two, five or ten-hour blocks, or in custom blocks at $3 an hour, or through a subscription starting at $10 a month for five hours with leftover credit carried forward. Wirecutter works the average out at 5 cents per minute of audio, one quarter of GoTranscript AI’s 20 cents, and roughly one thirtieth of a human transcript.
Handling is documented rather than implied. User data sits on EU-based servers and is RSA-encrypted, transcription runs on Vook.ai’s own infrastructure, and the team told Wirecutter that audio files and transcripts are not shared with any third party unless the user requests DeepL translation or GPT-powered summaries. Files are capped at three hours or 2 GB, against GoTranscript’s 4 GB, so longer recordings must be split or moved to the $50 Ultra plan.
The flaw is everything around the transcript. There is no sharing or collaboration beyond an email export, no universal search across your transcripts, and organisation amounts to read or unread status and tags. It is a transcription engine with a thin shell, priced accordingly.
Descript
Descript treats the transcript as the editing surface for the recording: delete a line of text and the corresponding audio disappears, and an overdub fills a gap with a synthetic voice. For anyone producing podcasts or video the transcript is a means rather than the deliverable, and that changes which tool is correct regardless of word error rate.
The published prices differ by billing model rather than by disagreement. Wirecutter lists free for one hour of transcripts a month and from $16 a month for ten hours; Zapier lists a free plan and from $16 per user a month; Media Copilot lists $24 a month for the tier it tested. Sonix’s competitor table puts its accuracy at about 95%.
What it adds beyond transcription is a genuine production workflow: filler-word removal from the audio itself rather than only from the text, AI voice generation to fix a misspoken line after the fact, XML export, subtitle options, and a collaborative editor Wirecutter compares to Notion and recommends specifically for teams working on transcripts together.
The flaw is fit. Media Copilot scores it 3.5 out of 5 and calls it overkill for basic transcription, with a steeper learning curve than the alternatives and lower accuracy on proper nouns and names. If your output is a document rather than a cut, you are paying for an editor you will not open.
Good Tape
Good Tape answers the question the accuracy tables ignore, which is where your audio goes after the transcript comes back. Media Copilot describes a Danish company processing audio on EU-based servers, storing nothing by default, and committing explicitly to never training its AI on customer files, and rates it 4 out of 5 overall with five stars for security, the highest security score in that test.
Pricing is flat and stated: $16.85 a month, with strong speaker identification included. There is no per-hour component and no enterprise negotiation, which suits an individual reporter or researcher rather than a team.
The case for it is specific rather than general. For anyone working with confidential material, a supplier that deletes by default and refuses to train on your files is not a bonus feature; it is the entry condition, and it is the one attribute in this whole category that cannot be fixed by editing the output afterwards.
The flaws are the price of that focus. Media Copilot records occasional glitches on noisy audio, no filler-word removal, no mobile app, and a limited feature set against competitors, which rules it out for field recording on a phone and for anyone who needs integrations. Choose it when the confidentiality question outranks the convenience one.
Google Pinpoint
Google Pinpoint makes one argument and makes it well: it is free, and for anyone already inside Google Workspace it is free in a way that requires no new procurement. Media Copilot rates it 3.5 out of 5, five stars on price, and calls it the best free option for budget-conscious reporters.
What it adds beyond transcription is a fact-check integration with Google Search and summaries linked back to the transcript, which Media Copilot notes the paid tools cannot match, plus Meet transcripts through Workspace itself.
The flaws are substantial and mostly structural. Accuracy lags the paid services, particularly on noisy audio and overlapping speakers. There is no speaker identification, which on a multi-person interview means a wall of text nobody is attributed in. There is no option to remove filler words.
The condition that decides it for professional work is the last one on Media Copilot’s list: human reviewers may access sample data. For a clean two-person interview on the record, that is a reasonable trade for a free tool. For anything a source expects to stay confidential, it is a disqualification, and it is the same trade you make with most free tiers in this category without being told.
Trint
Trint is built around newsroom volume and shared workflow rather than around a single accurate file. Wirecutter describes unlimited AI transcription from $60 a month in a collaborative editor with full-text search, sharing and comments, a full-featured mobile app for transcribing in the field, and export into a wide range of subtitle formats. Sonix’s table quotes from $80 a month and 40 or more languages; PCMag quotes $69 per person a month billed annually and calls it pricey but worthwhile for larger organisations that use the collaboration.
The team features are the product. Story folders, shared review, translation and export into editorial systems are what a desk buys, and none of them appear in the per-minute services further up this page.
The flaw is accuracy at that price, and Wirecutter is specific about it. Trint measured 95% in its test, with errors that changed meaning rather than reading as typos: “steady” as “study”, “nebulously” as “nebulous Lee”, “That’s always gonna” as “I was always going to”. At $60 to $80 a month, 95% with meaning-changing errors is a harder purchase than the same figure at 5 cents a minute.
Choose it where the workflow is worth more than the last few percent, which in practice means a desk with several people touching the same transcripts every week.
Fireflies
Fireflies sends a bot into meetings, captures the transcript and pushes summaries onward, and its natural home is a sales or customer team where the transcript ends up in a CRM rather than in a document. CleverX publishes 90 to 94% accuracy with speaker labels at $10 to $19 a month; Sonix’s competitor table puts it at about 85% with a free tier and Pro from $18; Wirecutter notes support for 100 languages.
The multilingual reach is the strongest published claim. TicNote’s meeting roundup lists it specifically for global teams and multilingual workflows, which is where the meeting tools with six-language support stop being usable.
The flaws are interface and search. Wirecutter calls its interface the most basic of the meeting apps tested, with transcript search that takes longer than others to index, which matters when the archive is the point of the product.
The consideration nobody in this search result writes down is consent. A visible bot joining a call is a disclosure event, and in several jurisdictions recording a conversation without notifying every participant is not permitted. Any meeting tool that joins calls carries that question, and it belongs in the purchase decision rather than in a footnote after the first complaint.
AssemblyAI
AssemblyAI is an API rather than an application, and it competes on integration rather than on interface. CleverX publishes 92 to 96% accuracy, support for more than 100 languages including strong handling of accented English, and usage pricing from $0.30 an hour, and recommends it for developer-friendly integration and multilingual research.
The comparison that decides it is not against Otter or Sonix but against running Whisper yourself. At $0.30 an hour, AssemblyAI costs roughly three times the $0.006 a minute of OpenAI’s hosted Whisper API and infinitely more than running the model locally, and what you buy for the difference is speaker labels, formatting, a maintained endpoint and support.
Usage-based pricing needs modelling before it is comparable to anything else on this page. At two hours a week it is under $3 a month and beats every subscription here. At an hour a day it approaches the price of a seat licence, without the editor a seat licence includes.
The flaw is that it is a component rather than a tool. There is no editor, no sharing, no meeting bot and no mobile app, so somebody has to build the part your colleagues actually touch. Choose it when transcription is a feature of your product rather than a step in your week.
OpenAI Whisper
Whisper is the free floor underneath this entire category and the reason the paid prices moved. It is open source under an MIT licence, runs locally at no cost on hardware you already own, and Wirecutter measured its Large model at 98.7% accuracy and the faster Small model at 97.7%, which puts a free model above several services charging monthly.
It is also the engine inside the products above it, openly in Vook.ai’s case, which describes added layers on Whisper run on serverless GPUs, and likely in others, as Wirecutter observes. Buying transcription in 2026 frequently means buying an interface around this model, and knowing that changes what the subscription is for.
Access comes in three forms. Run it locally through Python for nothing. Use a wrapper, and Wirecutter names MacWhisper on Mac, Aiko on iOS and Speech Translate on Windows machines with CUDA-compatible Nvidia GPUs. Or call OpenAI’s hosted API at $0.006 a minute, which is 5% of GoTranscript AI’s per-minute rate.
The flaw is that it is a model rather than a product. There are no speaker labels without a wrapper, no editor, no sharing, no summaries, and a command-line setup that Wirecutter notes less technically inclined users struggle with. The transcript is excellent and everything around the transcript is your problem.
What a transcript costs per minute of audio
Converting every pricing model into cents per minute of audio is the only way to compare a $17 subscription with a $1.60 per-minute service, and the conversion changes the ranking. The table below normalises the published prices already cited above, with subscriptions divided across a stated monthly volume of 10 hours, which is 600 minutes. Change the volume and the order changes with it, which is the point.
| Service and tier | Published price | Cents per minute at 10 hours a month |
|---|---|---|
| Human transcription, typical range | $1 to $3 per minute | 100 to 300 |
| GoTranscript, human, one-day | $1.60 per minute | 160 |
| GoTranscript, AI only | $0.20 per minute | 20 |
| Sonix, subscription plus hourly | $22 a month plus $5 an hour | About 12 |
| Vook.ai, average of its plans | About $0.05 per minute | 5 |
| Otter, annual billing | $8.33 per user a month | About 1.4 |
| AssemblyAI, usage | $0.30 an hour | 0.5 |
| Whisper via OpenAI API | $0.006 per minute | 0.6 |
| Whisper run locally | Free | 0 |
The conversion exposes the assumption buried in every subscription: a $17 plan is cheap at 20 hours a month and expensive at one. Wirecutter’s own framing is the same in reverse, noting that transcribing a minute of audio by hand takes about four minutes according to Rev’s published figure, which is what makes human work cost between $1 and $3 a minute in the first place.
When a human pass is still required
Three cases have not changed, and in all three the accepted route is a machine pass followed by human review rather than one or the other. A legal record is the first, where a single substituted word changes a fact and the transcript may be relied on by somebody who was not present. Medical documentation is the second, where terminology defeats a general model and the error rate on drug and procedure names is not the error rate on the rest of the file. Anything published verbatim is the third, where a misheard name becomes a correction.
The market has settled on the same answer. Wirecutter observes that today’s human transcription services usually run AI for the first pass and have humans clean up the copy, and that the most accurate transcripts in its testing came from exactly those tag-team efforts, which is what GoTranscript sells at over 99% and what Rev sells as its human tier. PCMag notes a further detail worth knowing for sensitive material: human services often split a recording across several transcriptionists so no single person hears the whole thing, and typically require non-disclosure agreements.
What the machine changed is the price of the first pass, not the requirement for the second.
What free tiers cost you
Free plans in this category differ from paid ones in two ways that never appear on the pricing page, and both are worth two minutes before uploading a client’s recording. The first is training rights, meaning whether your audio is used to improve the vendor’s models. The second is retention, meaning how long the audio and the transcript stay on their servers after you have the text.
The published positions differ sharply. Media Copilot records that Otter uses customer data to train its AI with an opt-out available, and that Google Pinpoint’s human reviewers may access sample data. The same test records Good Tape deleting recordings by default and committing never to train on customer files, and Wirecutter records Vook.ai storing on EU servers with nothing shared with third parties unless the user asks for DeepL translation or GPT summaries. Four suppliers, four different answers, none of them on the price list.
The free limits themselves are equally uneven, and they are usually the reason a trial ends rather than quality: Otter gives 300 minutes a month with a 30-minute session cap and three lifetime file uploads, Rev gives 45 AI minutes a month in English only, Sonix gives 30 minutes, Vook.ai gives the first 30 minutes, and Descript gives one hour a month. Whisper run locally has no limit at all, which is the trade it offers in exchange for the setup.
Which one should you pick?
By job rather than by rank, and the attribute that decides each one is named beside it. Meetings: Otter, or Fireflies where the transcript has to reach a CRM, deciding on integration rather than accuracy. Interviews and difficult audio: Sonix, on Media Copilot’s hard-audio result. Video and podcasts: Descript, because the transcript is the editing surface. Confidential material: Good Tape, on deletion by default and no training. Newsroom volume: Trint, for the shared workflow, with its 95% measured accuracy accepted knowingly. Regulated or published work: GoTranscript, or Rev’s human tier, both above 99% with a person in the loop. Large archives on a budget: Vook.ai at about 5 cents a minute. Building it into a product: AssemblyAI, or Whisper’s API at $0.006 a minute if you can build the rest. No budget: Whisper locally, or Google Pinpoint if a command line is not an option and the material is not sensitive.
Nothing on this list is the best tool for everything, and any page telling you otherwise is usually selling one of them.
How tools get into this list
Inclusion is editorial and free. Every tool here was chosen because it is in real use in this category or because an independent publication tested it, and no vendor has paid to be named, been offered a link, or been given a rating in exchange for anything. We take no affiliate revenue from any tool on this page, which is why no entry is a link.
We have not run a transcription benchmark, and nothing here claims we did. Every accuracy figure and every price on this page belongs to the publication or vendor named beside it in the text. If you build a transcription tool and think this page is wrong or out of date, write and say so, and corrections that come with a source get made. How we work with contributors is set out on our write for us page, and anything commercial is arranged separately and disclosed on the page it appears on.
Frequently asked questions
These questions come up alongside the main one and are answered here rather than in sections of their own, each answer self-contained.
What is the most accurate AI transcription tool?
On the published measurements, OpenAI’s Whisper Large model at 98.7% and Vook.ai at over 98% lead Wirecutter’s test of 15 services, and Sonix led Media Copilot’s test on difficult audio. Above 99% appears only where a human reviews the machine output, which is what GoTranscript delivers in under a day and what Rev sells as a per-minute add-on.
Is AI transcription good enough for interviews?
Yes, with a review pass, and the errors cluster in two predictable places rather than spreading evenly. Names and other proper nouns are the first, and any passage where two people talk over each other is the second. Both are fast to fix once you know to look for them, which is why a 96% transcript is usable and a 96% legal record is not.
Do these tools train on my audio?
Some do, and free plans are where it most often happens. Otter uses customer data to train its AI with an opt-out, and Google Pinpoint states that human reviewers may access sample data, while Good Tape commits to never training on customer files and deletes recordings by default. The answer for any given service is in its terms and its data-processing addendum rather than on the pricing page.
How long does transcription take?
The fastest AI services run under one minute per minute of audio, and Wirecutter transcribed its whole test batch through GoTranscript AI in six minutes, while Reduct took about three minutes per minute of audio and counts as slow. Human work is measured in hours: Rev sets a two-hour minimum on human review and GoTranscript returns in under a day. Doing it yourself takes about four minutes per minute of audio, which is Rev’s own published figure.
The accuracy numbers in this category are real, narrow and published by parties with different interests, which is why the same tool appears at 83% on a competitor’s page and four stars on an independent test in the same week. Read the figure with the publisher attached, convert the price into cents per minute of the audio you actually record, check the training and retention terms before uploading anything belonging to somebody else, and put a person in the loop wherever one wrong word costs more than the transcript did.