Google Cloud Speech-to-Text vs. Phrase

Overview
ProductRatingMost Used ByProduct SummaryStarting Price
Google Cloud Speech-to-Text
Score 6.1 out of 10
N/A
Speech-to-Text on Google Cloud is a tool used to convert speech into text using an API powered by Google’s AI technologies. The vendor states users can transcribe content in real time or from stored files; deliver a better user experience in products through voice commands; and, gain insights from customer interactions to improve service.
$0.02
per min
Phrase
Score 3.4 out of 10
Small Businesses (1-50 employees)
Phrase is a Language Intelligence provider. Its enterprise platform automates, manages, and delivers multilingual content. Global brands use Phrase across hundreds of languages to reduce time to market and deliver consistent brand experiences worldwide. The Phrase Platform brings together translation management, software localization, multimedia localization, machine translation, workflow automation, and language AI in a single integrated environment. From marketing…
$27
per month (billed annually)
Pricing
Google Cloud Speech-to-TextPhrase
Editions & Modules
Speech-to-Text V2 API
$0.016
per min
Speech-to-Text V1 API
$0.024
per min
Freelancer & LSP plans - Freelancer
$27
per month (billed annually)
Developer Plan - Software UI/UX
$525
per month (billed annually)
Freelancer & LSP plans - Professional
$525
per month (billed annually)
Business Plan - Team
$1,245
per month (billed annually)
Business & Enterprise Plans
Custom
Offerings
Pricing Offerings
Google Cloud Speech-to-TextPhrase
Free Trial
YesNo
Free/Freemium Version
YesNo
Premium Consulting/Integration Services
NoNo
Entry-level Setup FeeNo setup feeOptional
Additional DetailsSpeech-to-Text V1 API V1 offers data residency for multi region only. Models include short, long, phone call, and video. V1 does not include audit logging. New customers get $300 in free credits and 60 minutes for transcribing and analyzing audio free per month, not charged against your credits. Speech-to-Text V2 API V2 offers data residency for multi and single region. Models include short, long, telephony, video, and Chirp. V2 does include audit logging and support for customer managed encryption keys.
More Pricing Information
Community Pulse
Google Cloud Speech-to-TextPhrase
Considered Both Products
Google Cloud Speech-to-Text
Chose Google Cloud Speech-to-Text
Google low latency streaming api seems to be working best when compared to other cloud-supporting tools as this will help in realtime transcription for customer interactions. while comparing azure and amazon the google support more than 125 languages and ascents as we work with …
Chose Google Cloud Speech-to-Text
Earlier we were completely reliant on text pad or notepad, where we used to manually capture the information, which seems to be very hectic for the long-running meeting, because holding the information and capturing them and redocumenting it is very big process and it involves …
Chose Google Cloud Speech-to-Text
They just remind me of each other. Whenever I have a question, whether for personal or for professional reasons, I take out my smartphone, click the Gemini app, and then click the mic to ask my question and have the answer read back to me. I love Googles AI System.
Chose Google Cloud Speech-to-Text
Firefly is a great notetaker application that plus into meetings and organizes the data. However it comes presctured while the Google Cloud Speech-to-Text application you can better organized the data. Then have the option to plug it into other platforms to further organize the …
Chose Google Cloud Speech-to-Text
It delivered high accuracy in accented and noisy environments. Regarding its language support, it offers a variety of languages and dialects. Its's Api's are well-documented and easily integrated with our GCP-based stack. Also, its deployment is fast, and it is cost-effective. …
Chose Google Cloud Speech-to-Text
One major setback is the integration of multiple languages, where Google has support for more than 120 languages, and Amazon only supports approximately 30 languages. Regarding the transcription behaviour, Google Translate is very accurate, but Amazon Translate sometimes spells …
Chose Google Cloud Speech-to-Text
Descript is definetly less accurate than Google Cloud tool, while Google Cloud Speech-to-Text does have some troubles with overlapping and background noises, it still performs better than Descript. However Descript has some video editing functions which are not available in …
Chose Google Cloud Speech-to-Text
Otter is good for simple note taking and its UI is quite simple. But it seriously lacks in providing appropriate transcription as it is less accurate with Indian accents and being unresponsive in real time transcriptions. On the other hand Google Cloud Speech-to-Text excels in …
Chose Google Cloud Speech-to-Text
Google Cloud Speech-to-Text is more recommended by clients and also based on our research, we found that this is the best option for our application.
Chose Google Cloud Speech-to-Text
We use Google Speech to Text on the recommendation of a partner who uses it, in fact we do not evaluate other applications such as Amazon Transcript or similar
Chose Google Cloud Speech-to-Text
Google Cloud Speech to Text has a significantly cleaner and easier-to-use User Interface. If the user is already familiar with the Google Cloud product suite, then onboarding with this software will be an extremely smooth process. If a user has previously used other …
Chose Google Cloud Speech-to-Text
Much better and more accurate than integrated Microsoft dictate or translate.
Chose Google Cloud Speech-to-Text
I didn't see other options that are even competitive with using Google's Cloud Speech-to-Text in terms of cost and reliability.
Chose Google Cloud Speech-to-Text
I have not used other software at this time. But this is a great software and completely worth using.
Chose Google Cloud Speech-to-Text
I like Google Cloud Speech-to-Text the most when it comes to other apps I have used so far. It have reduced my work, saved lot of time and made me less stress in meetings. It has also helped us in taking requirement gathering, knowledge transfer important notes to further …
Chose Google Cloud Speech-to-Text
While both Speechify and Google Speech-to-text do the job, certain elements that I find missing on Speechify are: it only works on Desktop with Windows OS, the customizations aspect is missing, there is no mobile app support (people these days want everything on their mobile …
Chose Google Cloud Speech-to-Text
I've also trialed IBM Watson Speech to Text for similar use cases. While both are highly capable, I find the Google Cloud Speech-to-Text software's accuracy and integrations to be a cut above.​ Harnessing Google's speech recognition prowess has elevated our firm's value …
Chose Google Cloud Speech-to-Text
Google Cloud Speech-to-Text outperformed its competitors significantly in terms of accuracy, surpassing any other product available. Additionally, its support for multiple languages was unrivaled in the market. Moreover, for clients with robust bandwidth, Google Cloud …
Chose Google Cloud Speech-to-Text
Office 365 word document text to speech engine.
This is popular among office users, but less relevant for mobile devices.
Chose Google Cloud Speech-to-Text
1. It's an efficient tool for improving efficiency by saving a lot of time in typing. 2. It saves at least 40-50% of our time, thus increasing efficiency. The amazing thing I liked about it is the accuracy with multiple accents & multiple languages. 3. It also takes …
Chose Google Cloud Speech-to-Text
I did not compare to other providers.
Chose Google Cloud Speech-to-Text
The accuracy of Google Cloud Speech-to-Text is much better than any other tool. It has better API integration with 3rd party tools. The transcription is on at real-time basis with the best efficiency. It has good language support from across the globe. It provides better noise …
Chose Google Cloud Speech-to-Text
Google Cloud Speech-to-Text is better than these other services. The main driver is the cost for the service and what you get, the value proposition is very good. Also, the scalability of Google Cloud Speech-to-Text is great, so that down the line, as our needs change and …
Phrase
Chose Phrase
It is very expensive and more difficult to navigate. Compared to the other support is slow and ineffective. Dealing with anything more involved is all but impossible. It has become more prone to bugs since being bought by TMS. Support suggests things like "use incognito" mode …
Chose Phrase
I did not select Memsource. I use it because my clients use it, but I find it very useful especially due to a friendly interface and quick jumping between segments of different statuses during translations.
Chose Phrase
SDL Trados Studio is more robust but also much more bloated and resource-intensive and ultimately less flexible. I use both daily, but Memsource is my go-to in almost every situation.
Chose Phrase
Memsource is quicker and easier to use. You can even start translating on your mobile phone or tablet, as long as you have an internet connection. Plus Memsource shows me a live preview of my translation. I think Memsource is great for beginners as well, since you don't need a …
Chose Phrase
Memsource is easier and more flexible than other CAT-tools, and definitely cheaper than Trados.
Chose Phrase
Starting a translation is easier than with SDL Trados. TM is better than with Smartling and Crowdin. The overview of New and Accepted work is so easy.
Chose Phrase
Memsource is the best, because it does not have any unneeded functionality that would compromise its intuitive use, and those functionalities it includes are all easily accessible.
Chose Phrase
Memsource stands out because it can be used in ANY browser. Thanks to its mobile app, it is easy to accept and review the work that translators are assigned.
Chose Phrase
I also use Wordbee, but the thing I don't like it most about it is that the number of segments you can display at a time is limited to 100 and you must move around many pages if the job size is large. I use SDL Trados Studio as well, but not cloud-based, so not really …
Chose Phrase
While I liked working with SDL Trados Studio and memoQ, Memsource is just easier to use overall, allowing you to work faster and more efficiently.
Chose Phrase
Memsource is the easiest to use among all other CAT tools out there. Even when you use it for the first time, it will be easy for anyone to use. You don't even need much instruction or a tutorial process. Still, it allows you to do a lot. It's easy to find some features you …
Chose Phrase
There is no true comparison because Memsource consistently outperforms the competition because it is simpler to use and has greater TM and MT integration.
Chose Phrase
Memsource is way cheaper and easier to use. Other softwares seem really old-fashioned and out of date. They are very difficult to use for a new learner and the UI is not good looking. Also, the price for products like Trados is too much expensive for a new freelancer and many …
Chose Phrase
It was selected by the client. I would go for XTM.
Chose Phrase
Memsoucre has a more friendly user interface. It keeps offering new features (like machine translation add on) which wasn't there when we first started using Memsource. Very smooth and easy communication with customer service team. My tickets are addressed quickly and …
Chose Phrase
I have been using memoQ and SDL Trados Studio for a longer time, but the Web editor for Memsource is the easiest to use of all. I do like the functionality in memoQ better, but Memsource is easier for our translators.
Chose Phrase
Memsource has most of the features you may encounter in memoQ but at a cheaper price than you can afford.
Chose Phrase
What distinguishes Memsource from other translation tools is that it does not require massive specifications to run. Average computers can run it smoothly without needing 8 or 12 Gb ram. Being able to use the browser to do the translation task is also another excellent feature. …
Chose Phrase
Memsource is the platform chosen by one of my clients, Booking.com. I have direct contact with my client, and the machine translation service offered is better than the one offered by its competitors. It's more intuitive and easy to use as well.
Chose Phrase
Memsource includes all the basic features a freelance translator needs in a very user-friendly approach. All you need is a click away and provides quick results.
Chose Phrase
Memsource has a very good user interface that people can quickly learn and start using. Also, its analysis features are top class which helps me provide a detailed estimate to my client based on the file particulars. It has a QA feature that helps me do a very high-quality …
Chose Phrase
Memsource is more user friendly and innovative. The interface looks better and more accessible. I have not experienced any downtime with Memsource. The times that I used Memsource it didn't give me any errors or issues. But with the other tool, there were several times when …
Best Alternatives
Google Cloud Speech-to-TextPhrase
Small Businesses
RingCentral Contact Center
RingCentral Contact Center
Score 8.5 out of 10

No answers on this topic

Medium-sized Companies
Zoom Contact Center
Zoom Contact Center
Score 8.7 out of 10

No answers on this topic

Enterprises
Verint Speech and Text Analytics
Verint Speech and Text Analytics
Score 8.4 out of 10

No answers on this topic

All AlternativesView all alternativesView all alternatives
User Ratings
Google Cloud Speech-to-TextPhrase
Likelihood to Recommend
5.3
(0 ratings)
3.0
(0 ratings)
Likelihood to Renew
-
(0 ratings)
9.6
(0 ratings)
Usability
7.2
(0 ratings)
2.0
(0 ratings)
Support Rating
-
(0 ratings)
8.0
(0 ratings)
User Testimonials
Google Cloud Speech-to-TextPhrase
Likelihood to Recommend
So, I've had scenarios like when I collaborate with a team where the people are from around the world. So, I used it there, and we spoke to each other in their native language. That boosts everyone's confidence in our collaborative efforts. I've also utilized its model and the API in my projects, including a Virtual assistant and a multilingual application that allows us to learn languages from around the world. We tested it with a group of 12 people, and that's when it failed. I mean, it's not a failure, but it can't detect every person.
Read full review
I like Memsource because when I am not home and I don't have my laptop, I can borrow a computer, log in to my Memsource account and I'm ready to begin translating. I can even download the source files and check the TM and glossary there. It's not necessary to download the software and lose time on that.
Read full review
Pros
  • An amazing tool which helps a lot in a meetings.
  • It's an efficient tool for improving efficiency by saving a lot of time typing. It saves at least 40-50% of our time, thus increasing efficiency.
  • Incredible accuracy with multiple accents & multiple language.
  • It takes punctuation into consideration.
Read full review
  • The synching of the desktop app with the online app is great and very helpful.
  • It is easy to look for all segments containing a specific term and correct them in a batch.
  • Suggestions of other segments where a certain term already appears are very helpful.
Read full review
Cons
  • The software does occasionally get confused by confusing terminology.
  • Its web-based interface can also feel a tad hard to use compared to more appealing desktop apps.
  • I've experienced the occasional technical issue, though the provider's support team is quick to troubleshoot.
Read full review
  • In the case of EN to JA translation, we enter text and then convert it to get the correct final text (word or phrase) which uses correct Chinese character(s), as there are often multiple Chinese characters with the same reading but different meanings. When those conversion options are displayed, it is usually possible to change our selection among them by hitting the Tab key, but in Memsource, hitting the Tab key makes us leave the text conversion and move to the source segment, so we must make sure to use arrow keys to choose the right conversion option when using Memsource.
  • When there are tag elements used in the source text, the tags must exist in the target text of course, but the tag order also must be the same. The tags cannot be moved around in the segment, which causes problems in the case of the Japanese language, because the word order differs between EN and JA.
  • In the Japanese language, Italics are basically not used, so there must be no texts between the tags which specify Italic font. But then the segment cannot be confirmed and the job status cannot be changed to "Complete".
Read full review
Likelihood to Renew
No answers on this topic
I use and manage an Academic edition (approx. 15 students) and the corporate license (10 users). In both cases, the features are easy to find, set and use. My students and my translators understand the dynamics of the platform easily and get used to them quickly.
Read full review
Usability
The reasoning behind my 10 is that the UI is very intuitive; I didn't require any formal training to use it. Google's speech-to-text is not just a conversion tool; it helps automate mundane tasks, saves time, and has an almost human-like understanding.
Read full review
It is designed to be a lean system, but that means that things are clustered in weird groups and a lot of trial and error is involved getting the settings right. If something goes wrong first level support is very quick and has very fast answers which are quite typical of first level support and essentially equate to asking you if you have switched it off and on yet. Anything more involved is challenging for support to address.
Read full review
Support Rating
No answers on this topic
We rarely use support, but most questions were answered in a timely fashion, although we didn't exactly find them satisfactory. That's mostly the fault of the software and not the Support team because we asked for things that Memsource couldn't do.
Read full review
Alternatives Considered
Earlier we were completely reliant on text pad or notepad, where we used to manually capture the information, which seems to be very hectic for the long-running meeting, because holding the information and capturing them and redocumenting it is very big process and it involves human work we were looking for some automation which can fix this issue then we got Google Cloud Speech-to-Text which converts audio to text files easier and faster and also it support many languages there by helping to align with various different clients across the globe and make the discussion seamless.
Read full review
It is very expensive and more difficult to navigate. Compared to the other support is slow and ineffective. Dealing with anything more involved is all but impossible. It has become more prone to bugs since being bought by TMS. Support suggests things like "use incognito" mode to log in. Use a different browser to log-in. That is a concern given the customer data involved if things at that level are not being fixed. If that's what it's like in the showroom, what is the kitchen like?
Read full review
Return on Investment
  • Right now it is very costly for any small company to afford price can be reduced
  • Extension to skype/webex can increase connectivity to multiple systems and capture the data efficiently
  • Multi language features is an great asset to this tool as it will help us to transcribe text of any language. More than 100 plus languages support
  • Customer support is also great who assist us whenever we face issue and gets resolved very fastly
Read full review
  • Positive Impact due to saved time.
  • Positive Impact due to not worrying about the format of the file (it is automatically taken care of).
  • Positive Impact due to being able to follow the client's style guide clearly and without errors.
  • Positive Impact due to avoiding common errors through QA feature.
Read full review
ScreenShots

Google Cloud Speech-to-Text Screenshots

Screenshot of audio transcription creation -  Using the Speech-to-Text API from within the Cloud Console by creating an audio transcription is done in just a few steps. It can transcribe short, long, and streaming audio.Screenshot of creating subtitles for videos using AI -  Transcriptions with captions and subtitles can be added to existing content or in real time to streaming content. Google's video transcription model can be used for indexing or subtitling video and/or multispeaker content and uses similar machine learning technology as YouTube does for video captioning.Screenshot of adding Speech-to-Text to apps - The video pictures covers how to add AI to an application without extensive machine learning model experience. The pretrained Speech-to-Text API lets users enable AI for applications.Screenshot of Language, speech, text, and translation with Google Cloud API - The pictures displays a section of Google training course, where learners use the Speech-to-Text API to transcribe an audio file into a text file, translate with the Google Cloud Translation API, and create synthetic speech with Natural Language AI.

Phrase Screenshots

Screenshot of the Phrase dashboard.Screenshot of Phrase's machine translation options.Screenshot of Open ecosystem

Bring your own AI engine, connect with your preferred language service providers, and integrate seamlessly with 50+ tools across your tech stack to keep workflows aligned and scalable.Screenshot of the self-service portal providing on-demand machine translation.Screenshot of Phrase Studio, which is an AI-powered subtitling and dubbing to scale video and multimedia content with speed, consistency, and control.Screenshot of a customizable dashboard designed to surface the data that matters, enabling continuous optimization of workflows, cost, and performance.