Skip to Main Content
IBM Data and AI Ideas Portal for Customers


This portal is to open public enhancement requests against products and services offered by the IBM Data & AI organization. To view all of your ideas submitted to IBM, create and manage groups of Ideas, or create an idea explicitly set to be either visible by all (public) or visible only to you and IBM (private), use the IBM Unified Ideas Portal (https://ideas.ibm.com).


Shape the future of IBM!

We invite you to shape the future of IBM, including product roadmaps, by submitting ideas that matter to you the most. Here's how it works:


Search existing ideas

Start by searching and reviewing ideas and requests to enhance a product or service. Take a look at ideas others have posted, and add a comment, vote, or subscribe to updates on them if they matter to you. If you can't find what you are looking for,


Post your ideas

Post ideas and requests to enhance a product or service. Take a look at ideas others have posted and upvote them if they matter to you,

  1. Post an idea

  2. Upvote ideas that matter most to you

  3. Get feedback from the IBM team to refine your idea


Specific links you will want to bookmark for future use

Welcome to the IBM Ideas Portal (https://www.ibm.com/ideas) - Use this site to find out additional information and details about the IBM Ideas process and statuses.

IBM Unified Ideas Portal (https://ideas.ibm.com) - Use this site to view all of your ideas, create new ideas for any IBM product, or search for ideas across all of IBM.

ideasibm@us.ibm.com - Use this email to suggest enhancements to the Ideas process or request help from IBM for submitting your Ideas.

IBM Employees should enter Ideas at https://ideas.ibm.com


Status Not under consideration
Workspace Speech Services
Created by Guest
Created on Jan 18, 2022

Watson Speech to Text should return timestamps accurate to milliseconds for transcription

Real-life scenario:

Researchers from Brandeis University, Boston University, Harvard, Boston College, and Northeastern University are investigating cognitive aging and biomarkers of dementia. Currently, they have been hand-scoring cognitive interviews which is arduous. To automate some of the manual work, Watson Text to Speech service is being used to run audio transcriptions through the service while getting back transcriptions with timestamps of when a word starts to be spoken and when it is ended in speech. The timestamps returned from the service return with 2 decimal places which is not enough precision that is required to compare to prior work (which uses 3 decimal places -- precision is up to milliseconds).


Problem statement:

The issue that researchers mentioned above are running into is comparing transcription timestamps generated by Speech to Text service to hand-scoring done for interviews prior to using Watson. The precision mismatch does not allow the researchers to use Watson effectively.


Current workaround:

No workaround is possible since there is a mismatch in precision level returned by Watson Speech-to-text service.


Proposed solution:

Having timestamps returned with 3 decimal places (with millisecond precision) would enable the researchers to automate the longitudinal research with Watson Speech-to-Text service.


Benefits/Value:

Watson Speech-to-Text service returning timestamps with milliseconds precision will enable customers around the world to get more accurate transcriptions and use the service more confidently in research globally. With precision, researchers will be more likely to use the service and cite it in publications.


Users impacted:

Every user using Watson Speech-to-Text around the world will get more precise data points for transcription without their current downstream applications breaking due to this change.

Needed By Month
  • Admin
    Marco Noel
    Reply
    |
    May 9, 2022

    I'm not sure I understand the business value of adding an extra decimal to a timestamp. Please expand.

    Once we have more details, we will review but at this time, we are dealing with higher priorities.

  • Guest
    Reply
    |
    Mar 3, 2022
    Could I get an update on this please, this is a high priority request which will determine if the research will keep using IBM Watson or switch to a different Speech to text service.