Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]


Groups > sci.image.processing > #4200

Re: VOICE RECOGNITION BASED ON SPECTROGRAM

From Martin Leese <please@see.Web.for.e-mail.INVALID>
Newsgroups sci.image.processing
Subject Re: VOICE RECOGNITION BASED ON SPECTROGRAM
Date 2017-01-08 12:18 -0700
Organization A noiseless patient Spider
Message-ID <o4u379$i5i$1@dont-email.me> (permalink)
References <af02e83a-8247-417c-b287-9ab8517891a6@googlegroups.com>

Show all headers | View raw


manaswi.navin@gmail.com wrote:
> Can we find out total number of speakers and their duration by looking at/analysing spectrogram.!
> [image description]
> (https://drive.google.com/drive/folders/0B4rwzcsr5hevdEJlam9scTRodTg)

In general, no.  Different speakers can
use similar frequency ranges, so frequency
doesn't work for this.  The ear/brain uses
spatial processing (search for "cocktail
party effect").

>  By just looking at the image, I can see some pattern, but I am looking for right solution in terms of opencv code(python)

I can't, because I do not have permission
to view the file.

-- 
Regards,
Martin Leese
E-mail: please@see.Web.for.e-mail.INVALID
Web: http://members.tripod.com/martin_leese/

Back to sci.image.processing | Previous | NextPrevious in thread | Find similar | Unroll thread


Thread

VOICE RECOGNITION BASED ON SPECTROGRAM manaswi.navin@gmail.com - 2017-01-07 13:14 -0800
  Re: VOICE RECOGNITION BASED ON SPECTROGRAM Martin Leese <please@see.Web.for.e-mail.INVALID> - 2017-01-08 12:18 -0700

csiph-web