Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]
Groups > comp.compression > #550 > unrolled thread
| Started by | Nimo <azeez541@gmail.com> |
|---|---|
| First post | 2011-09-11 23:57 -0700 |
| Last post | 2011-09-18 14:08 +0200 |
| Articles | 20 — 10 participants |
Back to article view | Back to comp.compression
At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Nimo <azeez541@gmail.com> - 2011-09-11 23:57 -0700
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Jim Leonard <mobygamer@gmail.com> - 2011-09-13 07:52 -0700
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Nimo <azeez541@gmail.com> - 2011-09-13 19:36 -0700
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec tom st denis <tom@iahu.ca> - 2011-09-14 04:33 -0700
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Industrial One <industrial_one@hotmail.com> - 2011-09-14 06:01 -0700
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Industrial One <industrial_one@hotmail.com> - 2011-09-14 06:32 -0700
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Alex Mizrahi <alex.mizrahi@gmail.com> - 2011-09-14 18:00 +0300
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Industrial One <industrial_one@hotmail.com> - 2011-09-14 16:45 -0700
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec glen herrmannsfeldt <gah@ugcs.caltech.edu> - 2011-09-15 02:08 +0000
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Industrial One <industrial_one@hotmail.com> - 2011-09-15 07:28 -0700
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec glen herrmannsfeldt <gah@ugcs.caltech.edu> - 2011-09-15 17:43 +0000
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Industrial One <industrial_one@hotmail.com> - 2011-09-15 13:21 -0700
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Nimo <azeez541@gmail.com> - 2011-09-14 06:50 -0700
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Earl_Colby_Pottinger <earlcolby.pottinger@sympatico.ca> - 2011-09-14 12:22 -0700
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Willem <willem@toad.stack.nl> - 2011-09-14 19:39 +0000
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Peter Schepers <schepers@uwaterloo.ca> - 2011-09-14 15:47 -0400
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Jim Leonard <mobygamer@gmail.com> - 2011-09-14 13:26 -0700
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Nimo <azeez541@gmail.com> - 2011-09-14 06:51 -0700
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Jim Leonard <mobygamer@gmail.com> - 2011-09-14 07:12 -0700
Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Thomas Richter <thor@math.tu-berlin.de> - 2011-09-18 14:08 +0200
| From | Nimo <azeez541@gmail.com> |
|---|---|
| Date | 2011-09-11 23:57 -0700 |
| Subject | At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec |
| Message-ID | <d67ad7fd-1e5e-4b76-814b-77beaed51c57@l2g2000vbn.googlegroups.com> |
Hi there,
After a long time, back to my group.
well, will keep the stuff to the point, if you have any doubts,
I'm always here to help you..
// A New Speech Codec Based upon Advanced Tensor Basis & Galerkin
Techniques //
ALGORITHM BITRATE(s) MOS QUALITY SUBJECTIVE OPINION DELAY
*** 37 bps 5 Transparent
imperceptible 10th part delay in a sec
I'm getting CD quality "speech" at 37 bits / sec.
A checkmate to G.series stuff, AMR, speex etc.
Any Ideas, immediately, I mean as fast as possible I can go out with
this technology.
Important Links:-
Academic / Research
http://www.ircc.iitb.ac.in/IRCC-Webpage/patent273.jsp
Commercial
http://www.voiceage.com/index.php
http://www.sipro.com/
greetings
so long
nimo
This is to Thomas, pls shoot..
[toc] | [next] | [standalone]
| From | Jim Leonard <mobygamer@gmail.com> |
|---|---|
| Date | 2011-09-13 07:52 -0700 |
| Message-ID | <d9884c69-4047-4ef9-b449-2edd3a0e0e15@l4g2000vbv.googlegroups.com> |
| In reply to | #550 |
On Sep 12, 1:57 am, Nimo <azeez...@gmail.com> wrote: > I'm getting CD quality "speech" at 37 bits / sec. Functional example please?
[toc] | [prev] | [next] | [standalone]
| From | Nimo <azeez541@gmail.com> |
|---|---|
| Date | 2011-09-13 19:36 -0700 |
| Message-ID | <b3dbe73e-aabc-4d19-bafb-4409fc422f0e@j13g2000prj.googlegroups.com> |
| In reply to | #551 |
On Sep 13, 7:52 am, Jim Leonard <mobyga...@gmail.com> wrote: > On Sep 12, 1:57 am, Nimo <azeez...@gmail.com> wrote: > > > I'm getting CD quality "speech" at 37 bits / sec. > > Functional example please? Sorry, I didn't get you ...? so long nimo
[toc] | [prev] | [next] | [standalone]
| From | tom st denis <tom@iahu.ca> |
|---|---|
| Date | 2011-09-14 04:33 -0700 |
| Message-ID | <77448bd3-2dae-4d0f-b0f9-a5dab6cd1a30@d14g2000yqb.googlegroups.com> |
| In reply to | #553 |
On Sep 13, 10:36 pm, Nimo <azeez...@gmail.com> wrote: > On Sep 13, 7:52 am, Jim Leonard <mobyga...@gmail.com> wrote: > > > On Sep 12, 1:57 am, Nimo <azeez...@gmail.com> wrote: > > > > I'm getting CD quality "speech" at 37 bits / sec. > > > Functional example please? > > Sorry, I didn't get you ...? > > so long > nimo decoder + sample compressed stream please. Tom
[toc] | [prev] | [next] | [standalone]
| From | Industrial One <industrial_one@hotmail.com> |
|---|---|
| Date | 2011-09-14 06:01 -0700 |
| Message-ID | <2178b36c-71e6-4cac-baf2-f91cb1a4ace1@dq7g2000vbb.googlegroups.com> |
| In reply to | #554 |
There is no possible way speech can be encoded at any recognizable quality at only 37 kbps unless it was a text-to-speech routine because 37 kbps is barely enough to even encode flowing text losslessly.
[toc] | [prev] | [next] | [standalone]
| From | Industrial One <industrial_one@hotmail.com> |
|---|---|
| Date | 2011-09-14 06:32 -0700 |
| Message-ID | <c44284e1-c189-4094-84d2-420b6ea71650@x21g2000prd.googlegroups.com> |
| In reply to | #556 |
On Sep 14, 1:01 pm, Industrial One <industrial_...@hotmail.com> wrote: > There is no possible way speech can be encoded at any recognizable > quality at only 37 kbps unless it was a text-to-speech routine because > 37 kbps is barely enough to even encode flowing text losslessly. It appears I mispoke. That paragraph of mine above is exactly 200 bytes and takes about 14-15 seconds to recite aloud, that is about 14 bytes/s or 112 bits/s. Just where the hell do you expect to store the very complex information such as my voice, intonation, how often I pause, how fast I talk etc.?
[toc] | [prev] | [next] | [standalone]
| From | Alex Mizrahi <alex.mizrahi@gmail.com> |
|---|---|
| Date | 2011-09-14 18:00 +0300 |
| Message-ID | <4e70c196$0$302$14726298@news.sunsite.dk> |
| In reply to | #557 |
>> There is no possible way speech can be encoded at any recognizable >> quality at only 37 kbps unless it was a text-to-speech routine because >> 37 kbps is barely enough to even encode flowing text losslessly. > > It appears I mispoke. That paragraph of mine above is exactly 200 > bytes and takes about 14-15 seconds to recite aloud, that is about 14 > bytes/s or 112 bits/s. Just where the hell do you expect to store the > very complex information such as my voice, intonation, how often I > pause, how fast I talk etc.? We can put it in other way: 37 bits/second gives you 137*10^9 possible different seconds of speech. Does that match number of different sounds human can make in a second?
[toc] | [prev] | [next] | [standalone]
| From | Industrial One <industrial_one@hotmail.com> |
|---|---|
| Date | 2011-09-14 16:45 -0700 |
| Message-ID | <63d09335-e70a-42ed-a7c9-54bce6173069@y7g2000yqh.googlegroups.com> |
| In reply to | #560 |
On Sep 14, 3:00 pm, Alex Mizrahi <alex.mizr...@gmail.com> wrote: > >> There is no possible way speech can be encoded at any recognizable > >> quality at only 37 kbps unless it was a text-to-speech routine because > >> 37 kbps is barely enough to even encode flowing text losslessly. > > > It appears I mispoke. That paragraph of mine above is exactly 200 > > bytes and takes about 14-15 seconds to recite aloud, that is about 14 > > bytes/s or 112 bits/s. Just where the hell do you expect to store the > > very complex information such as my voice, intonation, how often I > > pause, how fast I talk etc.? > > We can put it in other way: 37 bits/second gives you 137*10^9 possible > different seconds of speech. Does that match number of different sounds > human can make in a second? Given there are 6 billion people on the planet each whom have their own unique voice and that a few words can fit into one second, where there are 50,000 common words in English alone, just one language out of many and that there are limitless different combinations of intonation, pauses, slurs and stutters, slowing down/speeding up speech I would say hell yeah there are way more than 137 billion different possible combinations in one second of speech. Text is not even possible to losslessly encode at 37 bps in typical cases and you believe it can be done with audio? Put the crackpipe down, nigga. You's hallucinatin'.
[toc] | [prev] | [next] | [standalone]
| From | glen herrmannsfeldt <gah@ugcs.caltech.edu> |
|---|---|
| Date | 2011-09-15 02:08 +0000 |
| Message-ID | <j4rmmr$fdu$1@speranza.aioe.org> |
| In reply to | #567 |
Industrial One <industrial_one@hotmail.com> wrote: (snip on audio voice compression to 37bits/s) > Given there are 6 billion people on the planet each whom have their > own unique voice and that a few words can fit into one second, where > there are 50,000 common words in English alone, just one language out > of many and that there are limitless different combinations of > intonation, pauses, slurs and stutters, slowing down/speeding up > speech I would say hell yeah there are way more than 137 billion > different possible combinations in one second of speech. Well, it only takes 33 bits to describe which of the 6 billion people is speaking. If some bits in the beginning describe the voice of the person speaking, those bits don't have to be resent for every word. So 37 is a little low, but if you only indicate phonemes, and previously the specifics of the voice of the specific person, it could be pretty low. > Text is not even possible to losslessly encode at 37 bps in typical > cases and you believe it can be done with audio? Put the crackpipe > down, nigga. You's hallucinatin'. Some people speak (and read) slower than others. -- glen
[toc] | [prev] | [next] | [standalone]
| From | Industrial One <industrial_one@hotmail.com> |
|---|---|
| Date | 2011-09-15 07:28 -0700 |
| Message-ID | <97525cf1-4205-4f7f-b4c6-8680acc1b44a@d14g2000yqb.googlegroups.com> |
| In reply to | #568 |
On Sep 15, 2:08 am, glen herrmannsfeldt <g...@ugcs.caltech.edu> wrote: > Industrial One <industrial_...@hotmail.com> wrote: > > (snip on audio voice compression to 37bits/s) > > > Given there are 6 billion people on the planet each whom have their > > own unique voice and that a few words can fit into one second, where > > there are 50,000 common words in English alone, just one language out > > of many and that there are limitless different combinations of > > intonation, pauses, slurs and stutters, slowing down/speeding up > > speech I would say hell yeah there are way more than 137 billion > > different possible combinations in one second of speech. > > Well, it only takes 33 bits to describe which of the 6 billion > people is speaking. You don't get it. 6 billion is not an upper limit, there could be 600 billion tomorrow and they would still have distinct voices. You can't compress contents just by indexing it for the same reason that you can't compress all 750,000 existing movies in the world to 20 bits. You would have to include the library containing the contents. In this case, you would need 6 billion 22 khz audio samples. That's about a 264 GB library + the 37 bits per second for whatever I wanna compress. Oh wait, it won't recognize the 6,000,000,001st person born tomorrow because his voice profile isn't in the library. Damn! Back to where we've started with a 22 khz mono .WAV recording which gives you the freedom to record whatever you want perfectly because it doesn't care about the content, only asks for a measly 352 kilobits of info per second. > If some bits in the beginning describe the > voice of the person speaking, those bits don't have to be resent > for every word. So 37 is a little low, but if you only indicate A voice profile would probably be at least 100 KB, just for the characteristics of the vocal chords. > phonemes, and previously the specifics of the voice of the specific > person, it could be pretty low. You would still be missing the intonation info of the person so they would end up sounding like a robotic text-to-speech program like Microsoft Sam. > > Text is not even possible to losslessly encode at 37 bps in typical > > cases and you believe it can be done with audio? Put the crackpipe > > down, nigga. You's hallucinatin'. > > Some people speak (and read) slower than others. > > -- glen At 37 bps it would take 43 seconds to read 40 words, thats about 3/4 of a second per syllable. Nobody except a retard talks that slow.
[toc] | [prev] | [next] | [standalone]
| From | glen herrmannsfeldt <gah@ugcs.caltech.edu> |
|---|---|
| Date | 2011-09-15 17:43 +0000 |
| Message-ID | <j4tdf4$4ed$1@speranza.aioe.org> |
| In reply to | #569 |
Industrial One <industrial_one@hotmail.com> wrote: (snip, and previous snip, on audio voice compression to 37bits/s) >> > Given there are 6 billion people on the planet each whom have their >> > own unique voice and that a few words can fit into one second, where >> > there are 50,000 common words in English alone, just one language out >> > of many and that there are limitless different combinations of >> > intonation, pauses, slurs and stutters, slowing down/speeding up >> > speech I would say hell yeah there are way more than 137 billion >> > different possible combinations in one second of speech. >> Well, it only takes 33 bits to describe which of the 6 billion >> people is speaking. > You don't get it. 6 billion is not an upper limit, there could be 600 > billion tomorrow and they would still have distinct voices. You can't > compress contents just by indexing it for the same reason that you > can't compress all 750,000 existing movies in the world to 20 bits. You do have to be careful as to what the problem is. You can compress the movies down if I happen to live next to a video store. Then you only need enough bits to tell me which DVD to grab. > You would have to include the library containing the contents. In this > case, you would need 6 billion 22 khz audio samples. That's about a > 264 GB library + the 37 bits per second for whatever I wanna compress. > Oh wait, it won't recognize the 6,000,000,001st person born tomorrow OK, lets ignore the OP's 37b/s and consider how close one can come with how many bits. >> If some bits in the beginning describe the >> voice of the person speaking, those bits don't have to be resent >> for every word. So 37 is a little low, but if you only indicate > A voice profile would probably be at least 100 KB, just for the > characteristics of the vocal chords. That sounds a little larger than I would have suggested, but it depends on how close you want to get. I will guess that you can get close enough for someone to recognize the person with less. >> phonemes, and previously the specifics of the voice of the specific >> person, it could be pretty low. > You would still be missing the intonation info of the person so they > would end up sounding like a robotic text-to-speech program like > Microsoft Sam. Or like Watson on Jeopardy! (rerun last night, if you missed it). So add some more bits for intonation. > At 37 bps it would take 43 seconds to read 40 words, thats about 3/4 > of a second per syllable. Nobody except a retard talks that slow. But it isn't off by a huge factor. Also, you can still do ordinary text compression on it. You can cache the vocal tract characteristics for future calls, too. -- glen
[toc] | [prev] | [next] | [standalone]
| From | Industrial One <industrial_one@hotmail.com> |
|---|---|
| Date | 2011-09-15 13:21 -0700 |
| Message-ID | <aec955c0-20f1-407f-b9df-eaf765d95b09@t29g2000vby.googlegroups.com> |
| In reply to | #570 |
On Sep 15, 5:43 pm, glen herrmannsfeldt <g...@ugcs.caltech.edu> wrote: > Industrial One <industrial_...@hotmail.com> wrote: > > (snip, and previous snip, on audio voice compression to 37bits/s) > > >> > Given there are 6 billion people on the planet each whom have their > >> > own unique voice and that a few words can fit into one second, where > >> > there are 50,000 common words in English alone, just one language out > >> > of many and that there are limitless different combinations of > >> > intonation, pauses, slurs and stutters, slowing down/speeding up > >> > speech I would say hell yeah there are way more than 137 billion > >> > different possible combinations in one second of speech. > >> Well, it only takes 33 bits to describe which of the 6 billion > >> people is speaking. > > You don't get it. 6 billion is not an upper limit, there could be 600 > > billion tomorrow and they would still have distinct voices. You can't > > compress contents just by indexing it for the same reason that you > > can't compress all 750,000 existing movies in the world to 20 bits. > > You do have to be careful as to what the problem is. You can > compress the movies down if I happen to live next to a video store. > Then you only need enough bits to tell me which DVD to grab. Doesn't change the fact that you've compressed nothing. The DVDs remain 4.7 gigs. > > You would have to include the library containing the contents. In this > > case, you would need 6 billion 22 khz audio samples. That's about a > > 264 GB library + the 37 bits per second for whatever I wanna compress. > > Oh wait, it won't recognize the 6,000,000,001st person born tomorrow > > OK, lets ignore the OP's 37b/s and consider how close one can > come with how many bits. You're still operating from a wrong premise. When you really get down to it you'll just end up where you've started and realize that losslessly you can only compress it by half and end up with 176 kbps. Think about it, how many samples per second do you need to reproduce high-quality sound for speech and catch even the highest-pitched queer voice? 22,050, that's already a bitrate in the kilobits. How many bits do you need per sample for a faithful amplitude resolution that will represent every possible loud or quiet element? 16 bits. 352,800 to index which of the possible 2^352800 combos our specific recorded sound is. These are some reality numbers for you, mang. There are many many people out there who you will never meet or listen to any of their speeches so naturally you wouldn't give a shit if your audio library wouldn't be able to compress their spoken words but the fact remains that they do exist, and 2^352800 potentially exist, not 2^37. The minute your compressor discriminates, the minute it fails. > >> phonemes, and previously the specifics of the voice of the specific > >> person, it could be pretty low. > > You would still be missing the intonation info of the person so they > > would end up sounding like a robotic text-to-speech program like > > Microsoft Sam. > > Or like Watson on Jeopardy! (rerun last night, if you missed it). > > So add some more bits for intonation. That would be a hell of a lot of bits. Intonation is highly complex, context-dependant and for the most part unique to each person. That Watson robot isn't even the best example of robotic speech as his voice was clearly programmed with modern techniques to make him sound as natural as possible, his intonation is noticeably reduced but not completely lacking. > > At 37 bps it would take 43 seconds to read 40 words, thats about 3/4 > > of a second per syllable. Nobody except a retard talks that slow. > > But it isn't off by a huge factor. Also, you can still do > ordinary text compression on it. Last I recall, 7-zip with maximum settings only compresses text by about half. 37 bps is a reading speed of less than one word per second, and even two words per second is slow bordering on legally retarded. > You can cache the vocal tract characteristics for future calls, too. Irrelevant.
[toc] | [prev] | [next] | [standalone]
| From | Nimo <azeez541@gmail.com> |
|---|---|
| Date | 2011-09-14 06:50 -0700 |
| Message-ID | <d2c44a81-3c0c-41bb-a8ef-b235921b4348@f24g2000prb.googlegroups.com> |
| In reply to | #557 |
On Sep 14, 6:32 am, Industrial One <industrial_...@hotmail.com> wrote: > On Sep 14, 1:01 pm, Industrial One <industrial_...@hotmail.com> wrote: > > > There is no possible way speech can be encoded at any recognizable > > quality at only 37 kbps unless it was a text-to-speech routine because > > 37 kbps is barely enough to even encode flowing text losslessly. > > It appears I mispoke. That paragraph of mine above is exactly 200 > bytes and takes about 14-15 seconds to recite aloud, that is about 14 > bytes/s or 112 bits/s. Just where the hell do you expect to store the > very complex information such as my voice, intonation, how often I > pause, how fast I talk etc.? 1. When a distinguished but elderly scientist states that something is possible, he is almost certainly right. When he states that something is impossible, he is very probably wrong. 2.The only way of discovering the limits of the possible is to venture a little way past them into the impossible. 3.Any sufficiently advanced technology is indistinguishable from magic. Clarke's three laws. wait few days(yes, not weeks just daysl. ICQ is going to be in history ..)
[toc] | [prev] | [next] | [standalone]
| From | Earl_Colby_Pottinger <earlcolby.pottinger@sympatico.ca> |
|---|---|
| Date | 2011-09-14 12:22 -0700 |
| Message-ID | <60ffb2f7-9d29-4e55-93d1-13a593128e78@k15g2000yqd.googlegroups.com> |
| In reply to | #561 |
On Sep 14, 9:50 am, Nimo <azeez...@gmail.com> wrote: > On Sep 14, 6:32 am, Industrial One <industrial_...@hotmail.com> wrote: > > > On Sep 14, 1:01 pm, Industrial One <industrial_...@hotmail.com> wrote: > > > > There is no possible way speech can be encoded at any recognizable > > > quality at only 37 kbps unless it was a text-to-speech routine because > > > 37 kbps is barely enough to even encode flowing text losslessly. > > > It appears I mispoke. That paragraph of mine above is exactly 200 > > bytes and takes about 14-15 seconds to recite aloud, that is about 14 > > bytes/s or 112 bits/s. Just where the hell do you expect to store the > > very complex information such as my voice, intonation, how often I > > pause, how fast I talk etc.? > > 1. When a distinguished but elderly scientist states that something is > possible, > he is almost certainly right. When he states that something is > impossible, he is very probably wrong. > 2.The only way of discovering the limits of the possible is to venture > a little way past them into the impossible. > 3.Any sufficiently advanced technology is indistinguishable from > magic. > > Clarke's three laws. > > wait few days(yes, not weeks just daysl. ICQ is going to be in > history ..) Clarke's Laws are over-ridden by the Idiom of 'Fool me once, shame on you; fool me twice, shame on me', lame claims like your's always fail - ALWAYS! Why if you really had something did you not prepare in advance and have it ready before announcing it? If you only needed days, why not wait those few days be posting. Like all the lamers before you, you hoped to get people praising you over something you never had. I predict that a week from now you will have some weak excuse why you can't release a test suite. Lamers like you always do that all the time, year after year. Words are just hot air, working code is KING!
[toc] | [prev] | [next] | [standalone]
| From | Willem <willem@toad.stack.nl> |
|---|---|
| Date | 2011-09-14 19:39 +0000 |
| Message-ID | <slrnj720o8.2c9q.willem@toad.stack.nl> |
| In reply to | #563 |
Earl_Colby_Pottinger wrote:
) On Sep 14, 9:50?am, Nimo <azeez...@gmail.com> wrote:
)> 1. When a distinguished but elderly scientist states that something is
)> possible,
)> ?he is almost certainly right. When he states that something is
)> impossible, he is very probably wrong.
)> 2.The only way of discovering the limits of the possible is to venture
)> a little way past them into the impossible.
)> 3.Any sufficiently advanced technology is indistinguishable from
)> magic.
)>
)> Clarke's three laws.
)>
)> ? ?wait few days(yes, not weeks just daysl. ICQ is going to be in
)> history ..)
)
) Clarke's Laws are over-ridden by the Idiom of 'Fool me once, shame on
) you; fool me twice, shame on me', lame claims like your's always fail
) - ALWAYS!
)
) Why if you really had something did you not prepare in advance and
) have it ready before announcing it? If you only needed days, why not
) wait those few days be posting.
)
) Like all the lamers before you, you hoped to get people praising you
) over something you never had. I predict that a week from now you will
) have some weak excuse why you can't release a test suite. Lamers
) like you always do that all the time, year after year.
)
) Words are just hot air, working code is KING!
Speech compression at 37kbps seems quite plausible, although
the number 37 seems a bit arbitrary. I would expect 32 or 40.
ADPCM speech compression does 12, 24, 32 or 40kbps, for example.
I'd worry more about the 'CD quality' claim, which is rather more
subjective.
SaSW, Willem
--
Disclaimer: I am in no way responsible for any of the statements
made in the above text. For all I know I might be
drugged or something..
No I'm not paranoid. You all think I'm paranoid, don't you !
#EOT
[toc] | [prev] | [next] | [standalone]
| From | Peter Schepers <schepers@uwaterloo.ca> |
|---|---|
| Date | 2011-09-14 15:47 -0400 |
| Message-ID | <j4r0c9$ajd$1@rumours.uwaterloo.ca> |
| In reply to | #564 |
On 14/09/2011 3:39 PM, Willem wrote: > Speech compression at 37kbps seems quite plausible, although > the number 37 seems a bit arbitrary. I would expect 32 or 40. But the claim is 37 _bits_ per second (as the subject header says), not kbps. PS
[toc] | [prev] | [next] | [standalone]
| From | Jim Leonard <mobygamer@gmail.com> |
|---|---|
| Date | 2011-09-14 13:26 -0700 |
| Message-ID | <2862cce2-8555-4c32-9471-085ac0a8fde0@i39g2000yqn.googlegroups.com> |
| In reply to | #561 |
On Sep 14, 8:50 am, Nimo <azeez...@gmail.com> wrote: > Clarke's three laws. Ah yes, the wonderful and talented science FICTION writer. Did you offer the above as proof your claims are fiction?
[toc] | [prev] | [next] | [standalone]
| From | Nimo <azeez541@gmail.com> |
|---|---|
| Date | 2011-09-14 06:51 -0700 |
| Message-ID | <33ba8f5d-9cbc-4507-a734-c20779630853@s15g2000pre.googlegroups.com> |
| In reply to | #554 |
On Sep 14, 4:33 am, tom st denis <t...@iahu.ca> wrote: > On Sep 13, 10:36 pm, Nimo <azeez...@gmail.com> wrote: > > > On Sep 13, 7:52 am, Jim Leonard <mobyga...@gmail.com> wrote: > > > > On Sep 12, 1:57 am, Nimo <azeez...@gmail.com> wrote: > > > > > I'm getting CD quality "speech" at 37 bits / sec. > > > > Functional example please? > > > Sorry, I didn't get you ...? > > > so long > > nimo > > decoder + sample compressed stream please. > > Tom wait baby, getting US patent and then game starts..
[toc] | [prev] | [next] | [standalone]
| From | Jim Leonard <mobygamer@gmail.com> |
|---|---|
| Date | 2011-09-14 07:12 -0700 |
| Message-ID | <6e90f5d2-5f02-4316-9965-48ca6d5a87f1@m38g2000vbn.googlegroups.com> |
| In reply to | #558 |
On Sep 14, 8:51 am, Nimo <azeez...@gmail.com> wrote: > > > decoder + sample compressed stream please. > > wait baby, getting US patent and then game starts.. I think you'll save yourself a lot of grief and heartache if you create a decompressor + sample compressed stream *before* you bother with patents. Until you do, you have nothing worth protecting with a patent.
[toc] | [prev] | [next] | [standalone]
| From | Thomas Richter <thor@math.tu-berlin.de> |
|---|---|
| Date | 2011-09-18 14:08 +0200 |
| Message-ID | <j54n0h$256$1@news.belwue.de> |
| In reply to | #550 |
On 12.09.2011 08:57, Nimo wrote: > Hi there, > > After a long time, back to my group. > > well, will keep the stuff to the point, if you have any doubts, > I'm always here to help you.. > > // A New Speech Codec Based upon Advanced Tensor Basis& Galerkin > Techniques // > > > ALGORITHM BITRATE(s) MOS QUALITY SUBJECTIVE OPINION DELAY > > *** 37 bps 5 Transparent > imperceptible 10th part delay in a sec > > > > I'm getting CD quality "speech" at 37 bits / sec. > > > A checkmate to G.series stuff, AMR, speex etc. > > > Any Ideas, immediately, I mean as fast as possible I can go out with > this technology. Does this refer to the link you gave? I don't see a MOS of 5 on this patent. But allow me to make a couple of comments: - If you provide a MOS score, you should specifically say what the task was you gave to the observers: Understand the words (text-to-speech would suffice), get the pronounciation (vocoding possible), recognize the speaker? - The MOS scores are very sloppy, no error bars. Unclear how many observers were used to perform the experiments. What is "CD speech" quality? The quality of a vocoder can be very high, yet it might be not what has been asked for. The G.series are for natural speech representation in a non-parametric way, allowing to identify the speaker etc... Of course parametric coding is known, and text-to-speech is known, but still, if I make a phone call, I would be very disappointed if all I get would be a vocoder I talk to, even if the quality is very high, and thus none of these techniques ended the ITU-T G-series of standards. Thus, at least, I can accuse the original "inventors" of making very unscientific comparisons, or not telling me what the actual intend of the work should be. Greetings, Thomas
[toc] | [prev] | [standalone]
Back to top | Article view | comp.compression
csiph-web