Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]


Groups > comp.compression > #550 > unrolled thread

At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec

Started byNimo <azeez541@gmail.com>
First post2011-09-11 23:57 -0700
Last post2011-09-18 14:08 +0200
Articles 20 — 10 participants

Back to article view | Back to comp.compression


Contents

  At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Nimo <azeez541@gmail.com> - 2011-09-11 23:57 -0700
    Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Jim Leonard <mobygamer@gmail.com> - 2011-09-13 07:52 -0700
      Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Nimo <azeez541@gmail.com> - 2011-09-13 19:36 -0700
        Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec tom st denis <tom@iahu.ca> - 2011-09-14 04:33 -0700
          Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Industrial One <industrial_one@hotmail.com> - 2011-09-14 06:01 -0700
            Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Industrial One <industrial_one@hotmail.com> - 2011-09-14 06:32 -0700
              Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Alex Mizrahi <alex.mizrahi@gmail.com> - 2011-09-14 18:00 +0300
                Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Industrial One <industrial_one@hotmail.com> - 2011-09-14 16:45 -0700
                  Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec glen herrmannsfeldt <gah@ugcs.caltech.edu> - 2011-09-15 02:08 +0000
                    Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Industrial One <industrial_one@hotmail.com> - 2011-09-15 07:28 -0700
                      Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec glen herrmannsfeldt <gah@ugcs.caltech.edu> - 2011-09-15 17:43 +0000
                        Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Industrial One <industrial_one@hotmail.com> - 2011-09-15 13:21 -0700
              Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Nimo <azeez541@gmail.com> - 2011-09-14 06:50 -0700
                Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Earl_Colby_Pottinger <earlcolby.pottinger@sympatico.ca> - 2011-09-14 12:22 -0700
                  Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Willem <willem@toad.stack.nl> - 2011-09-14 19:39 +0000
                    Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Peter Schepers <schepers@uwaterloo.ca> - 2011-09-14 15:47 -0400
                Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Jim Leonard <mobygamer@gmail.com> - 2011-09-14 13:26 -0700
          Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Nimo <azeez541@gmail.com> - 2011-09-14 06:51 -0700
            Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Jim Leonard <mobygamer@gmail.com> - 2011-09-14 07:12 -0700
    Re: At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec Thomas Richter <thor@math.tu-berlin.de> - 2011-09-18 14:08 +0200

#550 — At 37bits/sec, A Wide band ( upto 32KHz) Speech Codec

FromNimo <azeez541@gmail.com>
Date2011-09-11 23:57 -0700
SubjectAt 37bits/sec, A Wide band ( upto 32KHz) Speech Codec
Message-ID<d67ad7fd-1e5e-4b76-814b-77beaed51c57@l2g2000vbn.googlegroups.com>
Hi there,

   After a long time, back to my group.

well, will keep the stuff to the point, if you have any doubts,
I'm always here to help you..

// A New Speech Codec Based upon  Advanced Tensor Basis & Galerkin
Techniques //


ALGORITHM   BITRATE(s)  MOS   QUALITY    SUBJECTIVE OPINION  DELAY

   ***                 37 bps        5       Transparent
imperceptible    10th part delay in a sec



I'm getting CD quality "speech" at 37 bits / sec.


      A checkmate to G.series stuff, AMR, speex etc.


Any Ideas,  immediately, I mean as fast as possible I can go out with
this technology.


  Important Links:-

Academic / Research
 http://www.ircc.iitb.ac.in/IRCC-Webpage/patent273.jsp

Commercial
http://www.voiceage.com/index.php
http://www.sipro.com/

greetings
  so long
    nimo
This is to Thomas, pls shoot..

[toc] | [next] | [standalone]


#551

FromJim Leonard <mobygamer@gmail.com>
Date2011-09-13 07:52 -0700
Message-ID<d9884c69-4047-4ef9-b449-2edd3a0e0e15@l4g2000vbv.googlegroups.com>
In reply to#550
On Sep 12, 1:57 am, Nimo <azeez...@gmail.com> wrote:
> I'm getting CD quality "speech" at 37 bits / sec.

Functional example please?

[toc] | [prev] | [next] | [standalone]


#553

FromNimo <azeez541@gmail.com>
Date2011-09-13 19:36 -0700
Message-ID<b3dbe73e-aabc-4d19-bafb-4409fc422f0e@j13g2000prj.googlegroups.com>
In reply to#551
On Sep 13, 7:52 am, Jim Leonard <mobyga...@gmail.com> wrote:
> On Sep 12, 1:57 am, Nimo <azeez...@gmail.com> wrote:
>
> > I'm getting CD quality "speech" at 37 bits / sec.
>
> Functional example please?

Sorry, I didn't get you ...?


so long
  nimo

[toc] | [prev] | [next] | [standalone]


#554

Fromtom st denis <tom@iahu.ca>
Date2011-09-14 04:33 -0700
Message-ID<77448bd3-2dae-4d0f-b0f9-a5dab6cd1a30@d14g2000yqb.googlegroups.com>
In reply to#553
On Sep 13, 10:36 pm, Nimo <azeez...@gmail.com> wrote:
> On Sep 13, 7:52 am, Jim Leonard <mobyga...@gmail.com> wrote:
>
> > On Sep 12, 1:57 am, Nimo <azeez...@gmail.com> wrote:
>
> > > I'm getting CD quality "speech" at 37 bits / sec.
>
> > Functional example please?
>
> Sorry, I didn't get you ...?
>
> so long
>   nimo

decoder + sample compressed stream please.

Tom

[toc] | [prev] | [next] | [standalone]


#556

FromIndustrial One <industrial_one@hotmail.com>
Date2011-09-14 06:01 -0700
Message-ID<2178b36c-71e6-4cac-baf2-f91cb1a4ace1@dq7g2000vbb.googlegroups.com>
In reply to#554
There is no possible way speech can be encoded at any recognizable
quality at only 37 kbps unless it was a text-to-speech routine because
37 kbps is barely enough to even encode flowing text losslessly.

[toc] | [prev] | [next] | [standalone]


#557

FromIndustrial One <industrial_one@hotmail.com>
Date2011-09-14 06:32 -0700
Message-ID<c44284e1-c189-4094-84d2-420b6ea71650@x21g2000prd.googlegroups.com>
In reply to#556
On Sep 14, 1:01 pm, Industrial One <industrial_...@hotmail.com> wrote:
> There is no possible way speech can be encoded at any recognizable
> quality at only 37 kbps unless it was a text-to-speech routine because
> 37 kbps is barely enough to even encode flowing text losslessly.

It appears I mispoke. That paragraph of mine above is exactly 200
bytes and takes about 14-15 seconds to recite aloud, that is about 14
bytes/s or 112 bits/s. Just where the hell do you expect to store the
very complex information such as my voice, intonation, how often I
pause, how fast I talk etc.?

[toc] | [prev] | [next] | [standalone]


#560

FromAlex Mizrahi <alex.mizrahi@gmail.com>
Date2011-09-14 18:00 +0300
Message-ID<4e70c196$0$302$14726298@news.sunsite.dk>
In reply to#557
>> There is no possible way speech can be encoded at any recognizable
>> quality at only 37 kbps unless it was a text-to-speech routine because
>> 37 kbps is barely enough to even encode flowing text losslessly.
>
> It appears I mispoke. That paragraph of mine above is exactly 200
> bytes and takes about 14-15 seconds to recite aloud, that is about 14
> bytes/s or 112 bits/s. Just where the hell do you expect to store the
> very complex information such as my voice, intonation, how often I
> pause, how fast I talk etc.?

We can put it in other way: 37 bits/second gives you 137*10^9 possible 
different seconds of speech. Does that match number of different sounds 
human can make in a second?

[toc] | [prev] | [next] | [standalone]


#567

FromIndustrial One <industrial_one@hotmail.com>
Date2011-09-14 16:45 -0700
Message-ID<63d09335-e70a-42ed-a7c9-54bce6173069@y7g2000yqh.googlegroups.com>
In reply to#560
On Sep 14, 3:00 pm, Alex Mizrahi <alex.mizr...@gmail.com> wrote:
> >> There is no possible way speech can be encoded at any recognizable
> >> quality at only 37 kbps unless it was a text-to-speech routine because
> >> 37 kbps is barely enough to even encode flowing text losslessly.
>
> > It appears I mispoke. That paragraph of mine above is exactly 200
> > bytes and takes about 14-15 seconds to recite aloud, that is about 14
> > bytes/s or 112 bits/s. Just where the hell do you expect to store the
> > very complex information such as my voice, intonation, how often I
> > pause, how fast I talk etc.?
>
> We can put it in other way: 37 bits/second gives you 137*10^9 possible
> different seconds of speech. Does that match number of different sounds
> human can make in a second?

Given there are 6 billion people on the planet each whom have their
own unique voice and that a few words can fit into one second, where
there are 50,000 common words in English alone, just one language out
of many and that there are limitless different combinations of
intonation, pauses, slurs and stutters, slowing down/speeding up
speech I would say hell yeah there are way more than 137 billion
different possible combinations in one second of speech.

Text is not even possible to losslessly encode at 37 bps in typical
cases and you believe it can be done with audio? Put the crackpipe
down, nigga. You's hallucinatin'.

[toc] | [prev] | [next] | [standalone]


#568

Fromglen herrmannsfeldt <gah@ugcs.caltech.edu>
Date2011-09-15 02:08 +0000
Message-ID<j4rmmr$fdu$1@speranza.aioe.org>
In reply to#567
Industrial One <industrial_one@hotmail.com> wrote:

(snip on audio voice compression to 37bits/s)

> Given there are 6 billion people on the planet each whom have their
> own unique voice and that a few words can fit into one second, where
> there are 50,000 common words in English alone, just one language out
> of many and that there are limitless different combinations of
> intonation, pauses, slurs and stutters, slowing down/speeding up
> speech I would say hell yeah there are way more than 137 billion
> different possible combinations in one second of speech.

Well, it only takes 33 bits to describe which of the 6 billion
people is speaking.  If some bits in the beginning describe the
voice of the person speaking, those bits don't have to be resent
for every word.  So 37 is a little low, but if you only indicate
phonemes, and previously the specifics of the voice of the specific
person, it could be pretty low.

> Text is not even possible to losslessly encode at 37 bps in typical
> cases and you believe it can be done with audio? Put the crackpipe
> down, nigga. You's hallucinatin'.

Some people speak (and read) slower than others.

-- glen

[toc] | [prev] | [next] | [standalone]


#569

FromIndustrial One <industrial_one@hotmail.com>
Date2011-09-15 07:28 -0700
Message-ID<97525cf1-4205-4f7f-b4c6-8680acc1b44a@d14g2000yqb.googlegroups.com>
In reply to#568
On Sep 15, 2:08 am, glen herrmannsfeldt <g...@ugcs.caltech.edu> wrote:
> Industrial One <industrial_...@hotmail.com> wrote:
>
> (snip on audio voice compression to 37bits/s)
>
> > Given there are 6 billion people on the planet each whom have their
> > own unique voice and that a few words can fit into one second, where
> > there are 50,000 common words in English alone, just one language out
> > of many and that there are limitless different combinations of
> > intonation, pauses, slurs and stutters, slowing down/speeding up
> > speech I would say hell yeah there are way more than 137 billion
> > different possible combinations in one second of speech.
>
> Well, it only takes 33 bits to describe which of the 6 billion
> people is speaking.

You don't get it. 6 billion is not an upper limit, there could be 600
billion tomorrow and they would still have distinct voices. You can't
compress contents just by indexing it for the same reason that you
can't compress all 750,000 existing movies in the world to 20 bits.
You would have to include the library containing the contents. In this
case, you would need 6 billion 22 khz audio samples. That's about a
264 GB library + the 37 bits per second for whatever I wanna compress.
Oh wait, it won't recognize the 6,000,000,001st person born tomorrow
because his voice profile isn't in the library. Damn! Back to where
we've started with a 22 khz mono .WAV recording which gives you the
freedom to record whatever you want perfectly because it doesn't care
about the content, only asks for a measly 352 kilobits of info per
second.

>  If some bits in the beginning describe the
> voice of the person speaking, those bits don't have to be resent
> for every word.  So 37 is a little low, but if you only indicate

A voice profile would probably be at least 100 KB, just for the
characteristics of the vocal chords.

> phonemes, and previously the specifics of the voice of the specific
> person, it could be pretty low.

You would still be missing the intonation info of the person so they
would end up sounding like a robotic text-to-speech program like
Microsoft Sam.

> > Text is not even possible to losslessly encode at 37 bps in typical
> > cases and you believe it can be done with audio? Put the crackpipe
> > down, nigga. You's hallucinatin'.
>
> Some people speak (and read) slower than others.
>
> -- glen

At 37 bps it would take 43 seconds to read 40 words, thats about 3/4
of a second per syllable. Nobody except a retard talks that slow.

[toc] | [prev] | [next] | [standalone]


#570

Fromglen herrmannsfeldt <gah@ugcs.caltech.edu>
Date2011-09-15 17:43 +0000
Message-ID<j4tdf4$4ed$1@speranza.aioe.org>
In reply to#569
Industrial One <industrial_one@hotmail.com> wrote:

(snip, and previous snip, on audio voice compression to 37bits/s)

>> > Given there are 6 billion people on the planet each whom have their
>> > own unique voice and that a few words can fit into one second, where
>> > there are 50,000 common words in English alone, just one language out
>> > of many and that there are limitless different combinations of
>> > intonation, pauses, slurs and stutters, slowing down/speeding up
>> > speech I would say hell yeah there are way more than 137 billion
>> > different possible combinations in one second of speech.

>> Well, it only takes 33 bits to describe which of the 6 billion
>> people is speaking.

> You don't get it. 6 billion is not an upper limit, there could be 600
> billion tomorrow and they would still have distinct voices. You can't
> compress contents just by indexing it for the same reason that you
> can't compress all 750,000 existing movies in the world to 20 bits.

You do have to be careful as to what the problem is.  You can
compress the movies down if I happen to live next to a video store.
Then you only need enough bits to tell me which DVD to grab.

> You would have to include the library containing the contents. In this
> case, you would need 6 billion 22 khz audio samples. That's about a
> 264 GB library + the 37 bits per second for whatever I wanna compress.
> Oh wait, it won't recognize the 6,000,000,001st person born tomorrow

OK, lets ignore the OP's 37b/s and consider how close one can
come with how many bits.

>> If some bits in the beginning describe the
>> voice of the person speaking, those bits don't have to be resent
>> for every word.  So 37 is a little low, but if you only indicate

> A voice profile would probably be at least 100 KB, just for the
> characteristics of the vocal chords.

That sounds a little larger than I would have suggested, but it
depends on how close you want to get.  I will guess that you can
get close enough for someone to recognize the person with less.

>> phonemes, and previously the specifics of the voice of the specific
>> person, it could be pretty low.

> You would still be missing the intonation info of the person so they
> would end up sounding like a robotic text-to-speech program like
> Microsoft Sam.

Or like Watson on Jeopardy! (rerun last night, if you missed it).

So add some more bits for intonation.  

> At 37 bps it would take 43 seconds to read 40 words, thats about 3/4
> of a second per syllable. Nobody except a retard talks that slow.

But it isn't off by a huge factor.  Also, you can still do 
ordinary text compression on it.

You can cache the vocal tract characteristics for future calls, too.

-- glen

[toc] | [prev] | [next] | [standalone]


#571

FromIndustrial One <industrial_one@hotmail.com>
Date2011-09-15 13:21 -0700
Message-ID<aec955c0-20f1-407f-b9df-eaf765d95b09@t29g2000vby.googlegroups.com>
In reply to#570
On Sep 15, 5:43 pm, glen herrmannsfeldt <g...@ugcs.caltech.edu> wrote:
> Industrial One <industrial_...@hotmail.com> wrote:
>
> (snip, and previous snip, on audio voice compression to 37bits/s)
>
> >> > Given there are 6 billion people on the planet each whom have their
> >> > own unique voice and that a few words can fit into one second, where
> >> > there are 50,000 common words in English alone, just one language out
> >> > of many and that there are limitless different combinations of
> >> > intonation, pauses, slurs and stutters, slowing down/speeding up
> >> > speech I would say hell yeah there are way more than 137 billion
> >> > different possible combinations in one second of speech.
> >> Well, it only takes 33 bits to describe which of the 6 billion
> >> people is speaking.
> > You don't get it. 6 billion is not an upper limit, there could be 600
> > billion tomorrow and they would still have distinct voices. You can't
> > compress contents just by indexing it for the same reason that you
> > can't compress all 750,000 existing movies in the world to 20 bits.
>
> You do have to be careful as to what the problem is.  You can
> compress the movies down if I happen to live next to a video store.
> Then you only need enough bits to tell me which DVD to grab.

Doesn't change the fact that you've compressed nothing. The DVDs
remain 4.7 gigs.

> > You would have to include the library containing the contents. In this
> > case, you would need 6 billion 22 khz audio samples. That's about a
> > 264 GB library + the 37 bits per second for whatever I wanna compress.
> > Oh wait, it won't recognize the 6,000,000,001st person born tomorrow
>
> OK, lets ignore the OP's 37b/s and consider how close one can
> come with how many bits.

You're still operating from a wrong premise. When you really get down
to it you'll just end up where you've started and realize that
losslessly you can only compress it by half and end up with 176 kbps.

Think about it, how many samples per second do you need to reproduce
high-quality sound for speech and catch even the highest-pitched queer
voice? 22,050, that's already a bitrate in the kilobits. How many bits
do you need per sample for a faithful amplitude resolution that will
represent every possible loud or quiet element? 16 bits. 352,800 to
index which of the possible 2^352800 combos our specific recorded
sound is. These are some reality numbers for you, mang.

There are many many people out there who you will never meet or listen
to any of their speeches so naturally you wouldn't give a shit if your
audio library wouldn't be able to compress their spoken words but the
fact remains that they do exist, and 2^352800 potentially exist, not
2^37. The minute your compressor discriminates, the minute it fails.

> >> phonemes, and previously the specifics of the voice of the specific
> >> person, it could be pretty low.
> > You would still be missing the intonation info of the person so they
> > would end up sounding like a robotic text-to-speech program like
> > Microsoft Sam.
>
> Or like Watson on Jeopardy! (rerun last night, if you missed it).
>
> So add some more bits for intonation.  

That would be a hell of a lot of bits. Intonation is highly complex,
context-dependant and for the most part unique to each person. That
Watson robot isn't even the best example of robotic speech as his
voice was clearly programmed with modern techniques to make him sound
as natural as possible, his intonation is noticeably reduced but not
completely lacking.

> > At 37 bps it would take 43 seconds to read 40 words, thats about 3/4
> > of a second per syllable. Nobody except a retard talks that slow.
>
> But it isn't off by a huge factor.  Also, you can still do
> ordinary text compression on it.

Last I recall, 7-zip with maximum settings only compresses text by
about half. 37 bps is a reading speed of less than one word per
second, and even two words per second is slow bordering on legally
retarded.

> You can cache the vocal tract characteristics for future calls, too.

Irrelevant.

[toc] | [prev] | [next] | [standalone]


#561

FromNimo <azeez541@gmail.com>
Date2011-09-14 06:50 -0700
Message-ID<d2c44a81-3c0c-41bb-a8ef-b235921b4348@f24g2000prb.googlegroups.com>
In reply to#557
On Sep 14, 6:32 am, Industrial One <industrial_...@hotmail.com> wrote:
> On Sep 14, 1:01 pm, Industrial One <industrial_...@hotmail.com> wrote:
>
> > There is no possible way speech can be encoded at any recognizable
> > quality at only 37 kbps unless it was a text-to-speech routine because
> > 37 kbps is barely enough to even encode flowing text losslessly.
>
> It appears I mispoke. That paragraph of mine above is exactly 200
> bytes and takes about 14-15 seconds to recite aloud, that is about 14
> bytes/s or 112 bits/s. Just where the hell do you expect to store the
> very complex information such as my voice, intonation, how often I
> pause, how fast I talk etc.?

1. When a distinguished but elderly scientist states that something is
possible,
 he is almost certainly right. When he states that something is
impossible, he is very probably wrong.
2.The only way of discovering the limits of the possible is to venture
a little way past them into the impossible.
3.Any sufficiently advanced technology is indistinguishable from
magic.

Clarke's three laws.

   wait few days(yes, not weeks just daysl. ICQ is going to be in
history ..)

[toc] | [prev] | [next] | [standalone]


#563

FromEarl_Colby_Pottinger <earlcolby.pottinger@sympatico.ca>
Date2011-09-14 12:22 -0700
Message-ID<60ffb2f7-9d29-4e55-93d1-13a593128e78@k15g2000yqd.googlegroups.com>
In reply to#561
On Sep 14, 9:50 am, Nimo <azeez...@gmail.com> wrote:
> On Sep 14, 6:32 am, Industrial One <industrial_...@hotmail.com> wrote:
>
> > On Sep 14, 1:01 pm, Industrial One <industrial_...@hotmail.com> wrote:
>
> > > There is no possible way speech can be encoded at any recognizable
> > > quality at only 37 kbps unless it was a text-to-speech routine because
> > > 37 kbps is barely enough to even encode flowing text losslessly.
>
> > It appears I mispoke. That paragraph of mine above is exactly 200
> > bytes and takes about 14-15 seconds to recite aloud, that is about 14
> > bytes/s or 112 bits/s. Just where the hell do you expect to store the
> > very complex information such as my voice, intonation, how often I
> > pause, how fast I talk etc.?
>
> 1. When a distinguished but elderly scientist states that something is
> possible,
>  he is almost certainly right. When he states that something is
> impossible, he is very probably wrong.
> 2.The only way of discovering the limits of the possible is to venture
> a little way past them into the impossible.
> 3.Any sufficiently advanced technology is indistinguishable from
> magic.
>
> Clarke's three laws.
>
>    wait few days(yes, not weeks just daysl. ICQ is going to be in
> history ..)

Clarke's Laws are over-ridden by the Idiom of 'Fool me once, shame on
you; fool me twice, shame on me', lame claims like your's always fail
- ALWAYS!

Why if you really had something did you not prepare in advance and
have it ready before announcing it?  If you only needed days, why not
wait those few days be posting.

Like all the lamers before you, you hoped to get people praising you
over something you never had.  I predict that a week from now you will
have some weak excuse why you can't release a test suite.   Lamers
like you always do that all the time, year after year.

Words are just hot air, working code is KING!

[toc] | [prev] | [next] | [standalone]


#564

FromWillem <willem@toad.stack.nl>
Date2011-09-14 19:39 +0000
Message-ID<slrnj720o8.2c9q.willem@toad.stack.nl>
In reply to#563
Earl_Colby_Pottinger wrote:
) On Sep 14, 9:50?am, Nimo <azeez...@gmail.com> wrote:
)> 1. When a distinguished but elderly scientist states that something is
)> possible,
)> ?he is almost certainly right. When he states that something is
)> impossible, he is very probably wrong.
)> 2.The only way of discovering the limits of the possible is to venture
)> a little way past them into the impossible.
)> 3.Any sufficiently advanced technology is indistinguishable from
)> magic.
)>
)> Clarke's three laws.
)>
)> ? ?wait few days(yes, not weeks just daysl. ICQ is going to be in
)> history ..)
)
) Clarke's Laws are over-ridden by the Idiom of 'Fool me once, shame on
) you; fool me twice, shame on me', lame claims like your's always fail
) - ALWAYS!
)
) Why if you really had something did you not prepare in advance and
) have it ready before announcing it?  If you only needed days, why not
) wait those few days be posting.
)
) Like all the lamers before you, you hoped to get people praising you
) over something you never had.  I predict that a week from now you will
) have some weak excuse why you can't release a test suite.   Lamers
) like you always do that all the time, year after year.
)
) Words are just hot air, working code is KING!

Speech compression at 37kbps seems quite plausible, although
the number 37 seems a bit arbitrary.  I would expect 32 or 40.

ADPCM speech compression does 12, 24, 32 or 40kbps, for example.

I'd worry more about the 'CD quality' claim, which is rather more
subjective.


SaSW, Willem
-- 
Disclaimer: I am in no way responsible for any of the statements
            made in the above text. For all I know I might be
            drugged or something..
            No I'm not paranoid. You all think I'm paranoid, don't you !
#EOT

[toc] | [prev] | [next] | [standalone]


#565

FromPeter Schepers <schepers@uwaterloo.ca>
Date2011-09-14 15:47 -0400
Message-ID<j4r0c9$ajd$1@rumours.uwaterloo.ca>
In reply to#564
On 14/09/2011 3:39 PM, Willem wrote:
> Speech compression at 37kbps seems quite plausible, although
> the number 37 seems a bit arbitrary.  I would expect 32 or 40.

But the claim is 37 _bits_ per second (as the subject header says), not 
kbps.

PS

[toc] | [prev] | [next] | [standalone]


#566

FromJim Leonard <mobygamer@gmail.com>
Date2011-09-14 13:26 -0700
Message-ID<2862cce2-8555-4c32-9471-085ac0a8fde0@i39g2000yqn.googlegroups.com>
In reply to#561
On Sep 14, 8:50 am, Nimo <azeez...@gmail.com> wrote:
> Clarke's three laws.

Ah yes, the wonderful and talented science FICTION writer.

Did you offer the above as proof your claims are fiction?

[toc] | [prev] | [next] | [standalone]


#558

FromNimo <azeez541@gmail.com>
Date2011-09-14 06:51 -0700
Message-ID<33ba8f5d-9cbc-4507-a734-c20779630853@s15g2000pre.googlegroups.com>
In reply to#554
On Sep 14, 4:33 am, tom st denis <t...@iahu.ca> wrote:
> On Sep 13, 10:36 pm, Nimo <azeez...@gmail.com> wrote:
>
> > On Sep 13, 7:52 am, Jim Leonard <mobyga...@gmail.com> wrote:
>
> > > On Sep 12, 1:57 am, Nimo <azeez...@gmail.com> wrote:
>
> > > > I'm getting CD quality "speech" at 37 bits / sec.
>
> > > Functional example please?
>
> > Sorry, I didn't get you ...?
>
> > so long
> >   nimo
>
> decoder + sample compressed stream please.
>
> Tom

wait baby, getting US patent and then game starts..

[toc] | [prev] | [next] | [standalone]


#559

FromJim Leonard <mobygamer@gmail.com>
Date2011-09-14 07:12 -0700
Message-ID<6e90f5d2-5f02-4316-9965-48ca6d5a87f1@m38g2000vbn.googlegroups.com>
In reply to#558
On Sep 14, 8:51 am, Nimo <azeez...@gmail.com> wrote:
>
> > decoder + sample compressed stream please.
>
> wait baby, getting US patent and then game starts..

I think you'll save yourself a lot of grief and heartache if you
create a decompressor + sample compressed stream *before* you bother
with patents.  Until you do, you have nothing worth protecting with a
patent.

[toc] | [prev] | [next] | [standalone]


#573

FromThomas Richter <thor@math.tu-berlin.de>
Date2011-09-18 14:08 +0200
Message-ID<j54n0h$256$1@news.belwue.de>
In reply to#550
On 12.09.2011 08:57, Nimo wrote:
> Hi there,
>
>     After a long time, back to my group.
>
> well, will keep the stuff to the point, if you have any doubts,
> I'm always here to help you..
>
> // A New Speech Codec Based upon  Advanced Tensor Basis&  Galerkin
> Techniques //
>
>
> ALGORITHM   BITRATE(s)  MOS   QUALITY    SUBJECTIVE OPINION  DELAY
>
>     ***                 37 bps        5       Transparent
> imperceptible    10th part delay in a sec
>
>
>
> I'm getting CD quality "speech" at 37 bits / sec.
>
>
>        A checkmate to G.series stuff, AMR, speex etc.
>
>
> Any Ideas,  immediately, I mean as fast as possible I can go out with
> this technology.

Does this refer to the link you gave? I don't see a MOS of 5 on this patent.

But allow me to make a couple of comments:

- If you provide a MOS score, you should specifically say what the task 
was you gave to the observers: Understand the words (text-to-speech 
would suffice), get the pronounciation (vocoding possible), recognize 
the speaker?

- The MOS scores are very sloppy, no error bars. Unclear how many 
observers were used to perform the experiments.

What is "CD speech" quality? The quality of a vocoder can be very high, 
yet it might be not what has been asked for. The G.series are for 
natural speech representation in a non-parametric way, allowing to 
identify the speaker etc... Of course parametric coding is known, and 
text-to-speech is known, but still, if I make a phone call, I would be 
very disappointed if all I get would be a vocoder I talk to, even if the 
quality is very high, and thus none of these techniques ended the ITU-T 
G-series of standards.

Thus, at least, I can accuse the original "inventors" of making very 
unscientific comparisons, or not telling me what the actual intend of 
the work should be.

Greetings,
	Thomas

[toc] | [prev] | [standalone]


Back to top | Article view | comp.compression


csiph-web