Path: csiph.com!newsfeed.hal-mli.net!feeder3.hal-mli.net!newsfeed.hal-mli.net!feeder1.hal-mli.net!de-l.enfer-du-nord.net!feeder2.enfer-du-nord.net!fu-berlin.de!uni-berlin.de!individual.net!not-for-mail From: =?ISO-8859-15?Q?Niels_Fr=F6hling?= Newsgroups: comp.compression Subject: Re: Words as well as characters as encoding units Date: Tue, 08 May 2012 00:42:21 -0600 Lines: 8 Message-ID: References: <357e1db3-17e2-4911-9423-1cf394d894d5@ns1g2000pbc.googlegroups.com> Mime-Version: 1.0 Content-Type: text/plain; charset=ISO-8859-15; format=flowed Content-Transfer-Encoding: 7bit X-Trace: individual.net ts0NzRQGFiv2DRD2pTnwjwUy0Gsb66sL3QlAThwVc5ZDjtteKOkSV5AMgoEbO8U+0j Cancel-Lock: sha1:Iw91pCGz+uTNTizB2Xfh+0SaWVg= User-Agent: Mozilla/5.0 (Windows; U; Windows NT 5.0; de-DE; rv:1.7.5) Gecko/20041206 Thunderbird/1.0 Mnenhy/0.7.1 In-Reply-To: Xref: csiph.com comp.compression:1296 > My layman's impression could certainly be wrong. But to me this seems > to be that LZW normally attempts to find as much correlations between > words as possible, which to a large part might actually not be > profitable work in practice. Can you name a language without correlations? Can you even point at digital data which is uncorrelated? (I'd exclude SETI captures here :^), that is likely uncorrelated data)