Groups | Search | Server Info | Keyboard shortcuts | Login | Register [http] [https] [nntp] [nntps]
Groups > comp.lang.php > #3162
| Path | csiph.com!x330-a1.tempe.blueboxinc.net!usenet.pasdenom.info!news.albasani.net!news2.arglkargh.de!noris.net!newsfeed.arcor.de!newsspool3.arcor-online.net!news.arcor.de.POSTED!not-for-mail |
|---|---|
| Content-Type | text/plain; charset="UTF-8" |
| Message-ID | <1666924.NXiflum83G@PointedEars.de> (permalink) |
| From | Thomas 'PointedEars' Lahn <PointedEars@web.de> |
| Reply-To | Thomas 'PointedEars' Lahn <php@PointedEars.de> |
| Organization | PointedEars Software (PES) |
| Date | Wed, 14 Sep 2011 14:07:27 +0200 |
| User-Agent | KNode/4.4.11 |
| Content-Transfer-Encoding | 8Bit |
| Subject | Re: Trying to decode text that is supposed to be ISO-8859-1 |
| Newsgroups | comp.lang.php |
| References | <8739fzj397.fsf@gmail.com> <slrnj708l4.bek.hellsop@nibelheim.ninehells.com> |
| Followup-To | comp.lang.php |
| MIME-Version | 1.0 |
| Lines | 38 |
| NNTP-Posting-Date | 14 Sep 2011 14:07:27 CEST |
| NNTP-Posting-Host | 5741a9b6.newsspool3.arcor-online.net |
| X-Trace | DXC=dH?>4kbRmZcOKO]LCQ@0g`McF=Q^Z^V3h4Fo<]lROoRa8kF<OcfhCOkC?h^emQ26NkDZm8W4\YJNlUFDj8T]K0?cDmniOZG;K0kG;Ngc<[DDfj |
| X-Complaints-To | usenet-abuse@arcor.de |
| Xref | x330-a1.tempe.blueboxinc.net comp.lang.php:3162 |
Followups directed to: comp.lang.php
Show key headers only | View raw
Peter H. Coffin wrote:
> On Tue, 13 Sep 2011 19:56:20 -0600, Bart Kastermans wrote:
>> I have downloaded a file that claims to be ISO-8859-1. In it (among
>> many other stuff) are the bytes shown here (first column is the
>> character, the second is ord(character), the third and fourth are binary
>> respectively hexidecimal representations of the character.
>>
>> P / 80 / 01010000 / 50
>> l / 108 / 01101100 / 6c
>> z / 122 / 01111010 / 7a
>> e / 101 / 01100101 / 65
>> \303 / 195 / 11000011 / c3
>> \205 / 133 / 10000101 / 85
>> \313 / 203 / 11001011 / cb
>> \206 / 134 / 10000110 / 86
>>
>> This is supposed to be ISO-8859-1 encoded, and should encode the
>> character U+0148 (\v{n}; Latin small letter n with caron).
>>
>> Does anybody have any idea how I could decode this (or how it was
>> encoded in the first place)? Any suggestions would be greatly
>> appreciated.
>
> It's UTF-8 encoded representation of a false ISO-8859-1(? probably
> CP1251, actually) […]
Windows-125_2_ (Western) corresponds largely with ISO-8859-1. Windows-1251,
which is the proper name for that character set and encoding, is Cyrillic
above 0x7F, and corresponds largely with ISO-8859-5.
PointedEars
--
Anyone who slaps a 'this page is best viewed with Browser X' label on
a Web page appears to be yearning for the bad old days, before the Web,
when you had very little chance of reading a document written on another
computer, another word processor, or another network. -- Tim Berners-Lee
Back to comp.lang.php | Previous | Next — Previous in thread | Next in thread | Find similar | Unroll thread
Re: Trying to decode text that is supposed to be ISO-8859-1 "Peter H. Coffin" <hellsop@ninehells.com> - 2011-09-13 22:42 -0500
Re: Trying to decode text that is supposed to be ISO-8859-1 Thomas 'PointedEars' Lahn <PointedEars@web.de> - 2011-09-14 14:07 +0200
Re: Trying to decode text that is supposed to be ISO-8859-1 "Peter H. Coffin" <hellsop@ninehells.com> - 2011-09-14 08:37 -0500
Re: Trying to decode text that is supposed to be ISO-8859-1 Thomas 'PointedEars' Lahn <PointedEars@web.de> - 2011-09-21 00:14 +0200
csiph-web