Linux and UNIX Man Pages

Linux & Unix Commands - Search Man Pages

encode::arabic::parkinson(3pm) [debian man page]

Encode::Arabic::Parkinson(3pm)				User Contributed Perl Documentation			    Encode::Arabic::Parkinson(3pm)

NAME
Encode::Arabic::Parkinson - Dil Parkinson's transliteration of Arabic REVISION
$Revision: 179 $ $Date: 2007-01-14 01:23:25 +0100 (Sun, 14 Jan 2007) $ SYNOPSIS
use Encode::Arabic::Parkinson; # imports just like 'use Encode' would, plus more while ($line = <>) { # Dil Parkinson's mapping into the Arabic script print encode 'utf8', decode 'parkinson', $line; } # shell filter of data, e.g. in *n*x systems instead of viewing the Arabic script proper % perl -MEncode::Arabic::Parkinson -pe '$_ = encode "parkinson", decode "utf8", $_' # employing the modes of conversion for filtering and trimming Encode::Arabic::enmode 'parkinson', 'nosukuun', 'LWE xml'; Encode::Arabic::Parkinson->demode(undef, undef, 'strip _'); $decode = "AiqoraLo hRvaA Ol_n~a_S~a bi___OnotibaAhI."; $encode = encode 'parkinson', decode 'parkinson', $decode; # $encode eq "AiqraL hRvaA Aln~aS~a biAntibaAhI." DESCRIPTION
Dil Parkinson's notation is a one-to-one transliteration of the Arabic script for Modern Standard Arabic, using lower ASCII characters to encode the graphemes of the original script. IMPLEMENTATION Similar to that in Encode::Arabic::Buckwalter. EXPORTS & MODES The module exports as if "use Encode" also appeared in the package. The other "import" options are just delegated to Encode and imports performed properly. The conversion modes of this module allow to override the setting of the ":xml" option, in addition to filtering out diacritical marks and stripping off kashida. The modes and aliases relate like this: our %Encode::Arabic::Parkinson::modemap = ( 'default' => 0, 'undef' => 0, 'fullvocalize' => 0, 'full' => 0, 'nowasla' => 4, 'vocalize' => 3, 'nosukuun' => 3, 'novocalize' => 2, 'novowels' => 2, 'none' => 2, 'noshadda' => 1, 'noneplus' => 1, ); enmode ($obj, $mode, $xml, $kshd) demode ($obj, $mode, $xml, $kshd) These methods can be invoked directly or through the respective functions of Encode::Arabic. The meaning of the extra parameters follows from the examples of usage. SEE ALSO
Encode::Arabic, Encode, Encode::Encoding Xerox Arabic Home Page <http://www.arabic-morphology.com/> AUTHOR
Otakar Smrz, <http://ufal.mff.cuni.cz/~smrz/> eval { 'E<lt>' . ( join '.', qw 'otakar smrz' ) . "x40" . ( join '.', qw 'mff cuni cz' ) . 'E<gt>' } Perl is also designed to make the easy jobs not that easy ;) COPYRIGHT AND LICENSE
Copyright 2006-2007 by Otakar Smrz This library is free software; you can redistribute it and/or modify it under the same terms as Perl itself. perl v5.10.1 2010-01-18 Encode::Arabic::Parkinson(3pm)

Check Out this Related Man Page

Encode::TW(3pm) 					 Perl Programmers Reference Guide					   Encode::TW(3pm)

NAME
Encode::TW - Taiwan-based Chinese Encodings SYNOPSIS
use Encode qw/encode decode/; $big5 = encode("big5", $utf8); # loads Encode::TW implicitly $utf8 = decode("big5", $big5); # ditto DESCRIPTION
This module implements tradition Chinese charset encodings as used in Taiwan and Hong Kong. Encodings supported are as follows. Canonical Alias Description -------------------------------------------------------------------- big5-eten /big-?5$/i Big5 encoding (with ETen extensions) /big5-?et(en)?$/i /tca-?big5$/i big5-hkscs /big5-?hk(scs)?$/i /hk(scs)?-?big5$/i Big5 + Cantonese characters in Hong Kong MacChineseTrad Big5 + Apple Vendor Mappings cp950 Code Page 950 = Big5 + Microsoft vendor mappings -------------------------------------------------------------------- To find out how to use this module in detail, see Encode. NOTES
Due to size concerns, "EUC-TW" (Extended Unix Character), "CCCII" (Chinese Character Code for Information Interchange), "BIG5PLUS" (CMEX's Big5+) and "BIG5EXT" (CMEX's Big5e) are distributed separately on CPAN, under the name Encode::HanExtra. That module also contains extra China-based encodings. BUGS
Since the original "big5" encoding(1984) is not supported anywhere (glibc and DOS-based systems uses "big5" to mean "big5-eten"; Microsoft uses "big5" to mean "cp950"), a conscious decision was made to alias "big5" to "big5-eten", which is the de facto superset of the original big5. The "CNS11643" encoding files are not complete. For common "CNS11643" manipulation, please use "EUC-TW" in Encode::HanExtra, which contains planes 1-7. The ASCII region (0x00-0x7f) is preserved for all encodings, even though this conflicts with mappings by the Unicode Consortium. SEE ALSO
Encode perl v5.12.1 2010-04-26 Encode::TW(3pm)
Man Page