Splitting file based on pattern and first character Post: 302646523

10 More Discussions You Might Find Interesting

1. Shell Programming and Scripting

Splitting the file based on logic

Hello I have a requirement where i need to split the Input fixed width file which contains multiple invoices into multiple files with 2 invoices per file. Each invoice can be identified by its first line's second character which is "H" and sixth character is " " space and the invoice would...

2. Shell Programming and Scripting

Splitting a file based on two patterns

Hi there, I've an input file as follows: *START 1001 a1 1002 a2 1003 a3 1004 a4 *END *START 1001 b1 1002 b2 1004 b4 *END *START 1001 c1 1004 c4 *END

3. UNIX for Dummies Questions & Answers

Splitting a file based on first 8 chars

I have an input file of this format <Date><other data> For example, 20081213aaaaaaaaa 20081213bbbbbbbbb 20081220ccccccccc 20081220ddddddddd 20081220eeeeeeeee 20081227ffffffffffffff The first 8 chars are date in YYYYMMDD formT. I need to split this file into n files where n is the...

4. Shell Programming and Scripting

Splitting large file into multiple files in unix based on pattern

I need to write a shell script for below scenario My input file has data in format: qwerty0101TWE 12345 01022005 01022005 datainala alanfernanded 26 qwerty0101mXZ 12349 01022005 06022008 datainalb johngalilo 28 qwerty0101TWE 12342 01022005 07022009 datainalc hitalbert 43 qwerty0101CFG 12345...

5. Shell Programming and Scripting

Problem with splitting large file based on pattern

Hi Experts, I have to split huge file based on the pattern to create smaller files. The pattern which is expected in the file is: Master..... First... second.... second... third.. third... Master... First.. second... third... Master... First... second.. second.. second.....

6. Shell Programming and Scripting

File character adjustment based on specific character

i have a reqirement to adjust the data in a file based on a perticular character the sample data is as below 483PDEAN CORRIGAN 52304037528955WAGES 50000 89BP ABCD MASTER352 5434604223735428 4200 58BP SOUTHERN WA848 ...

7. Shell Programming and Scripting

Merging two special character separated files based on pattern matching

Hi. I have 2 files of below format. File1 AA~1~STEVE~3.1~4.1~5.1 AA~2~DANIEL~3.2~4.2~5.2 BB~3~STEVE~3.3~4.3~5.3 BB~4~TIM~3.4~4.4~5.4 File 2 AA~STEVE~AA STEVE WORKS at AUTO COMPANY AA~DANIEL~AA DANIEL IS A ELECTRICIAN BB~STEVE~BB STEVE IS A COOK I want to match 1st and 3rd...

8. Shell Programming and Scripting

Splitting based on occurence of a Character at fixed position

I have a requirement where i need to split a file based on occurence of a character which is present at a fixed position. Description is as below: 1. The file will be more than 1 Lakh records. 2. Each line will be of fixed length of 987 characters. 3. At position 28 in each line either 'C' or...

9. Shell Programming and Scripting

Splitting textfile based on pattern and name new file after pattern

Hi there, I am pretty new to those things, so I couldn't figure out how to solve this, and if it is actually that easy. just found that awk could help:(. so i have a textfile with strings and numbers (originally copy pasted from word, therefore some empty cells) in the following structure: SC...

10. UNIX for Beginners Questions & Answers

Splitting a file based on a pattern

Hi All, I am having a problem. I tried to extract the chunk of data and tried to fix I am not able to. Any help please Basically I need to remove the for , values after K, this is how it is now A,, B, C,C, D,D, 12/04/10,12/04/10, K,1,1,1,1,0,3.0, K,1,1,1,2,0,4.0,...

LEARN ABOUT DEBIAN

cdb

cdb(5)								File Formats Manual							    cdb(5)

NAME

       cdb - Constant DataBase file format

DESCRIPTION

       A  cdb database is a single file used to map `keys' to `values', having records of (key,value) pairs.  File consists of 3 parts: toc (table
       of contents), data and index (hash tables).

       Toc has fixed length of 2048 bytes, containing 256 pointers to hash tables inside index sections.  Every pointer consists of position of  a
       hash  table  in	bytes from the beginning of a file, and a size of a hash table in entries, both are 4-bytes (32 bits) unsigned integers in
       little-endian form.  Hash table length may have zero length, meaning that corresponding hash table is empty.

       Right after toc section, data section follows without any alingment.  It consists of series of records, each is a key length, value  (data)
       length,	key  and  value.  Again, key and value length are 4-byte unsigned integers.  Each next record follows previous without any special
       alignment.

       After data section, index (hash tables) section follows.  It should be looked to in conjunction with toc section, where	each  of  max  256
       hash tables are defined.  Index section consists of series of hash tables, with starting position and length defined in toc section.  Every
       hash table is a sequence of records each holds two numbers: key's hash value and record position inside data section (bytes from the begin-
       ning  of  a  file  to  first  byte of key length starting data record).	If record position is zero, then this is an empty hash table slot,
       pointed to nowhere.

       CDB hash function is
	 hv = ((hv << 5) + hv) ^ c
       for every single c byte of a key, starting with hv = 5381.

       Toc section indexed by (hv % 256), i.e. hash value modulo 256 (number of entries in toc section).

       In order to find a record, one should: first, compute the hash value (hv) of a key.  Second, look to hash table number hv modulo  256.	If
       it  is  empty,  then there is no such key exists.  If it is not empty, then third, loop by slots inside that hash table, starting from slot
       with number hv divided by 256 modulo length of that table, or ((hv / 256) % htlen), searching for this hv in hash table.   Stop	search	on
       empty  slot (if record position is zero) or when all slots was probed (note cyclic search, jumping from end to beginning of a table).  When
       hash value in question is found in hash table, look to key of corresponding record, comparing it with key in question.  If them of the same
       length  and equals to each other, then record is found, overwise, repeat with next hash table slot.  Note that there may be several records
       with the same key.

SEE ALSO

       cdb(1), cdb(3).

AUTHOR

       The tinycdb package written by Michael Tokarev <mjt@corpit.ru>, based on ideas and shares file format with  original  cdb  library  by  Dan
       Bernstein.

LICENSE

       Public domain.

								     Apr, 2005								    cdb(5)

10 More Discussions You Might Find Interesting

1. Shell Programming and Scripting

Splitting the file based on logic

Discussion started by: dsdev_123

2. Shell Programming and Scripting

Splitting a file based on two patterns

Discussion started by: kbirde

3. UNIX for Dummies Questions & Answers

Splitting a file based on first 8 chars

Discussion started by: paruthiveeran

4. Shell Programming and Scripting

Splitting large file into multiple files in unix based on pattern

Discussion started by: jimmy12

5. Shell Programming and Scripting

Problem with splitting large file based on pattern

Discussion started by: saisanthi

6. Shell Programming and Scripting

File character adjustment based on specific character

Discussion started by: pema.yozer

7. Shell Programming and Scripting

Merging two special character separated files based on pattern matching

Discussion started by: crypto87