Linux shell script to insert new lines based on delimiter count Post: 302986016

Sponsored Content

Top Forums Shell Programming and Scripting Linux shell script to insert new lines based on delimiter count Post 302986016 by digitalnirvana on Friday 18th of November 2016 09:09:11 AM

11-18-2016

Registered User

Apologies, my bad. I should've uploaded the file. Attached is a masked .dat file renamed as .txt for uploading.

On opening it with notepad++ in Windows, the null characters show up as boxes. In a Linux vi the nulls are ^@.

All the records in this file are in one row. This particular file has 2 records followed by the trailer record.

First record = starts at the beginning of the file 00000230 (this field gives the length of the record in bytes)
Second record = starts at the next 00000230 (it is a coincidence, here both records have same length)
Trailer record = starts at 0000096 (the trailer length is of 96 bytes and it also has 80 delimiters of ^@ or null characters. Ignore my earlier post saying trailer has 20 delimiters. It has 80 actually)

As the field lengths are variable so we cannot define a record in terms of total length of its fields or total bytes. This is why we are defining a record as effectively having length of 80 ^@ delimiters.

I require the 1st record in one row, 2nd record in next row and so on till the end of the file, with the trailer in the last row. If there is a way of adding a newline after every 80th ^@ from the beginning till the eof, then perhaps it will work?

The only unprintable character is null ^@, no TABS or other spaces, all other characters are alphanumeric.

Please let me know if any questions. Thanks for the help

MED_BIL_accmasked.DAT.txt (556 Bytes)

digitalnirvana

View Public Profile for digitalnirvana

Find all posts by digitalnirvana

10 More Discussions You Might Find Interesting

1. UNIX for Dummies Questions & Answers

Perl/shell script count the lines

Hi Guys, I want to write a perl/shell script do parse the following file input file content NPA-NXX SC 2084549 45 2084552 45 2084563 2007 2084572 45 2084580 45 3278411 45 3278430 45 3278493 530 3278507 530...

2. Shell Programming and Scripting

insert leading zeroes based on the character count

Hi, I need add leading zeroes to a field in a file based on the character count. The field can be of 1 character to 6 character length. I need to make the field 14bytes. eg: 8351,20,1 8351,234,6 8351,2,0 8351,1234,2 8351,123456,1 8351,12345,2 This should become. ...

3. Shell Programming and Scripting

Shell script to put delimiter for a no delimiter variable length text file

Hi, I have a No Delimiter variable length text file with following schema - Column Name Data length Firstname 5 Lastname 5 age 3 phoneno1 10 phoneno2 10 phoneno3 10 sample data - ...

4. Shell Programming and Scripting

Bash script to count and insert

Hi not sure if this is possible but I need some help with a bash script, I have a text file and on the first line that starts with 7150230 I need it to put a 1 at position 79 and a 2 at position 88, this is where it gets complicated, on the next line it finds that starts with 7150230 I then need it...

5. Shell Programming and Scripting

Insert Columns before the last Column based on the Count of Delimiters

Hi, I have a requirement where in I need to insert delimiters before the last column of the total delimiters is less than a specified number. Say if the delimiters is less than 139, I need to insert 2 columns ( with blanks) before the last field awk -F '�' '{ if (NF-1 < 139)} END { "Insert 2...

6. Shell Programming and Scripting

File Count Based on FileDate using Shell Script

I have file listed in my directory in following format -rwxrwxr-x+ 1 test test 4.9M Oct 3 16:06 test20141002150108.txt -rwxrwxr-x+ 1 test test 4.9M Oct 4 16:06 test20141003150108.txt -rwxrwxr-x+ 1 test test 4.9M Oct 5 16:06 test20141005150108.txt -rwxrwxr-x+ 1 test ...

7. Shell Programming and Scripting

Removing duplicate lines on first column based with pipe delimiter

Hi, I have tried to remove dublicate lines based on first column with pipe delimiter . but i ma not able to get some uniqu lines Command : sort -t'|' -nuk1 file.txt Input : 38376KZ|09/25/15|1.057 38376KZ|09/25/15|1.057 02006YB|09/25/15|0.859 12593PS|09/25/15|2.803...

8. Shell Programming and Scripting

awk joining multiple lines based on field count

Hi Folks, I have a file with fields as follows which has last field in multiple lines. I would like to combine a line which has three fields with single field line for as shown in expected output. Please help. INPUT hname01 windows appnamec1eda_p1, ...

9. Shell Programming and Scripting

Shell script count lines and sum numbers from multiple files

I want to count the number of lines, I need this result be a number, and sum the last numeric column, I had done to make this one at time, but I need to make this for a crontab, so, it has to be an script, here is my lines: It counts the number of lines: egrep -i String file_name_201611* |...

10. Shell Programming and Scripting

Split files based on row delimiter count

I have a huge file (around 4-5 GB containing 20 million rows) which has text like: <EOFD>11<EOFD>22<EORD>2<EOFD>2222<EOFD>3333<EORD>3<EOFD>44<EOFD>55<EORD>66<EOFD>888<EOFD>9999<EORD> Actually above is an extracted file from a Sql Server with each field delimited by <EOFD> and each row ends...

LEARN ABOUT DEBIAN

cdb

cdb(5)								File Formats Manual							    cdb(5)

NAME

       cdb - Constant DataBase file format

DESCRIPTION

       A  cdb database is a single file used to map `keys' to `values', having records of (key,value) pairs.  File consists of 3 parts: toc (table
       of contents), data and index (hash tables).

       Toc has fixed length of 2048 bytes, containing 256 pointers to hash tables inside index sections.  Every pointer consists of position of  a
       hash  table  in	bytes from the beginning of a file, and a size of a hash table in entries, both are 4-bytes (32 bits) unsigned integers in
       little-endian form.  Hash table length may have zero length, meaning that corresponding hash table is empty.

       Right after toc section, data section follows without any alingment.  It consists of series of records, each is a key length, value  (data)
       length,	key  and  value.  Again, key and value length are 4-byte unsigned integers.  Each next record follows previous without any special
       alignment.

       After data section, index (hash tables) section follows.  It should be looked to in conjunction with toc section, where	each  of  max  256
       hash tables are defined.  Index section consists of series of hash tables, with starting position and length defined in toc section.  Every
       hash table is a sequence of records each holds two numbers: key's hash value and record position inside data section (bytes from the begin-
       ning  of  a  file  to  first  byte of key length starting data record).	If record position is zero, then this is an empty hash table slot,
       pointed to nowhere.

       CDB hash function is
	 hv = ((hv << 5) + hv) ^ c
       for every single c byte of a key, starting with hv = 5381.

       Toc section indexed by (hv % 256), i.e. hash value modulo 256 (number of entries in toc section).

       In order to find a record, one should: first, compute the hash value (hv) of a key.  Second, look to hash table number hv modulo  256.	If
       it  is  empty,  then there is no such key exists.  If it is not empty, then third, loop by slots inside that hash table, starting from slot
       with number hv divided by 256 modulo length of that table, or ((hv / 256) % htlen), searching for this hv in hash table.   Stop	search	on
       empty  slot (if record position is zero) or when all slots was probed (note cyclic search, jumping from end to beginning of a table).  When
       hash value in question is found in hash table, look to key of corresponding record, comparing it with key in question.  If them of the same
       length  and equals to each other, then record is found, overwise, repeat with next hash table slot.  Note that there may be several records
       with the same key.

SEE ALSO

       cdb(1), cdb(3).

AUTHOR

       The tinycdb package written by Michael Tokarev <mjt@corpit.ru>, based on ideas and shares file format with  original  cdb  library  by  Dan
       Bernstein.

LICENSE

       Public domain.

								     Apr, 2005								    cdb(5)

10 More Discussions You Might Find Interesting

1. UNIX for Dummies Questions & Answers

Perl/shell script count the lines

Discussion started by: pistachio

2. Shell Programming and Scripting

insert leading zeroes based on the character count

Discussion started by: gpaulose

3. Shell Programming and Scripting

Shell script to put delimiter for a no delimiter variable length text file

Discussion started by: Gaurav Martha

4. Shell Programming and Scripting

Bash script to count and insert

Discussion started by: firefox2k2

5. Shell Programming and Scripting

Insert Columns before the last Column based on the Count of Delimiters

Discussion started by: arunkesi

6. Shell Programming and Scripting

File Count Based on FileDate using Shell Script

Discussion started by: krish2014

7. Shell Programming and Scripting

Removing duplicate lines on first column based with pipe delimiter

Discussion started by: parithi06