Split a huge 7 GB File Based on Pattern into 4 files Post: 302837085

10 More Discussions You Might Find Interesting

1. Shell Programming and Scripting

Split a file based on a pattern

Dear all, I have a large file which is composed of 8000 frames, what i would like to do is split the file into 8000 single files names file.pdb.1, file.pdb.2 etc etc each frame in the large file is seperated by a "ENDMDL" flag so my thinking is to use this flag a a point to split the files...

2. Shell Programming and Scripting

Split a file into multiple files based on the input pattern

I have a file with lines something like. ...... 123_start ...... ....... 123_end .... ..... 456_start ...... ..... 456_end .... ..... 789_start .... .... 789_end

3. Shell Programming and Scripting

Help- counting delimiter in a huge file and split data into 2 files

I’m new to Linux script and not sure how to filter out bad records from huge flat files (over 1.3GB each). The delimiter is a semi colon “;” Here is the sample of 5 lines in the file: Name1;phone1;address1;city1;state1;zipcode1 Name2;phone2;address2;city2;state2;zipcode2;comment...

4. Shell Programming and Scripting

split XML file into multiple files based on pattern

Hello, I am using awk to split a file into multiple files using command: nawk '{ if ( $1 == "<process" ) { n=split($2, arr, "\""); file=arr } print > file }' processes.xml <process name="Process1.process"> ...

5. Shell Programming and Scripting

Split a file based on pattern and size

Hello, I have a large file (2GB) that I would like to split based on pattern and size. I've used the following command to split the file (token is "HELLO") awk '/HELLO/{i++}{print > "file"i}' input.txt and the output is similar to the following (i included filesize in KB): 10 ...

6. Shell Programming and Scripting

Split the file based on pattern

Hi , I have huge files around 400 mb, which has clob data and have diffeent scenarios: I am trying to pass scenario number as parameter and and get required modified file based on the scenario number and criteria. Scenario 1: file name : scenario_1.txt ...

7. Shell Programming and Scripting

Help needed - Split large file into smaller files based on pattern match

Help needed urgently please. I have a large file - a few hundred thousand lines. Sample CP START ACCOUNT 1234556 name 1 CP END ACCOUNT CP START ACCOUNT 2224444 name 1 CP END ACCOUNT CP START ACCOUNT 333344444 name 1 CP END ACCOUNT I need to split this file each time "CP START...

8. Shell Programming and Scripting

Split Large Files Based On Row Pattern..

Hi all. I've tried searching the web but could not find similar problem to mine. I have one large file to be splitted into several files based on the matching pattern found in each row. For example, let's say the file content: ...

9. Shell Programming and Scripting

How to split a file based on pattern line number?

Hi i have requirement like below M <form_name> sdasadasdMklkM D ...... D ..... M form_name> sdasadasdMklkM D ...... D ..... D ...... D ..... M form_name> sdasadasdMklkM D ...... M form_name> sdasadasdMklkM i want split file based on line number by finding...

10. UNIX for Advanced & Expert Users

Split one file to many based on pattern

Hello All, I have records in a file in a pattern A,B,B,B,B,K,A,B,B,K Is there any command or simple logic I can pull out records into multiple files based on A record? I want output as File1: A,B,B,B,B,K File2: A,B,B,K

LEARN ABOUT NETBSD

split

SPLIT(1)						    BSD General Commands Manual 						  SPLIT(1)

NAME

     split -- split a file into pieces

SYNOPSIS

     split [-a suffix_length] [-b byte_count[k|m] | -l line_count -n chunk_count] [file [name]]

DESCRIPTION

     The split utility reads the given file and breaks it up into files of 1000 lines each.  If file is a single dash or absent, split reads from
     the standard input.  file itself is not altered.

     The options are as follows:

     -a      Use suffix_length letters to form the suffix of the file name.

     -b      Create smaller files byte_count bytes in length.  If 'k' is appended to the number, the file is split into byte_count kilobyte
	     pieces.  If 'm' is appended to the number, the file is split into byte_count megabyte pieces.

     -l      Create smaller files line_count lines in length.

     -n      Split file into chunk_count smaller files.

     If additional arguments are specified, the first is used as the name of the input file which is to be split.  If a second additional argument
     is specified, it is used as a prefix for the names of the files into which the file is split.  In this case, each file into which the file is
     split is named by the prefix followed by a lexically ordered suffix using suffix_length characters in the range ``a-z''.  If -a is not speci-
     fied, two letters are used as the suffix.

     If the name argument is not specified, 'x' is used.

STANDARDS

     The split utility conforms to IEEE Std 1003.1-2001 (``POSIX.1'').

HISTORY

     A split command appeared in Version 6 AT&T UNIX.

     The -a option was introduced in NetBSD 2.0.  Before that, if name was not specified, split would vary the first letter of the filename to
     increase the number of possible output files.  The -a option makes this unnecessary.

BSD
								   May 28, 2007 							       BSD

10 More Discussions You Might Find Interesting

1. Shell Programming and Scripting

Split a file based on a pattern

Discussion started by: Mish_99

2. Shell Programming and Scripting

Split a file into multiple files based on the input pattern

Discussion started by: abinash

3. Shell Programming and Scripting

Help- counting delimiter in a huge file and split data into 2 files

Discussion started by: lv99

4. Shell Programming and Scripting

split XML file into multiple files based on pattern

Discussion started by: chiru_h