I have gone through all the threads in the forum and tested out different things. I am trying to split a 3GB file into multiple files. Some files are even larger than this.
For example:
split -l 3000000 filename.txt
This is very slow and it splits the file with 3 million records in each... (10 Replies)
Hello all.
Sorry, I know this question is similar to many others, but I just can seem to put together exactly what I need.
My file is tab delimitted and contains approximately 1 million rows. I would like to send lines 1,4,& 7 to a file. Lines 2, 5, & 8 to a second file. Lines 3, 6, & 9 to... (11 Replies)
Dear all,
I have a large file which is composed of 8000 frames, what i would like to do is split the file into 8000 single files names file.pdb.1, file.pdb.2 etc etc
each frame in the large file is seperated by a "ENDMDL" flag so my thinking is to use this flag a a point to split the files... (4 Replies)
Hello All,
I have a file which is having below type of data,
Jul 19 2011 | 123456
Jul 19 2011 | 123456
Jul 20 2011 | 123456
Jul 20 2011 | 123456
Here I wanted to grep for date pattern as below, so that it should only grep "Jul 20" OR "Jul ... (9 Replies)
I have a binary (hex) file I need to parse to get some data which are encoded this way:
.* b4 . . . 01 12 .* af .* 83 L1 x1 x2 xL 84 L2 y1 y2 yL
By another words there is a stream of hexadecimal bytes (in my example separated by space for better readability). I need to get value stored in... (3 Replies)
Hello, I have a large file (2GB) that I would like to split based on pattern and size.
I've used the following command to split the file (token is "HELLO")
awk '/HELLO/{i++}{print > "file"i}' input.txt
and the output is similar to the following (i included filesize in KB):
10 ... (2 Replies)
Hi ,
I have huge files around 400 mb, which has clob data and have diffeent scenarios:
I am trying to pass scenario number as parameter and and get required modified file based on the scenario number and criteria.
Scenario 1:
file name : scenario_1.txt
... (2 Replies)
Hi
i have requirement like below
M <form_name> sdasadasdMklkM
D ......
D .....
M form_name> sdasadasdMklkM
D ......
D .....
D ......
D .....
M form_name> sdasadasdMklkM
D ......
M form_name> sdasadasdMklkM
i want split file based on line number by finding... (10 Replies)
Hi ,
I have a file where i have modifed certain things compared to original file . The difference of the original file and modified file is as follows.
# diff mir_lex.c.modified mir_lex.c.orig
3209c3209
< if(yy_current_buffer -> yy_is_our_buffer == 0) {
---
>... (5 Replies)
Hello All,
I have records in a file in a pattern A,B,B,B,B,K,A,B,B,K
Is there any command or simple logic I can pull out records into multiple files based on A record? I want output as
File1: A,B,B,B,B,K
File2: A,B,B,K (9 Replies)
Discussion started by: deal1dealer
9 Replies
LEARN ABOUT ULTRIX
csplit
csplit(1) General Commands Manual csplit(1)Name
csplit - context split
Syntax
csplit [ -s ] [ -k ] [ -f prefix ] file arg1 [ ...argn ]
Description
The command reads file and separates it into n+1 sections, as defined by the arguments arg1...argn. By default, the sections are placed in
xx00...xxn (n may not be greater than 99). The named file is sectioned in the following way:
00: From the start of file up to (but not including) the line referenced by arg1.
01: From the line referenced by arg1 up to the line referenced by arg2.
.
.
.
n: From the line referenced by argn to the end of file.
If the file argument is a minus (-) then standard input is used. A minus is an ASCII octal 055.
Options-s Suppresses the printing of all character counts. If the -s option is omitted, the command prints the character counts
for each file created.
-k Leaves previously created files intact. If the -k option is omitted, automatically removes created files if an error
occurs.
-fprefix Names the created files prefix00...prefixn. The default is xx00...xxn.
The arguments (arg1...argn) to can be a combination of the following:
/rexp/[offset] A file is created for the section from the current line up to (but not including) the line containing the regular
expression rexp. The current line becomes the line containing rexp. The optional offset is plus (+) or minus
(-) the number of lines. For example, /Page/-5.
%rexp%[offset] This argument is the same as /rexp/[offset], except that no file is created for the section.
lnno A file is created from the current line up to (but not including) lnno. The current line becomes lnno.
{num} Repeat argument. This argument may follow any of the above arguments. If it follows a rexp argument, that argu-
ment is applied num more times. If it follows lnno, the file will be split every lnno lines (num times) from
that point.
Enclose all rexp type arguments that contain blanks or other characters meaningful to the Shell in the appropriate quotes. Regular expres-
sions should not contain embedded new-lines. The command does not affect the original file; it is the user's responsibility to remove it.
Examples
csplit -f cobol file /procedure division/ /par5./ /par16./
This example creates four files, cobol00...cobol03. After editing the files that created, they can be recombined as follows:
cat cobol0[0-3] > file
Note that this example overwrites the original file.
csplit -k file 100 {99}
This example splits the file every 100 lines, up to 10,000 lines. The -k option causes the created files to be retained if there are less
than 10,000 lines; however, an error message would still be printed.
csplit -k prog.c '%main(%' '/^}/+1' {20}
Assuming that follows the normal C coding convention of ending routines with a right brace (}) at the beginning of the line, this example
creates a file containing each separate C routine (up to 21) in
Diagnostics
The diagnostics are self explanatory except for the following:
arg - out of range
This message means that the given argument did not reference a line between the current position and the end of the file.
See Alsoed(1), sh(1)csplit(1)