Get common lines from multiple files Post: 302437691

Sponsored Content

Top Forums Shell Programming and Scripting Get common lines from multiple files Post 302437691 by genehunter on Friday 16th of July 2010 01:20:07 AM

07-16-2010

Registered User

Get common lines from multiple files

FileA

Code:

chr1    31237964    NP_001018494.1    PUM1    M340L
chr1    31237964    NP_055491.1    PUM1    M340L
chr1    33251518    NP_037543.1    AK2    H191D
chr1    33251518    NP_001616.1    AK2    H191D
chr1    57027345    NP_001004303.2    C1orf168    P270S

FileB

Code:

                
chr1    116944164    NP_001533.2    IGSF3    R671W
chr1    33251518    NP_001616.1    AK2    H191D
chr1    57027345    NP_001004303.2    C1orf168    P270S
chr1    89606840    NP_940862.2    GBP6    R48C
chr1    110751878    NP_006393.2    HBXIP    P45L
chr1    246803952    NP_001001821.1    OR2T34    A244T

FileC

Code:

chr1    17164810    NP_055490.3    CROCC    G1471R
chr1    36323375    NP_055281.2    TEKT2    R61G
chr1    89606840    NP_940862.2    GBP6    R48C
chr1    40302534    NP_006358.1    CAP1    V115L
chr1    33251518    NP_001616.1    AK2    H191D
chr1    62026171    NP_795352.2    INADL    P336H

FileD

Code:

                
chr1    116944223    NP_001533.2    IGSF3    S651I
chr1    116944223    NP_001007238.1    IGSF3    S631I
chr1    150394079    XP_001724459.1    RPTN    E707G
chr1    36323375    NP_055281.2    TEKT2    R61G
chr1    150547095    NP_002007.1    FLG    E2297D
chr1    172075300    NP_060592.2    DARS2    G338E
chr1    222620225    NP_054903.1    CNIH4    G54S

I want a script (awk preferably or python) that will look for common lines in the 4 different files. Files are sorted on Col1, but can be resorted if necessary.
I want to have three output files
1) Commonlines in all 4 files
2) Common lines in any 3 files
3) Common lines in any 2 files. Getting which files have the common-line would be nice too.

Kindly help
~GH

genehunter

View Public Profile for genehunter

Find all posts by genehunter

10 More Discussions You Might Find Interesting

1. Shell Programming and Scripting

To find all common lines from 'n' no. of files

Hi, I have one situation. I have some 6-7 no. of files in one directory & I have to extract all the lines which exist in all these files. means I need to extract all common lines from all these files & put them in a separate file. Please help. I know it could be done with the help of...

2. Shell Programming and Scripting

Common lines from files

Hello guys, I need a script to get the common lines from two files with a criteria that if the first two columns match then I keep the maximum value of the 3rd column.(tab separated columns) Sample input: file1: 111 222 0.1 333 444 0.5 555 666 0.4 file 2: 111 222 0.7 555 666...

3. Shell Programming and Scripting

Common lines from files

Hello guys, I need a script to get the common lines from two files with a criteria that if the first two columns match then I keep the maximum value of the 5th column.(tab separated columns) . 3rd and 4th columns corresponds to the row which has highest value for the 5th column. Sample...

4. Shell Programming and Scripting

Merge multiple lines in same file with common key using awk

I've been a Unix admin for nearly 30 years and never learned AWK. I've seen several similar posts here, but haven't been able to adapt the answers to my situation. AWK is so damn cryptic! ;) I have a single file with ~900 lines (CSV list). Each line starts with an ID, but with different stuff...

5. Shell Programming and Scripting

Find common lines between multiple files

Hello everyone A few years Ago the user radoulov posted a fancy solution for a problem, which was about finding common lines (gene variation names) between multiple samples (files). The code was: awk 'END { for (R in rec) { n = split(rec, t, "/") if (n > 1) dup = dup ?...

6. Shell Programming and Scripting

Compare multiple files, and extract items that are common to ALL files only

I have this code awk 'NR==FNR{a=$1;next} a' file1 file2 which does what I need it to do, but for only two files. I want to make it so that I can have multiple files (for example 30) and the code will return only the items that are in every single one of those files and ignore the ones...

7. UNIX for Dummies Questions & Answers

Filter lines common in two files

Thanks everyone. I got that problem solved. I require one more help here. (Yes, UNIX definitely seems to be fun and useful, and I WILL eventually learn it for myself. But I am now on a different project and don't really have time to go through all the basics. So, I will really appreciate some...

8. Shell Programming and Scripting

Join common patterns in multiple lines into one line

Hi I have a file like 1 2 1 2 3 1 5 6 11 12 10 2 7 5 17 12 I would like to have an output as 1 2 3 5 6 10 7 11 12 17 any help would be highly appreciated Thanks

9. Shell Programming and Scripting

Join columns across multiple lines in a Text based on common column using BASH

10. Shell Programming and Scripting

Find common lines between all of the files in one folder

Could it be possible to find common lines between all of the files in one folder? Just like comm -12 . So all of the files two at a time. I would like all of the outcomes to be written to a different files, and the file names could be simply numbers - 1 , 2 , 3 etc. All of the file names contain...

LEARN ABOUT MOJAVE

sdiff

SDIFF(1)							   User Commands							  SDIFF(1)

NAME

       sdiff - side-by-side merge of file differences

SYNOPSIS

       sdiff [OPTION]... FILE1 FILE2

DESCRIPTION

       Side-by-side merge of file differences.

       -o FILE	--output=FILE
	      Operate interactively, sending output to FILE.

       -i  --ignore-case
	      Consider upper- and lower-case to be the same.

       -E  --ignore-tab-expansion
	      Ignore changes due to tab expansion.

       -b  --ignore-space-change
	      Ignore changes in the amount of white space.

       -W  --ignore-all-space
	      Ignore all white space.

       -B  --ignore-blank-lines
	      Ignore changes whose lines are all blank.

       -I RE  --ignore-matching-lines=RE
	      Ignore changes whose lines all match RE.

       --strip-trailing-cr
	      Strip trailing carriage return on input.

       -a  --text
	      Treat all files as text.

       -w NUM  --width=NUM
	      Output at most NUM (default 130) columns per line.

       -l  --left-column
	      Output only the left column of common lines.

       -s  --suppress-common-lines
	      Do not output common lines.

       -t  --expand-tabs
	      Expand tabs to spaces in output.

       -d  --minimal
	      Try hard to find a smaller set of changes.

       -H  --speed-large-files
	      Assume large files and many scattered small changes.

       --diff-program=PROGRAM
	      Use PROGRAM to compare files.

       -v  --version
	      Output version info.

       --help Output this help.

       If a FILE is `-', read standard input.

AUTHOR

       Written by Thomas Lord.

REPORTING BUGS

       Report bugs to <bug-gnu-utils@gnu.org>.

COPYRIGHT

       Copyright (C) 2002 Free Software Foundation, Inc.

       This  program  comes  with NO WARRANTY, to the extent permitted by law.	You may redistribute copies of this program under the terms of the
       GNU General Public License.  For more information about these matters, see the file named COPYING.

SEE ALSO

       The full documentation for sdiff is maintained as a Texinfo manual.  If the info and sdiff programs are properly installed  at  your  site,
       the command

	      info diff

       should give you access to the complete manual.

diffutils 2.8.1 						    April 2002								  SDIFF(1)

10 More Discussions You Might Find Interesting

1. Shell Programming and Scripting

To find all common lines from 'n' no. of files

Discussion started by: The Observer

2. Shell Programming and Scripting

Common lines from files

Discussion started by: jaysean

3. Shell Programming and Scripting

Common lines from files

Discussion started by: jaysean

4. Shell Programming and Scripting

Merge multiple lines in same file with common key using awk

Discussion started by: protosd

5. Shell Programming and Scripting

Find common lines between multiple files

Discussion started by: bibb

6. Shell Programming and Scripting

Compare multiple files, and extract items that are common to ALL files only

Discussion started by: castrojc

7. UNIX for Dummies Questions & Answers

Filter lines common in two files

Discussion started by: latsyrc

8. Shell Programming and Scripting

Join common patterns in multiple lines into one line

Discussion started by: Harrisham

9. Shell Programming and Scripting

Join columns across multiple lines in a Text based on common column using BASH

Discussion started by: nv186000

10. Shell Programming and Scripting

Find common lines between all of the files in one folder

Discussion started by: Eve

LEARN ABOUT MOJAVE

sdiff