Matching 2 files based on key Post: 303030720

10 More Discussions You Might Find Interesting

1. Shell Programming and Scripting

merging two files based on some key

I have to merge two files: The files are having the same format like A0this is first line TOlast line silmilarly other lines. I have to search for A0 line in the second file also and then put the data in the third file under A0 heading ,then for A1 and so on. A0 portion will be treminated...

2. Shell Programming and Scripting

Merge files based on key

Hi Friends, Can any one help me with merging these file based on two columns : File1: A|123|99|SAMS B|456|95|GEORGE D|789|85|HOVARD File2: S|123|99|NANcY|6357 S|123|99|GREGRO|83748 A|456|95|HARRY|827|somers S|456|95|ANTONY|546841|RUDOLPH|7263 B|456|95|SMITH|827|BOISE STATE|834...

3. Shell Programming and Scripting

joining files based on key column

Hi I have to join two files based on 1st column where 4th column of a2.txt=at and take 2nd column of a1.txt and 3rd column of a2.txt and check against source files ,if matches list those source file names. a1.txt a1|20090809|20090810 a2|20090907|20090908 a2.txt a1|d|file1.txt|at...

4. Shell Programming and Scripting

Matching 2 files based on one column

Hi, On a similar subject, the following. I have two files: file1.txt dbSNP_rsID,Chromosome,Position,Gene rs10399749,chr. 01,45162,? rs4030303,chr. 01,72434,? rs4030300,chr. 01,72515,? rs940550,chr. 01,78032,? rs13328714,chr. 01,81468,? rs11490937,chr. 01,222077,? rs6683466,chr....

5. Shell Programming and Scripting

Merge two files based on a 3rd key file

Hi, I want to merge the two files based on the key file's columns. The key file: DATE~DATE HOUSE~IN_HOUSE CUST~IN_CUST PRODUCT~PRODUCT ADDRESS~CUST_ADDR BASIS_POINTS~BASIS_POINTS ... The other 2 files are From_file & To_file - The From_file: DATE|date/time|29|9 ...

6. UNIX for Dummies Questions & Answers

How to fetch files right below based on some matching criteria?

I have a requirement where in i need to select records right below the search criteria qwertykeyboard white 10 20 30 30 40 50 60 70 80 qwertykeyboard black 40 50 60 70 90 100 qwertykeyboard and white are headers separated by a tab. when i execute my script..i would be searching...

7. UNIX for Dummies Questions & Answers

Merge selective columns from files based on common key

Hi, I am trying to selectively merge two files based on keys reported in the 1st column. File1: #file1-header1 file1-header2 111 qwe rtz uio 198 asd fgh jkl 165 yxc 789 poi uzt rew 89 lkj File2: #file2-header2 file2-header2 165 ghz nko2 ...

8. Shell Programming and Scripting

awk - Merge two files based on one key

Hi, I am struggling with the an awk command to merge two files based on a common key. I want to append the value from File2 ($2) onto the end of File1 where $1 from each file matches - If no match then nothing is apended File1 COL1|COL2|COL3|COL4|COL5|COL6|COL7...

9. Shell Programming and Scripting

Files summary using awk based on index key

Hello , I have several files which are looking similar to : file01.txt keyA001 350 X string001 value001 keyA001 450 X string002 value007 keyA001 454 X string002 value004 keyA001 500 X string003 value005 keyA001 255 X string004 value006 keyA001 388 X string005 value008 keyA001 1278 X...

10. UNIX for Beginners Questions & Answers

Match tab-delimited files based on key

I thought I had this figured out but was wrong so am humbly asking for help. The task is to add an additional column to FILE 1 based on records in FILE 2. The key is in COLUMN 1 for FILE 1 and in COLUMN 1 OR COLUMN 2 for FILE 2. I want to add the third column from FILE 2 to the beginning of...

LEARN ABOUT DEBIAN

ncdiff

NCDIFF(1)						      General Commands Manual							 NCDIFF(1)

NAME

       ncdiff - netCDF Differencer

SYNTAX

       ncdiff  [-3]  [-4]  [-6]  [-A]  [-C]  [-c]  [-D dbg] [-d dim,[ min][,[ max]]] [-F] [-h] [-L dfl_lvl] [-l path] [-O] [-p path] [-R] [-r] [-v
       var[,...]]  [-x] file_1 file_2 file_3

DESCRIPTION

       ncdiff subtracts variables in file_2 from the corresponding variables (those with the same name)  in  file_1  and  stores  the  results	in
       file_3.	 Variables in file_2 are broadcast to conform to the corresponding variable in file_1 if necessary.  Broadcasting a variable means
       creating data in non-existing dimensions from the data in existing dimensions.  For example, a two dimensional variable in  file_2  can	be
       subtracted  from  a four, three, or two (but not one or zero) dimensional variable (of the same name) in file_1.  This functionality allows
       the user to compute anomalies from the mean.  Note that variables in file_1 are not broadcast to  conform  to  the  dimensions  in  file_2.
       Thus,  ncdiff, the number of dimensions, or rank, of any processed variable in file_1 must be greater than or equal to the rank of the same
       variable in file_2.  Furthermore, the size of all dimensions common to both file_1 and file_2 must be equal.

       When computing anomalies from the mean it is often the case that file_2 was created by applying an averaging operator to a  file  with  the
       same  dimensions  as file_1, if not file_1 itself.  In these cases, creating file_2 with ncra rather than ncwa will cause the ncdiff opera-
       tion to fail.  For concreteness say the record dimension in file_1 is time.  If file_2 were created  by	averaging  file_1  over  the  time
       dimension with the ncra operator rather than with the ncwa operator, then file_2 will have a time dimension of size 1 rather than having no
       time dimension at all In this case the input files to ncdiff, file_1 and file_2, will have unequally sized  time  dimensions  which  causes
       ncdiff to fail.	To prevent this from occuring, use ncwa to remove the time dimension from file_2.  An example is given below.

       ncdiff will never difference coordinate variables or variables of type NC_CHAR or NC_BYTE.  This ensures that coordinates like (e.g., lati-
       tude and longitude) are physically meaningful in the output file, file_3.  This behavior is hardcoded.  ncdiff  applies	special  rules	to
       some NCAR CSM fields (e.g., ORO).  See NCAR CSM Conventions for a complete description.	Finally, we note that ncflint (ncflint netCDF File
       Interpolator) can be also perform file subtraction (as well as addition, multiplication and interpolation).

EXAMPLES

       Say files 85_0112.nc and 86_0112.nc each contain 12 months of data.  Compute the change in the monthly averages from 1985 to 1986:
	      ncdiff 86_0112.nc 85_0112.nc 86m85_0112.nc

       The following examples demonstrate the broadcasting feature of ncdiff.  Say we wish to compute the monthly anomalies of T from  the  yearly
       average of T for the year 1985.	First we create the 1985 average from the monthly data, which is stored with the record dimension time.
	      ncra 85_0112.nc 85.nc
	      ncwa -O -a time 85.nc 85.nc
       The  second command, ncwa, gets rid of the time dimension of size 1 that ncra left in 85.nc.  Now none of the variables in 85.nc has a time
       dimension.  A quicker way to accomplish this is to use ncwa from the beginning:
	      ncwa -a time 85_0112.nc 85.nc
       We are now ready to use ncdiff to compute the anomalies for 1985:
	      ncdiff -v T 85_0112.nc 85.nc t_anm_85_0112.nc
       Each of the 12 records in t_anm_85_0112.nc now contains the monthly deviation of T from the annual mean of T for each gridpoint.

       Say we wish to compute the monthly gridpoint anomalies from the zonal annual mean.  A zonal mean is a quantity that has been averaged  over
       the  longitudinal  (or  x) direction.  First we use ncwa to average over longitudinal direction lon, creating xavg_85.nc, the zonal mean of
       85.nc.  Then we use ncdiff to subtract the zonal annual means from the monthly gridpoint data:
	      ncwa -a lon 85.nc xavg_85.nc
	      ncdiff 85_0112.nc xavg_85.nc tx_anm_85_0112.nc
       Assuming 85_0112.nc has dimensions time and lon, this example only works if xavg_85.nc has no time or lon dimension.

       As a final example, say we have five years of monthly data (i.e., 60 months) stored in 8501_8912.nc and we wish to create a file which con-
       tains  the  twelve  month seasonal cycle of the average monthly anomaly from the five-year mean of this data.  The following method is just
       one permutation of many which will accomplish the same result.  First use ncwa to create the file containing the five-year mean:
	      ncwa -a time 8501_8912.nc 8589.nc
       Next use ncdiff to create a file containing the difference of each month's data from the five-year mean:
	      ncdiff 8501_8912.nc 8589.nc t_anm_8501_8912.nc
       Now use ncks to group the five January anomalies together in one file, and use ncra to create the average anomaly for  all  five  Januarys.
       These commands are embedded in a shell loop so they are repeated for all twelve months:
	      foreach idx (01 02 03 04 05 06 07 08 09 10 11 12)
	      ncks -F -d time,,,12 t_anm_8501_8912.nc foo.
	      ncra foo. t_anm_8589_.nc
	      end
       Note that ncra understands the stride argument so the two commands inside the loop may be combined into the single command
	      ncra -F -d time,,,12 t_anm_8501_8912.nc foo.
       Finally,  use  ncrcat  to  concatenate  the 12 average monthly anomaly files into one twelve-record file which contains the entire seasonal
       cycle of the monthly anomalies:
	      ncrcat t_anm_8589_??.nc t_anm_8589_0112.nc

AUTHOR

       NCO manual pages written by Charlie Zender and Brian Mays.

REPORTING BUGS

       Report bugs to <http://sf.net/bugs/?group_id=3331>.

COPYRIGHT

       Copyright (C) 1995-2011 Charlie Zender
       This is free software; see the source for copying conditions.  There is NO warranty; not even for MERCHANTABILITY or FITNESS FOR A PARTICU-
       LAR PURPOSE.

SEE ALSO

       The  full  documentation for NCO is maintained as a Texinfo manual called the NCO User's Guide.	Because NCO is mathematical in nature, the
       documentation includes TeX-intensive portions not viewable on character-based displays.	Hence the only complete and authoritative versions
       of   the   NCO	User's	 Guide	 are   the   PDF   (recommended),   DVI,   and	 Postscript   versions	 at   <http://nco.sf.net/nco.pdf>,
       <http://nco.sf.net/nco.dvi>,   and   <http://nco.sf.net/nco.ps>,   respectively.    HTML   and	 XML	versions    are    available	at
       <http://nco.sf.net/nco.html> and <http://nco.sf.net/nco.xml>, respectively.

       If the info and NCO programs are properly installed at your site, the command

	      info nco

       should give you access to the complete manual, except for the TeX-intensive portions.

HOMEPAGE

       The NCO homepage at <http://nco.sf.net> contains more information.

																	 NCDIFF(1)

10 More Discussions You Might Find Interesting

1. Shell Programming and Scripting

merging two files based on some key

Discussion started by: Vandana Yadav

2. Shell Programming and Scripting

Merge files based on key

Discussion started by: sbasetty

3. Shell Programming and Scripting

joining files based on key column

Discussion started by: akil

4. Shell Programming and Scripting

Matching 2 files based on one column

Discussion started by: swvanderlaan