Matching multiple fields from two files and then some? Post: 302657585

10 More Discussions You Might Find Interesting

1. Shell Programming and Scripting

AWK Matching Fields and Combining Files

Hello! I am writing a program to run through two large lists of data (~300,000 rows), find where rows in one file match another, and combine them based on matching fields. Due to the large file sizes, I'm guessing AWK will be the most efficient way to do this. Overall, the input and output I'm...

2. Shell Programming and Scripting

comparing two files for matching fields

I am newbie to unix and would please like some help to solve the task below I have two files, file_a.text and file_b.text that I want to evaluate. file_a.text 1698.74 1711.88 6576.25 899.41 3205.63 4187.98 697.35 1551.83 ...

3. Shell Programming and Scripting

How to merge two or more fields from two different files where there is non matching column?

Hi, Please excuse for often requesting queries and making R&D, I am trying to work out a possibility where i have two files field separated by pipe and another file containing only one field where there is no matching columns, Could you please advise how to merge two files. $more...

4. Shell Programming and Scripting

Print matching fields (if they exist) from two text files

Hi everyone, Given two files (test1 and test2) with the following contents: test1: 80263760,I71 80267369,M44 80274628,L77 80276793,I32 80277390,K05 80277391,I06 80279206,I43 80279859,K37 80279866,K35 80279867,J16 80280346,I14and test2: 80263760,PT18 80279867,PT01I need to do some...

5. UNIX for Beginners Questions & Answers

Awk: matching multiple fields between 2 files

Hi, I have 2 tab-delimited input files as follows. file1.tab: green A apple red B apple file2.tab: apple - A;Z Objective: Return $1 of file1 if, . $1 of file2 matches $3 of file1 and, . any single element (separated by ";") in $3 of file2 is present in $2 of file1 In order to...

6. Shell Programming and Scripting

awk to print fields that match using conditions and a default value for non-matching in two files

Trying to use awk to match the contents of each line in file1 with $5 in file2. Both files are tab-delimited and there may be a space or special character in the name being matched in file2, for example in file1 the name is BRCA1 but in file2 the name is BRCA 1 or in file1 name is BCR but in file2...

7. UNIX for Beginners Questions & Answers

Matching fields between two files, repeated records

In two previous posts (here) and (here), I received help from forum members comparing multiple fields across two files and selectively printing portions of each as output based upon would-be matches using awk. I had been fairly comfortable populating awk arrays with fields and using awk's special...

8. Shell Programming and Scripting

Matching two fields in two csv files, create new file and append match

I am trying to parse two csv files and make a match in one column then print the entire file to a new file and append an additional column that gives description from the match to the new file. If a match is not made, I would like to add "NA" to the end of the file Command that Ive been using...

9. Shell Programming and Scripting

Comparing two files by two matching fields

Long time listener first time poster. Hope someone can advise. I have two files, 1000+ lines in each, two fields in each file. After performing a sort, what is the best way to find exact matches where field $1 and $2 in file1 are also present in file2 on the same line, then output only those...

10. UNIX for Beginners Questions & Answers

awk for matching fields between files with repeated records

Hello all, I am having trouble with what should be an easy task, but seem to be missing something fundamental. I have two files, with File 1 consisting of a single field of many thousands of records. I also have File 2 with two fields and many thousands of records. My goal is that when $1 of...

LEARN ABOUT HPUX

nljust

nljust(1)						      General Commands Manual							 nljust(1)

NAME

       nljust - justify lines, left or right, for printing

SYNOPSIS

       digits] seq] just] mode] order] margin] width] ck] [file ...]

DESCRIPTION

       formats	for printing data written in languages with a right-to-left orientation.  It is designed to be used with the and the commands (see
       pr(1) and lp(1)).

       reads the concatenation of input files (or standard input if none are given) and produces on standard output a right-to-left formatted ver-
       sion of its input.  If appears as an input file name, reads standard input at that point.  Use to delimit the end of options.

       formats	input  files for all languages that are read from right to left.  For languages that have a left-to-right orientation, the command
       merely copies input files to standard output.

   Options
       recognizes the following options:

	      Justify data for all languages,
			  including those having a left-to-right text orientation.  By default only right-to-left language data is justified.  For
			  all other languages, input files are directly copied to standard output.

	      Select enhanced printer shapes for some Arabic characters.
			  With this option, two-character combinations of laam and alif are replaced by a single character.

	      Triggers ISO 8859-6 interpretation of the data.

	      Processes digits for output as hindi, western, or both.
			  digits can be or both.

	      Use	  seq as the escape sequence to select the primary character set.  This escape sequence is used by languages that have too
			  many characters to be accommodated by ASCII in a single 256-character set.  In these cases, the seq escape sequence  can
			  be  used  to	select	the non-ASCII character set.  The escape character itself(0x1b) is not given on the command line.
			  Hewlett-Packard escape sequences are used by default.

	      If	  just is left justify print lines.  If just is right-justify print lines starting from the (designated or default)  print
			  width column.  The default is right justification.

	      Replace leading spaces with alternative spaces.
			  Some	right-to-left character sets have a non-ASCII or alternative space.  This option can be useful when filtering out-
			  put (see pr(1)).  With right justification, the option causes line numbers to be placed immediately to the right of  the
			  tab  character.  Without the option, right justification causes line numbers to be placed at the print-width column.	By
			  default, leading spaces are not replaced by alternative spaces.

	      Indicate	  mode of any file to be formatted.  Mode refers to the text orientation of the file when it  was  created.   If  mode	is
			  assume  Latin  mode.	 If  mode is assume non-Latin mode.  By default, mode information is obtained from the environment
			  variable.

	      Do not terminate lines containing printable characters with a new-line.
			  By default, print lines are terminated by new-lines.

	      Indicate data
			  order of any file to be formatted.  The text orientation of a file can affect the way its data is arranged.  If order is
			  assume keyboard order.  If order is assume screen order.  By default, order information is obtained from the environment
			  variable.

	      Truncate print lines
			  that do not fit the designated or default line length.  Print lines are folded  (that  is,  wrapped  to  next  line)	by
			  default.

	      Expand input tabs to column positions
			  k+1,	2*k+1, 3*k+1, etc.  Tab characters in the input are expanded to the appropriate number of spaces.  If k is 0 or is
			  omitted, default tab settings at every eighth position is assumed.  If cd (any non-digit  character)	is  given,  it	is
			  treated  as  the  input tab character.  The default for c is the tab character.  always expands input tabs.  This option
			  provides a way to change the tab character and setting.  If this option is specified, at least one of the  parameters  c
			  or k must be given.

	      Designate a number as the print
			  margin.   The  print margin is the column where truncation or folding takes place.  The print margin determines how many
			  characters appear on a single line and can never exceed the print width.  The print margin is relative to the justifica-
			  tion.   If the print margin is 80, folding or truncation occurs at column 80 starting from the right during a right jus-
			  tification.  Similarly, folding or truncation occurs at column 80 starting from the left during  a  left  justification.
			  By default, the print margin is set to column 80.

	      Designates a number as the print
			  width.   The	print  width is the maximum number of columns in the print line.  Print width determines the start of text
			  during a right justification.  The larger the print width, the further to the right the text will start.  By default, an
			  80-column print width is used.

EXTERNAL INFLUENCES

   Environment Variables
       The  environment  variable  determines the mode and order of the file.  The syntax of is [mode][_order].  mode describes the mode of a file
       where represents Latin mode and represents non-Latin mode.  Non-Latin mode is assumed for values other than and order  describes  the  data
       order  of a file where is keyboard and is screen.  Keyboard order is assumed for values other than and Mode and order information in can be
       overridden from the command line.

       The environment variable determines the direction of a language (left-to-right or right-to-left) and whether context analysis of characters
       is necessary.

       The environment variable determines whether a language has alternative numbers.

       The environment variable determines the language in which messages are displayed.

   International Code Set Support
       Single-byte character code sets are supported.

EXAMPLES

       Right justify on a 132-column printer with a print margin at column 80 (the default):

       Right justify output of with line numbers on a 132-column printer with a print margin at column 132:

WARNINGS

       If with line numbers option) is piped to the separator character must be a tab(0x09).

       It is the user's responsibility to ensure that the environment variable accurately reflects the status of the file.

       Mode  and  justification must be consistent.  Only non-Latin-mode files can be right justified in a meaningful way.  Similarly, only Latin-
       mode files can be safely left justified.  If mode and justification do not match, the results are undefined.

       If present, alternative numbers always have a left-to-right orientation.

       The command is HP proprietary, not portable to other vendors' systems, and will not be provided in future HP-UX releases.

AUTHOR

       was developed by HP.

SEE ALSO

       forder(1), lp(1), pr(1), strord(3C).

																	 nljust(1)