Sort, Uniq, Duplicates Post: 302117708

10 More Discussions You Might Find Interesting

1. UNIX for Dummies Questions & Answers

sort/uniq

I have a file: Fred Fred Fred Jim Fred Jim Jim If sort is executed on the listed file, shouldn't the output be?: Fred Fred Fred Fred Jim Jim Jim

2. UNIX for Dummies Questions & Answers

Help with Last,uniq, sort and cut

Using the last, uniq, sort and cut commands, determine how many times the different users have logged in. I know how to use the last command and cut command... i came up with last | cut -f1 -d" " | uniq i dont know if this is right, can someone please help me... thanks

3. Shell Programming and Scripting

sort and uniq in perl

Does anyone have a quick and dirty way of performing a sort and uniq in perl? How an array with data like: this is bkupArr BOLADVICE_VN this is bkupArr MLT6800PROD2A this is bkupArr MLT6800PROD2A this is bkupArr BOLADVICE_VN_7YR this is bkupArr MLT6800PROD2A I want to sort it...

4. Shell Programming and Scripting

Removing duplicates [sort , uniq]

Hey Guys, I have file which looks like this, Contig201#numbPA Contig1452#nmdynD6PA dm022p15.r#CG6461PA dm005e16.f#SpatPA IGU001_0015_A06.f#CG17593PA I need to remove duplicates based on the chracter matching upto '#'. for example if we consider this.. Contig201#numbPA...

5. Shell Programming and Scripting

Help with Uniq and sort

The key is first field i want only uniq record for the first field in file. I want the output as or output as Appreciate help on this

6. Shell Programming and Scripting

sort | uniq question

Hello, I have a large data file: 1234 8888 bbb 2745 8888 bbb 9489 8888 bbb 1234 8888 aaa 4838 8888 aaa 3977 8888 aaa I need to remove duplicate lines (where the first column is the duplicate). I have been using: sort file.txt | uniq -w4 > newfile.txt However, it seems to keep the...

7. Shell Programming and Scripting

Sort and uniq after comparision

Hi All, I have a text file with the format shown below. Some of the records are duplicated with the only exception being date (Field 15). I want to compare all duplicate records using subscriber number (field 7) and keep only those records with greater date. ...

8. Shell Programming and Scripting

Sort uniq or awk

Hi again, I have files with the following contents datetime,ip1,port1,ip2,port2,number How would I find out how many times ip1 field shows up a particular file? Then how would I find out how many time ip1 and port 2 shows up? Please mind the file may contain 100k lines.

9. Shell Programming and Scripting

Uniq or sort -u or similar only between { }

Hi ! I am trying to remove doubbled entrys in a textfile only between delimiters. Like that example but i dont know how to do that with sort or similar. input: { aaa aaa } { aaa aaa } output: { aaa } {

10. UNIX for Dummies Questions & Answers

Uniq and sort -u

Hello all, Need to pick your brains, I have a 10Gb file where each row is a name, I am expecting about 50 names in total. So there are a lot of repetitions in clusters. So I want to do a sort -u file Will it be considerably faster or slower to use a uniq before piping it to sort...

LEARN ABOUT DEBIAN

tabmerge

TABMERGE(1p)						User Contributed Perl Documentation					      TABMERGE(1p)

NAME

       tabmerge - unify delimited files on common fields

SYNOPSIS

	 tabmerge [action] [options] file1 file2 [...]

       Actions:

	 --min		      Take only fields present in all files [DEFAULT]
	 --max		      Take all fields present
	 -f|--fields=f1[,f2]  Take only the fields mentioned in the
			      comma-separated list

       Options:

	 -l|--list	      List available fields
	 --fs=x 	      Use "x" as the field separator
			      (default is tab "	")
	 --rs=x 	      Use "x" as the record separator
			      (default is newline "
")
	 -s|--sort=f1[,f2]    Sort data ASCII-betically on field(s)
	 --stdout	      Print data in original delimited format
			      (i.e., not in a table format)

	 --help 	      Show brief help and quit
	 --man		      Show full documentation

DESCRIPTION

       This program merges the fields -- not the rows -- of delimited text files.  That is, if several files are almost but not quite entirely
       unlike each other in their structure (in their field names, numbers or orders), this script allows you to easily unify the files into one
       file with all the same fields.  The output can be based on fields as determined by the three "action" flags.

       For the following examples, consider three files that contain the following fields:

	 +------------+---------------------------------+
	 | File       | Fields				|
	 +------------+---------------------------------+
	 | merge1.tab | name, type, position		|
	 | merge2.tab | name, type, position, lod_score |
	 | merge3.tab | name, position			|
	 +------------+---------------------------------+

       To list all available fields in the files and the number of times they are present:

	 $ tabmerge --list merge*
	 +-----------+-------------------+
	 | Field     | No. Times Present |
	 +-----------+-------------------+
	 | lod_score | 1		 |
	 | name      | 3		 |
	 | position  | 3		 |
	 | type      | 2		 |
	 +-----------+-------------------+

       To merge the files on the minimum overlapping fields:

	 $ tabmerge merge*
	 +----------+----------+
	 | name     | position |
	 +----------+----------+
	 | RM104    | 2.30     |
	 | RM105    | 4.5      |
	 | TX5509   | 10.4     |
	 | UU189    | 19.0     |
	 | Xpsm122  | 3.3      |
	 | Xpsr9556 | 4.5      |
	 | DRTL     | 2.30     |
	 | ALTX     | 4.5      |
	 | DWRF     | 10.4     |
	 +----------+----------+

       To merge the files and include all the fields:

	 $ tabmerge --max merge*
	 +-----------+----------+----------+--------+
	 | lod_score | name	| position | type   |
	 +-----------+----------+----------+--------+
	 |	     | RM104	| 2.30	   | RFLP   |
	 |	     | RM105	| 4.5	   | RFLP   |
	 |	     | TX5509	| 10.4	   | AFLP   |
	 | 2.4	     | UU189	| 19.0	   | SSR    |
	 | 1.2	     | Xpsm122	| 3.3	   | Marker |
	 | 1.2	     | Xpsr9556 | 4.5	   | Marker |
	 |	     | DRTL	| 2.30	   |	    |
	 |	     | ALTX	| 4.5	   |	    |
	 |	     | DWRF	| 10.4	   |	    |
	 +-----------+----------+----------+--------+

       To merge and extract just the "name" and "type" fields:

	 $ tabmerge -f name,type merge*
	 +----------+--------+
	 | name     | type   |
	 +----------+--------+
	 | RM104    | RFLP   |
	 | RM105    | RFLP   |
	 | TX5509   | AFLP   |
	 | UU189    | SSR    |
	 | Xpsm122  | Marker |
	 | Xpsr9556 | Marker |
	 | DRTL     |	     |
	 | ALTX     |	     |
	 | DWRF     |	     |
	 +----------+--------+

       To merge the files on just the "name" and "lod_score" fields and sort on the name:

	 $ tabmerge -f name,lod_score -s name merge*
	 +----------+-----------+
	 | name     | lod_score |
	 +----------+-----------+
	 | ALTX     |		|
	 | DRTL     |		|
	 | DWRF     |		|
	 | RM104    |		|
	 | RM105    |		|
	 | TX5509   |		|
	 | UU189    | 2.4	|
	 | Xpsm122  | 1.2	|
	 | Xpsr9556 | 1.2	|
	 +----------+-----------+

       To do the same but mimic the original tab-delimited input:

	 $ tabmerge -f name,lod_score -s name --stdout merge*
	 name	 lod_score
	 ALTX
	 DRTL
	 DWRF
	 RM104
	 RM105
	 TX5509
	 UU189	 2.4
	 Xpsm122 1.2
	 Xpsr9556	 1.2

       Why would you want to do this?  Suppose you have several delimited text files with nearly the same structure and want to create just one
       file from them, but the fields may be in a different order in each file and/or some files may contain more or fewer fields than others.
       (As far-fetched as it may seem, it happens to the author more than he'd like.)

SEE ALSO

       o   Text::RecordParser

       o   Text::TabularDisplay

AUTHOR

       Ken Youens-Clark <kclark@cpan.org>.

LICENSE AND COPYRIGHT

       Copyright (C) 2006-10 Ken Youens-Clark.	All rights reserved.

       This program is free software; you can redistribute it and/or modify it under the terms of the GNU General Public License as published by
       the Free Software Foundation; version 2.

       This program is distributed in the hope that it will be useful, but WITHOUT ANY WARRANTY; without even the implied warranty of
       MERCHANTABILITY or FITNESS FOR A PARTICULAR PURPOSE.  See the GNU General Public License for more details.

perl v5.10.1							    2010-07-26							      TABMERGE(1p)

10 More Discussions You Might Find Interesting

1. UNIX for Dummies Questions & Answers

sort/uniq

Discussion started by: jimmyflip

2. UNIX for Dummies Questions & Answers

Help with Last,uniq, sort and cut

Discussion started by: jay1228

3. Shell Programming and Scripting

sort and uniq in perl

Discussion started by: reggiej

4. Shell Programming and Scripting

Removing duplicates [sort , uniq]

Discussion started by: sharatz83