Remove duplicates from a file Post: 302388967

Login or Register to Ask a Question and Join Our Community

Sponsored Content

Top Forums Shell Programming and Scripting Remove duplicates from a file Post 302388967 by EAGL� on Friday 22nd of January 2010 04:13:43 AM

Old

01-22-2010

Registered User

Hello,
i found this code on a web page which is said to be valid for only gnu linux and delete all lines except duplicate ones, i hope it works (sorry im using solaris10, couldnt try)

# delete all lines except duplicate lines (emulates "uniq -d").

Code:

sed '$!N; s/^\(.*\)\n\1$/\1/; t; D' infile

EAGL�

View Public Profile for EAGL�

Find all posts by EAGL�

10 More Discussions You Might Find Interesting

1. Shell Programming and Scripting

Remove duplicates from File from specific location

How can i remove the duplicate lines from a file, for example sample123456Sample testing123456testing XXXXX131323XXXXX YYYYY423432YYYYY fsdfdsf123456gsdfdsd all the duplicates from column 6-12 , must be deleted. I want to consider the first row, if same comes in the given range i want to...

2. Shell Programming and Scripting

remove duplicates within a block in a file..help required

hi.. i have a file in the following format :- name-a age -12 address-123 age-12 phone-22222 ============ name-ab age -11 address-123 age-11 phone-222223 ============= name-abc age -12 address-1234 age-12 phone-2222223 =============

3. Shell Programming and Scripting

Shell script to remove duplicates lines in a file

Hi, I am writing a shell script that needs to remove duplicate lines within a file by category. example: section a a c b a section b a b a c I need to remove the duplicates within th category with out removing the duplicates from the 2 different sections (one of the a's in section...

4. Shell Programming and Scripting

Remove duplicates from end of file

1/p ---- A B C A C o/p --- B A C From input file it should remove duplicates from end without changing order

5. Shell Programming and Scripting

Search based on 1,2,4,5 columns and remove duplicates in the same file.

Hi, I am unable to search the duplicates in a file based on the 1st,2nd,4th,5th columns in a file and also remove the duplicates in the same file. Source filename: Filename.csv "1","ccc","information","5000","temp","concept","new" "1","ddd","information","6000","temp","concept","new"...

6. Shell Programming and Scripting

How to remove duplicates from the .dat file

All, I have a file 1181CUSTOMER-L061411_003500.dat.Z having duplicate records in it. bash-2.05$ zcat 1181CUSTOMER-L061411_003500.dat.Z|grep "90876251S" 90876251S|ABG, AN ADAYANA COMPANY|3550 DEPAUW BLVD|||US|IN|INDIANAPOLIS||DAL|46268||||||GEN|||||||USD|||ABG, AN ADAYANA...

7. UNIX for Dummies Questions & Answers

Remove duplicates from a file

Can u tell me how to remove duplicate records from a file?

8. UNIX for Dummies Questions & Answers

Remove duplicates and keep them in a separate file

Hi, I have a tablular separated file and I want to remove all the rows that have duplicates. The diuplicates I need to check are in column 13. I have tried to use awk but I have no Idea how to keep the duplicate file. awk 'FNR==NR{a++;next}(a> 1)' tomodify.txt tomodify.txt > new.txt ...

9. Shell Programming and Scripting

To remove duplicates from pipe delimited file

Hi some one please help me to remove duplicates from a pipe delimited file based on first two columns. 123|asdf|sfsd|qwrer 431|yui|qwer|opws 123|asdf|pol|njio Here My first record and last record are duplicates.As per my requirement I want all the latest records into one file. I want the...

10. UNIX for Advanced & Expert Users

Remove duplicates in flat file

Hi all, I have a issues while loading a flat file to the DB. It is taking much time. When analyzed i found out that there are duplicates entry in the flat file. There are 2 type of Duplicate entry. 1) is entire row is duplicate. ( i can use sort | uniq) to remove the duplicated entry. 2) the...

LEARN ABOUT V7

uniq

UNIQ(1) 						      General Commands Manual							   UNIQ(1)

NAME

       uniq - report repeated lines in a file

SYNOPSIS

       uniq [ -udc [ +n ] [ -n ] ] [ input [ output ] ]

DESCRIPTION

       Uniq  reads  the  input file comparing adjacent lines.  In the normal case, the second and succeeding copies of repeated lines are removed;
       the remainder is written on the output file.  Note that repeated lines must be adjacent in order to be found; see sort(1).  If the -u  flag
       is  used, just the lines that are not repeated in the original file are output.	The -d option specifies that one copy of just the repeated
       lines is to be written.	The normal mode output is the union of the -u and -d mode outputs.

       The -c option supersedes -u and -d and generates an output report in default style but with each line preceded by a count of the number	of
       times it occurred.

       The n arguments specify skipping an initial portion of each line in the comparison:

       -n      The  first n fields together with any blanks before each are ignored.  A field is defined as a string of non-space, non-tab charac-
	       ters separated by tabs and spaces from its neighbors.

       +n      The first n characters are ignored.  Fields are skipped before characters.

SEE ALSO

       sort(1), comm(1)

																	   UNIQ(1)