In a huge file, Delete duplicate lines leaving unique lines Post: 302543885

10 More Discussions You Might Find Interesting

1. Shell Programming and Scripting

Delete lines from huge file

I have to delete 1st 7000 lines of a file which is 12GB large. As it is so large, i can't open in vi and delete these lines. Also I found one post here which gave solution using perl, but I don't have perl installed. Also some solutions were redirecting the o/p to a different file and renaming it....

2. Shell Programming and Scripting

delete semi-duplicate lines from file?

Ok here's what I'm trying to do. I need to get a listing of all the mountpoints on a system into a file, which is easy enough, just using something like "mount | awk '{print $1}'" However, on a couple of systems, they have some mount points looking like this: /stage /stand /usr /MFPIS...

3. UNIX for Dummies Questions & Answers

Delete duplicate lines and print to file

OK, I have read several things on how to do this, but can't make it work. I am writing this to a vi file then calling it as an awk script. So I need to search a file for duplicate lines, delete duplicate lines, then write the result to another file, say /home/accountant/files/docs/nodup ...

4. UNIX for Dummies Questions & Answers

How to delete or remove duplicate lines in a file

Hi please help me how to remove duplicate lines in any file. I have a file having huge number of lines. i want to remove selected lines in it. And also if there exists duplicate lines, I want to delete the rest & just keep one of them. Please help me with any unix commands or even fortran...

5. UNIX for Dummies Questions & Answers

Delete lines with duplicate strings based on date

Hey all, a relative bash/script newbie trying solve a problem. I've got a text file with lots of lines that I've been able to clean up and format with awk/sed/cut, but now I'd like to remove the lines with duplicate usernames based on time stamp. Here's what the data looks like 2007-11-03...

6. UNIX for Dummies Questions & Answers

How to delete partial duplicate lines unix

hi :) I need to delete partial duplicate lines I have this in a file sihp8027,/opt/cf20,1980182 sihp8027,/opt/oracle/10gRelIIcd,155200016 sihp8027,/opt/oracle/10gRelIIcd,155200176 sihp8027,/var/opt/ERP,10376312 and need to leave it like this: sihp8027,/opt/cf20,1980182...

7. Shell Programming and Scripting

Delete lines in file containing duplicate strings, keeping longer strings

The question is not as simple as the title... I have a file, it looks like this <string name="string1">RZ-LED</string> <string name="string2">2.0</string> <string name="string2">Version 2.0</string> <string name="string3">BP</string> I would like to check for duplicate entries of...

8. Shell Programming and Scripting

Delete duplicate lines... with a twist!

Hi, I'm sorry I'm no coder so I came here, counting on your free time and good will to beg for spoonfeeding some good code. I'll try to be quick and concise! Got file with 50k lines like this: "Heh, heh. Those darn ninjas. They're _____."*wacky The "canebrake", "timber" & "pygmy" are types...

9. UNIX for Beginners Questions & Answers

How to delete identical lines while leaving one undeleted?

Hi, I have a file as follows. file1 Hello Hi His Hi Hi Hungry hi so I want to delete identical lines while leaving one of them undeleted. So desired output will be Hello Hi

10. UNIX for Beginners Questions & Answers

Delete duplicate like pattern lines

Hi I need to delete duplicate like pattern lines from a text file containing 2 duplicates only (one being subset of the other) using sed or awk preferably. Input: FM:Chicago:Development FM:Chicago:Development:Score SR:Cary:Testing:Testcases PM:Newyork:Scripting PM:Newyork:Scripting:Audit...

LEARN ABOUT MINIX

sort

SORT(1) 						      General Commands Manual							   SORT(1)

NAME

       sort - sort a file of ASCII lines

SYNOPSIS

       sort [-bcdfimnru] [-tc]	[-o name] [+pos1] [-pos2] file ...

OPTIONS

       -b     Skip leading blanks when making comparisons

       -c     Check to see if a file is sorted

       -d     Dictionary order: ignore punctuation

       -f     Fold upper case onto lower case

       -i     Ignore nonASCII characters

       -m     Merge presorted files

       -n     Numeric sort order

       -o     Next argument is output file

       -r     Reverse the sort order

       -t     Following character is field separator

       -u     Unique mode (delete duplicate lines)

EXAMPLES

       sort -nr file	   # Sort keys numerically, reversed

       sort +2 -4 file	   # Sort using fields 2 and 3 as key

       sort +2 -t: -o out  # Field separator is :

       sort +.3 -.6	   # Characters 3 through 5 form the key

DESCRIPTION

       Sort  sorts  one or more files.	If no files are specified, stdin is sorted.  Output is written on standard output, unless -o is specified.
       The options +pos1 -pos2 use only fields pos1 up to but not including pos2 as the sort key, where a field is a string of	characters  delim-
       ited  by  spaces and tabs, unless a different field delimiter is specified with -t.  Both pos1 and pos2 have the form m.n where m tells the
       number of fields and n tells the number of characters.  Either m or n may be omitted.

SEE ALSO

       comm(1), grep(1), uniq(1).

																	   SORT(1)

10 More Discussions You Might Find Interesting

1. Shell Programming and Scripting

Delete lines from huge file

Discussion started by: rahulrathod

2. Shell Programming and Scripting

delete semi-duplicate lines from file?

Discussion started by: paqman

3. UNIX for Dummies Questions & Answers

Delete duplicate lines and print to file

Discussion started by: bfurlong

4. UNIX for Dummies Questions & Answers

How to delete or remove duplicate lines in a file

Discussion started by: reva