Sponsored Content
Full Discussion: Gsub function in awk
Top Forums Shell Programming and Scripting Gsub function in awk Post 302831721 by Don Cragun on Thursday 11th of July 2013 02:51:23 PM
Old 07-11-2013
Quote:
Originally Posted by yifangt
Thanks a lot Yoda, and all!

Yes, the condition and action part is what I missed. However, when I did:
Code:
awk 'gsub(/[[:punct:]_[:blank:]]/, " ", $0)' text.txt

which worked fine!
Code:
This is a test for gsub
I typed this random text file
which contains punctuation like           etc 
The script should remove all the punctuations

How come like this? Thank you again!
When condition evaluates to true (non-empty string or non-zero arithmetic value), the action defaults to printing the current line. Since gsub() returns the number of substitutions performed and all of your input lines contained a space character; changing each space (by [:blank:] matching a space and then changing it to a space), got you what you wanted. Having the underscore in your regular expression is redundant since underscore is a punctuation character.

Try your script again when your input file contains an empty line (nothing but a newline) and a line that contains a single word with no leading or trailing spaces or punctuation to see the difference.
This User Gave Thanks to Don Cragun For This Post:
 

10 More Discussions You Might Find Interesting

1. Shell Programming and Scripting

use var in gsub of awk

Hi all, This problem has cost me half a day, and i still do not know how to do. Any help will be appreciated. Thanks advance. I want to use a variable as the first parameters of gsub function of awk. Example: { ... arri]=gsub(i,tolower(i),$1) (which should be ambraced by //) ... } (1 Reply)
Discussion started by: summer_cherry
1 Replies

2. Shell Programming and Scripting

Help with AWK and gsub

Hello, I have a variable that displays the following results from a JVM.... 1602100K->1578435K I would like to collect the value of 1578435 which is the value after a garbage collection. I've tried the following command but it looks like I can't get the > to work. Any suggestions as... (4 Replies)
Discussion started by: npolite
4 Replies

3. Shell Programming and Scripting

awk gsub

Hi all I want to do a simple substitution in awk but I am getting unexpected output. My function accepts a time and then prints out a validation message if the time is valid. However some times may include a : and i want to strip this out if it exists before i get to the validation. I have shown... (4 Replies)
Discussion started by: pxy2d1
4 Replies

4. Shell Programming and Scripting

Awk Gsub Query

Hi, Can some one please explain the following line please throw some light on the ones marked in red awk '{print $9}' ${FTP_LOG} | awk -v start=${START_DATE} 'BEGIN { FS = "." } { old_line1=$0; gsub(/\-/,""); if ( $3 >= start ) print old_line1 }' | awk -v end=${END_DATE} 'BEGIN { FS="." } {... (3 Replies)
Discussion started by: crosairs
3 Replies

5. Shell Programming and Scripting

Awk gsub error.

I want to replace comma with space and "*646#" with space. I am using the following code: nawk -F"|" '{gsub(","," ",$3); gsub(/\*646\#/"," ",$3);print}' OFS="|" file I am getting following error: Help is appreciated (5 Replies)
Discussion started by: pinnacle
5 Replies

6. Shell Programming and Scripting

awk gsub with variables?

Hey, I would like to replace a string by a new one. Teh problem is that both strings should be variables to be flexible, because I am having a lot of files (with the same structure, but in different folders) for i in daysim_* do cd $i/5/ folder=`pwd |awk '{print $1}'` awk '{ if... (3 Replies)
Discussion started by: ergy1983
3 Replies

7. Shell Programming and Scripting

Awk; gsub in fields 3 and 4

I want to transform a log file into input for a database. Here's the log file: Tue Aug 4 20:17:01 PDT 2009 Wireless users: 339 Daily Average: 48.4285 = Tue Aug 11 20:17:01 PDT 2009 Wireless users: 295 Daily Average: 42.1428 = Tue Aug 18 20:17:01 PDT 2009 Wireless users: 294 Daily... (6 Replies)
Discussion started by: Bubnoff
6 Replies

8. Shell Programming and Scripting

Using of gsub function in AWK to replace space by underscore

I must design a UNIX script to monitor files whose size is over a threshold of 5 MB in a specific UNIX directory I meet a problem during the for loop in my script. Some file names contain spaces. ls -lrt | awk '$5>=5000000 && length($8)==5 {gsub(/ /,"_",$9); print};' -rw-r--r-- 1 was61 ... (2 Replies)
Discussion started by: Scofield38
2 Replies

9. Shell Programming and Scripting

awk gsub

Hi, I want to print the first column with original value and without any double quotes The output should look like <original column>|<column without quotes> $ cat a.txt "20121023","19301229712","100397" "20121023","19361629712","100778" "20121030A","19361630412","100838"... (3 Replies)
Discussion started by: ysrini
3 Replies

10. Shell Programming and Scripting

Using multiple gsub() function under a loop in awk

Hi ALL, I want to replace string occurrence in my file "Config" using a external file named "Mapping" using awk. $cat Config ! Configuration file for RAVI ! Configuration file for RACHANA ! Configuration file for BALLU $cat Mapping ravi:ram rachana:shyam ballu:hameed The... (5 Replies)
Discussion started by: useless79
5 Replies
SED(1)							      General Commands Manual							    SED(1)

NAME
sed - stream editor SYNOPSIS
sed [ -n ] [ -e script ] [ -f sfile ] [ file ] ... DESCRIPTION
Sed copies the named files (standard input default) to the standard output, edited according to a script of commands. The -f option causes the script to be taken from file sfile; these options accumulate. If there is just one -e option and no -f's, the flag -e may be omitted. The -n option suppresses the default output. A script consists of editing commands, one per line, of the following form: [address [, address] ] function [arguments] In normal operation sed cyclically copies a line of input into a pattern space (unless there is something left after a `D' command), applies in sequence all commands whose addresses select that pattern space, and at the end of the script copies the pattern space to the standard output (except under -n) and deletes the pattern space. An address is either a decimal number that counts input lines cumulatively across files, a `$' that addresses the last line of input, or a context address, `/regular expression/', in the style of ed(1) modified thus: The escape sequence ` ' matches a newline embedded in the pattern space. A command line with no addresses selects every pattern space. A command line with one address selects each pattern space that matches the address. A command line with two addresses selects the inclusive range from the first pattern space that matches the first address through the next pattern space that matches the second. (If the second address is a number less than or equal to the line number first selected, only one line is selected.) Thereafter the process is repeated, looking again for the first address. Editing commands can be applied only to non-selected pattern spaces by use of the negation function `!' (below). In the following list of functions the maximum number of permissible addresses for each function is indicated in parentheses. An argument denoted text consists of one or more lines, all but the last of which end with `' to hide the newline. Backslashes in text are treated like backslashes in the replacement string of an `s' command, and may be used to protect initial blanks and tabs against the stripping that is done on every script line. An argument denoted rfile or wfile must terminate the command line and must be preceded by exactly one blank. Each wfile is created before processing begins. There can be at most 10 distinct wfile arguments. (1)a text Append. Place text on the output before reading the next input line. (2)b label Branch to the `:' command bearing the label. If label is empty, branch to the end of the script. (2)c text Change. Delete the pattern space. With 0 or 1 address or at the end of a 2-address range, place text on the output. Start the next cycle. (2)d Delete the pattern space. Start the next cycle. (2)D Delete the initial segment of the pattern space through the first newline. Start the next cycle. (2)g Replace the contents of the pattern space by the contents of the hold space. (2)G Append the contents of the hold space to the pattern space. (2)h Replace the contents of the hold space by the contents of the pattern space. (2)H Append the contents of the pattern space to the hold space. (1)i text Insert. Place text on the standard output. (2)l List the pattern space on the standard output in an unambiguous form. Non-printing characters are spelled in two digit ascii, and long lines are folded. (2)n Copy the pattern space to the standard output. Replace the pattern space with the next line of input. (2)N Append the next line of input to the pattern space with an embedded newline. (The current line number changes.) (2)p Print. Copy the pattern space to the standard output. (2)P Copy the initial segment of the pattern space through the first newline to the standard output. (1)q Quit. Branch to the end of the script. Do not start a new cycle. (2)r rfile Read the contents of rfile. Place them on the output before reading the next input line. (2)s/regular expression/replacement/flags Substitute the replacement string for instances of the regular expression in the pattern space. Any character may be used instead of `/'. For a fuller description see ed(1). Flags is zero or more of g Global. Substitute for all nonoverlapping instances of the regular expression rather than just the first one. p Print the pattern space if a replacement was made. w wfile Write. Append the pattern space to wfile if a replacement was made. (2)t label Test. Branch to the `:' command bearing the label if any substitutions have been made since the most recent reading of an input line or execution of a `t'. If label is empty, branch to the end of the script. (2)w wfile Write. Append the pattern space to wfile. (2)x Exchange the contents of the pattern and hold spaces. (2)y/string1/string2/ Transform. Replace all occurrences of characters in string1 with the corresponding character in string2. The lengths of string1 and string2 must be equal. (2)! function Don't. Apply the function (or group, if function is `{') only to lines not selected by the address(es). (0): label This command does nothing; it bears a label for `b' and `t' commands to branch to. (1)= Place the current line number on the standard output as a line. (2){ Execute the following commands through a matching `}' only when the pattern space is selected. (0) An empty command is ignored. SEE ALSO
ed(1), grep(1), awk(1) SED(1)
All times are GMT -4. The time now is 01:41 PM.
Unix & Linux Forums Content Copyright 1993-2022. All Rights Reserved.
Privacy Policy