Awk based script to find the median of all individual columns in a data file Post: 302653173

10 More Discussions You Might Find Interesting

1. Shell Programming and Scripting

awk based script to print the "mode(statistics term)" for each column in a data file

Hi All, Thanks all for the continued support so far. Today, I need to find the most occurring string/number(also called mode in statistics terminology) for each column in a data file (.csv type). For one column of data(1.txt) like below Sample 1 2 2 3 4 1 1 1 2 I can find the mode...

2. Shell Programming and Scripting

awk based script to find the average of all the columns in a data file

Hi All, I need the modification for the below mentioned code (found in one more post https://www.unix.com/shell-programming-scripting/27161-script-generate-average-values.html) to find the average values for all the columns(but for a specific rows) and print the averages side by side. I have...

3. Shell Programming and Scripting

awk based script to ignore all columns from a file which contains character strings

Hello All, I have a .CSV file where I expect all numeric data in all the columns other than column headers. But sometimes I get the files (result of statistics computation by other persons) like below( sample data) SNO,Data1,Data2,Data3 1,2,3,4 2,3,4,SOME STRING 3,4,Inf,5 4,5,4,4 I...

4. Shell Programming and Scripting

Help with awk replacing identical columns based on another file

Hello, I am using Awk in UBUNTU 12.04. I have a file like following with three fields and 44706 rows. F1 A A F2 G G F3 A T I have another file like this: AL_1 F1 A A AL_2 F1 A T AL_3 F1 A A AL_1 F2 G G AL_2 F2 G A AL_3 F2 G G BO_1 F1 A A BO_2 F1 A T...

5. Shell Programming and Scripting

awk script to split file into multiple files based on many columns

So I have a space delimited file that I'd like to split into multiple files based on multiple column values. This is what my data looks like 1bc9A02 1 10 1000 FTDLNLVQALRQFLWSFRLPGEAQKIDRMMEAFAQRYCQCNNGVFQSTDTCYVLSFAIIMLNTSLHNPNVKDKPTVERFIAMNRGINDGGDLPEELLRNLYESIKNEPFKIPELEHHHHHH 1ku1A02 1 10...

6. UNIX for Dummies Questions & Answers

Median calculator based on id match

I am trying to calculate the median of a column of numbers if they match an ID type on a different column. The input file has 3 columns. The column that has the ID is column 1 and the column with the values I'd like to find the median for is column 3. The file does not need to be sorted. What I...

7. Shell Programming and Scripting

Find columns in a file based on header and print to new file

Hello, I have to fish out some specific columns from a file based on the header value. I have the list of columns I need in a different file. I thought I could read in the list of headers I need, # file with header names of required columns in required order headers_file=$2 # read contents...

8. Shell Programming and Scripting

In PErl script: need to read the data one file and generate multiple files based on the data

We have the data looks like below in a log file. I want to generat files based on the string between two hash(#) symbol like below Source: #ext1#test1.tale2 drop #ext1#test11.tale21 drop #ext1#test123.tale21 drop #ext2#test1.tale21 drop #ext2#test12.tale21 drop #ext3#test11.tale21 drop...

9. Shell Programming and Scripting

awk script to find data in three file and perform replace operation

Have three files. Any other approach with regards to file concatenation or splitting, etc is appreciated If column55(billngtype) of file1 contains YMNC or YPBC then pick the value of column13(documentnumber). Now find this documentnumber in column1(Billdoc) of file2 and grep the corresponding...

10. UNIX for Advanced & Expert Users

Need Optimization shell/awk script to aggreagte (sum) for all the columns of Huge data file

Optimization shell/awk script to aggregate (sum) for all the columns of Huge data file File delimiter "|" Need to have Sum of all columns, with column number : aggregation (summation) for each column File not having the header Like below - Column 1 "Total Column 2 : "Total ... ......

LEARN ABOUT DEBIAN

blockmedian

BLOCKMEDIAN(l)															    BLOCKMEDIAN(l)

NAME

       blockmedian - filter to block average (x,y,z) data by L1 norm.

SYNOPSIS

       blockmedian [ xyz[w]file(s) ] -Ix_inc[m|c][/y_inc[m|c]] -Rwest/east/south/north[r] [ -C ] [ -F ] [ -H[nrec] ] [ -L ] [ -Q ] [ -V ] [ -W[io]
       ] [ -: ] [ -bi[s][n] ] [ -bo[s][n] ]

DESCRIPTION

       blockmedian reads arbitrarily located (x,y,z) triples [or optionally weighted quadruples (x,y,z,w)] from standard input [or  xyz[w]file(s)]
       and  writes  to	standard output a median position and value for every non-empty block in a grid region defined by the -R and -I arguments.
       Either blockmean, blockmedian, or blockmode should be used as a pre-processor before running surface to avoid aliasing  short  wavelengths.
       These  routines	are  also  generally useful for decimating or averaging (x,y,z) data. You can modify the precision of the output format by
       editing the D_FORMAT parameter in your .gmtdefaults file, or you may choose binary input and/or output using  single  or  double  precision
       storage.

       xyz[w]file(s)
	      3  [or  4]  column ASCII file(s) [or binary, see -b] holding (x,y,z[,w]) data values. [w] is an optional weight for the data.  If no
	      file is specified, blockmedian will read from standard input.

       -I     x_inc [and optionally y_inc] is the grid spacing. Append m to indicate minutes or c to indicate seconds.

       -R     west, east, south, and north specify the Region of interest. To specify boundaries in degrees and minutes  [and  seconds],  use  the
	      dd:mm[:ss] format. Append r if lower left and upper right map coordinates are given instead of wesn.

OPTIONS

       -C     Use the center of the block as the output location [Default uses the median location (but see -Q)].  -C overrides -Q.

       -F     Block  centers  have pixel registration. [Default: grid registration.] (Registrations are defined in GMT Cookbook Appendix B on grid
	      file formats.) Each block is the locus of points nearest the grid value location. For example, with -R10/15/10/15 and and -I1:  with
	      the -F option 10 <= (x,y) < 11 is one of 25 blocks; without it 9.5 <= (x,y) < 10.5 is one of 36 blocks.

       -H     Input  file(s) has Header record(s). Number of header records can be changed by editing your .gmtdefaults file. If used, GMT default
	      is 1 header record.  Not used with binary data.

       -L     Indicates that the x column contains longitudes, which may differ from the region in -R  by  [multiples  of]  360  degrees  [Default
	      assumes no periodicity].

       -Q     (Quicker) Finds median z and (x, y) at that z [Default finds median x, median y, median z].

       -V     Selects verbose mode, which will send progress reports to stderr [Default runs "silently"].

       -W     Weighted	modifier[s].  Unweighted input and output has 3 columns x,y,z; Weighted i/o has 4 columns x,y,z,w.  Weights can be used in
	      input to construct weighted median values in blocks. Weight sums can be reported in output for later combining  several  runs,  etc.
	      Use -W for weighted i/o, -Wi for weighted input only, -Wo for weighted output only. [Default uses unweighted i/o]

       -:     Toggles  between	(longitude,latitude)  and  (latitude,longitude)  input/output. [Default is (longitude,latitude)].  Applies to geo-
	      graphic coordinates only.

       -bi    Selects binary input. Append s for single precision [Default is double].	Append n for the number of columns in the binary  file(s).
	      [Default is 3 (or 4 if -W is set) columns].

       -bo    Selects binary output. Append s for single precision [Default is double].

EXAMPLES

       To find 5 by 5 minute block medians from the double precision binary data in hawaii_b.xyg and output an ASCII table, try

       blockmedian hawaii_b.xyg -R198/208/18/25 -I5m -bi3 > hawaii_5x5.xyg

SEE ALSO

       blockmean(1gmt), blockmode(1gmt), gmt(1gmt), gmtdefaults(1gmt), nearneighbor(1gmt), surface(1gmt), triangulate(1gmt)

								    1 Jan 2004							    BLOCKMEDIAN(l)

10 More Discussions You Might Find Interesting

1. Shell Programming and Scripting

awk based script to print the "mode(statistics term)" for each column in a data file

Discussion started by: ks_reddy

2. Shell Programming and Scripting

awk based script to find the average of all the columns in a data file

Discussion started by: ks_reddy

3. Shell Programming and Scripting

awk based script to ignore all columns from a file which contains character strings

Discussion started by: ks_reddy

4. Shell Programming and Scripting

Help with awk replacing identical columns based on another file

Discussion started by: Homa

5. Shell Programming and Scripting

awk script to split file into multiple files based on many columns

Discussion started by: viored

6. UNIX for Dummies Questions & Answers

Median calculator based on id match

Discussion started by: verse123

7. Shell Programming and Scripting

Find columns in a file based on header and print to new file

Discussion started by: LMHmedchem

8. Shell Programming and Scripting

In PErl script: need to read the data one file and generate multiple files based on the data

Discussion started by: Sanjeev G

9. Shell Programming and Scripting

awk script to find data in three file and perform replace operation

Discussion started by: as7951

10. UNIX for Advanced & Expert Users

Need Optimization shell/awk script to aggreagte (sum) for all the columns of Huge data file

Discussion started by: kartikirans

LEARN ABOUT DEBIAN

blockmedian