I am trying to extract a news article from a web page. The sed I have written brings back a lot of Javascript code and sometimes advertisments too. Can anyone please help with this one ??? I need to fix this sed so it picks up the article ONLY (don't worry about the title or date .. i got those using a separate sed) ..
hi guys,
am required to prepare a report and mail it, to make it more appealing :p i wish to have content of mail in rich text format i.e html type
with mailx how to specify the content type of mail body as html?
Thanks in advance!!!
rishi (2 Replies)
I have a html file called myfile. If I simply put "cat myfile.html" in UNIX, it shows all the html tags like <a href=r/26><img src="http://www>. But I want to extract only text part.
Same problem happens in "type" command in MS-DOS.
I know you can do it by opening it in Internet Explorer,... (4 Replies)
Hi All,
I'm trying to extract some floating point numbers from within some HTML code like this:
<TR><TD class='awrc'>Parse CPU to Parse Elapsd %:</TD><TD ALIGN='right' class='awrc'> 64.50</TD><TD class='awrc'>% Non-Parse CPU:</TD><TD ALIGN='right' class='awrc'> ... (2 Replies)
Hello,
i try to extract urls from google-search-results, but i have problem with sed filtering of html-code.
what i wont is just list of urls thay apears between ........<p><a href=" and next following " in html code.
here is my code, i use wget and pipelines to filtering. wget works, but... (13 Replies)
I am attempting to extract weather data from the following website, but for the Victoria area only:
Text Forecasts - Environment Canada
I use this:
sed -n "/Greater Victoria./,/Fraser Valley./p"
But that phrasing does not sometimes get it all and think perhaps the website has more... (2 Replies)
Hi
I've searched for it for few hours now and i can't seem to find anything working like i want. I've got webpage, saved in file par with form like this:
<html><body><form name='sendme' action='http://example.com/' method='POST'>
<textarea name='1st'>abc123def678</textarea>
<textarea... (9 Replies)
Hi Expert,
Is there any other way to print and write to a same filename the content between two html tags?
Here the sample:
cat file.html
<div id="outline">
hello world<br>
</div>
<div id="container_faq">
test1<br>
</div>
<div class="widget_quick">
thead test<br>
</div>
... (3 Replies)
I'm extracting text between table tags in HTML
<th><a href="/wiki/Buick_LeSabre" title="Buick LeSabre">Buick LeSabre</a></th>
using this:
awk -F "</*th>" '/<\/*th>/ {print $2}' auto2 > auto3
then this (text between a href):
sed -e 's/\(<*>\)//g' auto3 > auto4
How to shorten this into one... (8 Replies)
Hi, Please see my code below i'm trying get an email send with attachment and html content in the body.
Using the code below will put the encoding for attachment in the body as well
SUBJECT="$(echo "XPI Monitoring "${tcnt}" transactions waiting \nContent-Type: text/html")"
cat... (3 Replies)
Hi
I have file like this:
jack black 104
daniel nick 75
lily harm 2
albert 5
and need to convert it into the html table like this:
NO.......name....family..... id
1...........jack.....black.....104
2..........daniel....nick.......75
3..........albert.................5
i mean... (5 Replies)
Discussion started by: indeed_1
5 Replies
LEARN ABOUT DEBIAN
news
NEWS(1) USER COMMANDS NEWS(1)NAME
news - display system news
SYNOPSIS
news [-adDeflnpvxs] [[article1] [article2] ..]
DESCRIPTION
The news command keeps you informed of news concerning the system. Each news item is contained in a separate file in the /var/lib/sysnews
directory. Anyone having write permission to this directory can create a news file.
If you run the news command without any flags, it displays every unread file in the /var/lib/sysnews directory.
Each file is preceded by an appropriate header. To avoid reporting old news, the news command stores a currency time. The news command con-
siders your currency time to be the date the $HOME/.news_time file was last modified. Each time you read the news, the modification time of
this file changes to that of the reading. Only news item files posted after this time are considered unread.
OPTIONS -a, --all
Display all news, also the already read news.
-d, --datestamp
Add a date stamp to each article name printed. this can only be used with the -nl flags.
-D, --datefmt <fmt>
Specify a date format, see the strftime(3) man page for more details. the default format is (%b %d %Y)
-f, --newsdir <dir>
Read news from an alternate newsdir.
-l, --oneperline
One article name per line.
-n, --names
Only show the names of news articles.
-p, --page
Pipe articles through $PAGER or more(1) if the $PAGER environment variable is not set.
-s, --articles
Reports the number of news articles.
MAINTAINER OPTIONS -e, --expire #
Expire news older than # days.
-x, --exclude a,b,c
A comma separated list of articles which may not be expired. if a file named .noexpire exists in the /var/lib/sysnews direcory,
filenames are read from it also. names in this file may be comma separated, and/or one per line.
AUTHOR
Charles, <int@link.xs4all.nl>
Linux 18 January 1995 NEWS(1)