05-15-2012 08:44 PM
I have almost 12000 html files. I want to convert these html files into text files. Is there any code to import html files in sas and convert it in to text file?
Thank you for your time.
05-15-2012 08:52 PM
You should just search for html to text converters on the web. Or perhaps html to xml.
If the format is very clean you might be able to read it with a data step.
You might also look into reading them into Excel and then importing into SAS from excel.
05-16-2012 12:56 AM
The html files are 10-k reports from SEC edgar. I want to convert these html files into text files. I also have url of these html files.
05-16-2012 02:13 AM
Oh. Maybe you should firstly download all these htmls at your local PC by using PROC HTTP or filename+url .Then using the method I posted convert it into SAS Datasets ,and use proc export to export txt files.
It is easy. I think.