Solved

extract data from web

Posted on 2011-03-18
8
177 Views
Last Modified: 2013-11-19
Hi,

collecting data from websites manually is very hard and time consuming, I'm looking for free application to extract data from web and put it in database or file. for example: I need to get the university information (faculties, department, members, contact information .. etc )

I found different application, but it is not free and hard to learn and customized. can you please guide me to find easy, free and powerful application that do the example mentioned above in a short time.

thanks
0
Comment
Question by:nmokhayesh
  • 4
  • 3
8 Comments
 
LVL 75

Expert Comment

by:Michel Plungjan
ID: 35170400
0
 
LVL 109

Expert Comment

by:Ray Paseur
ID: 35173465
Do you want to search the contents of web sites, or do you want to copy the web sites?
0
 

Author Comment

by:nmokhayesh
ID: 35173574
I need to copy selected contents
for instance
list of all professors names and their contacts (email , tel, website, research interests)
list of all departments and contact information
list of schools and programs discription in each one

thank
 
0
Master Your Team's Linux and Cloud Stack!

The average business loses $13.5M per year to ineffective training (per 1,000 employees). Keep ahead of the competition and combine in-person quality with online cost and flexibility by training with Linux Academy.

 
LVL 109

Expert Comment

by:Ray Paseur
ID: 35173641
I have used httrack and it worked fairly well to make a copy of the web site onto my hard drive.  Selection of the contents was still the major issue.  Although the local web site was faster than using the internet, you would still have to manually or programmatically isolate the information you wanted to keep.

It might be possible to get a copy of Wrensoft Zoom Indexer and use that to spider the site.  Caveat: I have never tried that on a site that I did not control.

One other possibility might be to contact the site owners and ask if they can isolate this information for you.  Educational institutions are often willing to help with requests like this.
0
 

Author Comment

by:nmokhayesh
ID: 35216825
OK I need to extract selected data from some university web pages to XML or excel sheet file using web scraping software

can you please tell me which free web scraping application can do this job in easy way. I search it but i got a lot of application but I do not know which one is useful/efficient

Thanks
Naif
0
 
LVL 109

Accepted Solution

by:
Ray Paseur earned 500 total points
ID: 35220096
There is no "easy way" because there is no clear vision of what you want to extract.  Each university web site is likely to be a bespoke application, so each such scraping and extraction algorithm will require custom programming.

That is why I recommended that you contact the site owners and ask if they can isolate this information for you.
0
 

Author Closing Comment

by:nmokhayesh
ID: 36710149
still not solved 100%
0
 
LVL 109

Expert Comment

by:Ray Paseur
ID: 36710198
No, it will never get solved 100%, full stop.  Here is what you asked for back in March (how many months ago was that?)

...find easy, free and powerful application that do the example mentioned above in a short time.

It would surprise me if you find easy, free and powerful all in the same package.  Those things are like Ohm's law.  Fix any two variables and the third is determined.
0

Featured Post

The New “Normal” in Modern Enterprise Operations

DevOps for the modern enterprise offers many benefits — increased agility, productivity, and more, but digital transformation isn’t easy, especially if you’re not addressing the right issues. Register for the webinar to dive into the “new normal” for enterprise modern ops.

Question has a verified solution.

If you are experiencing a similar issue, please ask a related question

Suggested Solutions

Title # Comments Views Activity
Help with query 3 31
Select record with the most recent date 14 67
Designing forms 3 19
Powershell Exchange mailboxsizes 3 11
This guide will walk you through the essential considerations and tech stack for building scalable websites. Know how to grow your business the smart way!
Today, the web development industry is booming, and many people consider it to be their vocation. The question you may be asking yourself is – how do I become a web developer?
The purpose of this video is to demonstrate how to set up an RSS Feed on a WordPress Website. This will be demonstrated using a Windows 8 PC. Feedburner will be used for this demonstration. Go to your WordPress login page. This will look like the…
In this seventh video of the Xpdf series, we discuss and demonstrate the PDFfonts utility, which lists all the fonts used in a PDF file. It does this via a command line interface, making it suitable for use in programs, scripts, batch files — any pl…

820 members asked questions and received personalized solutions in the past 7 days.

Join the community of 500,000 technology professionals and ask your questions.

Join & Ask a Question