RE: searching/indexing pdf files from a web page
by "Ann Ezzell" <amcbainezzell(at)alum.mit.edu>
|
| Date: |
Tue, 23 Jul 2002 18:32:45 -0700 |
| To: |
"'Hank Marquardt'" <hmarq(at)yerpso.net> |
| Cc: |
"'Johnson, Mark'" <JohnsonM(at)issaquah.wednet.edu>, <hwg-techniques(at)mail.hwg.org> |
| In-Reply-To: |
yerpso |
| |
todo: View
Thread,
Original
|
|
Yes, I believe it's MS-only, but Mark is running IIS - I checked ;-)
> -----Original Message-----
> From: Hank Marquardt [mailto:hmarq(at)yerpso.net]
> Sent: Tuesday, July 23, 2002 6:35 PM
> To: Ann Ezzell
> Cc: 'Hank Marquardt'; 'Johnson, Mark'; hwg-techniques(at)mail.hwg.org
> Subject: Re: searching/indexing pdf files from a web page
>
>
> Good to know there is something out there ... but it seems to
> be MS only
> which limits it's utility some for me anyway. I guess you
> could always
> use a windows box to do the indexing and then use the index
> on whatever
> platform you need ...
>
> On Tue, Jul 23, 2002 at 06:15:38PM -0700, Ann Ezzell wrote:
> >
> > As I pointed out to Mark off-list, there's an iFilter for
> PDFs that you
> > can install for Index Server / Indexing Services. Works
> like a charm.
> >
> > To see an example of this in action, go here:
> >
> > http://www.gpworldwide.com/_sitesearch/default.asp
> >
> > Search for biosparging.
> >
> > You should get 3 results, two of which are PDFs.
> >
> >
> > > -----Original Message-----
> > > From: owner-hwg-techniques(at)hwg.org
> > > [mailto:owner-hwg-techniques(at)hwg.org] On Behalf Of Hank Marquardt
> > > Sent: Tuesday, July 23, 2002 6:00 PM
> > > To: Johnson, Mark
> > > Cc: hwg-techniques(at)mail.hwg.org
> > > Subject: Re: searching/indexing pdf files from a web page
> > >
> > >
> > > I guess this will depend on your definition of 'search' ...
> > > so you just
> > > mean the title(filename)?, perhaps a database of info
> > > associated with the file, -or- do you mean search the
> > > contents of the pdf file itself?
> > >
> > > That last one doesn't seem plausible ... a quick google
> search didn't
> > > turn up anything useful, and running 'strings' on a couple
> > > pdf files on
> > > my machine yeilded nothing useful ...
> > >
> > > If it's just the filename or you have something in text form
> > > that can be
> > > searched, you can do this with any server side language
> you choose.
> > >
> > > Need more detail of what there is to work with.
> > >
> > >
> > >
> > > On Wed, Jul 24, 2002 at 12:33:47AM +0100, Johnson, Mark wrote:
> > > > I have an archive of 500 PDF documents. I need to create a
> > > web page that
> > > > searches the documents and creates an index of the results.
> > > What's the best
> > > > way to do this?
> > > >
> > > > Thanks, Mark
>
HWG hwg-techniques mailing list archives,
maintained by Webmasters @ IWA
This page is part of a preserved archive of archives.hwg.org. The site is no longer active and its content is not maintained. For enquiries about this archive, write to archive(at)iwanet.org.