Parallel text retrieval on PC clusters
buir.advisor | Aykanat, Cevdet | |
dc.contributor.author | Çatal, Aytül | |
dc.date.accessioned | 2016-07-01T10:58:57Z | |
dc.date.available | 2016-07-01T10:58:57Z | |
dc.date.issued | 2003 | |
dc.description | Cataloged from PDF version of article. | en_US |
dc.description.abstract | The inverted index partitioning problem is investigated for parallel text retrieval systems. The objective is to perform efficient query processing on an inverted index distributed across a PC cluster. Alternative strategies are considered and evaluated for inverted index partitioning, where index entries are distributed according to their document-ids or term-ids. The performance of both partitioning schemes depend on the total number of disk accesses and the total volume of communication in the system. In document-id partitioning, the total volume of communication is naturally minimum, whereas the total number of disk accesses may be larger compared to term-id partitioning. On the other hand, in term-id partitioning the total number of disk accesses is already equivalent to the lower bound achieved by the sequential algorithm, albeit the total communication volume may be quite large. The studies done so far perform these partitioning schemes in a round-robin fashion and compare the performance of them by simulation. In this work, a parallel text retrieval system is designed and implemented on a PC cluster. We adopted hypergraph-theoretical partitioning models and carried out performance comparison of round-robin and hypergraph-theoretical partitioning schemes on our parallel text retrieval system. We also designed and implemented a query interface and a user interface of our system. | en_US |
dc.description.provenance | Made available in DSpace on 2016-07-01T10:58:57Z (GMT). No. of bitstreams: 1 0002397.pdf: 559615 bytes, checksum: cf31075e54bdd2a82b4caaea36212692 (MD5) Previous issue date: 2003 | en |
dc.description.statementofresponsibility | Çatal, Aytül | en_US |
dc.format.extent | xi, 56 leaves, tables, graphics, 30 cm | en_US |
dc.identifier.itemid | BILKUTUPB072124 | |
dc.identifier.uri | http://hdl.handle.net/11693/29391 | |
dc.language.iso | English | en_US |
dc.rights | info:eu-repo/semantics/openAccess | en_US |
dc.subject | Parallel text retrieval | en_US |
dc.subject | system performance | en_US |
dc.subject | inverted index partitioning | en_US |
dc.subject | parallel query processing | en_US |
dc.subject | inverted index | en_US |
dc.subject.lcc | QA76.5 .C38 2003 | en_US |
dc.subject.lcsh | Parallel processing (Electronic computers). | en_US |
dc.title | Parallel text retrieval on PC clusters | en_US |
dc.type | Thesis | en_US |
thesis.degree.discipline | Computer Engineering | |
thesis.degree.grantor | Bilkent University | |
thesis.degree.level | Master's | |
thesis.degree.name | MS (Master of Science) |
Files
Original bundle
1 - 1 of 1
Loading...
- Name:
- 0002397.pdf
- Size:
- 546.5 KB
- Format:
- Adobe Portable Document Format
- Description:
- Full printable version