Publicação
Optimizing Memory Usage and Accesses on CUDA-Based Recurrent Pattern Matching Image Compression
| datacite.subject.fos | Ciências Naturais::Ciências da Computação e da Informação | |
| datacite.subject.fos | Engenharia e Tecnologia::Engenharia Eletrotécnica, Eletrónica e Informática | |
| datacite.subject.sdg | 09:Indústria, Inovação e Infraestruturas | |
| dc.contributor.author | Domingues, Patrício | |
| dc.contributor.author | Silva, João | |
| dc.contributor.author | Ribeiro, Tiago | |
| dc.contributor.author | Rodrigues, Nuno M. M. | |
| dc.contributor.author | Carvalho, Murilo B. de | |
| dc.contributor.author | Faria, Sérgio M. M. | |
| dc.date.accessioned | 2026-07-15T11:20:50Z | |
| dc.date.available | 2026-07-15T11:20:50Z | |
| dc.date.issued | 2014 | |
| dc.description | Computational Science and Its Applications - ICCSA 2014 14th International Conference, Guimarães, Portugal, June 30 - July 3, 204, Proceedings, Part IV. | |
| dc.description.abstract | This paper reports the adaptation of the Multidimensional Multiscale Parser (MMP) algorithm to CUDA. Specifically, we focus on memory optimization issues, such as the layout of data structures in memory, the type of GPU memory - shared, constant and global - and on achieving coalesced accesses. MMP is a demanding lossy compression algorithm for images. For example, MMP requires nearly 9000 seconds to encode the 512 x 512 Lenna image on a 2013's Intel Xeon. One of the main challenges to adapt MMP to manycore is related to the dependency over a pattern codebook which is built during the execution. This forces the input image to be processed sequentially. Nonetheless, CUDA-MMP achieves a 12x speedup over the sequential version when ran on an NVIDIA GTX 680. By further optimizing memory operations, the speedup is pushed to 17.1x. | eng |
| dc.description.sponsorship | Sponsors Associacao Portuguesa de Investigacao Operacional Kyushu Sangyo University (KSU) Monash University Universidade do Minho University of Basilicata University of Perugia | |
| dc.identifier.citation | Domingues, P., Silva, J., Ribeiro, T., Rodrigues, N. M. M., De Carvalho, M. B., & De Faria, S. M. M. (2014). Optimizing Memory Usage and Accesses on CUDA-Based Recurrent Pattern Matching Image Compression. In Computational Science and Its Applications – ICCSA 2014, LNCS vol. 8582, 560-575. Springer, Cham. https://doi.org/10.1007/978-3-319-09147-1_41 | |
| dc.identifier.doi | 10.1007/978-3-319-09147-1_41 | |
| dc.identifier.isbn | 9783319091464 | |
| dc.identifier.isbn | 9783319091471 | |
| dc.identifier.issn | 0302-9743 | |
| dc.identifier.issn | 1611-3349 | |
| dc.identifier.uri | http://hdl.handle.net/10400.8/16594 | |
| dc.language.iso | eng | |
| dc.peerreviewed | yes | |
| dc.publisher | Springer Nature | |
| dc.relation.hasversion | https://link.springer.com/chapter/10.1007/978-3-319-09147-1_41 | |
| dc.relation.ispartof | Lecture Notes in Computer Science | |
| dc.relation.ispartof | Computational Science and Its Applications – ICCSA 2014 | |
| dc.relation.ispartofseries | Lecture Notes in Computer Science (LNCS) | |
| dc.rights.uri | N/A | |
| dc.subject | CUDA | |
| dc.subject | image compression | |
| dc.subject | manycore computing | |
| dc.subject | memory optimization | |
| dc.title | Optimizing Memory Usage and Accesses on CUDA-Based Recurrent Pattern Matching Image Compression | eng |
| dc.type | book part | |
| dspace.entity.type | Publication | |
| oaire.citation.endPage | 575 | |
| oaire.citation.startPage | 560 | |
| oaire.citation.title | Computational Science and Its Applications - ICCSA 2014 | |
| oaire.citation.volume | 8582 | |
| oaire.version | http://purl.org/coar/version/c_970fb48d4fbd8a85 | |
| person.familyName | Domingues | |
| person.familyName | M. M. Rodrigues | |
| person.familyName | Faria | |
| person.givenName | Patrício | |
| person.givenName | Nuno | |
| person.givenName | Sergio | |
| person.identifier.ciencia-id | AA15-6185-C477 | |
| person.identifier.ciencia-id | 6917-B121-4E34 | |
| person.identifier.ciencia-id | 8815-4101-28DD | |
| person.identifier.orcid | 0000-0002-6207-6292 | |
| person.identifier.orcid | 0000-0001-9536-1017 | |
| person.identifier.orcid | 0000-0002-0993-9124 | |
| person.identifier.rid | C-5245-2011 | |
| person.identifier.scopus-author-id | 13411315400 | |
| person.identifier.scopus-author-id | 7006052345 | |
| person.identifier.scopus-author-id | 14027853900 | |
| relation.isAuthorOfPublication | b88ada5f-0d8b-4e55-ab0a-62aa82ea1388 | |
| relation.isAuthorOfPublication | b4ebe652-7f0e-4e67-adb0-d5ea29fc9e69 | |
| relation.isAuthorOfPublication | f69bd4d6-a6ef-4d20-8148-575478909661 | |
| relation.isAuthorOfPublication.latestForDiscovery | b88ada5f-0d8b-4e55-ab0a-62aa82ea1388 |
Ficheiros
Principais
1 - 1 de 1
A carregar...
- Nome:
- Optimizing Memory Usage and Accesses on CUDA-Based Recurrent Pattern Matching Image Compression.pdf
- Tamanho:
- 154.95 KB
- Formato:
- Adobe Portable Document Format
- Descrição:
- Preview-Abstrac
Licença
1 - 1 de 1
Miniatura indisponível
- Nome:
- license.txt
- Tamanho:
- 1.32 KB
- Formato:
- Item-specific license agreed upon to submission
- Descrição:
