Mostra el registre d'ítem simple

dc.contributor.authorBadia Sala, Rosa Maria
dc.contributor.authorHerrero Zaragoza, José Ramón
dc.contributor.authorLabarta Mancho, Jesús José
dc.contributor.authorPérez Cáncer, Josep Maria
dc.contributor.authorQuintana Ortí, Enrique Salvador
dc.contributor.authorQuintana Ortí, Gregorio
dc.contributor.otherUniversitat Politècnica de Catalunya. Departament d'Arquitectura de Computadors
dc.date.accessioned2010-03-08T12:11:15Z
dc.date.available2010-03-08T12:11:15Z
dc.date.created2009-12-25
dc.date.issued2009-12-25
dc.identifier.citationBadia, R. [et al.]. Parallelizing dense and banded linear algebra libraries using SMPSs. "Concurrency and computation: practice and experience", 25 Desembre 2009, vol. 21, núm. 18, p. 2438-2456.
dc.identifier.issn1532-0626
dc.identifier.urihttp://hdl.handle.net/2117/6565
dc.description.abstractThe promise of future many-core processors, with hundreds of threads running concurrently, has led the developers of linear algebra libraries to rethink their design in order to extract more parallelism, further exploit data locality, attain better load balance, and pay careful attention to the critical path of computation. In this paper we describe how existing serial libraries such as (C)LAPACK and FLAME can be easily parallelized using the SMPSs tools, consisting of a few OpenMP-like pragmas and a runtime system. In the LAPACK case, this usually requires the development of blocked algorithms for simple BLAS-level operations, which expose concurrency at a finer grain. For better performance, our experimental results indicate that column-major order, as employed by this library, needs to be abandoned in benefit of a block data layout. This will require a deeper rewrite of LAPACK or, alternatively, a dynamic conversion of the storage pattern at run-time. The parallelization of FLAME routines using SMPSs is simpler as this library includes blocked algorithms (or algorithms-by-blocks in the FLAME argot) for most operations and storage-by-blocks (or block data layout) is already in place.
dc.format.extent19 p.
dc.language.isoeng
dc.subjectÀrees temàtiques de la UPC::Informàtica::Arquitectura de computadors::Arquitectures paral·leles
dc.subject.lcshEmbedded computer systems
dc.subject.otherLinear algebra libraries
dc.subject.otherProgrammability
dc.subject.otherHigh performance
dc.subject.otherDynamic scheduling
dc.subject.otherMulti-core processors
dc.titleParallelizing dense and banded linear algebra libraries using SMPSs
dc.typeArticle
dc.subject.lemacOrdinadors immersos, Sistemes d'
dc.contributor.groupUniversitat Politècnica de Catalunya. CAP - Grup de Computació d'Altes Prestacions
dc.identifier.doi10.1002/cpe.1463
dc.description.peerreviewedPeer Reviewed
dc.relation.publisherversionhttps://onlinelibrary.wiley.com/doi/abs/10.1002/cpe.1463
dc.rights.accessRestricted access - publisher's policy
local.identifier.drac1613909
dc.description.versionPostprint (published version)
local.citation.authorBadia, R.; Herrero, J.; Labarta, J.; Pérez, J.; Quintana-Ortí, E.; Quintana-Ortí, G.
local.citation.publicationNameConcurrency and computation: practice and experience
local.citation.volume21
local.citation.number18
local.citation.startingPage2438
local.citation.endingPage2456


Fitxers d'aquest items

Imatge en miniatura

Aquest ítem apareix a les col·leccions següents

Mostra el registre d'ítem simple