Cultural advice

The Australian National University acknowledges, celebrates and pays our respects to the Ngunnawal and Ngambri people of the Canberra region and to all First Nations Australians on whose traditional lands we meet and work, and whose cultures are among the oldest continuing cultures in human history.

Aboriginal and Torres Strait Islander peoples are advised that ANU Library collections may include images, names, voices, and other representations of deceased persons.

Material in the collection may contain terms, language or views that reflect the period in which the item was created and may be considered inappropriate today.

How fast is -fast? Performance analysis of KDD applications using hardware performance counters on UltraSPARC-III

dc.contributor.authorCzezowski, Adamen_US
dc.contributor.authorChristen, Peteren_US
dc.date.accessioned2003-06-27en_US
dc.date.accessioned2004-05-19T12:16:49Zen_US
dc.date.accessioned2011-01-05T08:29:33Z
dc.date.available2004-05-19T12:16:49Zen_US
dc.date.available2011-01-05T08:29:33Z
dc.date.created2002en_US
dc.date.issued2002en_US
dc.description.abstractModern processors and computer systems are designed to be eÆcient and achieve high performance with applications that have regular memory access patterns. For example, dense linear algebra routines can be implemented to achieve near peak performance. While such routines have traditionally formed the core of many scientific and engineering applications, commercial workloads like database and web servers, or decision support systems (data warehouses and data mining) are one of the fastest growing market segments on high-performance computing platforms. Many of these commercial applications are characterised by more complex codes and irregular memory access patterns, which often result in a decrease of performance that is achieved. Due to their complexity and the lack of source code, performance analysis of commercial applications is not an easy task. Hardware performance counters allow detailed analysis of program behaviour, like number of instructions of various types, memory and cache access, hit and miss rates, or branch mispredictions. In this paper we describe experiments and present results conducted with various KDD applications on an UltraSPARC-III platform, and we compare these applications with an optimized dense matrix-matrix multiplication. We focus on compiler optimisations using the -fast ag and discuss di_erences in un-optimised and optimised codes.en_US
dc.format.extent302255 bytesen_US
dc.format.extent356 bytesen_US
dc.format.mimetypeapplication/pdfen_US
dc.format.mimetypeapplication/octet-streamen_US
dc.identifier.urihttp://hdl.handle.net/1885/40725en_US
dc.identifier.urihttp://digitalcollections.anu.edu.au/handle/1885/40725
dc.language.isoen_AUen_US
dc.subjectdata miningen_US
dc.subjectperformance analysisen_US
dc.subjectcompiler optimisationen_US
dc.subjectTR-CSen_US
dc.titleHow fast is -fast? Performance analysis of KDD applications using hardware performance counters on UltraSPARC-IIIen_US
dc.typeWorking/Technical Paperen_US
local.citationTR-CS-02-03en_US
local.contributor.affiliationDepartment of Computer Science, FEITen_US
local.contributor.affiliationANUen_US
local.description.refereednoen_US
local.identifier.citationmonthsepen_US
local.identifier.citationyear2002en_US
local.identifier.eprintid1527en_US
local.rights.ispublishedyesen_US

Downloads

Original bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
TR-CS-02-03.pdf
Size:
295.17 KB
Format:
Adobe Portable Document Format