Javascript must be enabled to continue!
Architectural support for high-performing hardware transactional memory systems
View through CrossRef
Parallel programming presents an efficient solution to exploit future multicore processors.
Unfortunately, traditional programming models depend on programmer’s skills for synchronizing
concurrent threads, which makes the development of parallel software a hard and errorprone
task. In addition to this, current synchronization techniques serialize the execution of
those critical sections that conflict in shared memory and thus limit the scalability of multithreaded
applications.
Transactional Memory (TM) has emerged as a promising programming model that solves
the trade-off between high performance and ease of use. In TM, the system is in charge of
scheduling transactions (atomic blocks of instructions) and guaranteeing that they are executed
in isolation, which simplifies writing parallel code and, at the same time, enables high concurrency
when atomic regions access different data. Among all forms of TM environments,
Hardware TM (HTM) systems is the only one that offers fast execution at the cost of adding
dedicated logic in the processor.
Existing HTMsystems suffer considerable delays when they execute complex transactional
workloads, especially when they deal with large and contending transactions because they lack
adaptability. Furthermore, most HTM implementations are ad hoc and require cumbersome
hardware structures to be effective, which complicates the feasibility of the design. This thesis
makes several contributions in the design and analysis of low-cost HTMsystems that yield good
performance for any kind of TM program.
Our first contribution, FASTM, introduces a novel mechanism to elegantly manage speculative
(and already validated) versions of transactional data by slightly modifying on-chip memory
engine. This approach permits fast recovery when a transaction that fits in private caches is discarded.
At the same time, it keeps non-speculative values in software, which allows in-place
x
memory updates. Thus, FASTM is not hurt from capacity issues nor slows down when it has to
undo transactional modifications.
Our second contribution includes two different HTM systems that integrate deferred resolution
of conflicts in a conventional multicore processor, which reduces the complexity of the
system with respect to previous proposals. The first one, FUSETM, combines different-mode
transactions under a unified infrastructure to gracefully handle resource overflow. As a result,
FUSETM brings fast transactional computation without requiring additional hardware nor extra
communication at the end of speculative execution. The second one, SPECTM, introduces a
two-level data versioning mechanism to resolve conflicts in a speculative fashion even in the
case of overflow.
Our third and last contribution presents a couple of truly flexible HTM systems that can
dynamically adapt their underlying mechanisms according to the characteristics of the program.
DYNTM records statistics of previously executed transactions to select the best-suited strategy
each time a new instance of a transaction starts. SWAPTM takes a different approach: it tracks
information of the current transactional instance to change its priority level at runtime. Both
alternatives obtain great performance over existing proposals that employ fixed transactional
policies, especially in applications with phase changes.
Title: Architectural support for high-performing hardware transactional memory systems
Description:
Parallel programming presents an efficient solution to exploit future multicore processors.
Unfortunately, traditional programming models depend on programmer’s skills for synchronizing
concurrent threads, which makes the development of parallel software a hard and errorprone
task.
In addition to this, current synchronization techniques serialize the execution of
those critical sections that conflict in shared memory and thus limit the scalability of multithreaded
applications.
Transactional Memory (TM) has emerged as a promising programming model that solves
the trade-off between high performance and ease of use.
In TM, the system is in charge of
scheduling transactions (atomic blocks of instructions) and guaranteeing that they are executed
in isolation, which simplifies writing parallel code and, at the same time, enables high concurrency
when atomic regions access different data.
Among all forms of TM environments,
Hardware TM (HTM) systems is the only one that offers fast execution at the cost of adding
dedicated logic in the processor.
Existing HTMsystems suffer considerable delays when they execute complex transactional
workloads, especially when they deal with large and contending transactions because they lack
adaptability.
Furthermore, most HTM implementations are ad hoc and require cumbersome
hardware structures to be effective, which complicates the feasibility of the design.
This thesis
makes several contributions in the design and analysis of low-cost HTMsystems that yield good
performance for any kind of TM program.
Our first contribution, FASTM, introduces a novel mechanism to elegantly manage speculative
(and already validated) versions of transactional data by slightly modifying on-chip memory
engine.
This approach permits fast recovery when a transaction that fits in private caches is discarded.
At the same time, it keeps non-speculative values in software, which allows in-place
x
memory updates.
Thus, FASTM is not hurt from capacity issues nor slows down when it has to
undo transactional modifications.
Our second contribution includes two different HTM systems that integrate deferred resolution
of conflicts in a conventional multicore processor, which reduces the complexity of the
system with respect to previous proposals.
The first one, FUSETM, combines different-mode
transactions under a unified infrastructure to gracefully handle resource overflow.
As a result,
FUSETM brings fast transactional computation without requiring additional hardware nor extra
communication at the end of speculative execution.
The second one, SPECTM, introduces a
two-level data versioning mechanism to resolve conflicts in a speculative fashion even in the
case of overflow.
Our third and last contribution presents a couple of truly flexible HTM systems that can
dynamically adapt their underlying mechanisms according to the characteristics of the program.
DYNTM records statistics of previously executed transactions to select the best-suited strategy
each time a new instance of a transaction starts.
SWAPTM takes a different approach: it tracks
information of the current transactional instance to change its priority level at runtime.
Both
alternatives obtain great performance over existing proposals that employ fixed transactional
policies, especially in applications with phase changes.
Related Results
619. Pharmacokinetic-Pharmacodynamic (PK-PD) Target Attainment Analyses to Support Epetraborole Dose Selection for the Treatment of Patients with Mycobacterium avium Complex (MAC) Lung Disease
619. Pharmacokinetic-Pharmacodynamic (PK-PD) Target Attainment Analyses to Support Epetraborole Dose Selection for the Treatment of Patients with Mycobacterium avium Complex (MAC) Lung Disease
Abstract
Background
Epetraborole (EBO) is an orally available, bacterial leucyl transfer RNA synthetase inhibitor that concentra...
LB2306. Population Pharmacokinetic (PPK), Pharmacokinetic/Pharmacodynamic attainment (PTA), and Clinical Pharmacokinetic/Pharmacodynamic (PK/PD) Analyses for Sulbactam-Durlobactam (SUL-DUR) to Support Dose Selection for the Treatment of Acinetobacter baum
LB2306. Population Pharmacokinetic (PPK), Pharmacokinetic/Pharmacodynamic attainment (PTA), and Clinical Pharmacokinetic/Pharmacodynamic (PK/PD) Analyses for Sulbactam-Durlobactam (SUL-DUR) to Support Dose Selection for the Treatment of Acinetobacter baum
Abstract
Background
SUL-DUR is a β-lactam/β-lactamase inhibitor combination in development for the treatment of ABC infections, ...
593. Population Pharmacokinetic Model Development for Epetraborole and Mycobacterium avium Complex (MAC) Lung Disease Patients Using Data from Phase 1 and 2 Studies
593. Population Pharmacokinetic Model Development for Epetraborole and Mycobacterium avium Complex (MAC) Lung Disease Patients Using Data from Phase 1 and 2 Studies
Abstract
Background
Epetraborole (EBO), an orally available bacterial leucyl transfer RNA synthetase inhibitor with potent activ...
592. Impact of Elevated MIC Values on Echinocandin Pharmacokinetic-Pharmacodynamic (PK-PD) Candida glabrata Target Attainment (TA)
592. Impact of Elevated MIC Values on Echinocandin Pharmacokinetic-Pharmacodynamic (PK-PD) Candida glabrata Target Attainment (TA)
Abstract
Background
Given the increasing prevalence of non-albicans Candida species, including C. glabrata and C. auris, which h...
Effects of Contextual Cues on False Memory: A Comparative Experimental Approach
Effects of Contextual Cues on False Memory: A Comparative Experimental Approach
Research on false memory formation using the Deese-Roediger-McDermott (DRM) paradigm has been extensively conducted in Western contexts. Yet, a significant gap remains in experimen...
Atomic dataflow model
Atomic dataflow model
With the recent switch in the design of general purpose processors from frequency scaling of a single processor core towards increasing the number of processor cores, parallel prog...
An HIV prevention intervention helps immigrants open up about transactional sex. The Makasi study
An HIV prevention intervention helps immigrants open up about transactional sex. The Makasi study
Abstract
Background
Transactional sex is known to be an exposure factor for HIV acquisition among immigrants in France. We analy...
Investigating sire fertility : the relationship between spermatozoa characteristics and early embryonic development
Investigating sire fertility : the relationship between spermatozoa characteristics and early embryonic development
[EMBARGOED UNTIL 12/1/2023] Early embryonic mortality in cattle occurs between days 1-27 of gestation and is one of the primary contributors to economic failure in the dairy indust...

