This is an archive of the discontinued LLVM Phabricator instance.

Paths

Table of Contentst

-
llvm/trunk/
-
trunk/
-
include/llvm/Target/
-
llvm/
-
Target/
-
TargetLowering.h
-
lib/
-
CodeGen/SelectionDAG/
-
SelectionDAG/
-
SelectionDAGBuilder.cpp
-
Target/
-
Mips/
-
MipsTargetMachine.cpp
-
Sparc/
-
SparcTargetMachine.cpp
-
XCore/
-
XCoreTargetMachine.cpp
-
test/CodeGen/XCore/
-
CodeGen/
-
XCore/
-
atomic.ll

Differential D5474

Erase fence insertion from SelectionDAGBuilder.cpp (NFC)
ClosedPublic

Authored by morisset on Sep 23 2014, 3:59 PM.

Download Raw Diff

Details

Reviewers

rkotler
t.p.northover
jfb

Commits

rGe2de06bef611: Erase fence insertion from SelectionDAGBuilder.cpp (NFC)
rL219957: Erase fence insertion from SelectionDAGBuilder.cpp (NFC)

Summary

Backends can use setInsertFencesForAtomic to signal to the middle-end that
montonic is the only memory ordering they can accept for
stores/loads/rmws/cmpxchg. The code lowering those accesses with a stronger
ordering to fences + monotonic accesses is currently living in
SelectionDAGBuilder.cpp. In this patch I propose moving this logic out of it
for several reasons:

There is lots of redundancy to avoid: extremely similar logic already exists in AtomicExpand.
The current code in SelectionDAGBuilder does not use any target-hooks, it does the same transformation for every backend that requires it
As a result it is plain *unsound*, as it was apparently designed for ARM. It happens to mostly work for the other targets because they are extremely conservative, but Power for example had to switch to AtomicExpand to be able to use lwsync safely (see r218331).
Because it produces IR-level fences, it cannot be made sound ! This is noted in the C++11 standard (section 29.3, page 1140):

Fences cannot, in general, be used to restore sequential consistency for atomic
operations with weaker ordering semantics.

It can also be seen by the following example (called IRIW in the litterature):

atomic<int> x = y = 0;
int r1, r2, r3, r4;
Thread 0:
  x.store(1);
Thread 1:
  y.store(1);
Thread 2:
  r1 = x.load();
  r2 = y.load();
Thread 3:
  r3 = y.load();
  r4 = x.load();

r1 = r3 = 1 and r2 = r4 = 0 is impossible as long as the accesses are all seq_cst.
But if they are lowered to monotonic accesses, no amount of fences can prevent it..

This patch does three things (I could cut it into parts, but then some of them
would not be tested/testable, please tell me if you would prefer that):

it provides a default implementation for emitLeadingFence/emitTrailingFence in

terms of IR-level fences, that mimic the original logic of SelectionDAGBuilder.
As we saw above, this is unsound, but the best that can be done without knowing
the targets well (and there is a comment warning about this risk).

it then switches Mips/Sparc/XCore to use AtomicExpand, relying on this default

implementation (that exactly replicates the logic of SelectionDAGBuilder, so no
functional change)

it finally erase this logic from SelectionDAGBuilder as it is dead-code.

Ideally, each target would define its own override for emitLeading/TrailingFence
using target-specific fences, but I do not know the Sparc/Mips/XCore memory model
well enough to do this, and they appear to be dealing fine with the ARM-inspired
default expansion for now (probably because they are overly conservative, as
Power was). If anyone wants to compile fences more agressively on these
platforms, the long comment should make it clear why he should first override
emitLeading/TrailingFence.

Diff Detail

Repository: rL LLVM

Event Timeline

morisset updated this revision to Diff 14022.Sep 23 2014, 3:59 PM

morisset retitled this revision from to Erase fence insertion from SelectionDAGBuilder.cpp (NFC).

morisset updated this object.

morisset edited the test plan for this revision. (Show Details)

morisset added reviewers: jfb, t.p.northover.

morisset added a subscriber: Unknown Object (MLST).

Herald added a subscriber: aemerson. · View Herald TranscriptSep 23 2014, 3:59 PM

jfb added a reviewer: rkotler.Oct 3 2014, 2:49 PM

I added @rkotler to the review for MIPS. I'm not sure if anyone on Sparc/XCore can comment. Overall this is NFC, so no big issue.

The change looks good overall.

include/llvm/Target/TargetLowering.h
979 ↗	(On Diff #14022)	The FIXME should probably say how to fix this (or maybe point it out in the code below), and then suggest that the asserts be added back.
987 ↗	(On Diff #14022)	You can create a doxygen group for these functions, start with `@{` in the comment for emitLeading, and after emitTrailing close with `}@`.
test/CodeGen/XCore/atomic.ll
26 ↗	(On Diff #14022)	Why does the barrier move?

morisset added inline comments.Oct 3 2014, 3:56 PM

include/llvm/Target/TargetLowering.h
979 ↗	(On Diff #14022)	I tried to give the solution just above "Backends should override this method to produce target-specific intrinsic for their fence". I do not see what more I can say, apart from maybe "look at what the ARM backend does".
987 ↗	(On Diff #14022)	Ok, thanks for the trick, I will do that.
test/CodeGen/XCore/atomic.ll
26 ↗	(On Diff #14022)	I am not entirely sure, but both placements appear correct. So while I agree it is surprising and slightly troubling, I decided to go ahead as I could not find any other case when it happens. I suspect it is just some the instruction scheduler that must pick differently between two valid choices, for some reason (The first lowering happens a bit earlier in the pipeline so maybe this cause some other pass in-between to change subtly its behaviour ?)

Rebase on trunk + use doxygen trick suggested by jfb

lgtm, though I'd like to make sure @t.p.northover agrees since he was involved in the discussion that led to this patch.

This revision is now accepted and ready to land.Oct 8 2014, 12:03 PM

Yep, this looks sensible to me too. Sorry I didn't get around to looking sooner.

Tim.

Closed by commit rL219957 (authored by @morisset).

Revision Contents

Path

Size

llvm/

trunk/

include/

llvm/

Target/

TargetLowering.h

51 lines

lib/

CodeGen/

SelectionDAG/

SelectionDAGBuilder.cpp

87 lines

Target/

Mips/

MipsTargetMachine.cpp

1 line

Sparc/

SparcTargetMachine.cpp

7 lines

XCore/

XCoreTargetMachine.cpp

7 lines

test/

CodeGen/

XCore/

atomic.ll

5 lines

Diff 15046

llvm/trunk/include/llvm/Target/TargetLowering.h

Show First 20 Lines • Show All 958 Lines • ▼ Show 20 Lines	virtual Value emitStoreConditional(IRBuilder<> &Builder, Value Val,
Value *Addr, AtomicOrdering Ord) const {		Value *Addr, AtomicOrdering Ord) const {
llvm_unreachable("Store conditional unimplemented on this target");		llvm_unreachable("Store conditional unimplemented on this target");
}		}

/// Inserts in the IR a target-specific intrinsic specifying a fence.		/// Inserts in the IR a target-specific intrinsic specifying a fence.
/// It is called by AtomicExpandPass before expanding an		/// It is called by AtomicExpandPass before expanding an
/// AtomicRMW/AtomicCmpXchg/AtomicStore/AtomicLoad.		/// AtomicRMW/AtomicCmpXchg/AtomicStore/AtomicLoad.
/// RMW and CmpXchg set both IsStore and IsLoad to true.		/// RMW and CmpXchg set both IsStore and IsLoad to true.
/// Backends with !getInsertFencesForAtomic() should keep a no-op here.
/// This function should either return a nullptr, or a pointer to an IR-level		/// This function should either return a nullptr, or a pointer to an IR-level
/// Instruction*. Even complex fence sequences can be represented by a		/// Instruction*. Even complex fence sequences can be represented by a
/// single Instruction* through an intrinsic to be lowered later.		/// single Instruction* through an intrinsic to be lowered later.
		/// Backends with !getInsertFencesForAtomic() should keep a no-op here.
		/// Backends should override this method to produce target-specific intrinsic
		/// for their fences.
		/// FIXME: Please note that the default implementation here in terms of
		/// IR-level fences exists for historical/compatibility reasons and is
		/// unsound ! Fences cannot, in general, be used to restore sequential
		/// consistency. For example, consider the following example:
		/// atomic<int> x = y = 0;
		/// int r1, r2, r3, r4;
		/// Thread 0:
		/// x.store(1);
		/// Thread 1:
		/// y.store(1);
		/// Thread 2:
		/// r1 = x.load();
		/// r2 = y.load();
		/// Thread 3:
		/// r3 = y.load();
		/// r4 = x.load();
		/// r1 = r3 = 1 and r2 = r4 = 0 is impossible as long as the accesses are all
		/// seq_cst. But if they are lowered to monotonic accesses, no amount of
		/// IR-level fences can prevent it.
		/// @{
virtual Instruction* emitLeadingFence(IRBuilder<> &Builder, AtomicOrdering Ord,		virtual Instruction* emitLeadingFence(IRBuilder<> &Builder, AtomicOrdering Ord,
bool IsStore, bool IsLoad) const {		bool IsStore, bool IsLoad) const {
assert(!getInsertFencesForAtomic());		if (!getInsertFencesForAtomic())
		return nullptr;

		if (isAtLeastRelease(Ord) && IsStore)
		return Builder.CreateFence(Ord);
		else
return nullptr;		return nullptr;
}		}

/// Inserts in the IR a target-specific intrinsic specifying a fence.
/// It is called by AtomicExpandPass after expanding an
/// AtomicRMW/AtomicCmpXchg/AtomicStore/AtomicLoad.
/// RMW and CmpXchg set both IsStore and IsLoad to true.
/// Backends with !getInsertFencesForAtomic() should keep a no-op here.
/// This function should either return a nullptr, or a pointer to an IR-level
/// Instruction*. Even complex fence sequences can be represented by a
/// single Instruction* through an intrinsic to be lowered later.
virtual Instruction* emitTrailingFence(IRBuilder<> &Builder, AtomicOrdering Ord,		virtual Instruction* emitTrailingFence(IRBuilder<> &Builder, AtomicOrdering Ord,
bool IsStore, bool IsLoad) const {		bool IsStore, bool IsLoad) const {
assert(!getInsertFencesForAtomic());		if (!getInsertFencesForAtomic())
		return nullptr;

		if (isAtLeastAcquire(Ord))
		return Builder.CreateFence(Ord);
		else
return nullptr;		return nullptr;
}		}
		/// @}

/// Returns true if the given (atomic) store should be expanded by the		/// Returns true if the given (atomic) store should be expanded by the
/// IR-level AtomicExpand pass into an "atomic xchg" which ignores its input.		/// IR-level AtomicExpand pass into an "atomic xchg" which ignores its input.
virtual bool shouldExpandAtomicStoreInIR(StoreInst *SI) const {		virtual bool shouldExpandAtomicStoreInIR(StoreInst *SI) const {
return false;		return false;
}		}

/// Returns true if the given (atomic) load should be expanded by the		/// Returns true if the given (atomic) load should be expanded by the
▲ Show 20 Lines • Show All 1,717 Lines • Show Last 20 Lines

llvm/trunk/lib/CodeGen/SelectionDAG/SelectionDAGBuilder.cpp

This file is larger than 256 KB, so syntax highlighting is disabled by default.

Show First 20 Lines • Show All 3,598 Lines • ▼ Show 20 Lines	for (unsigned i = 0; i != NumValues; ++i, ++ChainI) {
Chains[ChainI] = St;		Chains[ChainI] = St;
}		}

SDValue StoreNode = DAG.getNode(ISD::TokenFactor, getCurSDLoc(), MVT::Other,		SDValue StoreNode = DAG.getNode(ISD::TokenFactor, getCurSDLoc(), MVT::Other,
makeArrayRef(Chains.data(), ChainI));		makeArrayRef(Chains.data(), ChainI));
DAG.setRoot(StoreNode);		DAG.setRoot(StoreNode);
}		}

static SDValue InsertFenceForAtomic(SDValue Chain, AtomicOrdering Order,
SynchronizationScope Scope,
bool Before, SDLoc dl,
SelectionDAG &DAG,
const TargetLowering &TLI) {
// Fence, if necessary
if (Before) {
if (Order == AcquireRelease \|\| Order == SequentiallyConsistent)
Order = Release;
else if (Order == Acquire \|\| Order == Monotonic \|\| Order == Unordered)
return Chain;
} else {
if (Order == AcquireRelease)
Order = Acquire;
else if (Order == Release \|\| Order == Monotonic \|\| Order == Unordered)
return Chain;
}
SDValue Ops[3];
Ops[0] = Chain;
Ops[1] = DAG.getConstant(Order, TLI.getPointerTy());
Ops[2] = DAG.getConstant(Scope, TLI.getPointerTy());
return DAG.getNode(ISD::ATOMIC_FENCE, dl, MVT::Other, Ops);
}

void SelectionDAGBuilder::visitAtomicCmpXchg(const AtomicCmpXchgInst &I) {		void SelectionDAGBuilder::visitAtomicCmpXchg(const AtomicCmpXchgInst &I) {
SDLoc dl = getCurSDLoc();		SDLoc dl = getCurSDLoc();
AtomicOrdering SuccessOrder = I.getSuccessOrdering();		AtomicOrdering SuccessOrder = I.getSuccessOrdering();
AtomicOrdering FailureOrder = I.getFailureOrdering();		AtomicOrdering FailureOrder = I.getFailureOrdering();
SynchronizationScope Scope = I.getSynchScope();		SynchronizationScope Scope = I.getSynchScope();

SDValue InChain = getRoot();		SDValue InChain = getRoot();

const TargetLowering &TLI = DAG.getTargetLoweringInfo();
if (TLI.getInsertFencesForAtomic())
InChain =
InsertFenceForAtomic(InChain, SuccessOrder, Scope, true, dl, DAG, TLI);

MVT MemVT = getValue(I.getCompareOperand()).getSimpleValueType();		MVT MemVT = getValue(I.getCompareOperand()).getSimpleValueType();
SDVTList VTs = DAG.getVTList(MemVT, MVT::i1, MVT::Other);		SDVTList VTs = DAG.getVTList(MemVT, MVT::i1, MVT::Other);
SDValue L = DAG.getAtomicCmpSwap(		SDValue L = DAG.getAtomicCmpSwap(
ISD::ATOMIC_CMP_SWAP_WITH_SUCCESS, dl, MemVT, VTs, InChain,		ISD::ATOMIC_CMP_SWAP_WITH_SUCCESS, dl, MemVT, VTs, InChain,
getValue(I.getPointerOperand()), getValue(I.getCompareOperand()),		getValue(I.getPointerOperand()), getValue(I.getCompareOperand()),
getValue(I.getNewValOperand()), MachinePointerInfo(I.getPointerOperand()),		getValue(I.getNewValOperand()), MachinePointerInfo(I.getPointerOperand()),
0 /* Alignment */,		/Alignment=/ 0, SuccessOrder, FailureOrder, Scope);
TLI.getInsertFencesForAtomic() ? Monotonic : SuccessOrder,
TLI.getInsertFencesForAtomic() ? Monotonic : FailureOrder, Scope);

SDValue OutChain = L.getValue(2);		SDValue OutChain = L.getValue(2);

if (TLI.getInsertFencesForAtomic())
OutChain = InsertFenceForAtomic(OutChain, SuccessOrder, Scope, false, dl,
DAG, TLI);

setValue(&I, L);		setValue(&I, L);
DAG.setRoot(OutChain);		DAG.setRoot(OutChain);
}		}

void SelectionDAGBuilder::visitAtomicRMW(const AtomicRMWInst &I) {		void SelectionDAGBuilder::visitAtomicRMW(const AtomicRMWInst &I) {
SDLoc dl = getCurSDLoc();		SDLoc dl = getCurSDLoc();
ISD::NodeType NT;		ISD::NodeType NT;
switch (I.getOperation()) {		switch (I.getOperation()) {
Show All 10 Lines	void SelectionDAGBuilder::visitAtomicRMW(const AtomicRMWInst &I) {
case AtomicRMWInst::UMax: NT = ISD::ATOMIC_LOAD_UMAX; break;		case AtomicRMWInst::UMax: NT = ISD::ATOMIC_LOAD_UMAX; break;
case AtomicRMWInst::UMin: NT = ISD::ATOMIC_LOAD_UMIN; break;		case AtomicRMWInst::UMin: NT = ISD::ATOMIC_LOAD_UMIN; break;
}		}
AtomicOrdering Order = I.getOrdering();		AtomicOrdering Order = I.getOrdering();
SynchronizationScope Scope = I.getSynchScope();		SynchronizationScope Scope = I.getSynchScope();

SDValue InChain = getRoot();		SDValue InChain = getRoot();

const TargetLowering &TLI = DAG.getTargetLoweringInfo();		SDValue L =
if (TLI.getInsertFencesForAtomic())		DAG.getAtomic(NT, dl,
InChain = InsertFenceForAtomic(InChain, Order, Scope, true, dl, DAG, TLI);		getValue(I.getValOperand()).getSimpleValueType(),
		InChain,
SDValue L = DAG.getAtomic(		getValue(I.getPointerOperand()),
NT, dl, getValue(I.getValOperand()).getSimpleValueType(), InChain,		getValue(I.getValOperand()),
getValue(I.getPointerOperand()), getValue(I.getValOperand()),		I.getPointerOperand(),
I.getPointerOperand(), 0 /* Alignment */,		/* Alignment=*/ 0, Order, Scope);
TLI.getInsertFencesForAtomic() ? Monotonic : Order, Scope);

SDValue OutChain = L.getValue(1);		SDValue OutChain = L.getValue(1);

if (TLI.getInsertFencesForAtomic())
OutChain =
InsertFenceForAtomic(OutChain, Order, Scope, false, dl, DAG, TLI);

setValue(&I, L);		setValue(&I, L);
DAG.setRoot(OutChain);		DAG.setRoot(OutChain);
}		}

void SelectionDAGBuilder::visitFence(const FenceInst &I) {		void SelectionDAGBuilder::visitFence(const FenceInst &I) {
SDLoc dl = getCurSDLoc();		SDLoc dl = getCurSDLoc();
const TargetLowering &TLI = DAG.getTargetLoweringInfo();		const TargetLowering &TLI = DAG.getTargetLoweringInfo();
SDValue Ops[3];		SDValue Ops[3];
Show All 21 Lines	MachineMemOperand *MMO =
getMachineMemOperand(MachinePointerInfo(I.getPointerOperand()),		getMachineMemOperand(MachinePointerInfo(I.getPointerOperand()),
MachineMemOperand::MOVolatile \|		MachineMemOperand::MOVolatile \|
MachineMemOperand::MOLoad,		MachineMemOperand::MOLoad,
VT.getStoreSize(),		VT.getStoreSize(),
I.getAlignment() ? I.getAlignment() :		I.getAlignment() ? I.getAlignment() :
DAG.getEVTAlignment(VT));		DAG.getEVTAlignment(VT));

InChain = TLI.prepareVolatileOrAtomicLoad(InChain, dl, DAG);		InChain = TLI.prepareVolatileOrAtomicLoad(InChain, dl, DAG);
SDValue L = DAG.getAtomic(		SDValue L =
ISD::ATOMIC_LOAD, dl, VT, VT, InChain, getValue(I.getPointerOperand()),		DAG.getAtomic(ISD::ATOMIC_LOAD, dl, VT, VT, InChain,
MMO, TLI.getInsertFencesForAtomic() ? Monotonic : Order, Scope);		getValue(I.getPointerOperand()), MMO,
		Order, Scope);

SDValue OutChain = L.getValue(1);		SDValue OutChain = L.getValue(1);

if (TLI.getInsertFencesForAtomic())
OutChain = InsertFenceForAtomic(OutChain, Order, Scope, false, dl,
DAG, TLI);

setValue(&I, L);		setValue(&I, L);
DAG.setRoot(OutChain);		DAG.setRoot(OutChain);
}		}

void SelectionDAGBuilder::visitAtomicStore(const StoreInst &I) {		void SelectionDAGBuilder::visitAtomicStore(const StoreInst &I) {
SDLoc dl = getCurSDLoc();		SDLoc dl = getCurSDLoc();

AtomicOrdering Order = I.getOrdering();		AtomicOrdering Order = I.getOrdering();
SynchronizationScope Scope = I.getSynchScope();		SynchronizationScope Scope = I.getSynchScope();

SDValue InChain = getRoot();		SDValue InChain = getRoot();

const TargetLowering &TLI = DAG.getTargetLoweringInfo();		const TargetLowering &TLI = DAG.getTargetLoweringInfo();
EVT VT = TLI.getValueType(I.getValueOperand()->getType());		EVT VT = TLI.getValueType(I.getValueOperand()->getType());

if (I.getAlignment() < VT.getSizeInBits() / 8)		if (I.getAlignment() < VT.getSizeInBits() / 8)
report_fatal_error("Cannot generate unaligned atomic store");		report_fatal_error("Cannot generate unaligned atomic store");

if (TLI.getInsertFencesForAtomic())		SDValue OutChain =
InChain = InsertFenceForAtomic(InChain, Order, Scope, true, dl, DAG, TLI);		DAG.getAtomic(ISD::ATOMIC_STORE, dl, VT,
		InChain,
SDValue OutChain = DAG.getAtomic(		getValue(I.getPointerOperand()),
ISD::ATOMIC_STORE, dl, VT, InChain, getValue(I.getPointerOperand()),		getValue(I.getValueOperand()),
getValue(I.getValueOperand()), I.getPointerOperand(), I.getAlignment(),		I.getPointerOperand(), I.getAlignment(),
TLI.getInsertFencesForAtomic() ? Monotonic : Order, Scope);		Order, Scope);

if (TLI.getInsertFencesForAtomic())
OutChain =
InsertFenceForAtomic(OutChain, Order, Scope, false, dl, DAG, TLI);

DAG.setRoot(OutChain);		DAG.setRoot(OutChain);
}		}

/// visitTargetIntrinsic - Lower a call of a target intrinsic to an INTRINSIC		/// visitTargetIntrinsic - Lower a call of a target intrinsic to an INTRINSIC
/// node.		/// node.
void SelectionDAGBuilder::visitTargetIntrinsic(const CallInst &I,		void SelectionDAGBuilder::visitTargetIntrinsic(const CallInst &I,
unsigned Intrinsic) {		unsigned Intrinsic) {
▲ Show 20 Lines • Show All 3,968 Lines • Show Last 20 Lines

llvm/trunk/lib/Target/Mips/MipsTargetMachine.cpp

	Show First 20 Lines • Show All 172 Lines • ▼ Show 20 Lines
	} // namespace			} // namespace

	TargetPassConfig *MipsTargetMachine::createPassConfig(PassManagerBase &PM) {			TargetPassConfig *MipsTargetMachine::createPassConfig(PassManagerBase &PM) {
	return new MipsPassConfig(this, PM);			return new MipsPassConfig(this, PM);
	}			}

	void MipsPassConfig::addIRPasses() {			void MipsPassConfig::addIRPasses() {
	TargetPassConfig::addIRPasses();			TargetPassConfig::addIRPasses();
				addPass(createAtomicExpandPass(&getMipsTargetMachine()));
	if (getMipsSubtarget().os16())			if (getMipsSubtarget().os16())
	addPass(createMipsOs16(getMipsTargetMachine()));			addPass(createMipsOs16(getMipsTargetMachine()));
	if (getMipsSubtarget().inMips16HardFloat())			if (getMipsSubtarget().inMips16HardFloat())
	addPass(createMips16HardFloat(getMipsTargetMachine()));			addPass(createMips16HardFloat(getMipsTargetMachine()));
	}			}
	// Install an instruction selector pass using			// Install an instruction selector pass using
	// the ISelDag to gen Mips code.			// the ISelDag to gen Mips code.
	bool MipsPassConfig::addInstSelector() {			bool MipsPassConfig::addInstSelector() {
	▲ Show 20 Lines • Show All 43 Lines • Show Last 20 Lines

llvm/trunk/lib/Target/Sparc/SparcTargetMachine.cpp

	Show First 20 Lines • Show All 41 Lines • ▼ Show 20 Lines
	public:			public:
	SparcPassConfig(SparcTargetMachine *TM, PassManagerBase &PM)			SparcPassConfig(SparcTargetMachine *TM, PassManagerBase &PM)
	: TargetPassConfig(TM, PM) {}			: TargetPassConfig(TM, PM) {}

	SparcTargetMachine &getSparcTargetMachine() const {			SparcTargetMachine &getSparcTargetMachine() const {
	return getTM<SparcTargetMachine>();			return getTM<SparcTargetMachine>();
	}			}

				void addIRPasses() override;
	bool addInstSelector() override;			bool addInstSelector() override;
	bool addPreEmitPass() override;			bool addPreEmitPass() override;
	};			};
	} // namespace			} // namespace

	TargetPassConfig *SparcTargetMachine::createPassConfig(PassManagerBase &PM) {			TargetPassConfig *SparcTargetMachine::createPassConfig(PassManagerBase &PM) {
	return new SparcPassConfig(this, PM);			return new SparcPassConfig(this, PM);
	}			}

				void SparcPassConfig::addIRPasses() {
				addPass(createAtomicExpandPass(&getSparcTargetMachine()));

				TargetPassConfig::addIRPasses();
				}

	bool SparcPassConfig::addInstSelector() {			bool SparcPassConfig::addInstSelector() {
	addPass(createSparcISelDag(getSparcTargetMachine()));			addPass(createSparcISelDag(getSparcTargetMachine()));
	return false;			return false;
	}			}

	/// addPreEmitPass - This pass may be implemented by targets that want to run			/// addPreEmitPass - This pass may be implemented by targets that want to run
	/// passes immediately before machine code is emitted. This should return			/// passes immediately before machine code is emitted. This should return
	/// true if -print-machineinstrs should print out the code after the passes.			/// true if -print-machineinstrs should print out the code after the passes.
	Show All 28 Lines

llvm/trunk/lib/Target/XCore/XCoreTargetMachine.cpp

	Show All 35 Lines
	public:			public:
	XCorePassConfig(XCoreTargetMachine *TM, PassManagerBase &PM)			XCorePassConfig(XCoreTargetMachine *TM, PassManagerBase &PM)
	: TargetPassConfig(TM, PM) {}			: TargetPassConfig(TM, PM) {}

	XCoreTargetMachine &getXCoreTargetMachine() const {			XCoreTargetMachine &getXCoreTargetMachine() const {
	return getTM<XCoreTargetMachine>();			return getTM<XCoreTargetMachine>();
	}			}

				void addIRPasses() override;
	bool addPreISel() override;			bool addPreISel() override;
	bool addInstSelector() override;			bool addInstSelector() override;
	bool addPreEmitPass() override;			bool addPreEmitPass() override;
	};			};
	} // namespace			} // namespace

	TargetPassConfig *XCoreTargetMachine::createPassConfig(PassManagerBase &PM) {			TargetPassConfig *XCoreTargetMachine::createPassConfig(PassManagerBase &PM) {
	return new XCorePassConfig(this, PM);			return new XCorePassConfig(this, PM);
	}			}

				void XCorePassConfig::addIRPasses() {
				addPass(createAtomicExpandPass(&getXCoreTargetMachine()));

				TargetPassConfig::addIRPasses();
				}

	bool XCorePassConfig::addPreISel() {			bool XCorePassConfig::addPreISel() {
	addPass(createXCoreLowerThreadLocalPass());			addPass(createXCoreLowerThreadLocalPass());
	return false;			return false;
	}			}

	bool XCorePassConfig::addInstSelector() {			bool XCorePassConfig::addInstSelector() {
	addPass(createXCoreISelDag(getXCoreTargetMachine(), getOptLevel()));			addPass(createXCoreISelDag(getXCoreTargetMachine(), getOptLevel()));
	return false;			return false;
	Show All 19 Lines

llvm/trunk/test/CodeGen/XCore/atomic.ll

	Show All 16 Lines

	@pool = external global i64			@pool = external global i64

	define void @atomicloadstore() nounwind {			define void @atomicloadstore() nounwind {
	entry:			entry:
	; CHECK-LABEL: atomicloadstore			; CHECK-LABEL: atomicloadstore

	; CHECK: ldw r[[R0:[0-9]+]], dp[pool]			; CHECK: ldw r[[R0:[0-9]+]], dp[pool]
	; CHECK-NEXT: #MEMBARRIER
	%0 = load atomic i32* bitcast (i64* @pool to i32*) acquire, align 4

	; CHECK-NEXT: ldaw r[[R1:[0-9]+]], dp[pool]			; CHECK-NEXT: ldaw r[[R1:[0-9]+]], dp[pool]
				; CHECK-NEXT: #MEMBARRIER
	; CHECK-NEXT: ldc r[[R2:[0-9]+]], 0			; CHECK-NEXT: ldc r[[R2:[0-9]+]], 0
				%0 = load atomic i32* bitcast (i64* @pool to i32*) acquire, align 4

	; CHECK-NEXT: ld16s r3, r[[R1]][r[[R2]]]			; CHECK-NEXT: ld16s r3, r[[R1]][r[[R2]]]
	; CHECK-NEXT: #MEMBARRIER			; CHECK-NEXT: #MEMBARRIER
	%1 = load atomic i16* bitcast (i64* @pool to i16*) acquire, align 2			%1 = load atomic i16* bitcast (i64* @pool to i16*) acquire, align 2

	; CHECK-NEXT: ld8u r11, r[[R1]][r[[R2]]]			; CHECK-NEXT: ld8u r11, r[[R1]][r[[R2]]]
	; CHECK-NEXT: #MEMBARRIER			; CHECK-NEXT: #MEMBARRIER
	%2 = load atomic i8* bitcast (i64* @pool to i8*) acquire, align 1			%2 = load atomic i8* bitcast (i64* @pool to i8*) acquire, align 1
	▲ Show 20 Lines • Show All 55 Lines • Show Last 20 Lines