Go to the documentation of this file.
30 #include "llvm/IR/IntrinsicsNVPTX.h"
42 #define NVVM_REFLECT_FUNCTION "__nvvm_reflect"
43 #define NVVM_REFLECT_OCL_FUNCTION "__nvvm_reflect_ocl"
47 #define DEBUG_TYPE "nvptx-reflect"
56 NVVMReflect() : NVVMReflect(0) {}
71 cl::desc(
"NVVM reflection, enabled by default"));
75 "Replace occurrences of __nvvm_reflect() calls with 0/1",
false,
84 assert(
F.isDeclaration() &&
"_reflect function should not have a body");
85 assert(
F.getReturnType()->isIntegerTy() &&
86 "_reflect's return type should be integer");
125 Callee->getIntrinsicID() != Intrinsic::nvvm_reflect))
129 assert(Call->getNumOperands() == 2 &&
130 "Wrong number of operands to __nvvm_reflect function");
134 const Value *Str = Call->getArgOperand(0);
135 if (
const CallInst *ConvCall = dyn_cast<CallInst>(Str)) {
137 Str = ConvCall->getArgOperand(0);
141 Str = Str->stripPointerCasts();
142 assert(isa<Constant>(Str) &&
143 "Format of __nvvm_reflect function not recognized");
145 const Value *Operand = cast<Constant>(Str)->getOperand(0);
146 if (
const GlobalVariable *GV = dyn_cast<GlobalVariable>(Operand)) {
149 assert(GV->hasInitializer() &&
150 "Format of _reflect function not recognized");
151 const Constant *Initializer = GV->getInitializer();
152 Operand = Initializer;
155 assert(isa<ConstantDataSequential>(Operand) &&
156 "Format of _reflect function not recognized");
157 assert(cast<ConstantDataSequential>(Operand)->isCString() &&
158 "Format of _reflect function not recognized");
160 StringRef ReflectArg = cast<ConstantDataSequential>(Operand)->getAsString();
161 ReflectArg = ReflectArg.
substr(0, ReflectArg.
size() - 1);
165 if (ReflectArg ==
"__CUDA_FTZ") {
169 if (
auto *
Flag = mdconst::extract_or_null<ConstantInt>(
170 F.getParent()->getModuleFlag(
"nvvm-reflect-ftz")))
171 ReflectVal =
Flag->getSExtValue();
172 }
else if (ReflectArg ==
"__CUDA_ARCH") {
180 I->eraseFromParent();
A set of analyses that are preserved following a run of a transformation pass.
This is an optimization pass for GlobalISel generic memory operations.
This is a 'vector' (really, a variable-sized array), optimized for the case when the array is small.
static PreservedAnalyses none()
Convenience factory function for the empty preserved set.
#define NVVM_REFLECT_FUNCTION
raw_ostream & dbgs()
dbgs() - This returns a reference to a raw_ostream for debugging messages.
static PassRegistry * getPassRegistry()
getPassRegistry - Access the global registry object, which is automatically initialized at applicatio...
constexpr StringRef substr(size_t Start, size_t N=npos) const
Return a reference to the substring from [Start, Start + N).
Flag
These should be considered private to the implementation of the MCInstrDesc class.
PassRegistry - This class manages the registration and intitialization of the pass subsystem as appli...
static Constant * get(Type *Ty, uint64_t V, bool IsSigned=false)
If Ty is a vector type, return a Constant with a splat of the given value.
unsigned ID
LLVM IR allows to use arbitrary numbers as calling convention identifiers.
static cl::opt< bool > NVVMReflectEnabled("nvvm-reflect-enable", cl::init(true), cl::Hidden, cl::desc("NVVM reflection, enabled by default"))
inst_range instructions(Function *F)
This is an important base class in LLVM.
FunctionPass * createNVVMReflectPass(unsigned int SmVersion)
initializer< Ty > init(const Ty &Val)
assert(ImpDefSCC.getReg()==AMDGPU::SCC &&ImpDefSCC.isDef())
StringRef - Represent a constant reference to a string, i.e.
void initializeNVVMReflectPass(PassRegistry &)
amdgpu Simplify well known AMD library false FunctionCallee Callee
constexpr size_t size() const
size - Get the string size.
static bool runOnFunction(Function &F, bool PostInlining)
static PreservedAnalyses all()
Construct a special preserved set that preserves all passes.
SmallVector< Instruction *, 4 > ToRemove
INITIALIZE_PASS(NVVMReflect, "nvvm-reflect", "Replace occurrences of __nvvm_reflect() calls with 0/1", false, false) static bool runNVVMReflect(Function &F
A container for analyses that lazily runs them and caches their results.
FunctionPass class - This class is used to implement most global optimizations.
This class represents a function call, abstracting a target machine's calling convention.
PreservedAnalyses run(Function &F, FunctionAnalysisManager &AM)
LLVM Value Representation.
#define NVVM_REFLECT_OCL_FUNCTION