[flang][cuda] Delay descriptor alloc when addressed reused on host/device (#220534) CSE can share one fir.coordinate_of between the host-association capture store and a later fir.load. Treating that coordinate_of as a real use made cuf-alloc-delay think the movable group depended on an operand at the sink point, so the device descriptor stayed at function entry and cudaMallocManaged ran before cudaSetDevice. Count only users of the slot address that actually read it. Stores that populate the tuple still sink with the allocation group. GitOrigin-RevId: 81bf35d75a2925ff57a470b7a54c5d3730ade165
Flang is a ground-up implementation of a Fortran front end written in modern C++. It started off as the f18 project (https://github.com/flang-compiler/f18) with an aim to replace the previous flang project (https://github.com/flang-compiler/flang) and address its various deficiencies. F18 was subsequently accepted into the LLVM project and rechristened as Flang.
Please note that flang is not ready yet for production usage.
Read more about flang in the docs directory. Start with the compiler overview.
To better understand Fortran as a language and the specific grammar accepted by flang, read Fortran For C Programmers and flang's specifications of the Fortran grammar and the OpenMP grammar.
Treatment of language extensions is covered in this document.
To understand the compilers handling of intrinsics, see the discussion of intrinsics.
To understand how a flang program communicates with libraries at runtime, see the discussion of runtime descriptors.
If you're interested in contributing to the compiler, read the style guide and also review how flang uses modern C++ features.
If you are interested in writing new documentation, follow LLVM's Markdown style guide.
Consult the Getting Started with Flang for information on building and running flang.