Earlier quoted context omitted.
All I have to do is give up all hope of my code ever being portable to multiple architectures You know nothing about modern compiler atomics. Post 2011 compilers (LLVM, GCC, MSVC, ICC) standardized generic memory fences for C++ and C. These are fully portable as the compiler itself determines how the layout of Acquire/Releases needs to be changed on platform you are compiling too. Atomics are supported on ARM, x64, M…
Okay, what about HLLs? Memory fences don't work if you aren't writing C. And as I've said, forks are an easier way to get these guarantees, albeit at a moderate perf cost.
Are you just a professional contrarian?
Why are you concerned about system calls in a HLL?
Why are you concerned with concurrency guarantees in a HLL?
You care about none of these, you care about your run time. This is why you are working in a HLL.