Developer knowledge network · moderated exchange

UnreliableCode コミュニティ

開発者リサーチ、リバース エンジニアリング、コーディング コミュニティ

Discussion

Understanding std::memory_order_acquire and std::memory_order_release in lock-free queues [StackOverflow Architecture Guide]

cpp_concurrency_guru
C++ Standards Expert
MEMBER
担当者: 47
参加日: Feb 2018
投稿: 17
ありがとう: 75
2 か月前 · Jun 25, 2026 3:37 AM
#1

Why std::memory_order_seq_cst is often overkill for single-producer single-consumer queues:

When writing a lock-free queue:

  • Producer: Writes payload to buffer, then performs tail.store(newTail, std::memory_order_release).
  • Consumer: Loads tail.load(std::memory_order_acquire), then reads payload.

release guarantees that all prior memory writes (the buffer data) are visible to any thread that executes an acquire load on that same atomic variable. On x86/x64, hardware already enforces TSO (Total Store Order), so acquire/release loads/stores compile to plain mov instructions with zero mfence penalty!

memory_model_mook
Low-Level C Veteran
MEMBER
担当者: 163
参加日: Jan 2019
投稿: 11
ありがとう: 33
1 か月前 · Jun 25, 2026 10:10 AM
#2

Great explanation. On ARM64 (which is weakly ordered), acquire generates LDAR and release generates STLR instructions, avoiding full DMB ISH memory barriers.

profiler_pat
Performance Hunter
MEMBER
担当者: 146
参加日: Aug 2019
投稿: 33
ありがとう: 31
1 か月前 · Jun 25, 2026 4:43 PM
#3

Lock-free algorithms without acquire-release semantics are a ticking time bomb on modern mobile and server processors.