Unlocking LLM Inference: A Deep Dive into Advanced GPU Architectures | SWE180 Engineering Articles | SWE180