Sharing actual GPU core and VRAM utilization metrics for query on 10 LLM models | Hacker News Reader