Qwen3.8-27B at Q6 Held 60-63 Tokens/Sec Across a 3090 and a 3060 for a 20-Hour Unbroken Agentic Session
r/LocalLLaMA (461 upvotes, 145 comments)·medium signal
A r/LocalLLaMA post reports nearly 20 hours of continuous goal-oriented agentic coding on Qwen3.8-27B Q6 split across an RTX 3090 and an RTX 3060, sustaining 60-63 tok/s for the whole session. This is a duration claim rather than a throughput record, and duration is what usually breaks first on split-GPU local setups. Paired with 145 comments of setup detail, it is the clearest signal yet that a two-consumer-card rig can hold a workday-length agent loop.