I ran Laguna-S-2.1 through my private agentic eval vs Qwen3.5-122B on an RTX Pro 6000 (96GB). Fastest 100B+ I've tested and the best tool calling, but it invents facts under pressure.
2.55T2 sourcer/LocalLLaMA
Source record
Published by r/LocalLLaMA (T2 source). The original is at https://www.reddit.com/r/LocalLLaMA/comments/1v2ua8g/i_ran_lagunas21_through_my_private_agentic_eval/.
Pipeline notes
The summary and note below are generated by the signal pipeline — they are Beyond Desk’s reading, not quotations from the source.
SummaryUser compares Laguna-S-2.1 to Qwen3.5-122B on a private agentic evaluation suite run locally on an RTX Pro 6000 (96GB). Reports Laguna as the fastest 100B+ model tested with the best tool calling, but notes it fabricates facts under pressure.
Why it mattersFirsthand, hardware-specific benchmark of two open-weight models on agentic tasks, including a concrete failure mode (hallucination under pressure) that headline speed claims usually hide.
Cited by
No citations on record.
