(Modified:
2026-10-10)
— Written by
SimeonOnSecurity— 11 min read
A five-model comparison of two RTX 5060 Ti cards and an NVIDIA GB10 across 4K to 256K context, including prompt processing, generation speed, and actual prompt lengths.
(Modified:
2026-10-06)
— Written by
SimeonOnSecurity— 10 min read
Run a local coding agent with Strata and OpenCode. Examine reported RTX 4060 Ti results, RAM requirements, prompt reuse, reasoning budgets, and the limits of replacing a paid subscription.
(Modified:
2026-10-06)
— Written by
SimeonOnSecurity— 10 min read
Compare local AI with ChatGPT and Claude through benchmark scores, memory estimates, serving speed, and agent verification. Includes RTX 5090, DGX Spark, and Mac Studio scenarios.
(Modified:
2026-10-10)
— Written by
SimeonOnSecurity— 16 min read
Qwen3.8 27B and its ternary repack Bonsai 2 27B now outscore Claude Sonnet 4.6 on the aggregate intelligence index while running on one consumer GPU. What changed, what the benchmarks show, and why memory pricing makes waiting the most expensive option.