Gate live E2E tests behind explicit provider selection and add OpenAI-compatible live path
Add repeatable OpenAI-compatible provider backend configuration to a TypeScript runtime
Provider abstraction existed but boot path still hard-coded a single LLM provider
Clean-break standalone migration still had dual-mode config defaults and stale tests
Thinking-capable local Ollama agents produced empty scorable output in a [redacted:name] benchmark harness
Sandbox benchmark agents to prevent local answer-key leakage
CTF benchmark over-scored wrong-location findings and leaked answer hints in cold prompts
Silhouette-only k-means split refuses big-soupy clusters; forced-k fallback with cohesion-improvement gate splits the residue
Blender Python: multi-object edit mode silently corrupts other meshes' UVs/normals
PSX vertex snap on world-space UV materials causes texture "boiling" on flat surfaces
React snapshot-on-settle timer captures empty data when streaming flushes are throttled
[redacted:name] Qwen3 benchmark agents emit thinking-only output and schema-mismatched findings
[redacted:name] benchmark runner launched duplicate models because per-wave concurrency repeated the wave model
Add local Ollama-backed model trial to a TypeScript benchmark while preserving CLI agent tooling
Reconstruct a branch by removing one author's commits while preserving other authors' changes
CTF benchmark harness used local throwaway agents instead of provided real agent keys
Gemini benchmark agents hit 429 because free-tier daily project/model quota was exhausted, not just RPM burst traffic
React Three.js graph activity pulses kept renderer RAF loop alive indefinitely