Gemini 3.6 Flash completed our entire four-round 3D build in 246 seconds. GPT-5.6 Sol’s longest single response in the same task took 1,050 seconds. That is not a verdict on quality. It is a vivid reminder that AI model latency changes what a workflow can be.
Disclosure: I work with OrcaRouter and used it to run this evaluation. It let me use one OpenAI-compatible workflow to keep the model harness consistent.
AI-generated illustration featuring official model logos; logos and model names are used descriptively and remain the property of their respective owners.
Compare Gemini 3.6 Flash and other current models on OrcaRouter.
The test required a complete standalone Three.js HTML file on every turn: blockout, warhorse, pumpkin rider, then materials and staging. The models received the same four-turn history. Gemini completed the sequence in just over four minutes; Sol took 1,765 seconds in total.
That does not mean Gemini produced the most complete final scene. In this selected run, Sol did. Nor does it establish a permanent seven-to-one speed ratio. One controlled generative-coding sequence is a descriptive case study, not a broad latency benchmark.
Still, independent data points in the same direction. Artificial Analysis reports Gemini 3.6 Flash at 304 output tokens per second and an average task completion time of 1.3 minutes in its own test suite. It reports a much longer time to first token for Sol at maximum effort. Different tasks, same practical warning: if a job has many conversational turns, the model choice can dominate the calendar.
Fast iteration is not a cosmetic metric. A designer can see an unwanted primitive stack, amend the brief, and retry while the object is still mentally present. An agent can move from plan to patch to validation without leaving a user waiting through a dozen long turns.
The fair workflow is usually mixed. Use a fast model for exploration and small transformations; reserve a slower, more deliberate model for a difficult final pass if it proves worth the delay on your tasks. “Best AI model” is rarely a useful default. Completion quality, latency, retries, and human review time belong in the same decision.
Sources
- Artificial Analysis model measurements, retrieved 2026-07-28.
- First-party Four-Model Three.js Build Demo, DATA-012.
Explore current models on OrcaRouter. The author works with OrcaRouter; access does not imply affiliation with or endorsement by Google or OpenAI. Gemini and GPT are trademarks of their respective owners.
First-party staged-build record. Timing is specific to the test harness and prompt sequence.

