Accepts a plain-language animation prompt and infers the full search workflow without needing a detailed implementation spec.
What was measured
Input Handling
Accepts plain-language prompts without detailed specs.
decisive for this rankingcapability
The ranking is specifically about text inputs, so a tool must handle plain-language prompts without heavy specification to do the job well. (2 of 3 judges)
What was given, what came back
Test input: Search engines basics explainer · text
Input — what we sent
The exact prompt
Create an animation video explaining how search engines work.
A vague, linear educational animation prompt asking the tool to explain how search engines work with minimal guidance. Designed to stress inference, scene structuring, pacing, and coherent visual flow from an underspecified request.
Output — unretouched
Gemini Canvas — Screen Recording 2026-05-04 120930.mp4
Also checked on this input — same tool, 6 other criteria
Code Generation Quality✓ WorkedProduces syntactically correct HTML/CSS/JavaScript on the first attempt for a basic explainer animation.Code Generation Quality✓ WorkedGenerates syntactically correct HTML/CSS/JavaScript on the first attempt; the page ran immediately as a self-contained animation.Output Quality◐ MixedDoes not fully realize the requested developer-themed UI, leaving the search animation visually plain.Output Quality✓ WorkedIt can cover the full search-engine workflow—crawling, indexing, and ranking—without overlap or visual clutter.Output Quality◐ MixedKeeps the animation inside a macOS-style tab frame with a very small aspect ratio, which makes the result feel presentation-like rather than like a full motion-graphics video.Output Quality⚠ StruggledThe animation can be constrained to a small macOS-style tab frame, which makes it read more like a presentation than full-viewport motion graphics.
Provenance
- Observation
- bc790ae7-9b94-4532-b689-11b3d5bfa683
- Evidence run
- c751a76b-a105-4b3e-9ff3-6c4d6adc9f04
- Study
- Generate Code-based Animations from text inputs
- Research task
- 86b9nnzmy
- Tested at
- Jun 24, 2026
- Source
- first-party
- Evidence state
- observed
- Proof shown
- input + output shown
- Cost / latency
- not captured
- Repeat run
- not captured
- Tester
- not captured
The last three rows are honest blanks, not placeholders — our capture has no field for them yet.
Query this
get_evidence({
tool: "gemini-canvas"
})MCP · mcp.aidemos.com/api/mcp
Free with attribution.
Same input, same check — 3 other tools
measured on Input Handling
Antigravity✓ WorkedIt can turn a vague one-line educational prompt into a sensible linear explainer without needing detailed scene specs, inferring a four-stage search workflow on its own.Grok✓ WorkedAccepted a fully specified HTML/CSS/JavaScript animation prompt immediately, but still misread part of the request as interactive webpage behavior rather than animation-first motion graphics.RemotionVideo✓ WorkedAccepts a plain-language explainer prompt and infers a coherent 3-stage structure without needing detailed specs.
Real inputs and real outputs, no retouching · every cell queryable via API & MCP · aidemos.com