>
Tech News

Two AIs, One Website Prompt, Very Different Results

I gave Google Antigravity 2.0 and Cursor 3.0 the same prompt last weekend. The prompt was specific: build a single-page website for a fictional jewelry store called Diamond Vault, with a hero section, a products grid, a story section, and a contact form. Use placeholder images. Use placeholder copy. Make it look like a real small-business site, the kind a jewelry store in a small town would actually pay $2,000 to have built. Hand it back as a runnable project, not a description of what the project would look like.

Antigravity went home and built me a diamond. The hero section had a hand-drawn SVG (a vector image format that scales without losing quality) of a diamond, the products grid had per-product detail pages with simulated inventory state, the story section had a timeline, and the contact form had client-side validation, a thank-you state, and a fake API endpoint that returned a success message. The whole thing ran in 11 minutes. The code was clean, the design was opinionated, and the project was structured the way a senior frontend engineer would structure it. There were 4 unused files and a .env.example (a template file for environment variables) that I did not ask for. The unused files were a small price for the speed.

Cursor went home and checked boxes. The hero section was a heading and a paragraph. The products grid was a list of three placeholder cards with stock photos. The story section was a heading and a paragraph. The contact form was an unstyled HTML form that posted to #. The whole thing ran in 9 minutes. The code was boring, the design was minimal, and the project was structured the way a junior engineer would structure it: flat, no components, no design system.

By the spec I had given both tools, Cursor finished faster. By the spec of “the kind of small-business site a jewelry store would pay $2,000 for,” Antigravity finished better. The lesson is not “Antigravity is smarter.” The lesson is that the two tools optimize for different things, and the difference shows up the moment you give them a real assignment instead of a benchmark.

What Antigravity does that Cursor does not

Antigravity is a planning-first agent. Before it writes a single line of code, it produces a plan: a directory structure, a list of components, a list of state transitions, a list of assumptions, and a list of things it will not do. The plan is short, but it is the plan. The code that comes out of the plan is consistent with the plan. When Antigravity writes a products grid, it also writes a per-product detail page, because the plan said it would. When Antigravity writes a contact form, it also writes the form’s success state, because the plan said the form needed a success state.

Cursor is a generation-first tool. It writes the file you asked for. The product grid is a list of cards. The hero section is a heading and a paragraph. If you ask Cursor to add a per-product detail page, it will. If you do not ask, it will not. Cursor is faster because it does less. The cost is that the result is a list of files, not a project. The difference is the difference between a contractor who shows up with a plan and a contractor who shows up with a hammer.

For a single-file script or a small utility, Cursor is the right tool. The generation-first approach produces clean code fast, and the project is small enough that the plan is not needed. For a multi-file project with cross-cutting concerns (state, routing, validation, accessibility), the plan-first approach produces a project that holds together when you start adding to it. The deciding factor is not the prompt. The deciding factor is the project size.

What neither tool does well

Both tools failed at the same thing. The prompt said “make it look like a real small-business site.” Both tools produced a generic-looking site. The hero section had a stock photo of a diamond. The story section had placeholder copy that read like placeholder copy. Neither tool asked me who the customer is, what the price point is, what the store’s brand is. Neither tool could have asked, because the prompt did not say. The lesson is that the agent does not replace the brief. The agent executes the brief. If the brief is thin, the output is thin.

I have run this experiment four times in the last six weeks. The pattern is consistent. Antigravity produces more code, more files, more design opinion. Cursor produces less code, fewer files, less design opinion. Both tools hit the same ceiling: they cannot replace the brief. The agent is not a creative director. The agent is an executor. The brief is the creative director’s job.

How to use these tools in 2026

If you are a developer, the agent is a pair programmer (a coding partner that helps you write code in real time), not a replacement. The agent does the typing. You do the thinking. The plan-first tools (Antigravity, Claude Code, Codex in plan mode) are best for projects you would have spent a day planning by hand. The generation-first tools (Cursor, Copilot, Cody) are best for projects where the plan is already in your head. If you find yourself iterating on the same prompt five times in a row, the brief is the problem, not the tool.

If you are not a developer, the agent is a power tool that does not have a safety. The agent can produce a project that compiles, runs, and looks good. The agent cannot tell you whether the project is the right project. The cost of the agent is the same as the cost of any other power tool: you get out what you put in, and the put-in is the brief.

The single best use of these tools in 2026 is the one I have been doing all morning: writing a draft of an article, then handing the draft to the agent to tighten the prose. The agent is faster than I am at tightening prose. The agent is also faster than I am at writing prose. The difference is that I know what the article is supposed to say, and the agent does not. The agent is the typist. The brief is the writer. Neither is optional.

What I would tell past me

A few lessons that would have saved me the first afternoon of testing.

  • Run the same prompt on two tools before you commit. The differences between Antigravity and Cursor are not subtle. The plan-first approach produces a project. The generation-first approach produces a list of files. You will see the difference the first time you add a second page to the project.
  • Brief the brand, not just the spec. Both tools built a generic-looking site. The brief is the only place the brand can come from. The agent cannot invent the brand.
  • Set a time budget before you start. The 11 minutes vs 9 minutes is real, but the planning step on Antigravity is also a time cost. Give yourself 90 minutes per tool, then stop. The agent will keep generating if you let it.
  • Review the plan, not the code. The plan is the agent’s commitment to the project. If the plan is wrong, the code will be wrong. Reading the plan first saves a code review later.

Trade-offs

Antigravity is slower than Cursor on the same prompt. The 11 minutes vs 9 minutes is real. The plan that Antigravity produces is also a cost: you have to read the plan, confirm the plan, and adjust the plan. For a small project, the plan is overhead. For a large project, the plan is a forcing function that prevents the project from drifting.

Cursor produces fewer unused files. The .env.example and the four unused components that Antigravity shipped are not a deal-breaker, but they are noise. The noise is the cost of the planning-first approach. The trade-off is consistency for clutter.

Neither tool is a substitute for testing. Both produced code that ran, and both produced code that I would not have shipped to production without a code review. The agent does not replace the code review. The agent produces the code that the code review reviews. The review is the bottleneck. The agent is the fast part.

Both tools are still improving. The 2.0 release of Antigravity shipped a faster plan generator. The 3.0 release of Cursor shipped better file-tree awareness. The direction is clear: the agents are getting better at the planning-first approach and at the generation-first approach, and the gap between them is shrinking. In six months, the tools will probably converge on a hybrid that plans when the project is large and generates when the project is small. Until then, the choice is a real choice: which cost do you prefer to pay.

The 2,000 dollars the small-town jewelry store would have paid for a real site is still a real budget. The agent that can produce that site in 11 minutes is not yet a replacement for the developer who can produce that site in a week. The agent is a tool. The developer is a craftsperson. The craftsperson uses the tool. The tool does not replace the craftsperson. Not yet, anyway.

Leave a comment