Jules logo

Opinion page / measured product / rank 18

Jules Reviews and Ratings (2026)

Reader ratings for Jules, aggregated on the same 0 to 100 scale we use for the lab index, with our own commentary underneath. If you are looking for Jules opiniones, reviews or a rating you can check, the measurements behind every claim are on the Jules benchmark page.

Reader index

not rated yet

1 voters, 10 axis scores. Never blended into the lab index.

Lab index

68.69

Measured on the vibeOps spec

Strict pass rate

75.6%

34 clean and 4 failed executions

Median per prompt

146.5 s

p90 292.7 s

Reader index

One score per reader per axis on a 0 to 100 scale, editable at any time.

Reader index

not rated yet

Lab index

68.69

1 readers have scored 100% of the index weight so far. A reader index is published at 3 voters and 50% coverage.

Reader mean against the measured lab subscore for each axis
AxisWeightLabReadersVotesReader mean
Agent performance18%80.076.01
Reliability18%75.675.01
Scalability13%74.076.01
SEO and GEO12%55.051.01
API and MCP9%66.062.01
Integrations8%70.071.01
Design output7%62.065.01
Speed7%21.415.01
Value5%84.082.01
Code ownership3%96.093.01

The reader index is a weighted mean of reader submitted axis scores using the published index weights, renormalised over the axes readers have scored. It is never blended into the lab index. How this works.

Score Jules

Your scores feed the reader column only. Submitting an axis again edits your existing score rather than adding a second vote.

Reader scores need an account so one person cannot vote twice. Create a reader account or sign in.

Reader scores are published in a separate reader index column. They never move the measured lab index.

Reader lab notes on Jules

Written by readers with an account. Sorted by upvotes, then by date.

No reader lab notes yet. A note has to describe something you actually built, so the first one here is worth reading.

Write a lab note

What you built, what broke, what you measured. Short opinions belong in the open rating form further down.

Sign in to publish a lab note. Notes need at least 240 characters because a one line opinion is not evidence.

Open rating distribution

Open ratings submitted without an account. Neither the lab index nor the reader index is part of this average.

Reader rating distribution
BandShareRatings
90 to 1000
80 to 890
70 to 790
60 to 690
0 to 590

Lab commentary

Our reading of the numbers, kept separate from the reader ratings.

An asynchronous agent: you hand it a task against a repository and come back to a proposed change. Judged on wall clock that is expensive, and it posts one of the slower medians in the index, but the comparison is not quite like for like since nobody sits and watches it. Change quality was good and the tests it added were real tests rather than assertions that always pass.

On the measured axes, Jules is strongest at code ownership with a subscore of 96.0 and weakest at speed at 21.4. Because speed carries 7 of the 100 index points, that weakness costs it roughly 5.50 points against a perfect result on that axis alone.

Reader ratings and lab measurements answer different questions. A reader rating carries the thing a harness cannot capture: whether the product was pleasant to work with over weeks, how support behaved, whether the bill matched the plan. The lab index carries the thing an opinion cannot: 45 timed executions of the same specification with the failures counted. Read them side by side and treat a large gap between the two as the interesting signal.

0 open ratings

Newest first, no account needed. Moderated for spam and vendor astroturfing, not for sentiment.

No reader ratings published yet. Yours would be the first.

Add an open rating

No account needed. For a score that counts toward the reader index, use the axis form above.

Rate Jules

One rating per person per product, on the same 0 to 100 scale as the lab index. Ratings without a body are discarded. We publish the name you enter, so use what you are happy to see in public.