TestGorilla vs Vervoe
By Sergio Gualda · 6 min read · Updated
TestGorilla is a broad talent assessment platform: its site lists 350+ tests across programming, cognitive ability, personality and language, plus AI video interviews and resume scoring, assembled into a battery you pick. Vervoe takes the opposite approach — candidates complete immersive tasks such as spreadsheets, code, presentations or video answers, and AI grades the output with scores its site says hiring teams can trace back to how the candidate performed. The practical difference is what the candidate spends: a TestGorilla battery is a set of tests you can send to everyone who applied, while a Vervoe task asks for real work and is therefore usually reserved for people who have already made a cut.
Written by uxerhub, which sells a competing product for six product roles. Everything below about TestGorilla and Vervoe comes from their own public sites. Where we fit is at the end, in its own section, and you can skip it.
TestGorilla and Vervoe, side by side
| TestGorilla | Vervoe | |
|---|---|---|
| What the candidate does | Sits the tests you assembled from the library — multiple choice, coding, cognitive, personality, language. | Produces work: a spreadsheet, a piece of code, a presentation, a video answer, depending on the task built for the role. |
| What you get back | Scores per test, comparable against everyone else who sat the same battery. | The artefact plus an AI grade, with the scoring traceable to how the candidate performed. |
| Where it sits in the funnel | Early. Cheap enough in candidate time to send to a whole applicant pool. | Later. A real task costs real hours, so most teams use it once the pool is already cut. |
| Breadth of role coverage | Broad by design — one platform intended to cover hiring across an entire organisation. | Broad library too, but the value concentrates on roles whose work has an observable output in a short sitting. |
| AI in the scoring path | AI features across interviews, scoring and screening. | Central. AI grades the work the candidate produced, and the site is explicit that the reasoning is inspectable. |
| What it will not tell you | How someone performs on the actual job. A test score is a proxy and the platform does not claim otherwise. | Anything about the candidates you did not ask to do the task — which, at volume, is most of them. |
Which one you should pick
Choose TestGorilla if
- You hire across many functions and want one platform for all of them.
- The applicant pool is large and the immediate job is cutting it down.
- You want video interviews and resume screening in the same product.
- You need results that line up identically across a whole cohort.
Choose Vervoe if
- The role has work you can simulate in an hour and see clearly.
- You are choosing between a handful of finalists, not screening hundreds.
- You want evidence of output rather than a proxy score.
- Your candidates will invest the time — usually because the role is senior or the offer is strong.
The real question is where in the funnel you are
These two get compared as if they were competing answers to one question. They are competing answers to two, and which one you have depends entirely on how many people applied.
If four hundred CVs arrived, the expensive problem is that reading them tells you almost nothing and you cannot ask four hundred people for an afternoon of work. A test battery is the right instrument: it costs each candidate twenty or thirty minutes, everyone gets the same thing, and the output is a ranked list you can act on.
If you are down to five people who all look strong, a ranked list of test scores does not help much, because they will all clear the bar. What separates them is what they actually produce, and that is what a Vervoe task shows you.
Teams that run both are not being indecisive. They are using a cheap wide filter and then an expensive narrow one, in that order, which is the correct shape for almost any selection process.
What AI grading changes, and what it does not
Vervoe's grading of open work is the more ambitious engineering problem of the two, and its site's insistence that hiring teams can see how and why a candidate scored is the right instinct — traceability is what separates a defensible AI score from an unaccountable one.
It is worth being clear-eyed about what remains hard regardless of vendor. Grading open-ended work means a model is making a judgement about quality, and candidates now have the same class of model helping them produce the work. That is not an argument against the approach; it is the reason to look at the artefact yourself for anyone you are seriously considering, rather than treating the grade as the decision.
TestGorilla's exposure is different rather than smaller. Multiple-choice and coding tests are easier to score reliably and easier to prepare for, and a candidate who has sat similar batteries before has a real advantage that has nothing to do with the job.
Disclosure · our own product
Where uxerhub fits, if neither is quite it
uxerhub is neither a library nor a work sample. It covers six product roles — product designer, product manager, UX researcher, design engineer, front-end engineer and UX writer — with twelve written situations in which no option is correct, and scores which action a candidate reaches for and which they rule out.
The reason it exists between these two is time. A test battery is cheap enough for everyone but measures what someone knows; a work sample measures what they can do but costs hours. Fifteen written minutes is cheap enough to send to the whole pool while still measuring a decision rather than recall.
It is also narrower than both, deliberately, and it is not calibrated yet — we have not run enough hires through it to publish predictive validity, and we say so on our own site. If you hire across many functions, one of the two platforms above is the right purchase and this is not.
Frequently asked
- Is TestGorilla or Vervoe better for hiring designers?
- Neither is built specifically for design roles. TestGorilla's library covers many functions and Vervoe's strength is roles whose output can be produced and graded in a sitting, which is a harder fit for design work than for a spreadsheet or a piece of code. For design specifically, most teams end up supplementing either one with a portfolio review.
- Can you use TestGorilla and Vervoe together?
- Yes, and it is a common shape: a test battery to cut a large pool, then a work task on the survivors. The cost is that candidates who reach the final stage have now done two assessments, which is worth weighing against how much unpaid time the role justifies.
- Which one is cheaper?
- Neither publishes rates on its homepage, so any figure quoted elsewhere is worth checking against a current quote. The more useful comparison is candidate time: a test battery costs each applicant tens of minutes, a real work task costs hours, and at volume that difference dominates.
- Do candidates prefer one over the other?
- It depends on the stage. Candidates generally resent long unpaid tasks early in a process and accept them late, once there is a real prospect of an offer. A short battery is the easier ask at the top of a funnel; a substantial work sample is more defensible at the end.