Story Highlights
- Engagement metrics like session length and completion rate aren’t proof that a game actually teaches anything.
- This comparison ranks 7 studios by the evidence they can show, not just portfolio and price.
- The right kind of proof depends on the buyer: a school district, a hospital, and a corporate L&D team each need something different.
Six months after launch, the dashboard looks good: session length is up, completion is at 80%, and reviews are positive. Then a board member asks, “Did anyone learn anything?” The room goes quiet, because the team only measured engagement. That’s what happens before learning, or sometimes instead of it.
Most comparisons of top educational game development companies rank studios by portfolio and price. This one compares firms by the evidence they provide: what each studio can show, where that evidence was published, and whether an outside party verified it.
Some studios have peer-reviewed trials, registry listings, and peer-judged awards. Others have operational data from national-scale deployments. All of it matters, but each proves something different.
The Trouble With Proof In Educational Game Development

Educational game development companies rarely lie about outcomes. More often, they cite evidence that belongs to a different product, because results in this field are conditional.
One classroom study makes the point well. A Vanderbilt University study that evaluated more than 1,000 students across 7 states found statistically significant gains in performance and engagement from short, curriculum-aligned games. A later WestEd study of the same platform reported that students who played at least twice a week scored up to 23 percentile points higher on state science tests than peers who didn’t.
Those numbers apply to specific games, in specific subjects, over a specific length of time. If a studio presents them as proof that any educational game will work, that says more about its sales process than its product.
The practical move is to ask what kind of evidence a studio can actually produce, then match that evidence to your buyer. A hospital procurement team, a school district, and a head of L&D will each require different forms of proof.
Proofs Educational Game Development Companies Can Offer

A few categories of evidence are worth asking for directly:
- Peer-reviewed studies involve independent researchers, a control group, and a journal review process. They’re slow and expensive to produce, but they hold up best with academic or clinical audiences. Ask who funded the study, since industry funding is common and should be disclosed upfront.
- Registry and regulator listings come from government or professional bodies with published criteria, making them harder to influence than awards, and especially relevant in health, safety, and education contexts.
- Judged awards with published criteria, like the International Serious Play Awards, BAFTA, Games for Change, or the Japan Prize, aren’t proof of learning on their own. They do show that expert judges rated the work highly against known standards.
- Operational data from real deployments, including completion and error rates, time to competency, support tickets, and on-the-job retention, is weaker academically but often more persuasive internally, since it comes from your own users.
If a studio has none of these, that doesn’t automatically rule it out. It just means proving the game works becomes your responsibility, so budget for it accordingly.
The Top Educational Game Development Companies Compared

This list covers seven evidence-focused studios:
- Whimsy Games for custom builds with analytics that help generate your own evidence.
- Legends of Learning for curriculum-aligned science and math games backed by published efficacy research.
- Second Avenue Learning for custom learning software built with research partners.
- &ranj for behavior-change games and game-based assessment.
- Osso VR for surgical training tested in a randomized controlled trial.
- Level Ex for physician-focused medical games, and Strivr for immersive workplace training measured at enterprise scale.
| Company | Strongest Form Of Proof | Domain | Where It Was Checked | Best Fit |
|---|---|---|---|---|
| Whimsy Games | Operational data plus verified client reviews | Mobile, STEM, financial literacy, workplace gamification | Clutch and GoodFirms | You own the content and want evidence from your own users |
| Legends of Learning | Peer-reviewed efficacy research | K-8 science and math | Journal of the Learning Sciences, WestEd | Classroom deployment where test scores are the target |
| Second Avenue Learning | Judged award plus research collaboration | K-12, higher education, health training | International Serious Play Awards | Custom build for a publisher or institution |
| &ranj | Game-based assessment data | Corporate, healthcare, education | Japan Prize, European Innovative Games Award | Behavior change you need to measure, not just teach |
| Osso VR | Randomized controlled trial | Surgical and medical device training | Peer-reviewed orthopedic literature | Procedural skill where errors are expensive |
| Level Ex | Clinical adoption inside a medtech group | Physician and surgeon training | Acquired by Brainlab in 2020 | Specialty medical audiences and device education |
| Strivr | Enterprise-scale outcome data | Retail, logistics, safety, soft skills | Customer deployments, Stanford origins | Workforce training rolled out to thousands |
Whimsy Games

Proof: verified client reviews with named outcomes, plus the instrumentation to generate your own data.
Whimsy Games fits buyers who intend to prove results with their own data. The studio is Unity-first and full-cycle. Its Clutch profile lists 100+ mobile titles across casual, mid-core, and educational genres, a $25-$49 hourly rate, and a $10,000 minimum project size; it’s also listed on GoodFirms.
One GoodFirms review describes a gamification system with rewards, achievements, and progress tracking that improved staff engagement and retention. Another describes an educational game with an AI assistant designed to help learners understand complex material. These are client-reported outcomes from real deployments on a third-party platform, which is a stronger claim than citing someone else’s research.
For pilot-driven buyers, the value is instrumentation. Progress tracking, achievement states, and analytics are part of Whimsy’s educational and gamification work, not a phase-two add-on. The same team also handles design, art, engineering, QA, and live operations.
If you plan to run a controlled pilot and publish the results yourself, you need a partner that can build the game, measure the right signals, and iterate once the first data arrives. Bring the learning objectives and measurement plan, and Whimsy Games can execute the production side.
Legends of Learning

Proof: peer-reviewed research in a top-ranked education journal, plus an independent WestEd evaluation.
Legends of Learning grew out of a research question. Before launching the platform, its founder worked with a Vanderbilt University learning sciences professor on a study of more than 1,000 students across 7 states. The study tested whether short, curriculum-aligned games improved learning inside a real classroom curriculum.
Results showed statistically significant gains in both performance and engagement, and were published in the Journal of the Learning Sciences.
The platform now offers more than 1,000 science and math games, each lasting 5-25 minutes and aligned with elementary and middle school standards. Teachers can assign playlists and track comprehension through a dashboard. A later WestEd study reported that students using the platform at least twice a week scored up to 23 percentile points higher on state science tests than non-users.
Legends of Learning is a marketplace that commissions games from many independent developers, not a studio you’d hire for one bespoke build. For districts or publishers that need classroom-ready content with a research trail, it’s one of the strongest options available. For companies building a custom product, it’s better treated as a benchmark to measure against.
Second Avenue Learning

Proof: a gold medal at the 2023 International Serious Play Awards, and a research relationship with a university education school.
Second Avenue Learning has built custom learning software from Rochester, New York, since 2006. It’s a certified women-owned company and has worked with education publishers including Pearson, McGraw-Hill, and W. W. Norton.
The team combines learning designers, subject-matter experts, engineers, and artists, a structure that separates a true learning studio from a production shop.
Its evidence base is strongest in serious learning simulations. The company’s simulations supporting people with intellectual and developmental disabilities and Alzheimer’s, developed with the Alzheimer’s Association and supported by the Golisano Foundation, won gold at the 2023 International Serious Play Awards.
It has also presented research at the Games, Learning and Society conference with a University of Rochester education professor, adapting Bloom’s taxonomy into a framework for game-based design.
Second Avenue Learning was acquired by zSpace in 2025. That strengthens its access to immersive classroom hardware, but buyers should ask how independent the roadmap and project team remain after the acquisition.
&ranj

Proof: game-based assessments that measure behavior, backed by the Japan Prize and the European Innovative Games Award.
&ranj was founded in Rotterdam in 1999 and has been 50% owned by the consultancy &samhoud since 2015. The studio has delivered more than 400 games across corporate, healthcare, and education projects.
Its foundation is behavioral science. &ranj sells assessments and training that use play to measure what someone actually does in a simulated situation. That makes it a strong fit for evidence-focused projects. Self-reported confidence can rise after almost any training, but observed behavior in a scenario is a far stronger number for skeptical stakeholders.
&ranj offers both bespoke games and white-label titles tailored to a client’s branding, so smaller budgets don’t always have to start from scratch. Awards, including the Japan Prize and an Accenture Innovation Award, add external recognition, though they prove craft more than learning transfer.
Osso VR

Proof: a randomized controlled trial published in the orthopedic literature.
Osso VR trains surgical procedures in virtual reality and is one of the few companies in this market whose core claim has actually been tested in a randomized study.
In a blinded, randomized trial at UCLA, medical students trained on Osso VR’s module for intramedullary nailing of tibial shaft fractures were compared with students who used a published technique guide. The VR-trained group showed a large improvement in overall surgical performance.
The study was funded by Osso VR, which is common in early-stage medical technology but should still be disclosed. The participants were also medical students, so the result is meaningful evidence of skill acquisition, not proof of better patient outcomes down the line.
For device manufacturers, hospital systems, and training programs where procedural errors carry real cost, Osso VR is the kind of partner worth looking for: a company willing to test its product in a study it could have failed.
Level Ex

Proof: acquisition by a surgical technology company, and adoption inside clinical education.
Level Ex, founded in Chicago in 2015, builds video games for practicing physicians and surgeons. Its mobile, VR, and AR products place clinicians in realistic virtual patient scenarios and difficult decision-making situations. The team itself comes from the commercial games industry.
In 2020, Brainlab, a Munich-based digital medical technology company serving hospitals in more than 100 countries, acquired the 105-person studio. At the time, it was described as the first acquisition of a video game company by a major international healthcare company. Level Ex has continued operating under its own name and roadmap within the group.
That should be read as a signal of domain depth. If your audience is specialist clinicians and the content involves devices, procedures, or diagnostic judgment, Level Ex is a strong fit, since it already understands the level of medical accuracy these projects require.
Strivr

Proof: outcome data from deployments measured in the millions of learners.
Strivr began at Stanford, where its founder built VR practice for the university’s football team, before moving into workforce training. Walmart later adopted the platform across its training academies and sent headsets to stores nationwide. Employees used immersive modules covering situations such as Black Friday crowds and equipment handling, and more than one million people have now trained on the platform overall.
Strivr belongs on this list because of its measurement architecture. Its value is the data produced during practice. A headset can capture assessment data as employees train, letting companies evaluate performance across the workforce rather than relying solely on completion rates.
That makes Strivr a strong fit for large employers that need measurable training at scale. It isn’t suited to small deployments or classroom education, since the hardware logistics, content maintenance, and enterprise pricing all assume scale from the start.
Vetting A Studio’s Evidence Before You Sign

Have the studio name its most recent project where the learning outcome was actually measured, and who measured it. Find out what the game logs, how detailed that data is, and whether you can export it. Ask the team to identify which claim on its own website it would be comfortable defending to a researcher.
Then confirm what the studio would need from you to run a controlled pilot, how much of that work it can handle itself, and what it would do if the pilot showed no effect. A partner with a real answer to that has been through the process before.
What This Means For Your Next Pilot

Proof isn’t one thing. The right kind depends on who you need to convince: a district wants classroom research, a hospital wants a trial, and a CFO wants numbers pulled from your own workforce.
If that evidence doesn’t exist for your specific case yet, the job is to build a game that can produce it, run the pilot properly, and share the results internally before someone asks the question at the board meeting.
Thanks! Do share your feedback with us. ⚡
How can we make this post better? Your help would be appreciated. ✍





