TLDRocket
Sign in

Towards developing future-ready skills with generative AI

Google Research

Google Research launched Vantage, an experiment that uses AI avatars to test skills like collaboration and critical thinking. It grades soft skills the way we've long graded math, something schools have never really managed.

Based on reporting by Google Research — read the original for the full story.

Summary, retelling and take written by AI under human oversight; images are AI-generated illustrations. How we work · Report an error

Google Research just put a number on something educators have argued about for decades: how good are you, really, at resolving conflict or building on someone else's idea? The new tool, called Vantage, drops students into simulated conversations with AI avatars, then scores them on skills the OECD and the World Economic Forum keep flagging as essential for the next few decades, things like critical thinking and creative collaboration. It's live now in English through Google Labs, aimed at high school and college students.

The mechanics are more clever than they first sound. An "Executive LLM" runs the conversation and quietly steers it, according to Google, deliberately introducing friction, a pushback here, a disagreement there, so the student actually has to demonstrate the skill being measured rather than coast through a pleasant chat. Once the task wraps, a separate AI Evaluator reads the transcript against a rubric and produces a skill map: a score plus written feedback. It's essentially a testing engine built for things that don't fit neatly into a testing engine.

Google didn't just ship this and hope for the best. Working with New York University, the team ran a study with 188 people aged 18 to 25, checking two things: whether the steering actually worked, and whether the AI Evaluator's scores lined up with human raters. On the first question, steered conversations produced noticeably more usable evidence of the target skill than unsteered ones, without feeling scripted. On the second, the AI Evaluator's agreement with NYU's human graders matched the agreement between two human graders with each other. A second study with startup OpenMic, covering 180 students on creative writing tasks like character interviews, found similarly strong correlation between AI and human scoring.

The bigger ambition here is a

My take — AI-written commentary, not fact-checked reporting

I'll say the quiet part: this is Google building standardized testing for soft skills, which is either overdue or a little unsettling depending on your mood that day. The validation work against human raters is genuinely solid, not just a marketing claim, but I'd want to see how it handles kids outside the US 18-25 bracket before anyone puts this near a report card.

Read more about this at: Google Research

Related stories

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.