Google DeepMind's SL2T Sign-Language-to-Text Model Launches on Pixel 11
TECH

Google DeepMind's SL2T Sign-Language-to-Text Model Launches on Pixel 11

17+
Signals

Strategic Overview

  • 01.
    SL2T is a multilingual sign-language-to-text model that launched on August 12, 2026, powering sign-to-text dictation in Gboard and Live Transcribe on Pixel 11, starting with American Sign Language to English.
  • 02.
    The model was trained on 100,000+ hours of sign-language data across 50+ sign languages (about a quarter of it ASL), and uses on-device MediaPipe Holistic tracking of 130 key points on the face, body, and hands, sending only those coordinates - never raw video - off-device for translation.
  • 03.
    SL2T translates directly from pose-landmark sequences to text, skipping the intermediate 'gloss' annotation step used by most prior sign-language translation systems because glosses fail to capture non-manual markers and spatial grammar.
  • 04.
    SL2T scores 70 BLEURT zero-shot on the FLEURS-ASL benchmark (described as the most capable sign-language translation model to date on such benchmarks), 74 on a one-handed variant, 85 on PRESTO-ASL, and 64% exact-match on FSboard fingerspelling.

Deep Analysis

How SL2T Actually Works: Landmarks, Not Gloss, Not Raw Video

SL2T does not see sign language the way a video call does. An on-device model first converts a signer's hand, face, and body movement into geometric coordinate landmarks - MediaPipe Holistic tracks 130 key points across the face, body, and hands, frame by frame - and only those coordinates, never the original camera feed, are sent off for translation [1][2]. The camera feed itself is discarded immediately, a design Google frames explicitly as a privacy safeguard rather than an afterthought [1].

The more consequential design choice sits one layer deeper: SL2T translates directly from that landmark sequence to text, skipping the intermediate 'gloss' annotation step that most prior sign-language translation systems relied on [3]. Glosses - a word-for-word transcription convention - fail to capture rich, non-linear aspects of sign languages such as non-manual markers and spatial constructions [3]. Google also says the model is specifically optimized for left-handed and one-handed signing, noting roughly 10% of signers are left-handed [4], and includes hallucination-prevention mechanisms designed to avoid falsely transcribing non-signing movements as sign input [5]. These are the kinds of details that rarely make a press release but are exactly what determines whether a translation model is usable in daily life rather than just a lab demo.

The Benchmark Number That Matters: 70 BLEURT and the 'Beats a Human Interpreter' Claim

The Benchmark Number That Matters: 70 BLEURT and the 'Beats a Human Interpreter' Claim
SL2T's BLEURT scores across benchmarks, compared against a community-sourced human-interpreter baseline and previous best.

The number Google is leading with is 70 - a BLEURT score on the FLEURS-ASL benchmark in a zero-shot setting, which the company and independent coverage describe as the most capable sign-language translation result posted on that benchmark to date [1][6]. On other internal benchmarks, SL2T scores 74 BLEURT on a one-handed signing variant and 85 BLEURT on PRESTO-ASL, plus 64% exact-match accuracy on FSboard fingerspelling [1].

What makes 70 more than an incremental gain is a comparison that surfaced not in Google's own materials but in community discussion of the launch: against a widely cited previous-best score of 45 and a human-interpreter baseline reported around 64.6 on the same benchmark, SL2T's 70 would put it ahead of the average human interpreter on this specific test - a genuinely rare claim for any AI translation system. That framing deserves a caveat: it comes from third-party analysis of the benchmark rather than an official Google claim, and BLEURT is one metric on one academic benchmark, not a guarantee of real-world performance across dialects, lighting conditions, or unfamiliar signers. Google's own limitations list is candid about where the model still struggles: a 60-second maximum clip length, single-signer-only support, degraded accuracy in low light or extreme camera angles, and no training data from signers under 18 [1]. The company has also drawn a hard line around where SL2T should never be used - medical consultations, legal proceedings, police interactions, classroom instruction, job interviews, and government benefits determinations are all explicitly prohibited [1]- an unusually blunt admission that a benchmark-topping score does not equal a license for high-stakes deployment.

A Consumer Product First, Not a Research Demo: Built Into the Keyboard With the Deaf Community

Unlike most AI model announcements, SL2T did not ship as an API or a research paper - it shipped inside a keyboard. Because the feature lives in Gboard and Live Transcribe, a Deaf signer can dictate a web search, draft a message or document, or query Gemini simply by signing to the phone's camera instead of typing, and can sign responses back within a Live Transcribe conversation rather than typing replies to a hearing person [3]. That is a meaningfully different distribution strategy than shipping a standalone translation app: it puts the capability everywhere a Pixel 11 owner would normally reach for a keyboard.

The development process is also unusually specific about who shaped it. Google says the project was conceived by a Deaf Googler and formed an AI Sign Language Advisory Committee bringing together Deaf organizations and sign-language experts, who are co-authoring a report with Google on the model's capabilities and limitations [4]. That governance layer maps directly onto the product's guardrails: the explicit ban on medical, legal, policing, classroom, job-interview, and benefits-determination use cases [1]reads like the output of people who understand exactly how a mistranslated sign could hurt someone in those settings. The tradeoff is scope - at launch, SL2T supports only ASL-to-English, only on Pixel 11, leaving the roughly 200 sign languages used worldwide and every non-Pixel device unserved for now [3].

The Rare AI Launch Everyone Seems to Like - And the Strategy Debate It Reignited

AI product launches are rarely uncontroversial, which is part of what makes SL2T's reception notable. The top Reddit thread on r/singularity racked up roughly 2,200 upvotes and 144 comments with a dominant framing that this is a rare AI story people can just feel good about - commenters explicitly contrasted DeepMind's pursuit of 'cool and novel science' (weather forecasting, protein folding, and now sign language) with chatbot benchmark chasing elsewhere in the industry. On X, both DeepMind's own announcement and independent AI commentators drew celebratory replies rather than the skepticism that usually follows a big-tech AI reveal. The pushback that did surface was narrow: a sub-thread had to explain that ASL is a distinct language from English with its own grammar, so this is a translation feature for people with lower English literacy rather than a redundant convenience for people who could just type; another commenter doubted Google's long-term commitment to maintaining the feature, echoing a familiar pattern of product fatigue.

That goodwill fed directly into a second, more analytical debate: whether SL2T validates DeepMind's pattern of building specialized, narrow-purpose models for hard scientific and social problems - weather prediction, protein structure, now sign language - as a distinct strategic bet against the industry's focus on scaling general-purpose models. The Reddit debate over this point reached no consensus, but the fact that a keyboard feature triggered a company-strategy argument at all says something about how invested the AI community now is in reading every launch as evidence for one grand-strategy thesis or another.

Historical Context

2026-08-12
SL2T launched alongside the Pixel 11 device family, described in coverage as the first sign-language AI to ship inside a mainstream consumer product.

Power Map

Key Players
Subject

Google DeepMind's SL2T Sign-Language-to-Text Model Launches on Pixel 11

GO

Google DeepMind

Developer of SL2T; controls the model's architecture, training data, and rollout pace across languages and devices.

GO

Google Android / Pixel team

Ships SL2T inside Gboard and Live Transcribe exclusively on Pixel 11 at launch, controlling which devices and apps get access to the feature.

AI

AI Sign Language Advisory Committee (AISLAC)

External body of Deaf organizations and sign-language experts co-authoring a capabilities-and-limitations report with Google, giving it direct influence over how SL2T is positioned and restricted.

SA

Sam Sepah

Deaf Googler credited with conceptualizing the project, shaping its direction from the outset.

Fact Check

6 cited
  1. [1] Google DeepMind Brings Sign Language Translation to Phones With SL2T
  2. [2] Google's Sign Language AI Model Reads Body Landmarks
  3. [3] DeepMind's newest model allows Pixel 11 devices to transcribe sign language into text
  4. [4] Putting sign language AI into users' hands
  5. [5] Google debuts SL2T, an AI model that's designed to understand sign language
  6. [6] Google DeepMind brings sign language AI to Pixel 11

Source Articles

Top 5

THE SIGNAL.

Analysts
The Crowd

SL2T is our breakthrough sign language-to-text model powering new features for Deaf and hard of hearing users on @Android. Starting with American Sign Language-to-English on Pixel 11, people can sign directly into Gboard and Live Transcribe instead of typing.

@@GoogleDeepMind1124

Google DeepMind is doing some crazy work. This is so good and is going to help so many people. They just launched SL2T (Sign Language to Text), an AI model that translates sign language directly into text. Most spoken language dictation tools rely on audio to text, but sign...

@@ai_for_success168

We built SL2T with the Deaf community - guided by Deaf Googlers and our AI Sign Language Advisory Committee. Bringing ASL input to phones is just the beginning and we're working to expand this technology to more sign languages and applications. Find out more →

@@GoogleDeepMind91

DeepMind just released SL2T, sign language-to-text model, deaf users can now sign into their phones instead of typing, developed with heavy input from the Deaf community

@u/TorturedPoet302200
Broadcast
Google DeepMind Just Destroyed The Sign Language Barrier

Google DeepMind Just Destroyed The Sign Language Barrier

[8/12 15:00] Made by Google 2026 - Pixel 11, Gemini, and the SL2T sign language model / CodeRabbit raises $143M

[8/12 15:00] Made by Google 2026 - Pixel 11, Gemini, and the SL2T sign language model / CodeRabbit raises $143M

Pixel 11とGemini Intelligence全面刷新/Google手話AI/Suno×BMG提携ほか【Creative AI Digest No.149 2026.08.13】

Pixel 11とGemini Intelligence全面刷新/Google手話AI/Suno×BMG提携ほか【Creative AI Digest No.149 2026.08.13】

Google DeepMind's SL2T Sign-Language-to-Text Model Launches on Pixel 11 — AI News | Agentic Brew