Independent public-interest evaluations of how systems respect dignity and support agency.
HXR Evaluation
Google Nest Displayed the Request Correctly but Did Not Follow It
The device displayed the user’s request correctly, but it continued showing small text and speaking the time instead of restoring the large clock and staying silent. HXR is examining why a correctly captured request did not change the device’s behavior.
- Evaluation ID
- UXR-EVAL-0013
- Status
- Open — gathering evidence
- Evidence
- 1 supporting incident
- Last updated
- August 17, 2026
Technical record details
- Canonical analytical title
- Nest Voice-Control Mechanism: Recognized Intent to Display-State and Output-Modality Execution
- Evaluation type
- Mechanism
- Benchmark posture
- Developing Evidence
- Blue Score readiness
- Developmental Data Only
- Reform / positive practice
- Not Assessed
- Public revision
- 1
What changed in Public Revision 1
Public Revision 1 publishes the first recognized-intent execution mechanism Evaluation from UXR-2026-0817-0001 while preserving uncertainty about semantic understanding, action availability, assistant stack, and technical cause.
What this evaluation covers
These canonical Entity references identify the products, systems, organizations, units, and other durable subjects covered by this Evaluation. Inclusion does not by itself establish responsibility or credit.
Public-eligible organizations, products, services, departments, and workflows in this Evaluation's scope. Inclusion identifies subject and provider context only; it does not by itself establish responsibility, internal ownership, fault, motive, or prevalence.
- Subject entities
- Google Nest smart display— Smart display product family
Google— Technology Company / Digital Service Provider
Evaluation scope
The user-facing mechanism that maps recognized natural-language requests on a Google Nest smart display to executable display-state and spoken-output behavior. The exact backend implementation may be Google Assistant, Gemini for Home, a hybrid, or another Google control layer and is not established. The Evaluation therefore names the mechanism by observed function rather than inventing an internal component.
Question being evaluated
When a Nest smart display accurately receives a natural-language instruction that specifies both content and output modality, does the user-facing control mechanism preserve those constraints through to device execution?
Current findings
What works well or deserves recognition
The contributor reports visible recognition feedback showing that at least the speech/transcription layer received the wording. The device also possessed a large clock state outside the failed command path.
Difficulties and opportunities to improve
In the supporting Incident, recognized language did not demonstrate consequential control over two material user-specified variables: readable display presentation and suppression of spoken output. Rephrasing approximately eight to ten times did not produce the requested result.
Mixed or conditional findings
Correct visible transcription does not by itself prove full semantic understanding. Failure to execute does not prove that the needed action is absent; an available action could have been misselected, overridden, inaccessible, or constrained by another layer. Those hypotheses remain separate.
Evidence coverage and limits
Currently supported by one contributor-authorized public Incident Case and contributor testimony. No independent technical evidence establishes the active assistant stack, semantic interpretation, action availability, policy override, rendering behavior, speech-control limits, or root cause.
How the evidence connects
The same single Incident supports this mechanism Evaluation, the related Google organization Evaluation, and the Nest smart-display product Evaluation. These are three analytical lenses on one occurrence, not three independent incidents.
Supporting incidents
- UXR-2026-0817-0001: Google Nest Recognized a Visual-Only Time Request but Did Not Restore the Large Clock or Suppress Speech
Agency dimensions
Acceptance / benchmark test
A request such as “display the time on the screen and do not tell me the time” should, when supported, result in a clearly readable visual time display with no spoken announcement. Materially equivalent phrasings should preserve the same objective without undocumented magic wording. If the system cannot perform the requested action, it should state the limitation rather than ignore a material constraint while acting as though the request was satisfied.
Organization response
No formal Google response to UXR has been received.
Next observation or verification
Reproduce the interaction with controlled command variants; identify current device and assistant capabilities and documented commands; add organization-authored evidence or response; and test later software versions for newly exposed controls or improved modality compliance.