This is a 124M-parameter GPT-2 fine-tuned to read a smart-home command and emit the matching structured tool call. The model and the optional speech recognition run entirely in your browser via WebGPU — no server, no API, nothing leaves your device.

  1. 1 Pick an example command (or speak / type your own)
  2. 2 Load the model — first run streams ~330 MB from the Hugging Face Hub
  3. 3 Hit Generate and watch it produce the JSON tool call
1

Choose a command

Casual, lifelike phrasings — implied rooms, vague values, questions — for validating the assistant on real end-user input. Or edit the prompt in Advanced below.

2

Load the model

idle
3

Generate the tool call

Advanced — prompt & decoding options

The preset picker above rewrites this automatically. Edit it directly to test custom tool lists or phrasings.

Load a model first ↑
Benchmark & timing details
Run a generation to see prompt tokens, throughput and grammar overhead.