Local machine theatre · infinite act

Waiting to start

    WebLLM · on this device only

    Before you continue

    This experience runs a language model directly on your device. Nothing will start until you choose to proceed.

    • The first run downloads the model files and keeps them in the browser cache, using network bandwidth and local storage.
    • Loading and running the model can use significant GPU or shared memory, slow down other apps, warm the device, and consume battery.
    • Dialogue generation happens locally in your browser; no remote inference service receives the conversation.

    You can choose a lighter model from Setup before continuing.

    This session

    Full transcript

      Local generation

      Theatre setup

      Applying changes reloads the page and clears this session's transcript. Previously downloaded models remain in the browser cache.

      Direct the next turn

      Choose a topic

      The current exchange will end, and the next speaker will open this subject in character. 0/280