Packages LLM weights together with an inference engine into single-file executables. Users run large models locally across operating systems without installation, including speech transcription tools.