Use a link or an existing model
Inspect supported GGUF files, split sets, repository links, and local folders. Existing files remain in place.
Model sources →Open-source local inference / 0.4.6
Supply a model link or a local file. Aperture inspects your hardware with permission, explains a CPU/GPU configuration, and opens a local session through a supported runtime.
Your selected artifact and explicit context remain unchanged. Save the configuration, return to a local chat, or run a bounded experiment on the same machine.
From a chosen model to an explained method
Approve a read-only view of hardware and available resources.
Paste a model link or local path. Select the exact representation.
See the candidate backend, memory budgets, context, and limits.
Approve acquisition and runtime installation when required, then run.
This describes the terminal workflow. The website does not scan your computer.
Using the release
Inspect supported GGUF files, split sets, repository links, and local folders. Existing files remain in place.
Model sources →For compatible GGUF models, one GPU can execute some layers while the CPU handles the rest. Memory pools stay distinct.
Memory and placement →A scan does not approve downloads, installation, inference, or experiments. Results are not uploaded automatically.
Privacy and local state →Observed support
Native records cover Windows CPU, selected NVIDIA and Intel graphics paths, and Linux CPU. Version 0.4.1 includes a recorded RTX 3090 CUDA chat after the installed-library discovery correction.
Partial offload above available VRAM has been observed on a 4060. Successful execution above its physical capacity remains unverified. NPU inventory is not NPU execution.
Read the support and evidence matrix →Documentation