Local AI models

Pick the model you run. See whether it fits your GPU, how fast it should go, which workloads have verified kernels, and license the kernel on Sui.

Pair your agent to see which models fit your GPU.Get started

Decode speeds are estimates from your GPU's memory bandwidth; workload numbers on the model pages are sample data until a track measures them. Cover orbs: Orbkit by zzzzshawn (MIT).