Foundry Local enables local execution of Small Language Models using the hardware on your device.
DetailsLightweight, highly optimized CPU runtime for GGUF models and embeddings.
DetailsStandalone Docker Model Runner - run local AI models with no Docker Desktop or Engine required.
DetailsDesktop GUI for BitLlama LLM inference engine with Soul learning and model management
Details