open-source · since 2024
RamaLama
RamaLama community / Red Hat
An open-source tool for pulling, running and serving AI models through container-oriented workflows. It detects available accelerators, selects matching OCI runtime images, supports multiple model registries, and defaults local model execution to isolated rootless containers.
AImodel servingcontainersPodmanlocal inference
desk notes
Verified 2026-09-01 against the canonical site and repository. Latest GitHub release checked this shift: v0.24.0, published 2026-08-21. The project README still labels the software alpha/in development and warns users to expect breaking changes.