Skip to content
@makllama

MaKLlama

MaK(Mac+Kubernetes)llama: running and orchestrating large language models (LLMs) on Kubernetes with Mac nodes.

Popular repositories Loading

  1. makllama makllama Public

    MaK(Mac+Kubernetes)llama - Running and orchestrating large language models (LLMs) on Kubernetes with macOS nodes.

    Go 46 3

  2. llama.cpp llama.cpp Public

    Forked from ggml-org/llama.cpp

    LLM inference in C/C++

    C++ 3

  3. containerd containerd Public

    Forked from containerd/containerd

    An open and reliable container runtime

    Go 1

  4. cri cri Public

    Forked from virtual-kubelet/cri

    Go 1 1

  5. ktransformers ktransformers Public

    Forked from kvcache-ai/ktransformers

    A Flexible Framework for Experiencing Cutting-edge LLM Inference Optimizations

    Python 1

  6. ollama ollama Public

    Forked from ollama/ollama

    Get up and running with Llama 3, Mistral, Gemma, and other large language models.

    Go

Repositories

Showing 10 of 26 repositories
  • Mooncake Public Forked from kvcache-ai/Mooncake

    Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.

    makllama/Mooncake's past year of commit activity
    C++ 0 Apache-2.0 1,217 0 0 Updated Dec 17, 2025
  • llama.cpp Public Forked from ggml-org/llama.cpp

    LLM inference in C/C++

    makllama/llama.cpp's past year of commit activity
    C++ 3 MIT 23,480 0 0 Updated Nov 28, 2025
  • model-runner Public Forked from docker/model-runner

    Docker Model Runner

    makllama/model-runner's past year of commit activity
    Go 0 Apache-2.0 156 0 0 Updated Oct 29, 2025
  • MAD Public Forked from ROCm/MAD

    MAD (Model Automation and Dashboarding)

    makllama/MAD's past year of commit activity
    Shell 0 MIT 60 0 0 Updated Oct 28, 2025
  • gpustack Public Forked from gpustack/gpustack

    Manage GPU clusters for running LLMs

    makllama/gpustack's past year of commit activity
    Python 0 Apache-2.0 651 0 0 Updated Aug 4, 2025
  • ramalama Public Forked from containers/ramalama

    Ramalama is an open-source developer tool that simplifies the local serving of AI models from any source and facilitates their use for inference in production, all through the familiar language of containers.

    makllama/ramalama's past year of commit activity
    Python 0 MIT 368 0 0 Updated Jul 28, 2025
  • cozeloop Public Forked from coze-dev/coze-loop

    Next-generation AI Agent Optimization Platform: Cozeloop addresses challenges in AI agent development by providing full-lifecycle management capabilities from development, debugging, and evaluation to monitoring.

    makllama/cozeloop's past year of commit activity
    Go 0 Apache-2.0 799 0 0 Updated Jul 26, 2025
  • octotools Public Forked from octotools/octotools

    OctoTools: An agentic framework with extensible tools for complex reasoning

    makllama/octotools's past year of commit activity
    Python 0 MIT 190 0 0 Updated Jul 24, 2025
  • llama-box Public Forked from gpustack/llama-box

    LLM inference server implementation based on llama.cpp.

    makllama/llama-box's past year of commit activity
    C++ 0 MIT 29 0 0 Updated Jul 24, 2025
  • stable-diffusion.cpp Public Forked from leejet/stable-diffusion.cpp

    Stable Diffusion and Flux in pure C/C++

    makllama/stable-diffusion.cpp's past year of commit activity
    C++ 0 MIT 802 0 0 Updated Jul 24, 2025

Most used topics

Loading…