llama-cpp
llama.cpp local GGUF inference + HF Hub model discovery.
[llama.cppGGUFQuantizationHugging
PoorRican
0
modal-serverless-gpu
Serverless GPU cloud platform for running ML workloads. Use when you need on-demand GPU access without infrastructure management, deploying ML models as APIs, or running batch jobs with automatic scaling.
[InfrastructureServerlessGPUCloud
PoorRican
0