Here are
7 public repositories
matching this topic...
A manual for helping using tesla p40 gpu
Dashboard for AI Studio, Open Source Continuous Inference | Deepseek-R1, Qwen2.5, Llama3.1 | 4xRTX-5090 inside PRU2500, 2xH100 inside PRU2500, 8xMI210 in SuperMicro
Updated
Jul 22, 2026
TypeScript
Dynamic CUDA router for ExLlamaV2: auto-fallback from P40 (sm_61) to RTX 3050
Updated
Jun 23, 2026
Python
Real-time GPU/CPU telemetry dashboard for AI inference hardware with Telegram alerts
Updated
Jun 7, 2026
JavaScript
Command A Plus GGUF for llama.cpp enables 4-bit quantized 35B model on P40
Updated
Jun 16, 2026
Shell
Ultra-light Q1_S quant of GLM-5.2 MoE model optimized for Tesla P40 with llama.cpp
Updated
Jul 2, 2026
Shell
GGUF model for Cohere 30B A3B optimized for Tesla P40 deployment
Updated
Jun 11, 2026
Shell
Improve this page
Add a description, image, and links to the
p40
topic page so that developers can more easily learn about it.
Curate this topic
Add this topic to your repo
To associate your repository with the
p40
topic, visit your repo's landing page and select "manage topics."
Learn more
You can’t perform that action at this time.