|
Mila
Deep Neural Network Library
|
Files | |
| GemmaModel.ixx | |
| Gemma 4 inference model. | |
| GemmaModelConfig.ixx | |
| Deployment configuration for Gemma language models. | |
| GptModel.ixx | |
| GPT inference model. | |
| GptModelConfig.ixx | |
| Deployment configuration for Gpt2 language models. | |
| LlamaModel.ixx | |
| LLaMA inference model. | |
| LlamaModelConfig.ixx | |
| Deployment configuration for Llama language models. | |
| QuantizationDispatch.ixx | |
| The one place a runtime quantization setting becomes a compile-time policy. | |