Configure how a model loads into memory in Aspose.LLM for .NET — GPU layer offload, tensor split across GPUs, memory mapping, memory locking, tensor validation, and metadata overrides....The model is not usable for generation in this state — it is a tokenizer-only...boolean flags String StringValue general.architecture , llama.rope.scaling...