Skip to content
Repo

SDK reference

Types and functions in the modal-dojo Python SDK.

NameDescription
DeepSeek_V4_1_FlashDeepSeek-V4.1-Flash sparse-attention MoE model, 40 layers and 384 routed experts.
Gemma4_26B_A4BGoogle Gemma-4-26B-A4B-it multimodal MoE model with 26B total and 4B active parameters.
GLM_4_7Zhipu AI GLM-4.7 MoE model with 355B total and 32B active parameters.
HFModelConfigurationDownloads Hugging Face model weights with snapshot_download.
Inkling_SmallThinking Machines Lab Inkling-Small MoE model with 276B total and 12B active parameters.
Inkling_Small_LoRASelects Inkling_Small_LoRA_Recipe.
Kimi_K3Moonshot Kimi-K3 hybrid KDA/MLA MoE model, 93 layers and 896 routed experts.
ModelArchitectureMegatron transformer architecture parameters.
ModelConfigDefines model identity, weight download, and response parsing.
Moonlight_16B_A3B_InstructMoonshot AI Moonlight model with 16B total and 3B active parameters.
ParsedResponseStructured result of parsing raw model output.
Qwen3_0_6BAlibaba Qwen3-0.6B model.
Qwen3_1_7BAlibaba Qwen3-1.7B model.
Qwen3_30BAlibaba Qwen3-30B-A3B MoE model with 30B total and 3B active parameters.
Qwen3_4BAlibaba Qwen3-4B model.
Qwen3_5_0_8BAlibaba Qwen3.5-0.8B model.
Qwen3_5_2BAlibaba Qwen3.5-2B model.
Qwen3_5_4BAlibaba Qwen3.5-4B model.
Qwen3_5_9BAlibaba Qwen3.5-9B model.
Qwen3_6_27BQwen3.6-27B dense hybrid Gated DeltaNet/attention model.
Qwen3_6_35BAlibaba Qwen3.6-35B-A3B model.
Qwen3_8_27BAlibaba Qwen3.8-27B model.
Qwen3_8BAlibaba Qwen3-8B model.
Qwen3_ASR_1_7BAlibaba Qwen3-ASR-1.7B speech recognition model.
Qwen3_VL_8BAlibaba Qwen3-VL-8B-Instruct model.
ToolCallTool invocation parsed from model output.
NameDescription
DatasetConfigDataset fields and materialization behavior shared across training frameworks.
HarborDatasetA dataset loaded from Harbor tasks.
HuggingFaceDatasetA dataset loaded from a Hugging Face datasets repository.
MultimodalDatasetDataset of text prompts paired with image, audio, or video data.
OnlineRolloutPlaceholder rows that size a live generate batch.
NameDescription
DeepSeek_V4_1_Flash_RecipeDeepSeek-V4.1-Flash GRPO recipe for 8 nodes with 8 H200 GPUs each.
Gemma4_26B_A4B_RecipeGemma-4-26B-A4B recipe.
GLM_4_7_RecipeGLM-4.7 recipe.
Inkling_Small_LoRA_RecipeInkling-Small rank-32 LoRA recipe.
Inkling_Small_RecipeInkling-Small full-parameter recipe.
Kimi_K3_LoRA_RecipeKimi-K3 rank-32 LoRA recipe for 8 nodes with 8 B300 GPUs each.
MilesRecipeMiles training and Modal resource settings.
Moonlight_16B_A3B_RecipeMoonlight-16B-A3B recipe.
Qwen3_0_6B_RecipeQwen3-0.6B recipe.
Qwen3_1_7B_RecipeQwen3-1.7B recipe.
Qwen3_4B_RecipeQwen3-4B recipe.
Qwen3_5_0_8B_RecipeQwen3.5-0.8B recipe.
Qwen3_5_2B_RecipeQwen3.5-2B recipe.
Qwen3_5_4B_Miles_RecipeQwen3.5-4B recipe.
Qwen3_5_4B_RecipeQwen3.5-4B recipe.
Qwen3_5_9B_RecipeQwen3.5-9B recipe.
Qwen3_6_27B_RecipeQwen3.6-27B recipe.
Qwen3_6_35B_RecipeQwen3.6-35B-A3B recipe.
Qwen3_8_27B_RecipeQwen3.8-27B recipe.
Qwen3_8B_RecipeQwen3-8B recipe.
Qwen3_ASR_1_7B_RecipeQwen3-ASR-1.7B recipe.
Qwen3_VL_8B_RecipeQwen3-VL-8B recipe.
SlimeRecipeSlime training and Modal resource settings.
NameDescription
CheckpointA complete training checkpoint discovered on a Modal Volume.
CheckpointTypeWhether a checkpoint is Hugging Face or Megatron weights.
DashboardComponentSupported run-scoped dashboard component slots.
DashboardMetricConfigLog framework metrics to the Modal Dojo dashboard only.
DojoConfigErrorRaised when a training or deploy config is invalid.
DojoErrorBase error for Modal Dojo.
GpuAllocationErrorRaised when a recipe’s cluster or parallelism settings are invalid.
MetricConfigDefines metric tracker metadata, environment variables, and links.
ModalCaptureErrorRaised when a cloudpickled user callback captures live Modal state.
SampleA prompt, response, parsed structure, score, and metadata from one model call.
TrackioConfigTrackio logging configuration shared across all frameworks.
TrainConfigA dataset, model, and recipe for one training run.
TrainingGroupA parameter sweep over a base TrainConfig.
TrainingRunA launched training run that can be inspected, awaited, or loaded by ID.
WandbConfigWeights & Biases run metadata and credentials.
convert_megatron_checkpoint_to_hfConvert a Megatron checkpoint to Hugging Face format.
extract_codeExtract Python code from an LLM response.
score_in_sandboxRun code against test cases in a Modal sandbox.
NameDescription
CustomDeploymentA model deployed with an SGLang or vLLM recipe.
EndpointControls a Modal Endpoint that persists until stopped.
SandboxControls a Modal Sandbox in the app_name app.
SandboxResultOutcome of a single Sandbox.run invocation.
SglangRecipeSGLang server settings.
VllmRecipevLLM server settings.