Skills Plugins MCP Prompt Model 博客 我的中心

axolotl

Axolotl: YAML LLM fine-tuning (LoRA, DPO, GRPO).

DeepseekModel 官方收录技能 质量 优秀 · 90 v1.0.0

获取

https://deepseekmodel.com/api/download.php?id=nousresearch-hermes-agent-optional-skills-mlops-training-axolotl-skill-md&format=skill
下载 .skill 标准格式,含 system_prompt 与 model_config,导入任意 Agent 框架即可使用
.skill 文件中 system_prompt 字段的实际内容。
name axolotl description Axolotl: YAML LLM fine-tuning (LoRA, DPO, GRPO). version 1.0.0 author Orchestra Research license MIT dependencies ["axolotl","torch","transformers","datasets","peft","accelerate","deepspeed"] platforms ["linux","macos"] metadata {"hermes":{"tags":["Fine-Tuning","Axolotl","LLM","LoRA","QLoRA","DPO","KTO","ORPO","GRPO","YAML","HuggingFace","DeepSpeed","Multimodal"]}} Axolotl Skill What's inside Expert guidance for fine-tuning LLMs with Axolotl — YAML configs, 100+ models, LoRA/QLoRA, DPO/KTO/ORPO/GRPO, multimodal support. Assistance with axolotl development, generated from official documentation. When to Use This Skill This skill should be triggered when: Working with axolotl Asking about axolotl features or APIs Implementing axolotl solutions Debugging axolotl code Learning axolotl best practices Quick Reference Common Patterns Pattern 1: To validate that acceptable data transfer speeds exist for your training job, running NCCL Tests can help pinpoint bottlenecks, for example: ./build/all_reduce_perf -b 8 -e 128M -f 2 -g 3 Pattern 2: Configure your model to use FSDP in the Axolotl yaml. For example: fsdp_version: 2 fsdp_config: offload_params: true state_dict_type: FULL_STATE_DICT auto_wrap_policy: TRANSFORMER_BASED_WRAP transformer_layer_cls_to_wrap: LlamaDecoderLayer reshard_after_forward: true Pattern 3: The context_parallel_size should be a divisor of the total number of GPUs. For example: context_parallel_size Pattern 4: For example: - With 8 GPUs and no sequence parallelism: 8 different batches processed per step - With 8 GPUs and context_parallel_size=4: Only 2 different batches processed per step (each split across 4 GPUs) - If your per-GPU micro_batch_size is 2, the global batch size decreases from 16 to 4 context_parallel_size=4 Pattern 5: Setting save_compressed: true in your configuration enables saving models in a compressed format, which: - Reduces disk space usage by approximately 40% - Maintains compatibility with vLLM for accelerated inference - Maintains compatibility with llmcompressor for further optimization (example: quantization) save_compressed: true Pattern 6: Note It is not necessary to place your integration in the integrations folder. It can be in any location, so long as it’s installed in a package in your python env. See this repo for an example: https://github.com/axolotl-ai-cloud/diff-transformer integrations Pattern 7: Handle both single-example and batched data. - single example: sample[‘input_ids’] is a list[int] - batched data: sample[‘input_ids’] is a list[list[int]] utils.trainer.drop_long_seq(sample, sequence_len=2048, min_sequence_len=2) Example Code Patterns Example 1 (python): cli.cloud.modal_.ModalCloud(config, app= None ) Example 2 (python): cli.cloud.modal_.run_cmd(cmd, run_folder, volumes= None ) Example 3 (python): core.trainers.base.AxolotlTrainer( *_args, bench_data_collator= None , eval_data_collator= None , dataset_tags= None , **kwargs, ) Example 4 (python): core.trainers.base.AxolotlTrainer.log(logs, start_time= None ) Example 5 (python): prompt_strategies.input_output.RawInputOutputPrompter() Reference Files This skill includes comprehensive documentation in references/ : api.md - Api documentation dataset-formats.md - Dataset-Formats documentation other.md - Other documentation Use view to read specific reference files when detailed information is needed. Working with This Skill For Beginners Start with the getting_started or tutorials reference files for foundational concepts. For Specific Features Use the appropriate category reference file (api, guides, etc.) for detailed information. For Code Examples The quick reference section above contains common patterns extracted from the official docs. Resources references/ Organized documentation extracted from official sources. These files contain: Detailed explanations Code examples with language annotations Links to original documentation Table of contents for quick navigation scripts/ Add helper scripts here for common automation tasks. assets/ Add templates, boilerplate, or example projects here. Notes This skill was automatically generated from official documentation Reference files preserve the structure and examples from source docs Code examples include language detection for better syntax highlighting Quick reference patterns are extracted from common usage examples in the docs Updating To refresh this skill with updated documentation: Re-run the scraper with the same configuration The skill will be rebuilt with the latest information
Agent 识别该技能的关键词,点击任意一个即可复制。

该技能未提供触发词。

下载的 .skill 包内含以下字段。
字段 说明
format格式标识(skill/v1)
skill_id技能唯一 ID
name技能名称
version版本号
description技能描述
category所属分类(数组)
trigger_words触发词列表
tags标签列表
source来源标识
source_url来源链接(本页地址)
exported_at导出时间(每次下载生成)
system_prompt系统提示词正文
model_config模型参数:provider / model / temperature / max_tokens / top_p
examples示例
install_guide各平台导入说明(Coze / Dify / Claude / 自定义框架)
同一份技能可按不同平台格式导出。
.skill 标准格式,含 system_prompt 与 model_config,导入任意 Agent 框架即可使用 下载
.skillpro 增强格式,额外含脚本 / 工具 / 依赖 / 钩子占位 下载
.json 纯 JSON 导出,只含 system_prompt 与模型参数 下载
Coze 带 frontmatter 的 Markdown,Coze 平台导入用 下载
Dify Dify DSL,创建应用后直接导入 下载

每日精选 Skill 推荐,免费送到你邮箱

输入邮箱,每天接收一个精选 AI Agent 技能推荐。完全免费,持续更新。

验证码 --

提交后我们会发送一封确认邮件,点击邮件里的链接才会开始收信。

完全免费,取消任意时间。我们不会发送垃圾邮件。