r-collapse
Use when code loads or uses collapse (library(collapse), collapse::), performing fast grouped or weighted statistics in R, or seeking faster alternatives to dplyr aggregation
DeepseekModel
官方收录技能
质量 良好 · 64
v1.0.0
获取
https://deepseekmodel.com/api/download.php?id=arthurgailes-r-package-skills-skills-r-collapse-skill-md&format=skill
下载 .skill
标准格式,含 system_prompt 与 model_config,导入任意 Agent 框架即可使用
.skill 文件中 system_prompt 字段的实际内容。
name r-collapse description Use when code loads or uses collapse (library(collapse), collapse::), performing fast grouped or weighted statistics in R, or seeking faster alternatives to dplyr aggregation collapse: Fast Data Transformation Overview collapse provides C/C++-based high-performance grouped and weighted statistics. 50-100x faster than dplyr for grouped operations, matches data.table speed while working with any data frame type (tibbles, data.tables, xts). Core principle: Fast aggregation, transformation, and panel data operations through vectorized C code. References Read references/API.md before writing code. references/API.md - Complete function reference references/collapse-for-tidyverse-users.md - Migration guide and patterns references/collapse-documentation.md - Core concepts and usage references/collapse-and-sf.md - Working with spatial data references/collapse-object-handling.md - Data structure handling When to Use Use collapse when: Dataset >100k rows Weighted statistics required Panel data (between/within transformations) Time series lags/diffs/growth rates Performance bottleneck in dplyr pipeline Don't use: Small datasets (<10k rows) - dplyr is clearer Need arbitrary grouped functions (use dplyr) Working with sf (use sf and dplyr) Need reference semantics/in-place modification (use data.table) Complex joins (data.table's keyed/rolling/non-equi joins better) vs Alternatives: Scenario Use This Large grouped stats collapse Weighted computations collapse sf manipulation dplyr Reference semantics data.table Complex joins data.table Arbitrary group functions dplyr Quick Reference Task Function/Example Grouped stats fmean() , fsum() , fsd() , fmedian() Aggregation collap(df, ~ by, list(fmean, fsd)) Transform ftransform() , fmutate() Selection fselect() , fsubset() (~100x faster) Time series flag() , fdiff() , fgrowth() Panel data fwithin() , fbetween() , qsu() Grouping fgroup_by() , GRP() Core Pattern library ( collapse ) # Basic: grouped mean (50-100x faster than dplyr) data |> fgroup_by ( category ) |> fmean ( ) # Weighted aggregation data |> fgroup_by ( region ) |> fmean ( w = weight_col ) # Multiple stats at once collap ( data , ~ category , list ( fmean , fsd , fmedian ) ) # TRA transformations (key differentiator - single C pass) data |> fgroup_by ( id ) |> fmean ( TRA = "-" ) # Demean: subtract group mean data |> fgroup_by ( id ) |> fsd ( TRA = "/" ) # Scale: divide by group SD data |> fgroup_by ( id ) |> fmean ( TRA = "fill" ) # Fill: replace NA with group mean # See references/API.md for full TRA options ("-", "/", "fill", "-+", "replace") Common Mistakes Mistake Fix Using group_by() with collapse functions Use fgroup_by() or pass g = GRP(groupvar) collap() applies to ALL numeric columns Explicitly select columns before calling Expecting na.rm = FALSE default collapse defaults to na.rm = TRUE fwithin() / fbetween() collapse rows They return same # rows (centered/group means) Global options affect behavior Set arguments explicitly in package code Ignoring sort = FALSE speedup Add sort = FALSE when order doesn't matter (3x faster) Advanced See references/ for API reference, vignette content (tidyverse comparison, sf integration, object handling, development guidelines), and panel data patterns. Validator: lib/r-validators/numerical-validator.R Resources: Docs
Agent 识别该技能的关键词,点击任意一个即可复制。
该技能未提供触发词。
下载的 .skill 包内含以下字段。
| 字段 | 说明 |
|---|---|
| format | 格式标识(skill/v1) |
| skill_id | 技能唯一 ID |
| name | 技能名称 |
| version | 版本号 |
| description | 技能描述 |
| category | 所属分类(数组) |
| trigger_words | 触发词列表 |
| tags | 标签列表 |
| source | 来源标识 |
| source_url | 来源链接(本页地址) |
| exported_at | 导出时间(每次下载生成) |
| system_prompt | 系统提示词正文 |
| model_config | 模型参数:provider / model / temperature / max_tokens / top_p |
| examples | 示例 |
| install_guide | 各平台导入说明(Coze / Dify / Claude / 自定义框架) |