Browse Skills
841 skills across 8 categories
⚙️
1w ago
Axolotl
Provides expert assistance for configuring and running LLM fine-tuning with Axolotl, covering YAML configs, 100+ model architectures, LoRA/QLoRA, DPO/KTO/ORPO/GRPO, and multimodal support.
AI Engineering
+1%11.2K821
🧩
1w ago
SentencePiece
SentencePiece is a language-independent tokenizer that works on raw Unicode text without pre-tokenization. It supports BPE and Unigram algorithms for fast, deterministic vocabulary training, ideal for multilingual NLP models and CJK languages.
AI Engineering
+1%11.2K821
⚡
1w ago
HuggingFace Tokenizers
Rust-based tokenization library for NLP that processes 1GB of text in under 20 seconds. Supports BPE, WordPiece, and Unigram algorithms, custom vocabulary training, and seamless transformers integration for research and production.
AI Engineering
+1%11.2K821
🔥
1w ago
TorchTitan LLM Pretraining
Provides PyTorch-native distributed LLM pretraining using torchtitan with 4D parallelism (FSDP2, TP, PP, CP). Use when pretraining Llama 3.1, DeepSeek V3, or custom models at scale from 8 to 512+ GPUs with Float8, torch.compile, and distributed checkpointing.
AI Engineering
+1%11.2K821
🧠
1w ago
RWKV Architecture
A skill for using RWKV, the RNN-Transformer hybrid model that offers linear-time inference, infinite context, and constant memory. Covers installation, text generation, long context processing, fine-tuning, and comparison with Transformers.
AI Engineering
+1%11.2K821
🧠
1w ago
nanoGPT
A minimalist, educational GPT implementation in ~300 lines of code. Ideal for learning transformer architectures from scratch; supports training on Shakespeare (CPU) or reproducing GPT-2 on OpenWebText (multi-GPU).
AI Engineering
+1%11.2K821
🐍
1w ago
Mamba Architecture
Implements Mamba state-space models achieving O(n) complexity for sequence modeling. Enables 5× faster inference, million-token sequences, and no KV cache. Ideal for streaming and long-context applications.
AI Engineering
+1%11.2K821
⚡
1w ago
LitGPT LLM Implementation
Implements and trains LLMs using LitGPT with 20+ architectures like Llama, Gemma, Phi. Use for clean model implementations, understanding architectures, or production fine-tuning with LoRA/QLoRA. Single-file, no abstraction layers.
AI Engineering
+1%11.2K821
🔬
1w ago
Autoresearch
Orchestrates end-to-end autonomous AI research using a two-loop architecture for rapid experimentation and synthesis. Use for starting research projects, running autonomous experiments, or managing multi-hypothesis efforts.
AI Engineering
+1%11.2K821
🔢
2w ago
HQQ Quantization
Quantize large language models to 4/3/2-bit precision without calibration data, enabling faster quantization and deployment with vLLM or HuggingFace Transformers.
AI Engineering
10.7K796
🌐
2w ago
Web Access
Equips AI agents with full web browsing capabilities via a CDP proxy, including dynamic page interaction, media extraction, and site pattern learning.
Automation
+0%8.5K601
🚀
1w ago
Fluid Release
Automates the Fluid Framework client release process: minor and patch releases, branching, version bumps, changelogs, and type test updates. Supports interactive and autonomous modes with auto-detection of release state.
DevOps
+0%4.9K582