Labsco

Agent Skills

Instruction packs that give your AI agent know-how — some work anywhere, some only with the tool they came with.

automattic · Web Scraping

227 standalone skills
firecrawl logo

crewai-multi-agent

★ 11

by firecrawl

Multi-agent orchestration framework for autonomous AI collaboration. Use when building teams of specialized agents working together on complex tasks, when you need role-based agent collaboration with memory, or for production workflows requiring sequential/hierarchical execution. Built without LangChain dependencies for lean, fast execution.

🔥🔥🔥FreeQuick setup
firecrawl logo

constitutional-ai

★ 11

by firecrawl

Anthropic's method for training harmless AI through self-improvement. Two-phase approach - supervised learning with self-critique/revision, then RLAIF (RL from AI Feedback). Use for safety alignment, reducing harmful outputs without human labels. Powers Claude's safety system.

🔥🔥🔥✓ VerifiedFreeQuick setup
firecrawl logo

clip

★ 11

by firecrawl

OpenAI's model connecting vision and language. Enables zero-shot image classification, image-text matching, and cross-modal retrieval. Trained on 400M image-text pairs. Use for image search, content moderation, or vision-language tasks without fine-tuning. Best for general-purpose image understanding.

🔥🔥🔥FreeQuick setup
firecrawl logo

chroma

★ 11

by firecrawl

Open-source embedding database for AI applications. Store embeddings and metadata, perform vector and full-text search, filter by metadata. Simple 4-function API. Scales from notebooks to production clusters. Use for semantic search, RAG applications, or document retrieval. Best for local development and open-source projects.

🔥🔥🔥FreeQuick setup
firecrawl logo

axolotl

★ 11

by firecrawl

Expert guidance for fine-tuning LLMs with Axolotl - YAML configs, 100+ models, LoRA/QLoRA, DPO/KTO/ORPO/GRPO, multimodal support

🔥🔥FreeQuick setup
firecrawl logo

awq-quantization

★ 11

by firecrawl

Activation-aware weight quantization for 4-bit LLM compression with 3x speedup and minimal accuracy loss. Use when deploying large models (7B-70B) on limited GPU memory, when you need faster inference than GPTQ with better accuracy preservation, or for instruction-tuned and multimodal models. MLSys 2024 Best Paper Award winner.

🔥🔥🔥FreeQuick setup
firecrawl logo

autogpt-agents

★ 11

by firecrawl

Autonomous AI agent platform for building and deploying continuous agents. Use when creating visual workflow agents, deploying persistent autonomous agents, or building complex multi-step AI automation systems.

🔥🔥🔥FreeQuick setup
firecrawl logo

audiocraft-audio-generation

★ 11

by firecrawl

PyTorch library for audio generation including text-to-music (MusicGen) and text-to-sound (AudioGen). Use when you need to generate music from text descriptions, create sound effects, or perform melody-conditioned music generation.

🔥🔥🔥FreeQuick setup
firecrawl logo

moe-training

★ 11

by firecrawl

Train Mixture of Experts (MoE) models using DeepSpeed or HuggingFace. Use when training large-scale models with limited compute (5× cost reduction vs dense models), implementing sparse architectures like Mixtral 8x7B or DeepSeek-V3, or scaling model capacity without proportional compute increase. Covers MoE architectures, routing mechanisms, load balancing, expert parallelism, and inference optimization.

🔥🔥🔥FreeQuick setup
firecrawl logo

model-pruning

★ 11

by firecrawl

Reduce LLM size and accelerate inference using pruning techniques like Wanda and SparseGPT. Use when compressing models without retraining, achieving 50% sparsity with minimal accuracy loss, or enabling faster inference on hardware accelerators. Covers unstructured pruning, structured pruning, N:M sparsity, magnitude pruning, and one-shot methods.

🔥🔥🔥FreeQuick setup
firecrawl logo

nnsight-remote-interpretability

★ 11

by firecrawl

Provides guidance for interpreting and manipulating neural network internals using nnsight with optional NDIF remote execution. Use when needing to run interpretability experiments on massive models (70B+) without local GPU resources, or when working with any PyTorch architecture.

🔥🔥🔥FreeQuick setup
firecrawl logo

openrlhf-training

★ 11

by firecrawl

High-performance RLHF framework with Ray+vLLM acceleration. Use for PPO, GRPO, RLOO, DPO training of large models (7B-70B+). Built on Ray, vLLM, ZeRO-3. 2× faster than DeepSpeedChat with distributed architecture and GPU resource sharing.

🔥🔥🔥FreeQuick setup
firecrawl logo

phoenix-observability

★ 11

by firecrawl

Open-source AI observability platform for LLM tracing, evaluation, and monitoring. Use when debugging LLM applications with detailed traces, running evaluations on datasets, or monitoring production AI systems with real-time insights.

🔥🔥🔥FreeQuick setup
firecrawl logo

pinecone

★ 11

by firecrawl

Managed vector database for production AI applications. Fully managed, auto-scaling, with hybrid search (dense + sparse), metadata filtering, and namespaces. Low latency (<100ms p95). Use for production RAG, recommendation systems, or semantic search at scale. Best for serverless, managed infrastructure.

🔥🔥🔥FreeQuick setup
firecrawl logo

ml-paper-writing

★ 11

by firecrawl

Write publication-ready ML/AI papers for NeurIPS, ICML, ICLR, ACL, AAAI, COLM. Use when drafting papers from research repos, structuring arguments, verifying citations, or preparing camera-ready submissions. Includes LaTeX templates, reviewer guidelines, and citation verification workflows.

🔥🔥🔥FreeQuick setup
firecrawl logo

qdrant-vector-search

★ 11

by firecrawl

High-performance vector similarity search engine for RAG and semantic search. Use when building production RAG systems requiring fast nearest neighbor search, hybrid search with filtering, or scalable vector storage with Rust-powered performance.

🔥🔥🔥FreeQuick setup
firecrawl logo

training-llms-megatron

★ 11

by firecrawl

Trains large language models (2B-462B parameters) using NVIDIA Megatron-Core with advanced parallelism strategies. Use when training models >1B parameters, need maximum GPU efficiency (47% MFU on H100), or require tensor/pipeline/sequence/context/expert parallelism. Production-ready framework used for Nemotron, LLaMA, DeepSeek.

🔥🔥🔥FreeQuick setup
firecrawl logo

sparse-autoencoder-training

★ 11

by firecrawl

Provides guidance for training and analyzing Sparse Autoencoders (SAEs) using SAELens to decompose neural network activations into interpretable features. Use when discovering interpretable features, analyzing superposition, or studying monosemantic representations in language models.

🔥🔥🔥FreeQuick setup
firecrawl logo

speculative-decoding

★ 11

by firecrawl

Accelerate LLM inference using speculative decoding, Medusa multiple heads, and lookahead decoding techniques. Use when optimizing inference speed (1.5-3.6× speedup), reducing latency for real-time applications, or deploying models with limited compute. Covers draft models, tree-based attention, Jacobi iteration, parallel token generation, and production deployment strategies.

🔥🔥🔥FreeQuick setup
firecrawl logo

instructor

★ 11

by firecrawl

Extract structured data from LLM responses with Pydantic validation, retry failed extractions automatically, parse complex JSON with type safety, and stream partial results with Instructor - battle-tested structured output library

🔥🔥🔥FreeQuick setup
firecrawl logo

weights-and-biases

★ 11

by firecrawl

Track ML experiments with automatic logging, visualize training in real-time, optimize hyperparameters with sweeps, and manage model registry with W&B - collaborative MLOps platform

🔥🔥🔥FreeQuick setup
firecrawl logo

guidance

★ 11

by firecrawl

Control LLM output with regex and grammars, guarantee valid JSON/XML/code generation, enforce structured formats, and build multi-step workflows with Guidance - Microsoft Research's constrained generation framework

🔥🔥🔥FreeQuick setup
firecrawl logo

peft-fine-tuning

★ 11

by firecrawl

Parameter-efficient fine-tuning for LLMs using LoRA, QLoRA, and 25+ methods. Use when fine-tuning large models (7B-70B) with limited GPU memory, when you need to train <1% of parameters with minimal accuracy loss, or for multi-adapter serving. HuggingFace's official library integrated with transformers ecosystem.

🔥🔥🔥FreeQuick setup
firecrawl logo

model-merging

★ 11

by firecrawl

Merge multiple fine-tuned models using mergekit to combine capabilities without retraining. Use when creating specialized models by blending domain-specific expertise (math + coding + chat), improving performance beyond single models, or experimenting rapidly with model variants. Covers SLERP, TIES-Merging, DARE, Task Arithmetic, linear merging, and production deployment strategies.

🔥🔥🔥FreeQuick setup
firecrawl logo

nemo-evaluator-sdk

★ 11

by firecrawl

Evaluates LLMs across 100+ benchmarks from 18+ harnesses (MMLU, HumanEval, GSM8K, safety, VLM) with multi-backend execution. Use when needing scalable evaluation on local Docker, Slurm HPC, or cloud platforms. NVIDIA's enterprise-grade platform with container-first architecture for reproducible benchmarking.

🔥🔥🔥FreeQuick setup
firecrawl logo

tensorrt-llm

★ 11

by firecrawl

Optimizes LLM inference with NVIDIA TensorRT for maximum throughput and lowest latency. Use for production deployment on NVIDIA GPUs (A100/H100), when you need 10-100x faster inference than PyTorch, or for serving models with quantization (FP8/INT4), in-flight batching, and multi-GPU scaling.

🔥🔥🔥✓ VerifiedFreeNeeds API keys
firecrawl logo

long-context

★ 11

by firecrawl

Extend context windows of transformer models using RoPE, YaRN, ALiBi, and position interpolation techniques. Use when processing long documents (32k-128k+ tokens), extending pre-trained models beyond original context limits, or implementing efficient positional encodings. Covers rotary embeddings, attention biases, interpolation methods, and extrapolation strategies for LLMs.

🔥🔥🔥FreeQuick setup
firecrawl logo

optimizing-attention-flash

★ 11

by firecrawl

Optimizes transformer attention with Flash Attention for 2-4x speedup and 10-20x memory reduction. Use when training/running transformers with long sequences (>512 tokens), encountering GPU memory issues with attention, or need faster inference. Supports PyTorch native SDPA, flash-attn library, H100 FP8, and sliding window attention.

🔥🔥🔥FreeQuick setup
firecrawl logo

outlines

★ 11

by firecrawl

Guarantee valid JSON/XML/code structure during generation, use Pydantic models for type-safe outputs, support local models (Transformers, vLLM), and maximize inference speed with Outlines - dottxt.ai's structured generation library

🔥🔥🔥FreeQuick setup
firecrawl logo

pyvene-interventions

★ 11

by firecrawl

Provides guidance for performing causal interventions on PyTorch models using pyvene's declarative intervention framework. Use when conducting causal tracing, activation patching, interchange intervention training, or testing causal hypotheses about model behavior.

🔥🔥🔥FreeQuick setup
← PrevPage 7 of 8Next →