LabscoConnect MCP ↗
REPO PACKAGEOFFICIAL

astronomer agents

astronomer · publisher449 repository starsApache-2.0github.com/astronomer/agents

LABSCO SUMMARY

Summary

A 27-skill toolkit from Astronomer covering Airflow DAG development, dbt integration, data lineage, and warehouse queries — most of it needs a running Airflow instance or a warehouse connection before it does anything.

Twenty of the 27 skills are grouped in the README into five areas: data discovery and analysis (warehouse-init, analyzing-data, checking-freshness, profiling-tables), data lineage (the tracing skills plus custom OpenLineage extractors), DAG development (the airflow entrypoint plus authoring, testing, debugging, deploying, blueprint templates, and human-in-the-loop workflows), dbt integration through Astronomer's own Cosmos, and Airflow 2-to-3 migration. The other seven — including dag-factory, an Airflow-plugin builder, and delegating-to-otto, which hands a task to Astronomer's own hosted Otto agent — aren't in that table at all, which says the README has fallen behind the skill list rather than the reverse.

This is not something to sample for general coding help: every skill assumes a project already built on Apache Airflow, open-source or Astronomer's managed Astro, and the warehouse-focused skills need a configured connection to Snowflake, Postgres, BigQuery, or another supported source before they answer a single question. If your team already runs Airflow, this saves re-explaining DAG conventions and warehouse schema on every task; if it doesn't, none of it applies.

READ THE FULL ANALYSIS

What each skill needs before it runs. 18 of the 27 need a local tool already on your machine — the af CLI, a Jupyter kernel for warehouse queries, or the Astro CLI itself — 6 need an account key such as warehouse or Airflow API credentials, only 2 are ready with nothing configured, and one, dag-factory, we could not verify independently of the others.

The MCP server and CLI aren't part of the skill count above. astro-airflow-mcp, the Airflow REST API server, and its bundled af command-line tool ship in the same repository but install separately from the skills; the af CLI also collects anonymous command-name telemetry by default, opt out with af telemetry disable.

ALSO IN THIS PACKAGE

WHAT'S INSIDE

Skills

35 showing · 35 total
Ready to use

Nothing else to set up — install it and go.

GROUPS 6

35 of 35

SETUP

Showing all 35 skills

airflow-adapter

Airflow 2 and Airflow 3 answer requests differently, and this is the translation layer inside this codebase that hides the difference and picks the right one by itself.

NO EXTRA SETUP

Code you hand to Airflow

8 skills

airflow-hitl

Puts a human decision in the middle of an automated pipeline: it stops and waits for someone to approve, pick an option, or fill in a form before carrying on.

Code you hand to AirflowACCOUNT / KEY

airflow-plugins

Extends the Airflow dashboard itself — your own screens and your own code run inside it, instead of in a separate app you have to deploy and keep running.

Code you hand to AirflowLOCAL FILES

airflow-state-store

Lets an Apache Airflow pipeline remember where it got to, so a step that fails and retries carries on from its last checkpoint instead of starting over or launching the same outside job twice.

Code you hand to AirflowLOCAL FILES

authoring-dags

Writes and extends Airflow data pipelines in the shape this project already uses, and catches import errors before you run them.

Code you hand to AirflowLOCAL FILES

blueprint

An engineer defines the building blocks once, in Python; from then on a pipeline is described in a short settings file rather than written as code.

Code you hand to AirflowLOCAL FILES

cosmos-dbt-core

Takes a dbt project — a set of SQL transformations — and runs it on Airflow's schedule, with each transformation appearing as its own step you can watch and retry.

Code you hand to AirflowACCOUNT / KEY

cosmos-dbt-fusion

Running a dbt project on Airflow when that project uses dbt's newer Fusion engine, which has tighter rules than dbt Core about where and how it can run.

Code you hand to AirflowLOCAL FILES

dag-factory

Describes a whole Airflow pipeline in a configuration file rather than in Python — every step, its settings and what runs after what, listed in one file.

Code you hand to AirflowLOCAL FILES

Getting it running, and finding out why it stopped

9 skills

airflow

The command line for Apache Airflow, the system that runs a company's scheduled data pipelines — list what pipelines exist, start one, or find out why last night's run failed.

Getting it running, and finding out why it stoppedACCOUNT / KEY

debugging-dags

Works out why a data pipeline failed and writes it up: what actually broke, what it held up, the fix to run now, and how to stop it happening again.

Getting it running, and finding out why it stoppedLOCAL FILES

delegating-to-otto

Hands a job off to Otto, Astronomer's own Airflow specialist, which knows the version-by-version upgrade history that a general-purpose assistant would have to guess at.

Getting it running, and finding out why it stoppedACCOUNT / KEY

deploying-airflow

Gets your Airflow pipelines off your laptop and running for real — on Astronomer's managed service, or on your own servers with Docker or Kubernetes.

Getting it running, and finding out why it stoppedLOCAL FILES

managing-astro-deployments

Astronomer runs your data pipelines on its own servers; this is how you set those environments up from a terminal and push your code to them.

Getting it running, and finding out why it stoppedACCOUNT / KEY

managing-astro-local-env

Airflow is the software that runs data pipelines on a schedule, and this is the part that keeps a working copy of it on your own laptop instead of a server.

Getting it running, and finding out why it stoppedLOCAL FILES

setting-up-astro-project

Creates the starting folder for a new data-pipeline project and fills in what it needs to run — which packages to install, which databases to connect to.

Getting it running, and finding out why it stoppedLOCAL FILES

testing-dags

The run-it-and-see loop for a data pipeline: start it, and if it fails, work out what broke, fix that, and start it again until it passes.

Getting it running, and finding out why it stoppedLOCAL FILES

troubleshooting-astro-deployments

When pipelines are failing on your Astronomer deployment and nobody knows why, this works through its logs and settings in the order that usually finds the cause.

Getting it running, and finding out why it stoppedACCOUNT / KEY

Questions about the data itself

4 skills

analyzing-data

Ask a question about the company's data in plain English and get the answer back — it works out which tables hold it and runs the query for you.

Questions about the data itselfLOCAL FILES

checking-freshness

Tells you whether the data in a table is current enough to trust right now, or whether something has stopped refreshing it.

Questions about the data itselfNO EXTRA SETUP

profiling-tables

Someone hands you a database table you have never seen before, and this writes up what is in it — every column, how many rows, how fresh it is, and what looks wrong with the data.

Questions about the data itselfNO EXTRA SETUP

warehouse-init

Builds a one-page map of your company's data warehouse — which table actually holds the customers, which one holds the orders — so nobody has to go hunting for the right table again.

Questions about the data itselfLOCAL FILES

Steps written in Go or Java

6 skills

authoring-go-sdk-tasks

Writes the Go code behind an Apache Airflow pipeline step, compiled into one native program, while the pipeline itself stays defined in Python. The Go SDK is experimental and not yet production-ready.

Steps written in Go or JavaLOCAL FILES

authoring-java-sdk-tasks

Writes the Java code behind an Apache Airflow pipeline step, with each run starting a short-lived Java process, while the pipeline itself is still defined in Python. The Java SDK is in preview.

Steps written in Go or JavaLOCAL FILES

authoring-language-sdk-tasks

The groundwork shared by every Airflow language SDK: the pipeline's schedule, order and retries stay in Python, while chosen steps run code written in another language such as Java or Go.

Steps written in Go or JavaLOCAL FILES

configuring-airflow-language-sdks

Sets up Airflow so that steps written in Java or Go reach the right runner, using two settings: one that names each runner and one that sends a queue to it.

Steps written in Go or JavaLOCAL FILES

deploying-go-sdk-bundles

Builds a Go-written Airflow step into one self-contained program, puts it where Airflow looks for it, and ships it on Docker, Kubernetes or Astro.

Steps written in Go or JavaLOCAL FILES

deploying-java-sdk-bundles

Packages the Java code behind Airflow steps into JAR files with Gradle or Maven and gets them onto the workers, on plain Docker or Kubernetes or on Astro.

Steps written in Go or JavaLOCAL FILES

Surviving a version change, or a move

3 skills

migrating-ai-sdk-to-common-ai

An old add-on for calling AI models inside data pipelines has an official replacement now, and this does the swap across a whole project.

Surviving a version change, or a moveACCOUNT / KEY

migrating-airflow-2-to-3

Moving to Airflow 3 is not just a version bump — this is the code side of it: everything in your own pipelines that version 3 no longer accepts, found and fixed.

Surviving a version change, or a moveLOCAL FILES

migrating-dagster-to-airflow

Moves a data-pipeline project from Dagster to Airflow 3 on Astronomer's Astro one area at a time, with Dagster kept in charge until the new version matches it, and writes down plainly whatever is lost on the way.

Surviving a version change, or a moveACCOUNT / KEY

Where the data came from

4 skills

annotating-task-lineage

Labels each pipeline step with the tables and files it reads and writes, so Airflow can draw a map of where the data came from even when a step reports nothing itself.

Where the data came fromLOCAL FILES

creating-openlineage-extractors

Teaches a pipeline step to report its own inputs and outputs in code, so a step from someone else's library still shows up on the data-flow map.

Where the data came fromLOCAL FILES

tracing-downstream-lineage

Answers the question you ask before changing a table or a pipeline — who else is reading it, and how much damage a change would do.

Where the data came fromLOCAL FILES

tracing-upstream-lineage

Follows a table back to where its data actually came from — which pipeline wrote it, and which system or file that pipeline read.

Where the data came fromLOCAL FILES

HOW TO GET IT

Install

SINGLE SKILL

Install a skill by name

npx skills add astronomer/agents --skill <name> --full-depth

Pick the skill name from the Skills tab — each entry there installs independently.

PLUGIN MARKETPLACE

Add the marketplace, then install the plugin

/plugin marketplace add astronomer/agents
/plugin install astronomer-data@astronomer

Typed inside the agent's own prompt, not in a terminal. The marketplace is called astronomer, which is the part after the @.