Vaex skill: what it does and how to install it
Processes large scientific tabular datasets with Vaex using lazy expressions, filtered views, streamed statistics, binned plots and chunked file conversion.
Summary generated from the skill's documentation.
Install
$ npx skills add K-Dense-AI/scientific-agent-skills --skill vaexRun it in a terminal. If your agent is already running, start a new session so it picks the skill up.
About this skill
What it does. Guides Vaex workflows for larger-than-RAM HDF5, Arrow, CSV and Parquet data: inspect schemas, open files, define virtual columns, filter, run delayed reductions and grouped statistics, plot aggregated grids, and export in chunks. It highlights operations that may materialise data and validates counts, units, joins and reopened outputs.
When to use it. For single-machine columnar scientific analysis, especially repeated reductions and histograms. Out-of-core processing does not make sorting, joins, large group dictionaries or many estimator fits memory-bounded.
History
Repo stars
47.7kAbout +18.9k since 9 Jul 2026
Before 1 Oct 2026 the curve is estimated from public event data.
Stars are counted for the whole repository, which holds 74 skills.
Show as a table
| Date | Repo stars |
|---|---|
| 5 Oct 2026 | 47,652 |
| 4 Oct 2026 | 47,541 |
| 3 Oct 2026 | 47,444 |
| 2 Oct 2026 | 47,351 |
| 1 Oct 2026 | 47,271 |
| 24 Sept 2026 (estimated) | 46,976 |
| 17 Sept 2026 (estimated) | 46,257 |
| 10 Sept 2026 (estimated) | 41,836 |
| 3 Sept 2026 (estimated) | 30,545 |
| 27 Aug 2026 (estimated) | 29,347 |
| 20 Aug 2026 (estimated) | 29,160 |
| 13 Aug 2026 (estimated) | 29,027 |
| 6 Aug 2026 (estimated) | 28,947 |
| 30 Jul 2026 (estimated) | 28,894 |
| 23 Jul 2026 (estimated) | 28,894 |
| 16 Jul 2026 (estimated) | 28,841 |
| 9 Jul 2026 (estimated) | 28,708 |
Installs
1.7k
Tracking since . A chart appears once there are 7 days of data.
Installs via skills.sh
Similar skills
- PolarsReference for Polars, the Python DataFrame library: expressions, lazy queries, streaming and migration from pandas.
- Datalineage BigQuery Asset Impact AnalysisMaps downstream BigQuery tables, dashboards and processes to show the impact of a broken, stale or changed asset using Google Cloud Data Lineage.
- Exploratory Data AnalysisRuns exploratory analysis on scientific data files: profiles, missing-data audits and outlier checks.