Getting Started with Ciaren
For: anyone new to Ciaren. You get: the mental model in a few minutes, then a four-step path from installation to your first working flows.
Ciaren is a visual workflow builder for local data engineering and lightweight machine-learning work. You build flows — Ciaren's word for a saved pipeline — on a canvas, preview real data at each step, run them locally, and export readable pandas, polars, or lazy polars Python.
Alpha software
Ciaren is in early development. APIs, generated code, workflow files, and plugin contracts may change between releases. Use it for learning, experimentation, prototypes, and controlled internal workflows before relying on it for critical production jobs.
The Short Version
Ciaren sits between three familiar tools:
| If you know... | Ciaren feels like... |
|---|---|
| Spreadsheets | A repeatable, inspectable flow instead of a sequence of manual edits |
| Notebooks/scripts | A visual way to design pipelines that still exports normal Python |
| Orchestrators | A lighter local tool for building and running one flow without cluster setup |
It is not trying to hide Python. It is trying to make dataframe workflows easier to design, inspect, teach, share, and export.
What You Can Build
- InputBring data inCSV, Excel, Parquet, JSON, SQL, storage
- CleanClean and shape itnulls · types · dedupe · rename · filter
- TransformTransform itjoin · group · aggregate · pivot · window
- CleanValidate itnot-null · unique · ranges · expressions
- MLTrain or score modelssplit · train · evaluate · predict · MLflow
- OutputSend it outfile · SQL · storage · Python export
Each node has configuration, preview output, and generated code. That matters: you can explain a flow to a beginner, inspect it as an engineer, and move the result into a regular Python workflow when you need full control.
Why Ciaren Exists
Many data tools force an early tradeoff:
- spreadsheets are approachable, but hard to reproduce and review;
- notebooks are flexible, but can become fragile execution histories;
- orchestration systems are powerful, but heavy for local exploration;
- no-code tools can be fast, but often trap work inside a proprietary runtime.
Ciaren's answer is a local, plugin-first workflow model where the visual graph is the product experience and Python export is the escape hatch. You can start with the UI and still end with code.
Who It Is For
- Data analysts: clean, join, validate, and export datasets without writing every operation by hand.
- Python learners: see how visual dataframe operations become pandas and polars code.
- Data engineers: prototype repeatable flows locally, review generated code, and use the CLI/API for automation.
- ML practitioners: move from raw data to tracked lightweight ML workflows on the same canvas.
- Plugin authors: add custom nodes, connectors, and ML model types without changing core.
- Contributors: improve the editor, execution engine, transformations, examples, docs, tests, and plugin SDK.
What You Need First
Python 3.12+ (or Docker) and a browser. The Installation page lists the requirements for each install method. You do not need your own dataset: fresh installs seed a Demo project with sample datasets and working flows.
Choose Your Path
| Goal | Start here |
|---|---|
| "I just want to see it running" | Installation, install from PyPI or use Docker, then open the Demo project |
| "I want to build my first flow" | Quick Start |
| "I want to understand the UI" | Interface Tour |
| "I want a realistic example" | Sales Analysis or Data Quality Checks |
| "I care about generated code" | Engines and Python export |
| "I want to automate it" | CLI Reference, REST API, and Python SDK |
| "I want to extend it" | Plugins Overview and Build Your First Plugin |
| "I want to contribute" | CONTRIBUTING.md and Roadmap |
What Ciaren Is Not
Ciaren is designed for local, single-machine workflows. It is not currently:
- a distributed compute engine like Spark;
- a real-time streaming platform;
- a full Airflow/dbt replacement;
- a multi-user enterprise collaboration system;
- a tool for unbounded 100GB+ local datasets.
It does include a lightweight scheduler for running individual flows on a cron schedule. See Scheduling.
The Start Path

Read these in order. Each page ends with a link to the next one.
- Installation — get Ciaren running at
http://localhost:8055. - Quick Start — build, run, and export your first flow in about five minutes.
- Demo Project & Tutorials — follow four ready-made flows, from a linear cleanup to a three-input join.
- Next steps — learn the editor in the Interface Tour, follow an end-to-end example, or browse the Transformations reference.