Skip to content
Keboola Docs

Getting Started with Keboola

Build your first working data pipeline in Keboola: load 10,000 octopus sightings, join them with SQL, deliver a spreadsheet, run it on a schedule, and ship a map app.

Somebody claims that an octopus has been recorded near almost every coastline on Earth. You have the data to check — a century of ocean records — but it is spread across four files, none of which answers alone. That is the situation this guide starts from, and swap octopuses for orders or sensors and it is the ordinary one.

The question the arc answers:

How close to your coast does an octopus live — and how deep down do they really go?

You have 10,000 recorded octopus sightings from 1900 to 2026 — 201 species, each point with coordinates and, for almost half of them, a depth. The sightings file alone can plot dots. What it cannot tell you is who and where in any human sense: it names species by a numeric ID and says nothing about oceans or depth zones, so the moment the question becomes “which octopus, which ocean, how deep”, the file stops answering. Three small lookup tables — species names, depth zones, ocean basins — hold the missing halves, and joining them is the whole trick.

In under an hour you will have a table that answers all of it, delivered to a spreadsheet, rebuilt every morning without you, and an interactive map where you type your own coordinates and get the nearest recorded octopus — a working, scheduled pipeline with an app on top, not an exercise. Most steps are one Kai prompt: copy it, watch it build, check the result; every step also has the click-through version. Everything happens in the browser; nothing needs installing.

PhaseWhat happensKeboola calls it
Loadfour CSV files, fetched from a URL, become four tablesa data source connector
TransformSQL joins them into one wide tablea transformation
Deliverthat table appears in a Google Sheeta data destination connector
Automateall of it runs daily, in order, and emails you if it breaksa flow
Answera map of every sighting, with a “how close to me?” fielda data app

The joined table is the one someone would actually read: every sighting with its species by name, the ocean basin it sits in and — where a depth was recorded — its depth zone, down to a dumbo octopus (Grimpoteuthis challengeri) recorded 4,838 meters below the surface. It is not a demo — this is the mechanism a production project uses, just smaller.

  • A project. Step 1 gets you one. The Free Plan covers this guide’s work — the whole arc uses roughly six minutes of job runtime — as long as your project has runtime minutes available.
  • Basic SQL. One SELECT with a couple of JOINs. If you have never written SQL, the queries are given in full and you can paste them — in both Snowflake and BigQuery form, since which one you need depends on your project.
  • A Google account, for the delivery step. If you would rather not connect one, skip step 4 and build step 5’s flow with two phases instead of three — loading and transforming on a schedule is a real pipeline, just one that keeps its result inside Keboola.

Nothing needs installing. Everything below happens in the browser.

The building steps — load, transform, deliver, automate, and the app — can be done by asking Kai, Keboola’s built-in assistant, or by clicking through the UI yourself. Each of those pages puts both paths side by side: steps 2 through 6 open with a Do it with Kai tab, and a Do it yourself tab next to it.

Two things stay yours either way: creating the project in step 1, since Kai works inside a project, and the Google authorization in step 4 — that consent screen is in your own account. On step 4 the shared steps sit above the tabs, and the Kai tab picks up after them.

Kai is the default path: copy the prompt, watch it build, check the result. The clicking is still written out in full for when you want to see the mechanics — where a configuration, a mapping and a phase live — and it is worth doing by hand at least once. Each Kai block ends with what to check, and what to do when Kai’s version does not match — assume you will need that at least once.

In principle the whole arc fits into one long Kai request — data, atlas, sheet, flow and app in a single prompt. This guide keeps one prompt per step so every result stays checkable and you know what exists when you want to change it later; nothing here is hard, it is just five easy things in a row.

Building is only half of what Kai is for. Step 5 closes by pointing it at the table you just built and asking it the question at the top of this page; step 6 then turns that answer into an app anyone can open — with a field for your own coordinates, which is where the question stops being rhetorical.

Kai asks before it changes anything: project-modifying actions raise an approval dialog in the chat, so expect to confirm rather than watch it run unattended. Expect one prompt per object rather than one per request — asking for step 2’s whole configuration in a single sentence still raised six: the configuration, each of its four rows, and running the job. Questions that only read do not ask, because read-only tools are allowed by default. If the clicking gets tiring, click Always allow in the dialog, or pre-approve what you trust in tool permissions.

The Kai Agent button is visible to every user on a supported stack, but the feature has to be switched on — an organization admin can do it from the chat screen or in Settings → Features, or you can ask Keboola Support. There is also a monthly message allowance, which matters on the Free Plan. See Get started with Kai for both, and use cases for what else it does.

  1. Get a Project — create or join one, learn what a project and a stack are, find your way around.
  2. Get Your Data In — pull the four sample files into Storage with a connector, and understand buckets, tables and stages. One Kai prompt, run included.
  3. Transform Data — the SQL that turns species IDs into names and places every sighting in its ocean and depth zone. One Kai prompt builds and runs it.
  4. Deliver the Answer — push the result to a Google Sheet with a data destination connector. You authorize; one Kai prompt does the rest.
  5. Run It on a Schedule — run the whole thing in order, on a schedule, with notifications. Kai wires the flow; the schedule is two clicks.
  6. Build the App — a map of every sighting with a “how close to me?” field, described to Kai in one sentence and published with one button.
  7. Where to Go Next — what to learn next based on what you actually want to do, including how to drive Keboola from an AI assistant, an IDE, or your terminal.

Read them in order. Each step ends with a link to the next, and every page states what it assumes so you can also land on one directly and catch up.

Optional side trips, once the main path makes sense. None of them are needed to finish the arc:

If you are planning a rollout, not learning the tool

Section titled “If you are planning a rollout, not learning the tool”

This guide is for one person building one pipeline. For introducing Keboola to a team — project architecture, a data model, naming conventions, governance — start with Platform Onboarding instead.

Next: Get a project →

Ask Kai

Hi, I'm Kai — Keboola's AI assistant for the docs. Ask me anything and I'll answer from the documentation and cite the pages I use.

Kai is an AI and can make mistakes. Check the sources it links.