olOpsLyft Docs

Datasets

Bring business data into OpsLyft so cloud cost can be joined to revenue, usage, teams, and customers.

A dataset is data that lives outside your cloud bill: revenue by customer, product usage events, org hierarchy, LLM token counts, vendor invoices. Once you map its account, team, and date columns, it joins to cloud cost, so you can build unit costs (cost per customer, per request, per token) and allocate spend.

When a dataset's first sync finishes, you can pick it in the widget builder and use it in Dimension Studio.

The datasets list

The header shows the total and how many need attention. Tabs split datasets into All, Yours, and Platform (maintained by OpsLyft).

ColumnMeaning
DatasetName and identifier, such as Revenue by customer, monthly · revenue_monthly
SourceType and location, such as SQL database · postgres://db-prod.internal/finance
StatusSynced, Paused, Stale, and so on, with the time. See Dataset status.
RowsRow count from the last sync
OwnerWho's responsible for it

Datasets that need attention are listed first. Filter with the search box, Source, and Status, or select Export list.

Example datasets

DatasetSourceUsed for
Revenue by customer, monthlySQL databaseCost as a share of revenue, per customer
Customer usage eventsAmazon S3Cost per active customer or per event
Org hierarchyGoogle SheetsRolling teams up to departments for chargeback
Chargeback overridesGoogle SheetsManual allocation exceptions
LLM token usageAmazon S3Cost per token by team or feature
Region to entity mapFile uploadAllocating cost to legal entities
Vendor invoicesFile uploadAdding non-cloud vendors to the bill
Query volumePrometheusCost per query

Next steps

On this page