# US Airline Financial Performance — Quarterly, Per Record (`nexgensignal/us-airline-financial-records`) Actor

US DOT Airline Quarterly Financial Review (6mpe-e53g) as clean, per-record carrier economics - financial-statement line items by carrier, account and period. Public-domain, $0.05 per record.

- **URL**: https://apify.com/nexgensignal/us-airline-financial-records.md
- **Developed by:** [NexGen Signal](https://apify.com/nexgensignal) (community)
- **Categories:** Business, Developer tools, Other
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $33.50 / 1,000 airline financial records

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## US Airline Financial Performance — Quarterly, Per Record

Turn the U.S. DOT Airline Quarterly Financial Review into clean, per-record carrier economics — one row per financial-statement line item per carrier per period, ready to compare airline financial performance account by account.

Each source row becomes **one clean, flat record** with numeric fields coerced to real numbers, a stable
source-native `record_id`, and provenance stamped on every row: source, resource id, the licence notice,
the required attribution, a UTC retrieval timestamp and an interpretation caveat.

### What one record represents

The source is **`6mpe-e53g`** — *Airline Quarterly Financial Review (Main)* on the U.S. DOT open-data portal (data.transportation.gov). Each record is **one financial-statement line item (account) for one carrier in one period**: the reporting period, the carrier, the statement group, the account name and id, and the reported amount. Carrier-level rows and industry-group aggregate rows (e.g. 'Domestic Passenger Majors') both appear — the `group_name` tells you which.

For each record you get a composite `record_id` built from the source-native key, the analytic columns
listed below (reproduced verbatim, numbers as numbers), and the provenance block. Columns include `period_type`, `year`, `quarter`, `group_name` (statement group), `uniquecarrier`/`uniquecarriername` (the airline — named explicitly in the query because the source default view hides them), `item_id`/`item_name` (the account) and `val` (the amount).

### Coverage and volume

The live table holds **213,452** financial line-item records spanning U.S. carriers and industry groups across the quarters the DOT publishes — that is the record capacity of a full pull.

**Live count: 213,452 records — matches the Wave-3 index figure exactly.**

The Actor pages the source with keyless SODA `$query` requests ordered by the source-native key for a
stable total order, and stops as soon as your **Maximum records** cap is met. Crucially, the source default view hides `uniquecarrier`/`uniquecarriername`; this Actor issues a full SoQL `$query` that orders by carrier so every carrier-level record carries its airline (industry-aggregate rows have no single carrier by design).

### Licence and attribution

This is a **public-domain U.S. Government work** (17 U.S.C. §105) — free to use, redistribute and build on. The full notice travels on every record:

> U.S. DOT / Bureau of Transportation Statistics. Public-domain U.S. Government work (17 U.S.C. 105). Reproduced verbatim; no third-party content or logos.

The required attribution — `U.S. DOT / Bureau of Transportation Statistics` — travels on every record.

### Interpretation caveat

Each record is one financial-statement line item for one carrier (or industry group) in one period, in the source's reported units. Carriers are airline organisations, not persons. Figures are as-reported to the DOT, not audited or restated; some rows are industry-group aggregates.

Values are reproduced verbatim: the Actor never rescales, re-derives or editorialises a number.

### Person-data policy

There are no name, email, phone, personal-address or personal-identifier fields — the only name field is the carrier (airline) name, an organisation. A per-record assertion enforces the person-field allow-list at write time.

### Data quality and freshness

Numeric fields are coerced from the source's string encoding into real numbers (integers where whole,
floats otherwise); genuinely missing cells are delivered as `null`, never as zero. Text is passed
through verbatim. Every run re-reads the live source, so the data is as fresh as the portal itself, and
each record's `observed_at` stamp records exactly when the row was retrieved. Delivery order is fixed by
the source-native key, so a capped sample and a later full pull agree on their overlap and a repeated
run returns rows in the same order. The `RUN_RECEIPT` reports source rows scanned and records delivered
and charged for a per-run reconciliation.

### Provenance, licensing and compliance

Every run begins with a live source-preflight: the Actor reads the exact host's `robots.txt` at runtime
and refuses to proceed if the crawl policy disallows the data path. The gate result — URL, HTTP status,
byte length and a SHA-256 of the policy — is written to the run's `RUN_RECEIPT`, so each run carries its
own audit trail. The Actor identifies itself with a transparent, non-impersonating User-Agent and never
bypasses a block, solves a challenge, or fetches through a cache or mirror. When the door is genuinely
unavailable the run fails loudly and bills nothing.

### Inputs

- **Maximum records** (`maxRecords`) — hard cap on records delivered and billed. Raise it to pull the
  full set; lower it to sample cheaply. Records arrive in a stable, source-native order.

### Output

Records land in the Actor's default dataset and export as JSON, CSV, Excel or via the Apify API. A
tabular **overview view** surfaces the most useful columns for quick inspection while the full record
retains every selected field and provenance stamp.

### Fields in detail

The record leads with `record_id` — a stable composite key drawn from the source's own grain — followed
by the analytic columns described above and closed by a provenance block: `source`, `source_dataset`
(the Socrata resource id), `licence`, `attribution`, `caveat` and `observed_at`. Every one of those
provenance fields is present on every record, so a single row is self-describing: hand it to a colleague
or a downstream system and it carries its own origin, licence and retrieval time without reference back
to this page. Because delivery is ordered by the source-native key, the same record always carries the
same `record_id` across runs, which makes the dataset safe to diff, deduplicate, or upsert into a
warehouse. Nothing in the record is computed or inferred beyond the explicit count where one is stated —
every other value is the source's own, reproduced byte-for-byte.

### Sibling Actors

This Actor is the **economics** view of U.S. air carriers. Its siblings measure different things: **`us-carrier-traffic-performance-records`** is operational traffic (departures, passengers, load factor) and **`us-airport-pair-fare-records`** is market fares. Financials, traffic and fares — three different questions about the same industry. This Actor also shares its engineering — the runtime robots gate, push-then-charge
billing and verbatim-value discipline — with the fleet's other public-data records Actors.

### Pricing

This Actor uses Apify's pay-per-event model: a flat **$0.05 per record** actually delivered to the
dataset, and nothing else — no monthly rental, no per-run base fee, no compute charge. Deliver 40
records and you pay $2.00; deliver 10,000 and you pay $500.00. Billing is wired *after* delivery — each
record is pushed first and only then does the per-record event fire — so a mid-run failure can only
ever under-charge you, never over-charge. Use **Maximum records** to cap spend precisely.

### Scaling and limits

Set **Maximum records** low to sample the leading slice cheaply, or high to pull the full set. The Actor
paginates server-side and delivers incrementally, so memory stays flat regardless of how many records
you request, and you are billed only for what is actually delivered. Because the source is a live public
API, extremely deep pagination is ultimately bounded by the source's own paging behaviour; for the vast
majority of uses — sampling, a full refresh, or a scheduled top-up — the default paging is more than
sufficient. Schedule the Actor on Apify to keep a downstream table current: each run re-reads the live
source and re-stamps `observed_at`, so a nightly or weekly run gives you a dated, reproducible snapshot.

### Typical uses

Compare airline financial performance account by account; track operating revenue and expense by carrier over time; screen carriers by profitability; benchmark a carrier against its industry group; or feed an aviation-investor, credit or diligence model with clean financial-statement records.

### What this Actor does not do

It does not forecast or model, does not merge multiple source tables into one record, and it does not include any personal data — the only name is the airline's. It
gives you faithful, analysis-ready records — with a provenance trail you can audit on every run.

# Actor input Schema

## `maxRecords` (type: `integer`):

Maximum records delivered and billed. You are billed only for records actually delivered. Raise it to pull the full set.

## Actor input object example

```json
{
  "maxRecords": 500
}
```

# Actor output Schema

## `results` (type: `string`):

The delivered US airline financial record.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "maxRecords": 500
};

// Run the Actor and wait for it to finish
const run = await client.actor("nexgensignal/us-airline-financial-records").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "maxRecords": 500 }

# Run the Actor and wait for it to finish
run = client.actor("nexgensignal/us-airline-financial-records").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "maxRecords": 500
}' |
apify call nexgensignal/us-airline-financial-records --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,nexgensignal/us-airline-financial-records"
        }
    }
}

```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/AspDXedPTiYdtaF3n/builds/BW5vbG65pzLqxkPLe/openapi.json
