https://api.narrative.io with your API key.
1. Create the dataset
The dataset must carry at least one identifier column that Reddit can match on. Required columns lists them.id in the response.
Required columns
Both interfaces take the same schema. The dataset needs at least one of these identifier columns:
Each identifier column is named after its Rosetta Stone attribute, and the connector matches on that column name. A column holds either a plain string or an object. In the dataset schema, declare a string column with
"type": "string", and declare an object column with "type": "object" and its properties:
2. Activate the dataset
201 with the dataset, now "status": "active". Activation locks the schema, so activate only once the shape is settled.
3. Upload and ingest your file
Each line of the file is one row that matches the schema you declared in step 1. For the columns Reddit accepts and their shapes, see Supported identifiers and Required columns. You can load a file into a dataset over the API in two ways:- Signed-URL upload. Your integration uploads one file of up to 3 GB, then asks Narrative to ingest it.
- Managed S3 bucket. You write files to an S3 bucket that Narrative manages, and Narrative ingests each batch on its own. Use a managed bucket for files larger than 3 GB, or for files that another system delivers on a schedule.
Choose a file format
The dataset’sfile_config.type sets the format of every file you load into it. Parquet and JSON Lines both hold object columns, which nest properties inside one column:
- Parquet (
parquet) is the most compatible format for connector datasets. It stores nested struct columns and their types natively, and Narrative matches columns to the schema by name at every level. - JSON Lines (
json) holds one JSON object per line. Narrative matches each nested object to the schema by name.
flat) hold only scalar columns, so they can’t carry object columns.
Upload a file with a signed URL
Request an upload URL, then send the file straight to storage:PUT.
Then ingest the file into the dataset, passing that path as source_file:
Write files to a managed S3 bucket
A managed bucket is an S3 bucket that Narrative creates for your company. You write each batch of files into its own folder under the dataset’s path in the bucket, then write an empty_NIO_COMMIT file into that batch folder. Narrative ingests every file in the batch folder when the commit file appears, so you make no upload or ingest request. A file can be as large as S3 accepts.
See Ingesting Files from a Managed S3 Bucket to create the bucket, grant your AWS account access, and lay out the folders.
4. Confirm Reddit accepts the dataset
Before you create a connection, ask Narrative which connector interfaces the dataset satisfies:acceptedlists each interface you can connect the dataset to, by the connector’sapp_idand theinterface_id.errorslists each interface the schema does not satisfy. Itsdetailshold the reason, such as"required property '<column>' not found".
?tags=<tag> to check only the interfaces that carry that tag.
When the interface you want is under errors, the dataset’s schema doesn’t meet what the interface needs, for example a missing column or property. Activation locks the schema, so create a new dataset that fixes what the error names.
For Reddit, look for the Reddit Connector ("app_id": 23) in accepted. Add ?tags=reddit to check only the Reddit interfaces. audience_new delivers to a new audience and audience_existing delivers to one you already have:
5. Create the connection
To deliver to a new Custom Audience, create a connection with theaudience_new interface. The connector creates the Custom Audience in Reddit before it answers, so the audience exists in Reddit Ads Manager once you get 201.
type fields are required. The outer one identifies what you are connecting, and the one inside quick_settings selects the delivery interface.
What the connection does describes every field, its allowed values, and its default.
Record the connection id. You need it to stop the delivery.
Replace the audience on each delivery
Setdataset_write_mode to overwrite when each delivery should replace the members from the previous delivery. Use a dataset created with "write_mode": "overwrite", so each snapshot holds the whole audience. In the request below, 12346 is an example ID for an existing dataset of that kind. Steps 1 to 3 create and load one; set write_mode to overwrite in step 1:
Remove members after a number of days
Addttl_days to an append connection so people leave the audience once your data stops including them:
Deliver to an existing audience
To add a dataset to an audience you already have in Reddit, copy the audience ID from Reddit Ads Manager. Reddit audience IDs start withca.. Then create the connection with the audience_existing interface:
6. Confirm the connection
"status": "active" and the quick_settings you sent. The second lists every Reddit connection in your company. For an audience_new connection, the audience appears in Reddit Ads Manager under the audience_name you set.
To hear when each delivery finishes, subscribe to delivery notifications. The connector sends audience.delivery.completed when a delivery to the audience finishes.
Keeping the audience current
Write new data to the dataset with the same three calls as in step 3. The connection keeps running, so new rows reach the audience without further calls.Stopping delivery
DELETE /datasets/{dataset_id}.
Troubleshooting
Getting help
Contact your Narrative relationship manager with your company ID, the dataset ID, and the failing request and response.Related content
Inviting a partner via the API
Let a partner connect their Reddit Ads account so you can deliver to it
Reddit Connector
Supported identifiers, profiles, and invites
Connector Interfaces
Why a dataset connects to an interface rather than a connector
API Keys
Create and rotate keys for programmatic access

