Cloudflare R2
Synopsis
Creates a target that writes log messages to Cloudflare R2 buckets with support for various file formats and authentication methods. The target handles large file uploads efficiently with configurable rotation based on size or event count. Cloudflare R2 provides zero egress fees and S3-compatible object storage.
Schema
- name: <string>
description: <string>
type: cloudflarer2
pipelines: <pipeline[]>
status: <boolean>
properties:
key: <string>
secret: <string>
session: <string>
region: <string>
endpoint: <string>
part_size: <numeric>
bucket: <string>
buckets:
- bucket: <string>
name: <string>
format: <string>
compression: <string>
extension: <string>
schema: <string>
name: <string>
format: <string>
compression: <string>
extension: <string>
schema: <string>
max_size: <numeric>
batch_size: <numeric>
timeout: <numeric>
field_format: <string>
debug:
status: <boolean>
dont_send_logs: <boolean>
Configuration
The following fields are used to define the target:
| Field | Required | Default | Description |
|---|---|---|---|
name | Y | Target name | |
description | N | - | Optional description |
type | Y | Must be cloudflarer2 | |
pipelines | N | - | Optional post-processor pipelines |
status | N | true | Enable/disable the target |
Cloudflare R2 Credentials
| Field | Required | Default | Description |
|---|---|---|---|
key | Y | - | Cloudflare R2 access key ID |
secret | Y | - | Cloudflare R2 secret access key |
session | N | - | STS temporary session token (for short-lived credentials) |
region | N | - | R2 region. Left unset, no region is passed and the AWS SDK resolves one from the environment; R2 expects auto, so set it explicitly |
endpoint | Y | - | R2 endpoint URL (format: https://<account-id>.r2.cloudflarestorage.com) |
Connection
| Field | Required | Default | Description |
|---|---|---|---|
part_size | N | 5 | Multipart upload part size in megabytes (minimum 5MB) |
timeout | N | 30 | Connection timeout in seconds |
field_format | N | - | Data normalization format. See applicable Normalization section |
Files
| Field | Required | Default | Description |
|---|---|---|---|
bucket | N* | - | Default R2 bucket name (used if buckets not specified) |
buckets | N* | - | Array of bucket configurations for file distribution |
buckets.bucket | Y | - | R2 bucket name |
buckets.name | Y | - | File name template |
buckets.format | N | "json" | Output format: json, multijson, avro, parquet |
buckets.compression | N | - | Compression algorithm. See the Compression section below |
buckets.extension | N | Matches format | File extension override |
buckets.schema | N* | - | Schema reference (required for Avro and Parquet formats) |
name | N | "vmetric.{{.Timestamp}}.{{.Extension}}" | Default file name template when buckets not used |
format | N | "json" | Default output format when buckets not used |
compression | N | zstd | Default compression when buckets not used |
extension | N | Matches format | Default file extension when buckets not used |
schema | N | - | Default schema reference when buckets not used |
max_size | N | 33554432 | Maximum file size in bytes before rotation (default 32MB) |
batch_size | N | 100000 | Maximum number of messages per file |
* = Either bucket or buckets must be specified. When using buckets, schema is conditionally required for Avro and Parquet formats.
When max_size is reached, the current file is uploaded to R2 and a new file is created. Setting max_size to 0 does not disable rotation — an explicit 0 is replaced by the 32 MB default (MustInt64), so there is no unlimited setting. Raise the value instead.
Scheduling
See Scheduling and Pool Behavior for interval and cron fields shared by all targets.
Debug Options
| Field | Required | Default | Description |
|---|---|---|---|
debug.status | N | false | Enable debug logging |
debug.dont_send_logs | N | false | Process logs but don't send to target (testing) |
Details
The Cloudflare R2 target uploads each rotated file to the bucket it is routed to, in any of the supported file formats, and R2 charges no egress fees. R2 is Cloudflare's object storage service designed for high-performance data storage with global accessibility.
Authentication
Requires R2 access credentials obtained from the Cloudflare dashboard. Access keys are scoped to specific accounts and can be restricted to individual buckets for enhanced security.
Endpoint Configuration
The endpoint URL follows the pattern https://<account-id>.r2.cloudflarestorage.com where <account-id> is your Cloudflare account identifier found in the R2 dashboard.
File Formats
| Format | Description |
|---|---|
json | Each log entry is written as a separate JSON line (JSONL format) |
multijson | All log entries are written as a single JSON array |
avro | Apache Avro format with schema |
parquet | Apache Parquet columnar format with schema |
Compression
Some formats support built-in compression to reduce storage costs and transfer times. When supported, compression is applied at the file/block level before upload.
| Format | Default | Compression Codecs |
|---|---|---|
| JSON | - | Not supported |
| MultiJSON | - | Not supported |
| Avro | zstd | deflate, snappy, zstd |
| Parquet | zstd | gzip, snappy, zstd, brotli, lz4 |
File Management
Files are rotated based on size (max_size parameter) or event count (batch_size parameter), whichever limit is reached first. Template variables in file names enable dynamic file naming for time-based partitioning.
Bucket Routing
The target supports flexible bucket routing through pipeline configuration or explicit bucket settings:
Configuration-based routing: Define multiple buckets in the target configuration, each with its own format, compression, and schema settings. Logs are routed to specific buckets based on configuration.
Pipeline-based routing: Use the bucket field in pipeline processors to dynamically route logs to different buckets at runtime. This enables conditional routing based on log content, source, or other attributes.
Catch-all routing: When a log doesn't match any specific bucket configuration or when no bucket field is set in the pipeline, logs are routed to the catch-all bucket (configured via the bucket field in target properties).
Routing priority:
- Pipeline
bucketfield (highest priority) - Configured buckets in
bucketsarray (if bucket name matches) - Default
bucketfield (catch-all, lowest priority)
This multi-level routing enables flexible data distribution strategies, such as routing different log types to different buckets based on content analysis, source system, severity level, or any other runtime decision.
Templates
The following template variables can be used in file names:
| Variable | Description | Example |
|---|---|---|
{{.Year}} | Current year | 2024 |
{{.Month}} | Current month | 01 |
{{.Day}} | Current day | 15 |
{{.Timestamp}} | Current timestamp in nanoseconds | 1703688533123456789 |
{{.Format}} | File format | json |
{{.Extension}} | File extension | json |
{{.Compression}} | Compression type | zstd |
{{.TargetName}} | Target name | my_logs |
{{.TargetType}} | Target type | cloudflarer2 |
{{.Table}} | Bucket name | logs |
{{.Thread}} | Writer thread index. The sender also appends this automatically when two threads would otherwise produce the same path, so an explicit token is only needed to control WHERE it lands | 3 |
{{.ServiceRoot}} | The service root directory | /opt/vmetric |
Multipart Upload
Large files automatically use multipart upload protocol with configurable part size (part_size parameter). Default 5MB part size balances upload efficiency and memory usage.
Multiple Buckets
Single target can write to multiple R2 buckets with different configurations, enabling data distribution strategies (e.g., raw data to one bucket, processed data to another).
Schema Requirements
Avro and Parquet formats require a schema. The schema value can be a Library schema name, a built-in schema name, or an inline JSON definition. Parquet also accepts a schema file deployed under the schemas directory. Avro has no file lookup, so an Avro schema must be a name or inline JSON. See Avro and Parquet for the JSON definition format.
Examples
Basic Configuration
The minimum configuration for a JSON R2 target:
targets:
- name: basic_r2
type: cloudflarer2
properties:
key: "4f3e2a1b0c9d8e7f6a5b4c3d2e1f0a9b"
secret: "9b8a7c6d5e4f3a2b1c0d9e8f7a6b5c4d3e2f1a0b"
endpoint: "https://abc123def456.r2.cloudflarestorage.com"
region: "auto"
bucket: "datastream-logs"
Multiple Buckets
Configuration for distributing data across multiple R2 buckets with different formats:
targets:
- name: multi_bucket_export
type: cloudflarer2
properties:
key: "4f3e2a1b0c9d8e7f6a5b4c3d2e1f0a9b"
secret: "9b8a7c6d5e4f3a2b1c0d9e8f7a6b5c4d3e2f1a0b"
endpoint: "https://abc123def456.r2.cloudflarestorage.com"
region: "auto"
buckets:
- bucket: "raw-data-archive"
name: "raw-{{.Year}}-{{.Month}}-{{.Day}}-{{.Timestamp}}.json"
format: "multijson"
compression: "gzip"
- bucket: "analytics-data"
name: "analytics-{{.Year}}/{{.Month}}/{{.Day}}/data_{{.Timestamp}}.parquet"
format: "parquet"
schema: "<schema definition>"
compression: "snappy"
Multiple Buckets with Catch-All
Configuration for routing different log types to specific buckets with a catch-all for unmatched logs:
targets:
- name: multi_bucket_routing
type: cloudflarer2
properties:
key: "4f3e2a1b0c9d8e7f6a5b4c3d2e1f0a9b"
secret: "9b8a7c6d5e4f3a2b1c0d9e8f7a6b5c4d3e2f1a0b"
endpoint: "https://abc123def456.r2.cloudflarestorage.com"
region: "auto"
buckets:
- bucket: "security-logs"
name: "security-{{.Year}}-{{.Month}}-{{.Day}}-{{.Timestamp}}.json"
format: "json"
- bucket: "application-logs"
name: "app-{{.Year}}-{{.Month}}-{{.Day}}-{{.Timestamp}}.json"
format: "json"
bucket: "general-logs"
name: "general-{{.Timestamp}}.json"
format: "json"
Parquet Format
Configuration for daily partitioned Parquet files:
targets:
- name: parquet_analytics
type: cloudflarer2
properties:
key: "4f3e2a1b0c9d8e7f6a5b4c3d2e1f0a9b"
secret: "9b8a7c6d5e4f3a2b1c0d9e8f7a6b5c4d3e2f1a0b"
endpoint: "https://abc123def456.r2.cloudflarestorage.com"
region: "auto"
bucket: "analytics-lake"
name: "events/year={{.Year}}/month={{.Month}}/day={{.Day}}/part-{{.Timestamp}}.parquet"
format: "parquet"
schema: "<schema definition>"
compression: "snappy"
max_size: 536870912
High Reliability
Configuration with enhanced settings:
targets:
- name: reliable_r2
type: cloudflarer2
pipelines:
- checkpoint
properties:
key: "4f3e2a1b0c9d8e7f6a5b4c3d2e1f0a9b"
secret: "9b8a7c6d5e4f3a2b1c0d9e8f7a6b5c4d3e2f1a0b"
endpoint: "https://abc123def456.r2.cloudflarestorage.com"
region: "auto"
bucket: "critical-logs"
name: "logs-{{.Timestamp}}.json"
format: "json"
timeout: 60
part_size: 10
With Field Normalization
Using field normalization for standard format:
targets:
- name: normalized_r2
type: cloudflarer2
properties:
key: "4f3e2a1b0c9d8e7f6a5b4c3d2e1f0a9b"
secret: "9b8a7c6d5e4f3a2b1c0d9e8f7a6b5c4d3e2f1a0b"
endpoint: "https://abc123def456.r2.cloudflarestorage.com"
region: "auto"
bucket: "normalized-logs"
name: "logs-{{.Timestamp}}.json"
format: "json"
field_format: "cim"
Debug Configuration
Configuration with debugging enabled:
targets:
- name: debug_r2
type: cloudflarer2
properties:
key: "4f3e2a1b0c9d8e7f6a5b4c3d2e1f0a9b"
secret: "9b8a7c6d5e4f3a2b1c0d9e8f7a6b5c4d3e2f1a0b"
endpoint: "https://abc123def456.r2.cloudflarestorage.com"
region: "auto"
bucket: "test-logs"
name: "test-{{.Timestamp}}.json"
format: "json"
debug:
status: true
dont_send_logs: true
Troubleshooting
The cloudflarer2 target behaves the same way as the Amazon S3 target and reports the same errors for the same causes: refused credentials, a bucket that does not exist, region and endpoint mistakes, upload timeouts, throttling, and buckets that stay empty. Use the Troubleshooting section of Amazon S3 for the full list of errors, causes, and fixes. See Target Delivery Errors for how Director logs and retries target failures.
Log lines and the connection status carry this target's name, so match on the cause text, which is the part after Reason: or after the last colon, rather than on the target name shown in the examples there.
What differs for Cloudflare R2
-
Startup probes the account, not one bucket. Before anything is sent, the target makes one list-buckets call, so the credentials need an account-wide list permission. An R2 API token scoped to a single bucket typically fails that probe, even though the same token can write objects into that bucket. You get
api error AccessDeniedafterListBuckets, and nothing is sent until you issue an account-scoped token and use its access key ID and secret access key. -
keyandsecretmust both be set. They are used only when both resolve to a non-empty value. There is no instance-role or ambient credential fallback that can work here, so an empty value leaves the target with no credentials at all and it reportsno EC2 IMDS role foundinstead. A${VAR}or$secret{...}reference that resolves to an empty string fails the same way while the configuration still looks complete. -
regionmust beauto, andendpointmust be the full URL. R2 has no AWS-style regions, but a region is still used to sign the request. Left unset, the target reportsA region must be set when sending requests to S3.Setendpointtohttps://<account-id>.r2.cloudflarestorage.com, with the scheme included, using the account identifier shown for R2 in the Cloudflare dashboard. -
Name resolution and certificates cause most connection failures. Read the entry on the primary page that covers
no such host,connection refused, andcertificate signed by unknown authority. This target reads no TLS options, so a certificate issued by a private certificate authority, such as the one an inspecting proxy presents, has to be installed in the Director host's own trust store. Leaveuse_path_styleunset so the custom endpoint keeps path-style addressing.
No bucket name is checked at startup. The probe only confirms that the credentials can list the account, so a misspelled name under bucket or buckets passes startup and fails on the first upload routed to it.
If a bucket holds only the last batch, check that every name template contains {{.Timestamp}}. A name built only from {{.Year}}, {{.Month}}, and {{.Day}} produces the same object key on every flush, and each upload replaces the one before it.
Nothing is lost while any of these lasts. Incoming data stays queued and the target retries until you fix the cause.