Kinesis and Google BigQuery Integration
Powerful performance with an easy integration, powered by Telegraf, the open source data connector built by InfluxData.
5B+
Telegraf downloads
#1
Time series database
Source: DB Engines
1B+
Downloads of InfluxDB
2,800+
Contributors
Table of Contents
Powerful Performance, Limitless Scale
Collect, organize, and act on massive volumes of high-velocity data. Any data is more valuable when you think of it as time series data. with InfluxDB, the #1 time series platform built to scale with Telegraf.
See Ways to Get Started
Input and output integration overview
<p>The Kinesis plugin enables you to read from Kinesis data streams, supporting various data formats and configurations.</p>
<p>The Google BigQuery plugin allows Telegraf to write metrics to Google Cloud BigQuery, enabling robust data analytics capabilities for telemetry data.</p>
Integration details
Kinesis
<p>The Kinesis Telegraf plugin is designed to read from Amazon Kinesis data streams, enabling users to gather metrics in real-time. As a service input plugin, it operates by listening for incoming data rather than polling at regular intervals. The configuration specifies various options including the AWS region, stream name, authentication credentials, and data formats. It supports tracking of undelivered messages to prevent data loss, and users can utilize DynamoDB for maintaining checkpoints of the last processed records. This plugin is particularly useful for applications requiring reliable and scalable stream processing alongside other monitoring needs.</p>
Google BigQuery
<p>The Google BigQuery plugin for Telegraf enables seamless integration with Google Cloud’s BigQuery service, a popular data warehousing and analytics platform. This plugin facilitates the transfer of metrics collected by Telegraf into BigQuery datasets, making it easier for users to perform analyses and generate insights from their telemetry data. It requires authentication through a service account or user credentials and is designed to handle various data types, ensuring that users can maintain the integrity and accuracy of their metrics as they are stored in BigQuery tables. The configuration options allow for customization around dataset specifications and handling metrics, including the management of hyphens in metric names, which are not supported by BigQuery for streaming inserts. This plugin is particularly useful for organizations leveraging the scalability and powerful query capabilities of BigQuery to analyze large volumes of monitoring data.</p>
Configuration
Kinesis
Google BigQuery
Input and output integration examples
Kinesis
<ol> <li> <p><strong>Real-Time Data Processing with Kinesis</strong>: This use case involves integrating the Kinesis plugin with a monitoring dashboard to analyze incoming data metrics in real-time. For instance, an application could consume logs from multiple services and present them visually, allowing operations teams to quickly identify trends and react to anomalies as they occur.</p> </li> <li> <p><strong>Serverless Log Aggregation</strong>: Utilize this plugin in a serverless architecture where Kinesis streams aggregate logs from various microservices. The plugin can create metrics that help detect issues in the system, automating alerting processes through third-party integrations, enabling teams to minimize downtime and improve reliability.</p> </li> <li> <p><strong>Dynamic Scaling Based on Stream Metrics</strong>: Implement a solution where stream metrics consumed by the Kinesis plugin could be used to adjust resources dynamically. For example, if the number of records processed spikes, corresponding scale-up actions could be triggered to handle the increased load, ensuring optimal resource allocation and performance.</p> </li> <li> <p><strong>Data Pipeline to S3 with Checkpointing</strong>: Create a robust data pipeline where Kinesis stream data is processed through the Telegraf Kinesis plugin, with checkpoints stored in DynamoDB. This approach can ensure data consistency and reliability, as it manages the state of processed data, enabling seamless integration with downstream data lakes or storage solutions.</p> </li> </ol>
Google BigQuery
<ol> <li> <p><strong>Real-Time Analytics Dashboard</strong>: Leverage the Google BigQuery plugin to feed live metrics into a custom analytics dashboard hosted on Google Cloud. This setup would allow teams to visualize performance data in real-time, providing insights into system health and usage patterns. By using BigQuery’s querying capabilities, users can easily create tailored reports and dashboards to meet their specific needs, thus enhancing decision-making processes.</p> </li> <li> <p><strong>Cost Management and Optimization Analysis</strong>: Utilize the plugin to automatically send cost-related metrics from various services into BigQuery. Analyzing this data can help businesses identify unnecessary expenses and optimize resource usage. By performing aggregation and transformation queries in BigQuery, organizations can create accurate forecasts and manage their cloud spending efficiently.</p> </li> <li> <p><strong>Cross-Team Collaboration on Monitoring Data</strong>: Enable different teams within an organization to share their monitoring data using BigQuery. With the help of this Telegraf plugin, teams can push their metrics to a central BigQuery instance, fostering collaboration. This data-sharing approach encourages best practices and cross-functional awareness, leading to collective improvements in system performance and reliability.</p> </li> <li> <p><strong>Historical Analysis for Capacity Planning</strong>: By using the BigQuery plugin, companies can collect and store historical metrics data essential for capacity planning. Analyzing trends over time can help anticipate system needs and scale infrastructure proactively. Organizations can create time-series analyses and identify patterns that inform their long-term strategic decisions.</p> </li> </ol>
Feedback
Thank you for being part of our community! If you have any general feedback or found any bugs on these pages, we welcome and encourage your input. Please submit your feedback in the InfluxDB community Slack.
Powerful Performance, Limitless Scale
Collect, organize, and act on massive volumes of high-velocity data. Any data is more valuable when you think of it as time series data. with InfluxDB, the #1 time series platform built to scale with Telegraf.
See Ways to Get Started
Related Integrations
Related Integrations
HTTP and InfluxDB Integration
The HTTP plugin collects metrics from one or more HTTP(S) endpoints. It supports various authentication methods and configuration options for data formats.
View IntegrationKafka and InfluxDB Integration
This plugin reads messages from Kafka and allows the creation of metrics based on those messages. It supports various configurations including different Kafka settings and message processing options.
View IntegrationKinesis and InfluxDB Integration
The Kinesis plugin allows for reading metrics from AWS Kinesis streams. It supports multiple input data formats and offers checkpointing features with DynamoDB for reliable message processing.
View Integration