Integrating Apache Kafka® with other systems in a reliable and scalable way is often a key part of a streaming platform. Fortunately, Apache Kafka includes the Connect API that enables streaming integration both in and out of Kafka. Like any technology, understanding its architecture and deployment patterns is key to successful use, as is knowing where to go looking when things aren't working.
This talk discusses the key design concepts within Apache Kafka Connect and the pros and cons of standalone vs distributed deployment modes. We do a live demo of building pipelines with Apache Kafka Connect for streaming data in from databases, and out to targets including Elasticsearch. With some gremlins along the way, we go hands-on in methodically diagnosing and resolving common issues encountered with Apache Kafka Connect. The talk finishes off by discussing more advanced topics including Single Message Transforms, and deployment of Kafka Connect in containers.