This message was deleted.
# troubleshooting
s
This message was deleted.
r
you can use some CDC tool to import into kafka, then ingest in 'real-time'
there's also some option to import via jdbc_connector, but there's always some issues doing that way
d
Druid is opinionated when it comes to ingestion. Maybe one day, when multi-stage-query is matured enough, it can ingest directly using SQL.
a
Thanks for your responds. My point is why we have to push data to csv or tsv first then indexing them? @Abhishek Agarwal it is great to have this new solution now. But, this tool still will store data into disk then it will read data from local disk on druid to segment them. The following phrase from druid documentation:
During indexing, each sub-task would execute one of the SQL queries and the results are stored locally on disk. The sub-tasks then proceed to read the data from these local input files and generate segments.
a
all batch ingestion do generate intermediate data though some thinking will be needed to write a list of queries that lets you parallelize the execution.