Sagar Sumit
aa857beee0
[HUDI-2225] Add a compaction job in hudi-examples ( #3347 )
2021-08-03 11:31:56 +08:00
vinoth chandar
b21ae68e67
[MINOR] Improving runtime of TestStructuredStreaming by 2 mins ( #3382 )
2021-08-02 13:42:46 -07:00
Sivabalan Narayanan
fe508376fa
[HUDI-2177][HUDI-2200] Adding virtual keys support for MOR table ( #3315 )
2021-08-02 09:45:09 -04:00
zhangyue19921010
dde57b293c
[HUDI-2164] Let users build cluster plan and execute this plan at once using HoodieClusteringJob for async clustering ( #3259 )
...
* add --mode schedule/execute/scheduleandexecute
* fix checkstyle
* add UT testHoodieAsyncClusteringJobWithScheduleAndExecute
* log changed
* try to make ut success
* try to fix ut
* modify ut
* review changed
* code review
* code review
* code review
* code review
Co-authored-by: yuezhang <yuezhang@freewheel.tv >
2021-08-02 08:07:59 +08:00
Gary Li
6353fc865f
[HUDI-2218] Fix missing HoodieWriteStat in HoodieCreateHandle ( #3341 )
2021-07-30 02:36:57 -07:00
swuferhong
f7f5d4cc6d
[HUDI-2184] Support setting hive sync partition extractor class based on flink configuration ( #3284 )
2021-07-30 17:24:00 +08:00
Danny Chan
c4e45a0010
[HUDI-2254] Builtin sort operator for flink bulk insert ( #3372 )
2021-07-30 16:58:11 +08:00
swuferhong
8b19ec9ca0
[HUDI-2252] Default consumes from the latest instant for flink streaming reader ( #3368 )
2021-07-30 14:25:05 +08:00
Sivabalan Narayanan
7bdae69053
[HUDI-2253] Refactoring few tests to reduce runningtime. DeltaStreamer and MultiDeltaStreamer tests. Bulk insert row writer tests ( #3371 )
...
Co-authored-by: Sivabalan Narayanan <nsb@Sivabalans-MBP.attlocal.net >
2021-07-29 22:22:26 -07:00
pengzhiwei
c2370402ea
[HUDI-2251] Fix Exception Cause By Table Name Case Sensitivity For Append Mode Write ( #3367 )
2021-07-29 17:36:56 -04:00
Shawy Geng
44e41dc9bb
[HUDI-2117] Unpersist the input rdd after the commit is completed to … ( #3207 )
...
Co-authored-by: Vinoth Chandar <vinoth@apache.org >
2021-07-29 08:16:58 -07:00
pengzhiwei
f109c6cb0d
[MINOR] fix check style error ( #3365 )
2021-07-29 17:29:10 +08:00
pengzhiwei
bbadac7de1
[HUDI-1425] Performance loss with the additional hoodieRecords.isEmpty() in HoodieSparkSqlWriter#write ( #2296 )
2021-07-28 21:30:18 -07:00
Danny Chan
efbbb67420
[HUDI-2241] Explicit parallelism for flink bulk insert ( #3357 )
2021-07-29 09:57:37 +08:00
swuferhong
7739518879
[HUDI-2228] Add option 'hive_sync.mode' for flink writer ( #3352 )
2021-07-28 19:45:50 +08:00
swuferhong
eedfadeb46
[HUDI-2244] Fix database alreadyExists exception while hive sync ( #3361 )
2021-07-28 19:40:16 +08:00
Danny Chan
91c2213412
[HUDI-2245] BucketAssigner generates the fileId evenly to avoid data skew ( #3362 )
2021-07-28 19:26:37 +08:00
davehagman
8105cf588e
[HUDI-2230] Make codahale times transient to avoid serializable exceptions ( #3345 )
2021-07-28 14:45:09 +08:00
rmahindra123
8fef50e237
[HUDI-2044] Integrate consumers with rocksDB and compression within External Spillable Map ( #3318 )
2021-07-28 01:31:03 -04:00
mincwang
00cd35f90a
[HUDI-2215] Add rateLimiter when Flink writes to hudi. ( #3338 )
...
Co-authored-by: wangminchao <wangminchao@asinking.com >
2021-07-28 08:23:23 +08:00
Danny Chan
60758b36ea
[HUDI-2227] Only sync hive meta on successful commit for flink batch writer ( #3351 )
2021-07-27 20:10:08 +08:00
pengzhiwei
59ff8423f9
[HUDI-2223] Fix Alter Partitioned Table Failed ( #3350 )
2021-07-27 20:01:04 +08:00
Gary Li
925873bb3c
[HUDI-2217] Fix no value present in incremental query on MOR ( #3340 )
2021-07-27 17:30:01 +08:00
Danny Chan
ab2e0d0ba2
[HUDI-2219] Fix NPE of HoodieConfig ( #3342 )
2021-07-27 15:18:05 +08:00
Danny Chan
9d2a65a6a6
[HUDI-2209] Bulk insert for flink writer ( #3334 )
2021-07-27 10:58:23 +08:00
xiang2102
024cf01f02
[MINOR] Correct the words accroding in the comments to according ( #3343 )
...
Correct the words 'accroding' in the comments to 'according'
2021-07-27 08:48:58 +08:00
Sivabalan Narayanan
61148c1c43
[HUDI-2176, 2178, 2179] Adding virtual key support to COW table ( #3306 )
2021-07-26 17:21:04 -04:00
xiarixiaoyao
5353243449
[HUDI-2214]residual temporary files after clustering are not cleaned up ( #3335 )
2021-07-26 10:26:20 -07:00
Gary Li
a5638b995b
[MINOR] Close log scanner after compaction completed ( #3294 )
2021-07-26 17:39:13 +08:00
董可伦
a91296f14a
[HUDI-2216] Correct the words fiels in the comments to fields ( #3339 )
2021-07-25 12:15:57 +08:00
rmahindra123
a14b19fdd5
[HUDI-1241] Automate the generation of configs webpage as configs are added to Hudi repo ( #3302 )
2021-07-23 21:33:34 -07:00
Xuedong Luan
b2f7fcb8c8
[MINOR] Replace deprecated method isDir with isDirectory ( #3319 )
2021-07-24 10:02:24 +08:00
jsbali
66207ed91a
[HUDI-1848] Adding support for HMS for running DDL queries in hive-sy… ( #2879 )
...
* [HUDI-1848] Adding support for HMS for running DDL queries in hive-sync-tool
* [HUDI-1848] Fixing test cases
* [HUDI-1848] CR changes
* [HUDI-1848] Fix checkstyle violations
* [HUDI-1848] Fixed a bug when metastore api fails for complex schemas with multiple levels.
* [HUDI-1848] Adding the complex schema and resolving merge conflicts
* [HUDI-1848] Adding some more javadocs
* [HUDI-1848] Added javadocs for DDLExecutor impls
* [HUDI-1848] Fixed style issue
2021-07-23 09:03:15 -07:00
Xuedong Luan
71e14cf866
[HUDI-2213] Remove unnecessary parameter for HoodieMetrics constructor and fix NPE in UT ( #3333 )
2021-07-23 19:57:35 +08:00
pengzhiwei
2c910ee3af
[HUDI-2212] Missing PrimaryKey In Hoodie Properties For CTAS Table ( #3332 )
2021-07-23 15:21:57 +08:00
Xuedong Luan
6d592c5896
[HUDI-2211] Fix NullPointerException in TestHoodieConsoleMetrics ( #3331 )
2021-07-23 11:22:54 +08:00
pengzhiwei
5a2f3d439e
[HUDI-2139] MergeInto MOR Table May Result InCorrect Result ( #3230 )
2021-07-23 10:19:43 +08:00
Danny Chan
c89bf1de20
[HUDI-2205] Rollback inflight compaction for flink writer ( #3320 )
2021-07-22 22:56:51 +08:00
swuferhong
fe5d2e7f53
[HUDI-2206] Fix checkpoint blocked because getLastPendingInstant() action after than restoreWriteMetadata() action ( #3326 )
2021-07-22 16:35:07 +08:00
pengzhiwei
151f22e43a
[HUDI-2195] Sync Hive Failed When Execute CTAS In Spark2 And Spark3 ( #3299 )
2021-07-22 15:33:38 +08:00
Danny Chan
2370a9facb
[HUDI-2204] Add marker files for flink writer ( #3316 )
2021-07-22 13:34:15 +08:00
Vinay Patil
5a94b6bf54
[HUDI-2192] Clean up Multiple versions of scala libraries detected Warning ( #3292 )
2021-07-21 00:33:27 -07:00
satishkotha
4f1350f7c1
[MINOR] Disable codecov ( #3314 )
2021-07-20 22:07:22 -07:00
Sivabalan Narayanan
d58a8348dc
[HUDI-2007] Fixing hudi_test_suite for spark nodes and adding spark bulk_insert node ( #3074 )
2021-07-21 00:11:01 -04:00
Danny Chan
858e84b5b2
[HUDI-2198] Clean and reset the bootstrap events for coordinator when task failover ( #3304 )
2021-07-21 10:13:05 +08:00
yuzhaojing
634163a990
[HUDI-2145] Create new bucket when NewFileAssignState filled ( #3258 )
...
Co-authored-by: 喻兆靖 <yuzhaojing@bilibili.com >
2021-07-20 17:46:45 +08:00
Samrat
a086d255c8
[HUDI-1860] Add INSERT_OVERWRITE and INSERT_OVERWRITE_TABLE support to DeltaStreamer ( #3184 )
2021-07-19 21:49:43 -04:00
Sivabalan Narayanan
d5026e9a24
[HUDI-2161] Adding support to disable meta columns with bulk insert operation ( #3247 )
2021-07-19 20:43:48 -04:00
喻兆靖
2099bf41db
[HUDI-2193] Remove state in BootstrapFunction
2021-07-19 18:14:06 +08:00
pengzhiwei
572a214412
[HUDI-1884] MergeInto Support Partial Update For COW ( #3154 )
2021-07-17 12:59:18 +08:00