1. e7db7cf [core] Avoid retaining rejected cache pages (#9658) by yuzelin · 15 hours ago master
  2. 475be56 [spark] Fix batch delete for partial update sequence groups (#9546) by Arnav Balyan · 3 days ago
  3. 0289f3d [spark] Disable create table in default db (#9549) by junmuz · 3 days ago
  4. 70d98e5 [spark] Reject variant extraction paths containing the metadata delimiter (#9569) by jackylee · 3 days ago
  5. 1d5cc31 [format] Fail fast when skipping DELTA_LENGTH_BYTE_ARRAY data past end of page (#9571) by YangJie · 3 days ago
  6. dfacf42 [format] Close the Avro stats extractor's stream on a corrupt file (#9573) by YangJie · 3 days ago
  7. 5a702ac [format] Close the ParquetFileReader when post-construction setup fails (#9576) by YangJie · 3 days ago
  8. 32a0ed5 [common] Fix pre-epoch timestamp casts truncating toward zero (#9577) by jackylee · 3 days ago
  9. fe8b591 [format] Close the text stream when line-reader construction fails (#9579) by YangJie · 3 days ago
  10. 5110d36 [format] Build fresh ORC writer options per created writer (#9583) by YangJie · 3 days ago
  11. c10da57 [core] Clean up data file extra files (#9586) by Zhang Jiawei · 3 days ago
  12. 09e7a7a [format] Report unresolvable JSON string casts instead of a bare NPE (#9588) by YangJie · 3 days ago
  13. de99e88 [api] Trim whitespace in the bucket-key option value (#9591) by jackylee · 3 days ago
  14. 4242ff2 [common] Fail on a short remote read instead of caching zero padding (#9593) by jackylee · 3 days ago
  15. ffe8e51 [python] Persist LeRobot dataset metadata in Paimon (#9529) by XiaoHongbo · 3 days ago
  16. bdd6e41 [parquet] Push down StartsWith as Parquet binary range (#9594) by sanshi · 3 days ago
  17. f947686 [python] Notify callbacks for successful retried commits (#9589) by XiaoHongbo · 3 days ago
  18. 77c5689 [python] Make multimodal table creation race-safe (#9590) by XiaoHongbo · 3 days ago
  19. 85c8d60 [docs] Regenerate the configuration reference (#9592) by YangJie · 3 days ago
  20. b71ebe5 [docs] Update docs for options by JingsongLi · 3 days ago
  21. ea306e2 [api][core][spark] Support custom partition locations (#9540) by Dapeng Sun(孙大鹏) · 3 days ago
  22. 9109de4 refactor(core): make SchemaManager an interface (#9174) by baiyangtx · 3 days ago
  23. 0ce6a13 [core][python][spark] Optimize data evolution write column metadata (#9574) by Jingsong Lee · 4 days ago
  24. ba5801b [spark] Use the field index instead of the field id for TopN pushdown (#9566) by jackylee · 4 days ago
  25. f80d348 [format] Compare array wrapper-group names case-insensitively in ParquetReaderUtil (#9568) by YangJie · 4 days ago
  26. c331aeb [vortex] Fix predicate pushdown literals for DATE and second-precision TIMESTAMP (#9558) by Eunbin Son · 4 days ago
  27. fde1868 [format] Read BINARY-encoded decimals at scratch index 0 in parquet updater (#9563) by YangJie · 4 days ago
  28. e3cd4df [format] Keep parsing remaining fields after a malformed one in PERMISSIVE CSV mode (#9556) by YangJie · 4 days ago
  29. 14040b1 [format] Support sep as fallback key of csv.field-delimiter (#9565) by jackylee · 4 days ago
  30. 46ba090 [core][spark][flink] Fix nested data evolution isolation (#9564) by Jingsong Lee · 4 days ago
  31. aa98e58 [python] Parse BlobDescriptor v1/v2 bytes without misclassifying inline payload (#9539) by Wenchao Wu · 4 days ago
  32. ec922f6 [format] Give each row-format reader its own projected-row wrapper (#9561) by YangJie · 4 days ago
  33. ce01e64 [core][spark][flink] Support sub-field-level data evolution for nested columns (#8334) by Xiangyi Zhu · 4 days ago
  34. 2788fe5 [python] Infer multimodal vector columns by query dimension (#9553) by zhigang · 4 days ago
  35. 8c61c18 [python] Cache decoded BLOB indexes (#9547) by Yann Byron · 4 days ago
  36. e50bc46 Bump fast-uri from 3.1.5 to 3.1.7 in /docs (#9554) by dependabot[bot] · 4 days ago
  37. 87bdb2d [spark] Support action predicate pruning for data evolution self-merge (#9544) by zhoulii · 5 days ago
  38. e0b3027 [python] Expose Arrow batch reader for multimodal scans (#9537) by XiaoHongbo · 5 days ago
  39. 8d67f8c [format] Fix positional schema evolution for ORC reads (#9464) by Arnav Balyan · 5 days ago
  40. 2e29b9b [flink] Fix schema inference for multi partition Kafka topics (#9478) by Arnav Balyan · 5 days ago
  41. ee9e8b1 [format] Prevent text tables from accepting unsupported schemas (#9479) by Arnav Balyan · 5 days ago
  42. 58bc214 [spark] Fix analyze failure for timestamp ntz columns (#9485) by Arnav Balyan · 5 days ago
  43. 60381f4 [format] Preserve csv values matching the configured null literal (#9492) by Arnav Balyan · 5 days ago
  44. 710b6b4 [Bug] Fix debezium-json cannot infer primary keys from standard Debezium message key (#9495) by Arvin · 5 days ago
  45. 4fff146 [rest] Make RESTTokenFileIO cache maximum size configurable (#9528) by tonymtu · 5 days ago
  46. 87763fc [docs] Fix misleading DROP COLUMN doc for hive catalog (#9538) by Pei Yu · 5 days ago
  47. b579f38 [core] make stats model work while migrating (#9541) by weijie · 5 days ago
  48. c323585 Bump browserslist from 4.28.2 to 4.28.8 in /docs (#9542) by dependabot[bot] · 5 days ago
  49. 10fee64 [format] Introduce byte stream split for writing to parquet (#9499) by Arnav Balyan · 5 days ago
  50. dbb5423 [Bug] Fix concurrent LocalTableQuery lookups returning null across files (#9500) by Arvin · 5 days ago
  51. 8052d3b [docs] Fix stale paimon-vindex version in vector index documentation (#9505) by Eunbin Son · 5 days ago
  52. 3912c08 [pvfs] Set working directory to the catalog root in PaimonVirtualFileSystem (#9507) by Eunbin Son · 5 days ago
  53. 4a119b7 [test-util] Set accessible on @Parameters provider and @Parameter field (#9510) by Eunbin Son · 5 days ago
  54. 0997597 [common] Fix permit accounting of SemaphoredDelegatingExecutor under interruption and rejection (#9511) by YangJie · 5 days ago
  55. 19f35f5 [common] Fix partition getter indexing in InternalRowPartitionComputer (#9513) by YangJie · 5 days ago
  56. ae6816c [common][spark] Fix boolean hilbert value colliding with the null sentinel (#9515) by YangJie · 5 days ago
  57. ba7fec9 [common] Reject non-struct inner fields in variant shredding schema (#9518) by YangJie · 5 days ago
  58. 29a8cad [common] Fix double position advance in DataOutputSerializer.writeBytes (#9520) by YangJie · 5 days ago
  59. a26481b [spark] Let MSCK see the null partition of a value-only format table (#9521) by Dapeng Sun(孙大鹏) · 5 days ago
  60. b4812ef [common] Grow from one byte in MemorySliceOutput when the segment is empty (#9523) by YangJie · 5 days ago
  61. 923e694 [common] Make BinaryRow.anyNull respect the row offset (#9525) by YangJie · 5 days ago
  62. 09ed3bb [common][spark] Fix z-order boolean FALSE colliding with the null sentinel (#9527) by YangJie · 5 days ago
  63. 606f230 [python] Add ROSBag ingestion support (#9530) by Yann Byron · 5 days ago
  64. 684c2e5 [core] Support changelog-producer.ignore-update-before and changelog-producer.ignore-delete options (#9531) by junmuz · 5 days ago
  65. 3e510cf [python][ray] Cover conditional self-merge compaction rebase (#9516) by Jingsong Lee · 6 days ago
  66. da71a66 [iceberg] Leave the nanosecond refusal to the validator that knows the table (#9503) by Jiajia Li · 6 days ago
  67. 9c7deeb [python] Add distributed RoboMIND action backfill (#9496) by Yann Byron · 6 days ago
  68. a008390 [spark] Show format table partition statistics (#9501) by Dapeng Sun(孙大鹏) · 6 days ago
  69. 0ff874b [spark] Skip time-travel self-merge test on Spark 3.2 (#9502) by Dapeng Sun(孙大鹏) · 6 days ago
  70. 2c87f4a [spark] Support timestamp_ntz and variant in generic row access (#9493) by shyjsarah · 7 days ago
  71. 1367e46 [spark] Support partition pruning for data evolution self-merge (#9489) by zhoulii · 7 days ago
  72. ae83ef3 [python] Reject overlapping row-id update batches (#9484) by Yann Byron · 7 days ago
  73. f99d008 [spark] Support parallel sorted index building across Paimon partitions (#9491) by liangjie · 7 days ago
  74. f06ebd6 [python][torch] Make map-style reads lazy (#9486) by XiaoHongbo · 7 days ago
  75. d7b1bf9 [core] Rename FM index identifier to fm (#9487) by Jingsong Lee · 7 days ago
  76. fbafbbe [format] Support nested variant column pruning for shredded parquet files (#9389) by Juntao Zhang · 7 days ago
  77. 2e3b41c [spark] Clamp Format Table ANALYZE parallelism (#9482) by Dapeng Sun(孙大鹏) · 7 days ago
  78. d557d05 [python] Add LeRobot Dataset v3 import (#9446) by XiaoHongbo · 7 days ago
  79. bfeb7cf [multimodal] Support aligned multi-video file groups (#9480) by Jingsong Lee · 7 days ago
  80. 5e3609e [core][python] Support multiple packed video fields (#9473) by XiaoHongbo · 7 days ago
  81. 4708c73 [spark] Fix flaky row tracking concurrent compaction test (#9475) by Jingsong Lee · 7 days ago
  82. 963528c [python][torch] Support distributed iterable dataset sharding (#9429) by XiaoHongbo · 8 days ago
  83. 0e974c8 [iceberg] Refuse time precisions the mirror cannot publish (#9467) by Jiajia Li · 8 days ago
  84. 9c39e0d [core] Read the table branch schemas when rewriting file indexes (#9468) by Jiajia Li · 8 days ago
  85. 7277b77 [flink] Fix decimal column analysis with null statistics (#9469) by Arnav Balyan · 8 days ago
  86. e2e7b40 [format][python] Refactor video storage internals (#9470) by Jingsong Lee · 8 days ago
  87. d2ddbee [core] Fix column masking correctness in query-auth reads (#8570) by Jiajia Li · 8 days ago
  88. df31961 [core][python] Reuse PK sorted indexes with retired source files (#9460) by wangyong9999 · 8 days ago
  89. 5371a4f [iceberg] Allow publishing VARIANT columns with Iceberg format version 3 (#9246) by Victor Babenko · 8 days ago
  90. a51754c [Fix] MongoDB CDC: handle null fullDocument in transaction to prevent NullNode exception (#9332) (#9336) by Arvin · 8 days ago
  91. 5280516 [flink][cdc] Fix missing primary keys for Debezium records (#9465) by Arnav Balyan · 8 days ago
  92. 2bf35fb [format][python] Add packed video frame storage (#9461) by Jingsong Lee · 8 days ago
  93. d53103a [python] Add RoboMIND AgileX HDF5 pipeline (#9445) by Yann Byron · 8 days ago
  94. a74daf8 [hive] Fix query failure when dynamic filters are unavailable (#9448) by Arnav Balyan · 8 days ago
  95. 69406b5 [metrics] Add totalFileCount and minAvgFileSize metrics to detect small file problem (#9449) by junmuz · 8 days ago
  96. 72caec1 [api] Fix race condition causing queued tasks to execute after SequentialBatchIterator close (#9452) by Juntao Zhang · 8 days ago
  97. 42b39f1 [format] Fix ORC zstd compression option inconsistent across writers (#9453) by Arnav Balyan · 8 days ago
  98. 7873971 [api] Stop reading a server error message as a format string (#9454) by Jiajia Li · 8 days ago
  99. 22daae7 [iceberg] Refuse nanosecond timestamps while Iceberg metadata is enabled (#9455) by Jiajia Li · 8 days ago
  100. 09997f2 [core] rename FileStoreCommitImpl var name (#9457) by weijie · 8 days ago