Skip to content
Open
Show file tree
Hide file tree
Changes from all commits
Commits
Show all changes
10000 commits
Select commit Hold shift + click to select a range
4e9a743
[XLA:GPU] Skip collective symmetric buffer tests on non-Hopper archit…
Sep 30, 2026
29b9c9f
[XLA:GPU] Add PerDeviceState container for per-device runtime state.
Sep 30, 2026
e1d06e3
Merge pull request #127780 from vishwakt:fix-framework-llvm-symbol-vi…
Sep 30, 2026
9402391
Merge pull request #128227 from MaddipatlaChetan24:patch-43
Sep 30, 2026
9d72633
Merge pull request #128143 from kaivalya-cyber:fix/np-transpose-axes-…
Sep 30, 2026
40bfb8d
Automated Code Change
Sep 30, 2026
c3f7240
[XLA:GPU] Run the cuDNN frontend graph warmup on a private stream to …
Sep 30, 2026
2816dca
PR #49717: Export control_predecessors() to python
Sep 30, 2026
bbdf033
PR #49741: [ROCm] Remove dead ROCm version checks from GemmRewriter
Sep 30, 2026
17bf4fb
[XLA:GPU] Add REDUCE_SCATTER to xla_gpu_experimental_use_collective_k…
Sep 30, 2026
0d920ad
PR #49300: [XLA:GPU] Inline scan computation to_apply calls
Sep 30, 2026
e1d6d57
[XLA:GPU] tighter constraint checks in indexing map
Sep 30, 2026
25acadd
[NFC] Compute GPU instruction annotation titles and metadata once per…
Sep 30, 2026
8ebd54a
Merge pull request #123864 from vishwakt:fix-crop-and-resize-empty-crash
Sep 30, 2026
480682d
Merge pull request #127968 from RohithPariki:fix/87457-gpu-conv-filte…
Sep 30, 2026
c443ce9
Merge pull request #126161 from vishwakt:fix-splitv-size-overflow
Sep 30, 2026
9fe191d
[XLA:GPU] Support disabled VMM API case.
Sep 30, 2026
645e3ca
[XLA:GPU]: Move memcpy inside all-gather
Sep 30, 2026
5a1f0fd
[StableHLO] Fix windows link error in evalRunParallel
Sep 30, 2026
270a245
PR #49719: [tsl:concurrency] Deduplicate AsyncValue type ids by type …
Sep 30, 2026
b02620e
PR #47649: [XLA:GPU] Keep scatter window writes coalesced in ScatterS…
Sep 30, 2026
25daff1
Address review: name_scope, named conversion, SortEigenValues, comple…
Sep 30, 2026
717f958
PR #49767: [ROCm] Restrict usage of system env variables for rbe buil…
Sep 30, 2026
54a6c3a
Skip NVTX-only annotation work when no profiler domain is attached.
Sep 30, 2026
4706a67
Removing deprecated macros in XLA:GPU E2E tests and autotuner
Sep 30, 2026
e1e8f86
Migrate deprecated TF assertion macros in XLA:GPU codegen and Triton
Sep 30, 2026
a3165c2
Automated Code Change
Sep 30, 2026
4d67a2d
PR #49603: [ROCM] bug-fixing matmul plan cache behaviour
Sep 30, 2026
b685b4d
Clean up clang-tidy issues with unchecked .value() call of status_or …
Sep 30, 2026
6598f6e
[XLA:GPU][NFC] Make the PriorityFusion decision chain non-blocking
Sep 30, 2026
905d2aa
[XLA:GPU/CPU] Allow constraints in apply_indexing op.
Sep 30, 2026
44077b4
[XLA:GPU]: Relax constraints on all gather kernel.
Sep 30, 2026
ae061bd
[XLA:CPU] Enable Eigen cpp-gen intrinsics on macOS and support 128-bi…
Sep 30, 2026
7d4259b
Merge pull request #128179 from twelfthlabor:fix/readme-dead-build-links
Sep 30, 2026
bcb25f0
Migrate deprecated TF assertion macros in XLA:GPU runtime and StreamE…
Sep 30, 2026
a4da07c
Migrate deprecated TF assertion macros in XLA:GPU transforms and SPMD
Sep 30, 2026
6072d1f
PR #48195: [ROCm] Add a per-arch core info table
Sep 30, 2026
20739e9
[XLA:GPU] Fix high-pad masking in emitter
Sep 30, 2026
182f30e
[XLA] Remove dead instructions in outlined block body from caller com…
Sep 30, 2026
dc1c146
[XLA:CPU] Enable Eigen cpp-gen intrinsics on macOS and support 128-bi…
Sep 30, 2026
2f18a6d
[XLA:GPU/CPU] Allow constraints in apply_indexing op.
Sep 30, 2026
49b3f98
PR #49772: Bump github/codeql-action/upload-sarif from 4.38.0 to 4.38.1
Sep 30, 2026
e51d919
PR #48712: Support reporting the median execution time in --append_pr…
Sep 30, 2026
3ab41bc
Remove redundant macros include from cwise functors
Sep 30, 2026
00b5904
PR #49773: Bump astral-sh/setup-uv from 10.1.0 to 10.2.0
Sep 30, 2026
439414b
Merge remote-tracking branch 'origin/master' into fix/np-split-axis-v…
Sep 30, 2026
a920904
Simplify split axis bounds check and expand out-of-bounds test matrix
Sep 30, 2026
04a5e38
PR #49771: Bump lit from 23.1.1 to 23.1.2
Sep 30, 2026
6a2d4fb
Merge pull request #125949 from AshiteshSingh:digamma-only
Sep 30, 2026
fd130ef
Merge pull request #128116 from Abhirup0:fix-module-name-scope
Sep 30, 2026
82e6f60
Automated Code Change
Sep 30, 2026
e9e7e83
PR #46625: [XLA:CPU] Clamp materialized dynamic dimension sizes to th…
Sep 30, 2026
37b6a53
Merge pull request #128068 from bodapatisaikrishna:fix/76730-gpu-lina…
Sep 30, 2026
d42aabc
Merge pull request #127963 from RohithPariki:fix/116933-softmax-singl…
Sep 30, 2026
dd8864c
Merge pull request #128161 from vishwakt:fix-tf-function-literal-equa…
Sep 30, 2026
bb3dc25
Merge pull request #128141 from kaivalya-cyber:fix/np-expand-dims-val…
Sep 30, 2026
c799edd
Add unit tests to show non-flat support for ReshapeMover pass.
Sep 30, 2026
db3f72d
[xla:cpu] custom_call_thunk: unpoison result buffers after custom cal…
Sep 30, 2026
8d75832
Add SharedMemRefSliceOp to slice the VMEM belonging to a specific sub…
Sep 30, 2026
9ed524a
Clarify list-sections comment: the early static bounds check raises f…
Sep 30, 2026
700e5b7
Merge master into fix/np-split-axis-validation
Sep 30, 2026
64e2c0a
Remove unused XLA error check macros and internal helpers.
Sep 30, 2026
10b104c
PR #49658: [XLA:GPU] Use polymorphic CUDA graph node APIs in CudaComm…
Sep 30, 2026
7e4d7f8
Merge pull request #128002 from yqtian-se:fix-gpu-prim-diagnostic-scope
Sep 30, 2026
6b003a1
[XLA:GPU] Evaluate tiling candidates in parallel from PriorityFusion
Sep 30, 2026
833645d
Do not treat empty Range objects as bounded in value range analysis.
Sep 30, 2026
d9c7ca0
Add an optional linearization throttler to CommonPjRtClient to track …
Sep 30, 2026
2dc30d6
Add MegaScale performance tuning flags documentation
Sep 30, 2026
6a26801
Fix ResizeBilinear GPU work-count overflows
Sep 30, 2026
733f9ec
PR #49436: [xla:gpu] Support table offsets for dynamic slices
Sep 30, 2026
385ce38
Handle single-device-sharding related reshard during StableHLO export.
Sep 30, 2026
8a4ca64
[XLA] Memoize stack frame locations per frame id on HLO to StableHLO …
Sep 30, 2026
a9b5f68
Internal build fix
Sep 30, 2026
ec08601
Integrate LLVM at llvm/llvm-project@e40e0bc36b14
Sep 30, 2026
050cd71
[XLA] Fix fusion users and aliasing with duplicate operands.
Sep 30, 2026
8797bd8
Refactor AllGatherPadDsSimplifier to use Abseil Span and idiomatic C++.
Sep 30, 2026
1f338db
ci: fix ARM64 gVisor unit test incompatibilities and add Tests_v2 CI job
Sep 30, 2026
296119c
Pass `shape.layout()` when present to `MakeDefaultShapeForMemorySpace…
Sep 30, 2026
b0f0256
Pass kOverrideDefault in SigtermNotifier to prevent premature SIGTERM…
Sep 30, 2026
900903c
Fix vector width detection in eigen_unary_test.cc on ARM.
Oct 1, 2026
c38f546
[XLA] Fix off-by-one boundary check for GT in WhileLoopSimplifier::Tr…
Oct 1, 2026
a36c94d
[XLA] Outline cold explanation formatting and avoid stringstream allo…
Oct 1, 2026
9eae6bb
Pass `shape.layout()` when present to `MakeDefaultShapeForMemorySpace…
Oct 1, 2026
fec9c94
Treat CUPTI_ERROR_NOT_COMPATIBLE as recoverable in V2 subscriber paths
Oct 1, 2026
5c3996e
Fix collaborator wheel build environment value
Oct 1, 2026
28541de
Reuse pip package builder target in regression test
Oct 1, 2026
1a6e205
TF macro deprecation in StreamExecutor ROCm delay kernel test
Oct 1, 2026
cd668fb
Address review: add batched and 64-bit cases to testAcceptsNonTensorI…
Oct 1, 2026
f3b4e75
Disallow GEMM fusions that cannot be tiled with any default Triton co…
Oct 1, 2026
8540120
[XLA:GPU] Add a helper to measure GPU device-to-device memory bandwidth.
Oct 1, 2026
4f5d7c2
Compare DebugOptions in XLA GPU AOT golden verification except a deny…
Oct 1, 2026
d203267
[XLA:GPU] Tie one-shot all-gather size threshold to output dimension …
Oct 1, 2026
a9b70c2
PR #48643: make sure the live range of buffers in async ops span acro…
Oct 1, 2026
1617f57
PR #49803: [XLA:GPU][oneAPI] Enable TF32_TF32_F32 dot algorithm presets
Oct 1, 2026
2085fdb
PR #49815: [DOCS] Add hlo_isolation.md to TOC
Oct 1, 2026
e20d228
PR #48845: [XLA:SPMD] Preserve batch sharding for FFTs
Oct 1, 2026
0181c0f
Reverts 4f5d7c26f3c81c7c71f8d50658a877af386e7c69
Oct 1, 2026
8ceefe4
PR #48666: PJRT: Use per-device execution profiles.
Oct 1, 2026
4f2a3ba
Fix get_broadcast_user comment.
Oct 1, 2026
731bc44
Automated Code Change
Oct 1, 2026
063e60b
PR #49276: [XLA:GPU] Separate host and device execution timeouts
Oct 1, 2026
08437c5
[XLA] Relay control dependencies in MergeFusionInstructionIntoMultiOu…
Oct 1, 2026
e251662
Automated Code Change
Oct 1, 2026
d6a4437
Rename aot_compatibility_experimental to aot_compatibility
Oct 1, 2026
722283e
[XLA:GPU] Query max PTX ISA version from `CompilationProvider::GetLat…
Oct 1, 2026
7e3222b
Add InternableKernelLoaderSpec and a deduplicated table of custom ker…
Oct 1, 2026
c9277a9
[mosaic] Reject core types on HBM, host and VMEM_SHARED memory spaces
Oct 1, 2026
5b821ce
[XLA:GPU] Own all TiledHloInstructions in a single deque per TiledHlo…
Oct 1, 2026
7508492
[XLA:GPU] disable pad outside of gemm fusions
Oct 1, 2026
2003790
[XLA:GPU] HLO reproducer for pad out of bound reads
Oct 1, 2026
7ebc53c
Add MIG GPU target configs for H100 SXM and B200.
Oct 1, 2026
53b8381
[XLA:CPU] Enable Eigen cpp-gen intrinsics on macOS and support 128-bi…
Oct 1, 2026
f088aee
Set cuDNN graph SM version from device info
Oct 1, 2026
dc63149
[mosaic] Added a new `memref_memory_space_is` operation
Oct 1, 2026
f1d899c
[XLA:GPU] Use lock-free SymbolicExpr uniquing on single-threaded MLIR…
Oct 1, 2026
da5fdf0
Keep XLines sorted when MergeXSpace merges them.
Oct 1, 2026
06d5b71
Add XLA:GPU ahead-of-time compilation and compatibility guides
Oct 1, 2026
54a111a
[XLA:GPU] Move SymmetricMemoryTypeProto to collective_types.proto.
Oct 1, 2026
0b99cfb
PR #49876: [ROCm] Remove leftover pre-7.1 version guards
Oct 1, 2026
602ede1
Fix OSS build breakages: strict_properties in tf_runtime and repo_met…
Oct 1, 2026
bc3ca1c
Add unit tests for NaN and Infinity handling in HistogramSummary op a…
Oct 1, 2026
7f54df3
Merge pull request #128318 from somtri:docs-128064-data-format-cpu
Oct 1, 2026
b315383
Short-circuit user fusion checks and forbid concatenate epilogues in …
Oct 1, 2026
1fa6c20
Merge pull request #127684 from kaivalya-cyber:fix/np-split-axis-vali…
Oct 1, 2026
3587da0
Avoid linking XNNPACK in BuiltinOpResolverWithoutDefaultDelegates
Oct 1, 2026
bd8446e
Merge pull request #126876 from GodlyDonuts:codex/fix-complex-zero-po…
Oct 1, 2026
30cc4a3
Keep metadata IDs consistent when merging XSpaces.
Oct 1, 2026
cd6ffc1
Enable new XTile lowering by default in XLA:CPU.
Oct 1, 2026
af014f4
Optimize XNNPACK MoE expert weight dequantization, single-token GEMV,…
Oct 1, 2026
871f6a5
[XLA] Clamp outer index when simplifying nested dynamic-update-slice(…
Oct 1, 2026
4133a41
[XLA:GPU] Add collective fusion and block-level config support for Re…
Oct 1, 2026
e4a3d4b
Set cuDNN graph SM version from device info
Oct 1, 2026
bd7ff71
Simplify GQA graph construction in YNNPACK and tighten SDPA decode1 t…
Oct 1, 2026
c73d97d
Reverts f1d899c197504ede179f82e0c93b3266f4dee30a
Oct 1, 2026
a534b8e
Update version number for TFLite-in-play-services GPU Maven package i…
Oct 1, 2026
aa856d2
[XLA] Avoid full-module traversal in HloExtractor
Oct 1, 2026
a33484c
Fix -Wpass-failed build break in the XNNPACK MoE kernel under sanitiz…
Oct 1, 2026
68e6390
xla(profiler): Guard empty CUDA graph node map in the graph-node API …
Oct 1, 2026
51142ae
Share pip package builder library between CLI and test
Oct 1, 2026
e99a0a3
Merge remote-tracking branch 'origin/master' into fix-collaborator-wh…
Oct 1, 2026
6cc0abc
Update CommonPjrtClient to use the linearization throttler and merge in
Oct 1, 2026
f3670f5
PR #49808: [XLA:GPU][oneAPI] Align XPU Triton hook with upstream LLVM
Oct 1, 2026
14afa11
Preserve argument and result attributes when cloning IFRT program.
Oct 1, 2026
b38999c
[XLA] Allow recovering cloned computations/instructions by their orig…
Oct 1, 2026
6046b02
[XLA] Add `tsl::kIsDebugBuild` and replace `#ifndef NDEBUG` with `if …
Oct 1, 2026
cc29162
[XLA] Check that HloComputation::set_root_instruction does not drop a…
Oct 1, 2026
fb91074
**[XLA:CPU] Fix GlooCommunicator teardown race in ReduceScatterMinMax…
Oct 1, 2026
856044b
Fix hang in multi_client_input_util_test after SigtermNotifier change.
Oct 1, 2026
75394cc
[XLA:GPU] Emit fork/join dependencies for async thunks in kLHS comman…
Oct 1, 2026
70bf929
Handle rank-agnostic sharding forms in StableHLO import.
Oct 2, 2026
b2a6a88
Add repro and benchmark tests for threadpool_listener.cc
Oct 2, 2026
74bed23
Allow the YNNPACK delegate to rebuild its subgraph outside of delegat…
Oct 2, 2026
582a8bb
[Mosaic] Split min reduction kind into minf/minsi/minui and similarly…
Oct 2, 2026
4f99b54
Unify backend-specific ShouldPerformZeroCopyLinearize implementations…
Oct 2, 2026
98f9ea5
[XLA] Generalize 2ba2030
Oct 2, 2026
3bbf197
[XLA] Add native `exp2` and `log2` operations.
Oct 2, 2026
c0e65bb
[IFRT IR] Do not use deprecated useStrictPropertiesInAssemblyFormat
Oct 2, 2026
37852c1
[IFRT] Add Client::Load, LoadOptions, and XlaLoadOptions
Oct 2, 2026
ed97813
[IFRT IR] Do not use deprecated useStrictPropertiesInAssemblyFormat i…
Oct 2, 2026
5595c92
Fix node name and update expected counts in tests
Oct 2, 2026
d48131b
Automated Code Change
Oct 2, 2026
ac2c96f
Fix two issues in exporting NamedSharding with empty meshes.
Oct 2, 2026
39a2428
[Mosaic] Strengthen EraseLayoutOp verifier
Oct 2, 2026
7588a16
Changes to internal Google code
Oct 2, 2026
7b5c431
Automated Code Change
Oct 2, 2026
0fea3da
Automated Code Change
Oct 2, 2026
de0e646
[Mosaic] Add fold hook to fold divui(remui(x, c), c) -> 0 in TPUDialect.
Oct 2, 2026
50711d0
Match merged event metadata on name and stats, not name alone.
Oct 2, 2026
21fa321
Automated Code Change
Oct 2, 2026
21b3f3e
[XLA:GPU] Emit stablehlo.reduce_scatter in XTile fusion emitter
Oct 2, 2026
1934649
Automated Code Change
Oct 2, 2026
74bf108
PR #49923: [XLA:GPU][oneAPI[Bug-fix] Disable vector optimizations tha…
Oct 2, 2026
e5b4d05
PR #49879: Fix NVTX extended payload schemas.
Oct 2, 2026
6643097
PR #48865: Add nv_maxtext_llama3_8b_1n8g presubmit benchmark
Oct 2, 2026
421bfb8
Automated Code Change
Oct 2, 2026
3d7181f
Support fusing transposes above broadcasts in GEMM fusion.
Oct 2, 2026
c29e4c4
Changes to internal Google code
Oct 2, 2026
ce109c3
Update Triton device tests to parameterize and test GemmFusionV2.
Oct 2, 2026
e34001a
PR #48709: [XLA:GPU] add cuDNN ragged dot wgrad support
Oct 2, 2026
a43622a
Migrate several GemmFusion tests to GemmFusionTestVersioned.
Oct 2, 2026
0792e4d
[XLA:GPU] Bind one/two shots heuristics on the amount of read memory,…
Oct 2, 2026
3da7f6f
Improve tarfile extraction security
Oct 2, 2026
948c3e1
[XLA:GPU] Delete RTX6000PRO cross-compilation test on GB200
Oct 2, 2026
ad8fb66
[XLA:GPU][XLA:CPU] add constraints to Tile of concat
Oct 2, 2026
2343f12
PR #49495: [ROCm] Move shim script logic upstream
Oct 2, 2026
021f9f1
[XLA:GPU] Add TritonXLACollapseContiguousMinorDimsPass.
Oct 2, 2026
117b32f
[XLA:GPU] Enable one-shot Triton reduce-scatter.
Oct 2, 2026
020380e
Fix heap corruption in detection_postprocess multi-threaded NMS merge.
Oct 2, 2026
937eafd
Normalize collaborator flag and clear disabled wheel environment
Oct 2, 2026
3c6550f
Use Python 3.9-compatible collaborator argument annotation
Oct 2, 2026
3a6a6b4
[XLA:GPU]: Add GpuTopology in GpuExecutable and plumb it through dese…
Oct 2, 2026
15466d0
Keep pip package builder module as a library
Oct 2, 2026
d84b24f
Make the XNNPACK MoE dot-product vectorize hint robust to sanitizer b…
Oct 2, 2026
1a99ab4
[XLA:GPU] Derive collective kernel scratch memory type from the GPU t…
Oct 2, 2026
f970b70
Add a PjRtTopologyDescription overload of FunctionalHloRunner::Create…
Oct 2, 2026
5f81707
Reuse IFRT executables for identical clusters across compilations and…
Oct 2, 2026
502d81c
Integrate LLVM at llvm/llvm-project@018a9e4a74ba
Oct 2, 2026
efaebd8
Allow HloCSE to deduplicate ragged-all-to-all ops with fresh Allocate…
Oct 2, 2026
661ea39
Support full [-8, 7] INT4 weight range in TFLite optimized_4bit Fully…
Oct 2, 2026
16b2108
Pack mhlo.spmd_parameters_shardings into a tuple sharding when use_tu…
Oct 2, 2026
9a9906f
Decouple hlo_dump_ui_bin from JSCompiler's ES5 polyfills/runtime inje…
Oct 2, 2026
0d64832
Merge latest upstream master after review
Oct 2, 2026
cfdd974
Sanitize inherited collaborator environment on Windows
Oct 2, 2026
22ad7c5
Validate integer bounds in nn_ops _get_sequence.
Oct 2, 2026
9c0413b
[XLA:CPU] Peel trailing broadcasts, reshapes, and copies from library…
Oct 2, 2026
270f8eb
Update TableGen definition for WriteTrainingPredictions op to include…
Oct 2, 2026
d65f8d5
Pass incarnation ID to the MegascaleContext through PjRT C API interf…
Oct 2, 2026
51ec68a
Roll back of Support full [-8, 7] INT4 weight range in TFLite optimiz…
Oct 3, 2026
5d8e2af
Move PjRtStreamExecutorClient::ShouldDoDirectTransfer into
Oct 3, 2026
f7a10ff
Also reset XLA LocalClient and PJRT client in GPU Device TestOnly reset.
Oct 3, 2026
4920226
Convert sdy.parameters_shardings and sdy.output_shardings on ModuleOp…
Oct 3, 2026
170cb1a
Integrate Triton up to 2074a1b8a5cd283ce584f8ebca825c4f0b215fc1
Oct 3, 2026
6b749b0
Automated Code Change
Oct 3, 2026
3773189
Automated Code Change
Oct 3, 2026
9cdc70c
[PJRT:GPU] Record allocation events before deferring transfers.
Oct 3, 2026
6dec10d
Merge pull request #127116 from kokol16:fix/resize-bilinear
Oct 3, 2026
d2f4d5b
Support GPU Confidential Computing mode in PJRT GPU and StreamExecutor.
Oct 3, 2026
95fd718
Automated Code Change
Oct 3, 2026
8786bc3
[XLA] Build HloReachabilityMap without zero filling the matrix first
Oct 3, 2026
399f149
Enhance tests for array_ops.fold with invalid inputs
Oct 4, 2026
4125098
Update array_ops_test.py
Oct 4, 2026
75fe2c9
Call `GetOrExportFabricHandle` with the underlying range, rather than…
Oct 4, 2026
5e0dc2c
[PJRT:GPU] Fix null dereference in ScheduleRemoteSend for error buffers.
Oct 4, 2026
031dd1d
Relanding: [XLA:GPU] Use lock-free SymbolicExpr uniquing on single-th…
Oct 4, 2026
f01f1b4
[NFC] Use the MlirContextPool alias for existing MLIRContext pools.
Oct 4, 2026
3919b87
Add accuracy tests for additional trig functions.
Oct 4, 2026
ce2538d
Adjust module meta data for inputs and outputs when Shardy generates …
Oct 5, 2026
e0f4473
[XLA:GPU] Handle transposed WGRAD convolution in ConvKindAssignment.
Oct 5, 2026
a636085
Rolling back due to verifier changes breaking old exports with stride…
Oct 5, 2026
0d624fb
Automated Code Change
Oct 5, 2026
77c29ec
[IFRT IR] Stop setting sdy front end attributes
Oct 5, 2026
6f7d9ef
Automated Code Change
Oct 5, 2026
2b14a6f
Address review: run testAcceptsNonTensorInput in graph mode, check ei…
Oct 5, 2026
dd3eaac
[XLA:GPU] Build the cuDNN graph of a CuDnnThunk once per device.
Oct 5, 2026
166a60b
Harden tarfile extraction: use filter='data' plus fail-closed member …
Oct 5, 2026
32640d4
Harden tarfile extraction: use filter='data' plus fail-closed member …
Oct 5, 2026
6113212
Harden tarfile extraction: use filter='data' plus fail-closed member …
Oct 5, 2026
591c8a2
Revert data_utils.py changes — legacy keras folder is not accepted fo…
Oct 5, 2026
930d88b
Automated Code Change
Oct 5, 2026
caf9286
PR #49942: [xla:gpu] Support multiple execution timeout handlers and …
Oct 5, 2026
47c2751
Bump consan.mlir lit test to medium size to fix ASAN timeouts.
Oct 5, 2026
808354d
Fix coverage errors
Oct 5, 2026
9031160
Merge pull request #126166 from madib06ops:sparsecount-batch-index-tr…
Oct 5, 2026
2d93384
[TSL] Wait for queued and running work in default `UnboundedWorkQueue…
Oct 5, 2026
599738a
PR #49990: [XLA:GPU][oneAPI] Align XPU Triton hook with upstream Triton
Oct 5, 2026
dd47c27
Merge pull request #128414 from MaddipatlaChetan24:patch-47
Oct 5, 2026
120d342
Merge pull request #128367 from mchdich:fix-collaborator-wheel-env
Oct 5, 2026
accbdf0
Merge pull request #128467 from MaddipatlaChetan24:patch-51
Oct 5, 2026
4dcb51d
Merge pull request #128426 from MinhNhat2710:patch-3
Oct 5, 2026
1d968a6
Merge pull request #128323 from Nishuuzz:eig-convert-before-dtype
Oct 5, 2026
File filter

Filter by extension

Filter by extension


Conversations
Failed to load comments.
Loading
Jump to
The table of contents is too big for display.
Diff view
Diff view
  •  
  •  
  •  
The diff you're trying to view is too large. We only load the first 3000 changed files.
14 changes: 14 additions & 0 deletions .bazelignore
Original file line number Diff line number Diff line change
@@ -0,0 +1,14 @@
# Copyright 2023 The TensorFlow Authors. All Rights Reserved.
#
# Licensed under the Apache License, Version 2.0 (the "License");
# you may not use this file except in compliance with the License.
# You may obtain a copy of the License at
#
# http://www.apache.org/licenses/LICENSE-2.0
#
# Unless required by applicable law or agreed to in writing, software
# distributed under the License is distributed on an "AS IS" BASIS,
# WITHOUT WARRANTIES OR CONDITIONS OF ANY KIND, either express or implied.
# See the License for the specific language governing permissions and
# limitations under the License.
# ==============================================================================
1,378 changes: 928 additions & 450 deletions .bazelrc

Large diffs are not rendered by default.

3 changes: 2 additions & 1 deletion .bazelversion
Original file line number Diff line number Diff line change
@@ -1 +1,2 @@
3.7.2
7.7.0
# NOTE: Update Bazel version in tensorflow/tools/ci_build/release/common.sh.oss
4 changes: 4 additions & 0 deletions .clang-format
Original file line number Diff line number Diff line change
@@ -0,0 +1,4 @@
# Run manually to reformat a file:
# clang-format -i --style=file <file>
BasedOnStyle: Google
DerivePointerAlignment: false
179 changes: 179 additions & 0 deletions .gemini/styleguide.md
Original file line number Diff line number Diff line change
@@ -0,0 +1,179 @@
# TensorFlow PR Review Guidelines (Gemini Code Assist)

## Objective

Provide high-signal, technically rigorous, and actionable feedback on pull
requests. Prioritize correctness, API stability, performance, security, and
long-term maintainability while minimizing unnecessary or low-value comments.

## Alignment with TensorFlow Contribution Guidelines

This style guide is aligned with TensorFlow’s official contribution guidelines:
https://github.com/tensorflow/tensorflow/blob/master/CONTRIBUTING.md

The following rules are derived from TensorFlow contribution requirements and
are used by Gemini Code Assist to guide PR review feedback.

These include requirements such as mandatory test coverage, adherence to coding
standards, API stability, and consistent behavior across supported environments.

## Core Review Mindset

### 1. Evaluate Necessity and Scope

* Is this change essential?
* Does it solve a real problem for TensorFlow users?
* Is it aligned with TensorFlow’s scope and design goals?
* Does it introduce unnecessary scope or complexity?
* Does the benefit justify the long-term maintenance cost?
* **Action:** If not, clearly question the need for the change.

### 2. Challenge the Implementation

* Actively identify edge cases, failure scenarios, and incorrect assumptions.
* Validate tensor shapes, dtypes, and execution paths.
* Do not assume correctness based solely on passing tests.

### 3. Ensure Robustness

* Flag fragile or environment-dependent logic.
* Identify risks across different hardware targets (CPU, GPU, TPU).
* Ensure proper error handling and defensive checks.

### 4. Detect Low-Quality or Non-Idiomatic Code

* Identify overly generic, verbose, or context-insensitive code.
* Flag patterns that do not align with established TensorFlow practices.
* Highlight inconsistencies with existing repository patterns.

### 5. Respect Existing Design Patterns

* Follow established TensorFlow APIs and architectural conventions.
* Avoid suggesting unnecessary abstractions or structural changes.
* Maintain consistency with similar modules.

## Review Priorities

### 1. Test Coverage and Reliability (Mandatory)

* Verify that unit tests are included for any new logic, feature, or bug fix.
* Flag PRs where source code is modified without adding or updating
corresponding tests.
* Ensure tests cover edge cases and expected behavior.
* Tests must be small, fast, and deterministic.
* Flag flaky tests or reliance on external systems (e.g., network, file
system).
* Ensure tests run reliably across supported platforms (Linux, macOS,
Windows).

### 2. Code Quality and Formatting

* Enforce standard Python and C++ formatting conventions.
* Flag issues such as inconsistent indentation, poor variable naming, and
missing docstrings.
* Ensure code readability and consistency with repository standards.
* Formatting issues should be flagged, but feedback should remain concise and
avoid overwhelming developers with low-value comments.

### 3. API Stability (High Priority)

* Flag breaking changes to public APIs (`tf.*`).
* Ensure backward compatibility is preserved.
* Verify adherence to the official deprecation lifecycle.
* Maintain consistency in naming and argument structure.

### 4. Correctness and Numerical Behavior

* Validate tensor operations, shapes, and broadcasting logic.
* Identify potential numerical instability (e.g., overflow, underflow,
division by zero).

### 5. Performance and Efficiency

* Flag Python-side loops over tensors and suggest vectorized TensorFlow
operations.
* Ensure compatibility with `tf.function` and avoid retracing issues.
* Identify unnecessary memory usage or redundant computations.
* Flag hardcoded device placement (e.g., `/gpu:0`).

### 6. Security and Safe Execution

* Flag unsafe data handling, deserialization, or file operations.
* Highlight potential memory safety issues, especially in C++ kernels or
Python–C++ boundaries.
* Ensure code avoids patterns that could introduce security risks.

### 7. Dependencies and Scope

* Flag the introduction of unnecessary or heavy external dependencies.
* Ensure the change aligns with TensorFlow’s scope and does not increase
maintenance burden unnecessarily.
* Flag large pull requests that combine multiple unrelated changes (e.g., bug
fix + refactor) and recommend splitting them into smaller, focused PRs for
easier review and maintainability.

### 8. Readability and Maintainability

* Flag issues only when they affect clarity or long-term maintainability.
* Avoid suggesting purely subjective stylistic preferences that do not impact
readability or consistency.

### 9. Idiomatic TensorFlow & Model Training

* **Data Pipelines:** Flag raw in-memory tensor training when dataset size or
scalability is a concern. \
Prefer `tf.data.Dataset` with `.batch()`, `.prefetch(tf.data.AUTOTUNE)`, and
optional `.cache()` where it fits in memory to improve performance.

* **Batching:** Ensure training uses appropriate batching (`batch_size` or
dataset batching). \
Avoid feeding the entire dataset at once unless it is trivially small.

* **Vectorization:** Avoid Python loops over tensors (training or inference).
\
Prefer batched/model-level operations (`model(x)` instead of per-sample
calls).

* **Reproducibility:** Require explicit seeds in data creation and model
initialization:

```python
np.random.seed(0)
tf.random.set_seed(0)
```

* **Observability:** Require validation data (`validation_split` or validation
dataset) and meaningful metrics (e.g., `mae`, `accuracy`) to monitor
training and detect overfitting.

* **Callbacks:** Encourage use of `tf.keras.callbacks.EarlyStopping` and
`tf.keras.callbacks.ModelCheckpoint` for stable and efficient training.

## Out of Scope (Do Not Comment)

* Minor subjective stylistic preferences (e.g., personal formatting opinions)
that do not impact readability, consistency, or correctness.
* Trivial or non-impactful differences.

## Feedback Standards

* Be clear, concise, and actionable.
* Provide concrete suggestions where possible.
* Avoid vague or non-specific feedback.
* **Bad:** "This might be slow."
* **Good:** "This loop introduces Python overhead; consider vectorized
TensorFlow operations to improve performance and enable better graph
optimization."

## Review Guardrails

* Avoid speculative or uncertain feedback.
* Do not comment if no meaningful issue is identified.
* Avoid duplicate or redundant comments.
* Prioritize high-impact issues over minor observations.

## Behavior

* Provide suggestions only (do not block pull requests).
* Focus on correctness, performance, security, and API stability.
* Maintain a high signal-to-noise ratio in all feedback.
45 changes: 0 additions & 45 deletions .github/ISSUE_TEMPLATE/00-bug-issue.md

This file was deleted.

30 changes: 0 additions & 30 deletions .github/ISSUE_TEMPLATE/10-build-installation-issue.md

This file was deleted.

60 changes: 0 additions & 60 deletions .github/ISSUE_TEMPLATE/20-documentation-issue.md

This file was deleted.

23 changes: 0 additions & 23 deletions .github/ISSUE_TEMPLATE/30-feature-request.md

This file was deleted.

14 changes: 0 additions & 14 deletions .github/ISSUE_TEMPLATE/50-other-issues.md

This file was deleted.

Loading