-
Notifications
You must be signed in to change notification settings - Fork 607
Pull requests: AI-Hypercomputer/maxtext
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
Add smoke_test_parallelism test for MaxText parallelism and mesh verification.
#5210
opened Sep 12, 2026 by
copybara-service
Bot
Loading…
Add skills for maxtext developers
#5209
opened Sep 12, 2026 by
hengtaoguo
Collaborator
•
Draft
4 tasks done
feat(quantization): add native FP8 inference (serve_fp8_weight) for DenseGeneral and GMM v2
#5207
opened Sep 11, 2026 by
Shuwen-Fang
Collaborator
Loading…
4 tasks done
[NNX] Remove remaining Linen module code and dead references
#5203
opened Sep 11, 2026 by
ecnal-cienet
Collaborator
•
Draft
4 tasks done
Optimize MaxText cross-entropy loss memory, vocabulary tiling, and PyTree engine caching:
#5201
opened Sep 11, 2026 by
copybara-service
Bot
Loading…
PR #4360: [NNX] Delete Linen 4/5: remove the Linen decoder/attention layers and *_as_linen model wrappers
#5197
opened Sep 11, 2026 by
copybara-service
Bot
Loading…
4 tasks done
Migrate MaxText Docker images to Artifact Registry
#5196
opened Sep 11, 2026 by
SurbhiJainUSC
Collaborator
Loading…
4 tasks done
Add Cosmos 3 Core Attention Block to MaxText.
#5192
opened Sep 10, 2026 by
copybara-service
Bot
Loading…
Optimize JAX flash attention for long sequences with a specialized path.
#5191
opened Sep 10, 2026 by
copybara-service
Bot
Loading…
[MaxText] Eliminate XLA backward adjoint interior padding in reorder_sequence
#5188
opened Sep 10, 2026 by
copybara-service
Bot
Loading…
Fix: Numeric divergence between the mHC layer and its Pallas kernel
#5179
opened Sep 9, 2026 by
denis-mil
Loading…
4 tasks done
Reverts f7f720404a0c6cd47133419edd1a8ff4c5effa3c
#5174
opened Sep 9, 2026 by
copybara-service
Bot
Loading…
Share one RNG stream across the bridged NVFP4 quantizers
#5160
opened Sep 8, 2026 by
ecnal-cienet
Collaborator
•
Draft
4 tasks done
Share one CUSTOM_REMAT_TENSORS list between MaxTextConfig and RLConfig
#5159
opened Sep 8, 2026 by
lokic233
Loading…
Train & Eval on entire dataset when pdbs < 1
#5158
opened Sep 7, 2026 by
muskansh-google
Contributor
•
Draft
[Stacked PR 4/5] Enable GDN backward pass kernel and hybrid precision in Qwen3 model
#5154
opened Sep 6, 2026 by
Rohan-Bierneni
Collaborator
•
4/5
Loading…
3 tasks done
Previous Next
ProTip!
Find all pull requests that aren't related to any open issues with -linked:issue.