Skip to content

Pull requests: PrimeIntellect-ai/prime-rl

Author
Filter by author
Loading
Label
Filter by label
Loading
Use alt + click/return to exclude labels
or + click/return for logical OR
Projects
Filter by project
Loading
Milestones
Filter by milestone
Loading
Reviews
Assignee
Filter by who’s assigned
Assigned to nobody Loading
Sort

Pull requests list

chore(trainer): unify context-parallel setup across models
#3533 opened Sep 11, 2026 by garrett361 Contributor Loading…
fix(inference): preserve configured worker extension
#3531 opened Sep 11, 2026 by samjiawng Loading…
feat(trainer): add optional IcePop loss
#3526 opened Sep 10, 2026 by samsja Member Loading…
Persist per-token policy mask decisions
#3525 opened Sep 10, 2026 by samsja Member Loading…
Add Total Router Recall for MoE training
#3523 opened Sep 10, 2026 by samsja Member Draft
fix(orchestrator): restore zero-advantage filtering flag
#3515 opened Sep 9, 2026 by S1ro1 Collaborator 2/4 Draft
chore(inference): update vLLM and use its native FP32 head
#3513 opened Sep 9, 2026 by S1ro1 Collaborator 3/4 Draft
feat: add NVFP4 MoE training with optional 4over6 quantization
#3511 opened Sep 9, 2026 by S1ro1 Collaborator 4/4 Draft
feat(rl): add experimental Total Router Recall
#3509 opened Sep 9, 2026 by faresobeid Contributor Draft
Feat/dequant triton kernel
#3506 opened Sep 8, 2026 by dhanyamd Draft
feat: accept max reasoning effort for evaluation
#3499 opened Sep 7, 2026 by samsja Member Draft
feat(inference): launch managed Dynamo workers
#3493 opened Sep 6, 2026 by biswapanda Contributor Loading…
fix(inference): support vLLM skip-gather logits
#3492 opened Sep 5, 2026 by biswapanda Contributor Loading…
fix(wandb): use compatible GraphQL transport
#3489 opened Sep 4, 2026 by LOGO127 Loading…
ProTip! Follow long discussions with comments:>50.