Default NVFP4 DFlash to 256K context
#2
by baranowskiadam - opened
Reparameterize the equivalent YaRN correction ramp for an 8192-token origin and cap max_position_embeddings at 262144. Model weights are unchanged.
joerowell changed pull request status to merged