Default NVFP4 DFlash to 256K context

#2

Reparameterize the equivalent YaRN correction ramp for an 8192-token origin and cap max_position_embeddings at 262144. Model weights are unchanged.

joerowell changed pull request status to merged

Sign up or log in to comment