So, while doing a search to read the model card again, and because I didn't memorize the huggingface URL, I saw an AI "answer" at the top of the search, stating that while it natively supports a ~244k context, it could go to 1M using YaRN. That's the first I have heared of both, the 1M context, and YaRN. Has anyone used that method? I want to use this model for a Hermes agent, so a long context could be good...I think. But I am still waiting for my mobo to ship, so for now all I can do is ask. x) Thanks!   submitted by   /u/IngwiePhoenix [link]   [comments]